Pith. sign in

Paper Citation Record · LEDGER

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models

As of 3 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 2 inbound Pith citation observations for arXiv:2605.03821.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.03821 v1

Coverage vector

measured 16 of 16 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-07T15:37:35.652145Z

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T18:20:38.852755Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

16 of 16 outbound references displayed

  • verified exact0
  • verified fuzzy4
  • unresolved11
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7d7a09d8-7ec9-47b6-ac8f-b0bcb3137cdf · outbound

This paper cites an unresolved cited work.

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-05-27T03:08:46.128342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T15:37:35.652145Z digest=sha256:848e6ec62253654e44263f5f98326235f9ee8adb7b1983a44f92b5a057c0927c

Observation 8f95d322-8526-45ea-a0cd-c76b71c740e9 · outbound

This paper cites an unresolved cited work.

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-05-27T03:08:46.125079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T15:37:35.652145Z digest=sha256:e1b97c9ecaabef1580b30140c6da01bdab492256771589da9ef31092252575f6

Observation f37df0ba-dd27-49e7-8d54-bc4d6f2d7ff8 · outbound

This paper cites an unresolved cited work.

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-05-27T03:08:46.099535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T15:37:35.652145Z digest=sha256:937e32b698bfd1803ef9c82c0ac3d260b6651bdb0d2b6202c81e09968bc12fa4

Observation 7c131c60-e663-4288-b0f4-f4f0e4950ddf · outbound

This paper cites an unresolved cited work.

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-05-27T03:08:46.102855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T15:37:35.652145Z digest=sha256:9432925ba3eb3db03d4c961febc35937633f46e934d3b005b15c490bfde633d4

Observation f8a97480-cffb-4044-ad3f-758b2447b6ae · outbound

This paper cites Stage 1 consumes only the pixels{Is+k}T−1 k=0; Stage 2 additionally consumes the paired actions to construct the joint token sequence of §B.3.

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models Stage 1 consumes only the pixels{Is+k}T−1 k=0; Stage 2 additionally consumes the paired actions to construct the joint token sequence of §B.3

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T03:08:46.118043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T15:37:35.652145Z digest=sha256:26fd8ebfb21ef4e0b09eb395d8d12970cedd4abbe301452b9442ee9a804ba6fd

Observation 7eeeef1e-8c5a-4b3b-9620-898ae2cf95da · outbound

This paper cites After passing through the 4 Transformer layers, the output corresponding to the [CLS] token is used as the instruction embeddingt∈R256.

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models After passing through the 4 Transformer layers, the output corresponding to the [CLS] token is used as the instruction embeddingt∈R256

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T03:08:46.106901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T15:37:35.652145Z digest=sha256:15ccf9112dbd3708084402c224820e4a386609209a88545136d2e3c04b557b7c

Observation 83c71b15-6fbf-470e-8a3f-fa4e95ee7d13 · outbound

This paper cites close top drawer.

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models close top drawer

Reference 7

Resolution
malformed identifier
raw_fallback, observed 2026-05-27T03:08:46.087971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T15:37:35.652145Z digest=sha256:3124216a71dd657f1a7345abb9ecf56a1c1c7e9365fd931b61cf6b1ccd0c55fd

Observation d9882bf7-a644-45eb-bd01-16c0c05d74cc · outbound

This paper cites an unresolved cited work.

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-05-27T03:08:46.110299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T15:37:35.652145Z digest=sha256:024f2ff900b43d88bade004205f10cf723dcb4dc76d6efd6960e3068910ab8a8

Observation bc9d48a0-9d6c-4137-8c58-684bb9e56912 · outbound

This paper cites an unresolved cited work.

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-05-27T03:08:46.113931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T15:37:35.652145Z digest=sha256:65184a2f16c0895927a7bccee5b589982e3f4e07eebbcb76fb4fb6ad9efb49b8

Observation 5c4735c9-5b23-4f56-9587-62f17b550d47 · outbound

This paper cites Score each dimension carefully.

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models Score each dimension carefully

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T03:08:46.091993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T15:37:35.652145Z digest=sha256:0a46f4949f62bc26c39a26c54e0423abe24eb3e8e8cf64a6013c0bcd9554f4b1

Observation 84f41d9d-cc0a-45e1-9d6d-ecd368b58be8 · outbound

This paper cites an unresolved cited work.

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-05-27T03:08:46.121537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T15:37:35.652145Z digest=sha256:4d3d2d4ad93d7bf6e9582929a6bc588d7d5b3f779ba8f4c95c260c6db799151a

Observation 5c4fd555-3309-43ee-a838-e133e540fe90 · outbound

This paper cites an unresolved cited work.

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-05-27T03:08:46.077187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T15:37:35.652145Z digest=sha256:c474ac6e6adc0172d938248cc7dc16e7123396c0818de016342caa9963bf60a9

Observation 608e5def-d549-470a-a1d0-ad363bca76c0 · outbound

This paper cites an unresolved cited work.

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-05-27T03:08:46.073266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T15:37:35.652145Z digest=sha256:4426c30ee980922ca80345c5185ada444e766506f1ecc5e087b86210663b3708

Observation 02867b4e-f9e6-4eef-884c-e3ca5771aaff · outbound

This paper cites an unresolved cited work.

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-05-27T03:08:46.080784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T15:37:35.652145Z digest=sha256:ccfccf55ec6df301e5b187dc655ed64a6a8b55022d9fc68a3c3db5df314b963c

Observation d2776c39-77c1-4033-9c32-eb4d60fd5466 · outbound

This paper cites an unresolved cited work.

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-05-27T03:08:46.084252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T15:37:35.652145Z digest=sha256:a9afb623b3e759dd5ce0ac30d076e1c9faf7690a109c7bdaf862806539f4a83b

Observation 1bd7c672-9d21-4579-a7e3-4d27e56d55e5 · outbound

This paper cites reasoning.

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models reasoning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T03:08:46.095641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T15:37:35.652145Z digest=sha256:2ff71b078c1216e448a1457972e16f6bd0701f93495b4213548cb264233d158e

Pith citing papers

Observation fa51d136-edce-4373-90d4-fe273253d6b0 · inbound

MIS-HCC: Hierarchical Channel Clustering for Efficient Medical Image Segmentation cites this paper.

MIS-HCC: Hierarchical Channel Clustering for Efficient Medical Image Segmentation RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T18:20:38.852755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:20:38.852755Z digest=sha256:14aaa8d679f8dded09afeb5f1a0d61a2ac3a8d3ff97b9eb16ac86a1d1a33fcc0

Observation 681bab52-da7f-4f80-8208-920d355eba84 · inbound

Thinking in Video: Can Video Generators Really Reason About the Real World? cites this paper.

Thinking in Video: Can Video Generators Really Reason About the Real World? RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T17:47:09.268892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T17:47:09.268892Z digest=sha256:a2cc798a7aee6f8ace824e497bbde0a0f7ebe8f42d1802785dd756c6b986cba5