Pith. sign in

Paper Citation Record · LEDGER

Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 44 inbound Pith citation observations for arXiv:2501.14818.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.14818 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 44 of 44 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:04:31.774244Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T00:19:12.971623Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 643b9fc3-cbf4-43c7-81a6-f09073fdd162 · inbound

GR00T N1: An Open Foundation Model for Generalist Humanoid Robots cites this paper.

GR00T N1: An Open Foundation Model for Generalist Humanoid Robots Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-10T19:09:10.226821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T19:09:10.112304Z digest=sha256:57931e2e2edb82b1b2987d260ae22de5fdc1c8f0a42101ac177015a9daf5678d

Observation fb435987-89ee-4b47-8080-6a7c1e77b393 · inbound

SmolVLM: Redefining small and efficient multimodal models cites this paper.

SmolVLM: Redefining small and efficient multimodal models Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:23:51.730104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T20:23:50.552549Z digest=sha256:74b4f70f6a8e8b91353a2a0b926b39b9976efafaac3e286c4b040578beaea8fa

Observation 5c37d1f1-2132-4333-af88-222f777f6598 · inbound

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models cites this paper.

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:41:08.369391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T13:41:07.991012Z digest=sha256:076d163f44589342e7c47ab6b77122dad9c8cd537e84c64a4486aa176949b86f

Observation a12d6416-3e90-484c-9582-34b747b27507 · inbound

FLARE: Robot Learning with Implicit World Modeling cites this paper.

FLARE: Robot Learning with Implicit World Modeling Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:59:08.936657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T15:59:08.846629Z digest=sha256:e418128c06541aac49fbaf570284109a058c897a89d13af434b840ff9dc19826

Observation f7161beb-6561-4e9a-8243-92954042ffef · inbound

MEgoHand: Multimodal Egocentric Hand-Object Interaction Motion Generation cites this paper.

MEgoHand: Multimodal Egocentric Hand-Object Interaction Motion Generation Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:31.774244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:31.774244Z digest=sha256:49d315b423608f844e60613c64ab0de3e856029ef9003d51f8eef92493249588

Observation 69202f30-68ad-453a-98c2-d64ef031cf78 · inbound

Rex-Thinker: Grounded Object Referring via Chain-of-Thought Reasoning cites this paper.

Rex-Thinker: Grounded Object Referring via Chain-of-Thought Reasoning Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T10:55:14.714418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:55:14.714418Z digest=sha256:b37896465b4d5328b01923b822ef432084596d57bbca74c88236adc4ead07b16

Observation 9c95f372-a496-4127-bad7-6d43e929ea48 · inbound

ReXVQA: A Large-scale Visual Question Answering Benchmark for Generalist Chest X-ray Understanding cites this paper.

ReXVQA: A Large-scale Visual Question Answering Benchmark for Generalist Chest X-ray Understanding Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T10:47:46.925421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:47:46.925421Z digest=sha256:5af648ce9a4995fe76ba7e54e316288c10be6b7d498a6f92f6d6f629080e6c34

Observation 8ee657c0-142e-46e5-abb7-0f73ee36fc1f · inbound

AV-Reasoner: Improving and Benchmarking Clue-Grounded Audio-Visual Counting for MLLMs cites this paper.

AV-Reasoner: Improving and Benchmarking Clue-Grounded Audio-Visual Counting for MLLMs Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:05.611836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:05.611836Z digest=sha256:dd5d06517fc910b87afd80ddcf72489b5e5a020aeb6b2c9da4fcebd4f19419d6

Observation dd1c2e60-2032-4a6d-9fd3-542584e3016c · inbound

Bridging Perspectives: A Survey on Cross-view Collaborative Intelligence with Egocentric-Exocentric Vision cites this paper.

Bridging Perspectives: A Survey on Cross-view Collaborative Intelligence with Egocentric-Exocentric Vision Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 214

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:58.645840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:58.645840Z digest=sha256:f82c8b14cf9ff37000198a5ab2c1f98e14ca7e41cd9dce40e164e1466ffb8088

Observation bf36f4bd-1fae-47c5-85ba-3a755db47a5a · inbound

Native Visual Understanding: Resolving Resolution Dilemmas in Vision-Language Models cites this paper.

Native Visual Understanding: Resolving Resolution Dilemmas in Vision-Language Models Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:48:30.849435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:48:30.849435Z digest=sha256:0a9cde989572ca9f2a232cb7fd2d8ff51b2e28a34b22cd92c184153983b98731

Observation efc8c371-3195-4ec9-8a13-0156566cf750 · inbound

Llama Nemoretriever Colembed: Top-Performing Text-Image Retrieval Model cites this paper.

Llama Nemoretriever Colembed: Top-Performing Text-Image Retrieval Model Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T19:30:14.990336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:30:14.990336Z digest=sha256:cff2e586d204b998a97c283995e7c72412a77672019749703b6256e2a763cc2c

Observation 80b8bca0-a559-4b5b-82a1-f870a84d7c70 · inbound

High-Resolution Visual Reasoning via Multi-Turn Grounding-Based Reinforcement Learning cites this paper.

High-Resolution Visual Reasoning via Multi-Turn Grounding-Based Reinforcement Learning Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-19T06:12:07.077961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T06:10:57.219445Z digest=sha256:f2087c3f2075eeb48e7c9593138e6cd721d45f4c9187e64590b9ac236a865da6

Observation 89a3eceb-d827-4ab1-a1f0-98bd559e4348 · inbound

ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning cites this paper.

ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-19T03:22:00.910099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T03:18:14.655384Z digest=sha256:73dc1bc5c3e2ecb3b485ba875cbf5c83f92f812d6b7383d39ff092913df72534

Observation a8a66cdb-7bab-4703-a7d3-0cd63a9a7abc · inbound

FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model cites this paper.

FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-04T10:18:43.278717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:18:43.278717Z digest=sha256:7fe7c8234019634d91e8dfb8c812d97dcc649157c9758882f4ecc06aa496a9b0

Observation 9838a444-74c6-4038-9c95-699c7b9eb291 · inbound

RefBench-PRO: Perceptual and Reasoning Oriented Benchmark for Referring Expression Comprehension cites this paper.

RefBench-PRO: Perceptual and Reasoning Oriented Benchmark for Referring Expression Comprehension Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T18:15:05.687979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:15:05.687979Z digest=sha256:8540e55e8cbff61b0ebcd77737d9e0e3f5b3c7f192d9e1363f7bcd2c5b176ecc

Observation d1917e2a-55b9-4e1f-86c6-14e4f91c05fa · inbound

Adaptive Action Chunking at Inference-time for Vision-Language-Action Models cites this paper.

Adaptive Action Chunking at Inference-time for Vision-Language-Action Models Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:08:01.427266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T16:57:52.595165Z digest=sha256:dd58351f2d5efdfb9c883c673e31a5115a0bd60368e7a93f055fab84f7211b40

Observation a93d5529-c6e7-4497-9232-c89ed5af791d · inbound

MLLM-as-a-Judge Exhibits Model Preference Bias cites this paper.

MLLM-as-a-Judge Exhibits Model Preference Bias Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T10:01:00.015653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T15:42:39.226895Z digest=sha256:8d8e242b32a7048d243995cd8f8d40cf385d75c9ef321d8ffe3310ea44780972

Observation 6281c86a-55f1-408d-9855-7556fc46e6a0 · inbound

Unmasking the Illusion of Embodied Reasoning in Vision-Language-Action Models cites this paper.

Unmasking the Illusion of Embodied Reasoning in Vision-Language-Action Models Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:56:29.370055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T04:27:18.284698Z digest=sha256:f69cf417390d15b5e5f32fc968f37ff5947db8b8faf3310a4cf09675e4a05102

Observation aca1b3fb-d015-4f82-8a16-1f6255b734c2 · inbound

PLaMo 2.1-VL Technical Report cites this paper.

PLaMo 2.1-VL Technical Report Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:41:05.415340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T03:09:05.436900Z digest=sha256:d6ee2bb857a6e4829bf8cdb63c90cd78f96c75636a03b7e9dcad1a58f7a759eb

Observation 24eb9b59-7a43-497b-8d15-255c9959a614 · inbound

EmbodiedMidtrain: Bridging the Gap between Vision-Language Models and Vision-Language-Action Models via Mid-training cites this paper.

EmbodiedMidtrain: Bridging the Gap between Vision-Language Models and Vision-Language-Action Models via Mid-training Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:11:03.790974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T02:16:08.687340Z digest=sha256:38b1e6dd164e72504ff9deb55d0f8921a9cb6615673f0bd74afdf8416529252f

Observation beaa7510-805d-4dc7-8495-b697f14e6869 · inbound

$M^2$-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills cites this paper.

$M^2$-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:11:15.902154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T03:13:36.437080Z digest=sha256:180ee5138563267a2c46580f43112daca2b80419005c6df5091e97fe423c0d27

Observation 0567caf1-a2b9-4fbb-bd52-d46b7b614b62 · inbound

$M^2$-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills cites this paper.

$M^2$-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-12T18:17:43.564400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T18:17:43.564400Z digest=sha256:8165595f085eb62732c1ba3d56cbc813d9d9f7aa0b2520c5d346d764aec00781

Observation 795b5e01-06e5-4f67-800a-b9173d643d11 · inbound

Being-H0.7: A Latent World-Action Model from Egocentric Videos cites this paper.

Being-H0.7: A Latent World-Action Model from Egocentric Videos Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:01:04.535757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-09T20:48:01.461993Z digest=sha256:417c01eea0839beb2d8c6c0712b59400decf60e0f9f58e55bd305d0ce3e315a0

Observation ab5c764d-f318-4ed5-8c03-84fde4eb26fb · inbound

Cambrian-P: Pose-Grounded Video Understanding cites this paper.

Cambrian-P: Pose-Grounded Video Understanding Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-22T05:51:09.039957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T05:47:13.493461Z digest=sha256:87bf8e624ea485ba5b7c098d792c741721ce898af30dd6020699934741840e79

Observation f00be046-e1a3-4d5d-b731-f67b0dd3bc6c · inbound

Cambrian-P: Pose-Grounded Video Understanding cites this paper.

Cambrian-P: Pose-Grounded Video Understanding Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T13:29:37.689045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:29:37.689045Z digest=sha256:d390534be391990de40b58522379a91d373202be844f70ae9539e84cffad5da8

Observation a1dccfdb-2fc5-499e-aa8e-0039dfc9a9a4 · inbound

Gaze2Act: Gaze-Conditioned Vision-Language-Action Policies for Interactive Robot Manipulation cites this paper.

Gaze2Act: Gaze-Conditioned Vision-Language-Action Policies for Interactive Robot Manipulation Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:13:16.123173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T07:12:17.310331Z digest=sha256:90c60be721557fd2a93b233b63a4231081fea66880720531b98e8917db8097c3

Observation 259ad005-488c-469b-b9ca-d1fceaa58521 · inbound

Argus-Retriever: Vision-LLM Late-Interaction Retrieval with Region-Aware Query-Conditioned MoE for Visual Document Retrieval cites this paper.

Argus-Retriever: Vision-LLM Late-Interaction Retrieval with Region-Aware Query-Conditioned MoE for Visual Document Retrieval Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T10:46:52.601436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T04:55:04.299049Z digest=sha256:c946f1648db5476e0c889292b179ebb8ce5506aebf0c9e5df501167be8d14fa2

Observation a4df8f36-caaa-4b12-b784-cc83d247761f · inbound

LARA: Latent Action Representation Alignment for Vision-Language-Action Models cites this paper.

LARA: Latent Action Representation Alignment for Vision-Language-Action Models Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:47:09.560724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T22:25:17.522240Z digest=sha256:1fdd83bc058e42627b5108e570aaf09c282012a3f040b030e277c78d171db6fc

Observation 0d094686-c576-419d-b534-cf8f0dc5eba9 · inbound

LARA: Latent Action Representation Alignment for Vision-Language-Action Models cites this paper.

LARA: Latent Action Representation Alignment for Vision-Language-Action Models Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T08:35:34.490432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-01T07:17:42.045939Z digest=sha256:357035c7966c6ddffc23f5b7ef7c4b13aee39e9014ccea496397646cabcff94a

Observation a8d772a3-ca21-4a45-827c-67057ce3c2fe · inbound

RhinoVLA Technical Report cites this paper.

RhinoVLA Technical Report Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:17:17.961610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T21:36:15.806974Z digest=sha256:fb5d97d493715882cff19483cda55df95d18b0271eb89906068004151a6d33e9

Observation e796f94f-66b2-473b-b200-e7143df6901f · inbound

RhinoVLA Technical Report cites this paper.

RhinoVLA Technical Report Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-01T07:15:29.790701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-01T07:14:14.918962Z digest=sha256:9225036832ef88b6259bd6ab98fb163439921e49b6932f9ddb29ce18d3620387

Observation e8fd853e-026a-450a-83e8-87e71c406e24 · inbound

RhinoVLA Technical Report cites this paper.

RhinoVLA Technical Report Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-12T14:51:51.552221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T14:51:51.552221Z digest=sha256:7232c4ae8b8468d58db97707c77777ef70ae0edec33a1030fa689aaea0bdf5c4

Observation eb234e85-e580-44b9-9b56-f51854879b04 · inbound

RhinoVLA Technical Report cites this paper.

RhinoVLA Technical Report Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T12:15:09.422092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:15:09.422092Z digest=sha256:34b90f67209a920215b815b2a6c17a5f824a57480270fd318de5692cdfb7e6e8

Observation 171f1d79-6b59-477e-a1a0-3e6a776128c1 · inbound

Hy-Embodied-0.5-VLA: From Vision-Language-Action Models to a Real-World Robot Learning Stack cites this paper.

Hy-Embodied-0.5-VLA: From Vision-Language-Action Models to a Real-World Robot Learning Stack Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T11:31:34.532750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:31:34.532750Z digest=sha256:7a6bf584536412e324afa88051517ec34f5a7b1baba99927f34e838a5c9ca335

Observation 313f98a4-7fe9-46e9-9c18-f6f97753a115 · inbound

Contrastive Action-Image Pre-training for Visuomotor Control cites this paper.

Contrastive Action-Image Pre-training for Visuomotor Control Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-03T18:08:46.890780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T03:16:33.702802Z digest=sha256:443c44e8f0b78f0f4bb4ee7077c7065971b5dec5c3664654ecc6efe0abed7b8d

Observation f7e7ce41-e1b1-4cd2-9f29-0f8129e4e3ae · inbound

AMALIA-VL: A Native European Portuguese Open-Source Vision and Language Model cites this paper.

AMALIA-VL: A Native European Portuguese Open-Source Vision and Language Model Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:19:12.973201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T21:22:53.792778Z digest=sha256:ebccfb7d68f99974ae832c3c0da9a17746421195ef613580f2104e581047817f

Observation f50f10c1-97d0-44ab-b35b-e4b0aceedd0e · inbound

AMALIA-VL: A Native European Portuguese Open-Source Vision and Language Model cites this paper.

AMALIA-VL: A Native European Portuguese Open-Source Vision and Language Model Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-01T07:35:28.776187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-01T07:34:37.974692Z digest=sha256:db54dce2fa2c3837ca54f5329ca6d7f166ea06659d21f6ff93ae40270886fac8

Observation 07013e9d-7be1-4c34-be0f-fb8793c908e6 · inbound

AMALIA-VL: A Native European Portuguese Open-Source Vision and Language Model cites this paper.

AMALIA-VL: A Native European Portuguese Open-Source Vision and Language Model Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:17:24.983536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-02T22:15:14.561596Z digest=sha256:6ca07f84c9d6720d0c165c3f0abfdfd06a6c59c144d9e026ea0c5b5bbcea676c

Observation 6216653f-a7f8-4ae9-841c-fe29cc09c9ee · inbound

Consistency as Inductive Bias: Learning Cross-View Invariance for Robust Multimodal Reasoning cites this paper.

Consistency as Inductive Bias: Learning Cross-View Invariance for Robust Multimodal Reasoning Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-30T06:14:19.226773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T06:10:41.786328Z digest=sha256:f10a02e38028168e49a1aef8bd6d416f5e56d97ad79a5582e0f36597d5f5ddc4

Observation ec7f7cdf-df4a-49c9-801e-0dfa5d771377 · inbound

SigLIP-HD by Fine-to-Coarse Supervision cites this paper.

SigLIP-HD by Fine-to-Coarse Supervision Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-13T02:39:58.984385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T02:39:58.984385Z digest=sha256:38e6706fb712060e424312055d2028ad3d190317f52e5f3bcdebb8db92437372

Observation 8932aa1e-f882-4ecb-a368-4cadb38513a2 · inbound

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation cites this paper.

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 248

Resolution
unresolved
no resolver link, observed 2026-08-01T14:39:52.460914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T14:39:52.460914Z digest=sha256:f5661f5748a95b6cd075722c8e5948e9493f756d5e21af05b786b7adee81eddf

Observation f2480c1d-4f46-46f1-8207-fa855b5c8fbd · inbound

LeapBot-WA: World-Anchor Action Models via Predictive Latent Alignments cites this paper.

LeapBot-WA: World-Anchor Action Models via Predictive Latent Alignments Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-31T23:29:10.232317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T23:29:10.232317Z digest=sha256:b8aa52f3e75f5d6c68a9c6896d5de9570f0c446800ef9276d933aaa919699e09

Observation 7603044c-0348-4171-bf87-e0a8335079a6 · inbound

Hand-Object Interaction in the Age of Large Foundation Models:Reconstruction, Generation, and Embodied Transfer cites this paper.

Hand-Object Interaction in the Age of Large Foundation Models:Reconstruction, Generation, and Embodied Transfer Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 182

Resolution
unresolved
no resolver link, observed 2026-07-31T08:51:24.633040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T08:51:24.633040Z digest=sha256:d8b1a0fae63bfdc60ee45cc7c03be3cf917cc53920480c9f72b0c3446f30c17d

Observation e127f1b8-c197-423c-b864-fa7806f336bb · inbound

Hand-Object Interaction in the Age of Large Foundation Models:Reconstruction, Generation, and Embodied Transfer cites this paper.

Hand-Object Interaction in the Age of Large Foundation Models:Reconstruction, Generation, and Embodied Transfer Eagle 2: Building Post-Training Data Strategies from Scratch for Frontier Vision-Language Models

Reference 168

Resolution
unresolved
no resolver link, observed 2026-08-04T01:23:07.841255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:23:07.841255Z digest=sha256:b3ba1549f5294937e3fd49afc559b66c6285cd37168be95a2acec2357012358c