Pith. sign in

Paper Citation Record · LEDGER

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning

As of 17 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 0 inbound Pith citation observations for arXiv:2608.13026.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.13026 v1

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:21:34.060079Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

34 of 34 outbound references displayed

  • verified exact0
  • verified fuzzy13
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 869da127-0603-4a4c-ab34-7203f8d9e544 · outbound

This paper cites 2026 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2026 , eprint=

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:35.458290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:33.676420Z digest=sha256:4bcfee8d380602ba95fdddebc4120f79f5d65f869c12d3ea18a0623fa8cc59ef

Observation a65ec208-8e12-4bb8-b13e-7542844a638f · outbound

This paper cites 2025 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2025 , eprint=

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:35.442159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:33.694079Z digest=sha256:5ce55ae9567ba5cad7fb44819056c2aa599cb9eea1f683c3d1b5a326433e9d31

Observation edf3e6e3-82cb-4d7f-8356-2fbacf0a710c · outbound

This paper cites 2025 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2025 , eprint=

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:35.418731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:33.708722Z digest=sha256:1d656d5491ff18d236c0e110595cf6787fdf9db220450a1eebea86aa0ad2888e

Observation 3a46a949-f91e-4bc1-b018-29db7b10cf98 · outbound

This paper cites 2023 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2023 , eprint=

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.717518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.717518Z digest=sha256:a8b3e49f85e5cda73c3bf502520586b9d83645f85a56eb47dfa252ab23218fa2

Observation 9d75a46c-61e7-41d8-af41-bbd2cc248713 · outbound

This paper cites 2025 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2025 , eprint=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.732528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.732528Z digest=sha256:bc45605e71872e8d3e5c54ed4d1d1ed5f9a26c0ea9e2126e029a59e5ec50e34e

Observation 8c46405b-e01b-4803-a7fb-eff70999bb0b · outbound

This paper cites 2025 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2025 , eprint=

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.748196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.748196Z digest=sha256:45c2e30fa42ff50ad9e8dfe1bd128d970796f290d9dc30c6618863915e3c642f

Observation 8b708809-6225-4c7f-8ff2-b9e994a6bd60 · outbound

This paper cites 2024 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2024 , eprint=

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:35.318982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:33.758361Z digest=sha256:169584db4dbc8cad59ad6e3dc203bae49f88ec67a202aeb7689371cb05b7cc5b

Observation 682b361c-f657-4973-8262-36bfb41fcbbf · outbound

This paper cites 2026 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2026 , eprint=

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.770823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.770823Z digest=sha256:c2a0f9048b93111babf2cb62dd02a9c6c253b2984d351b9916e6c170178fdd95

Observation a5f7eb8f-9713-4f83-b856-48b733bc1ed5 · outbound

This paper cites 2025 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2025 , eprint=

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.783172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.783172Z digest=sha256:ab7155dacbcd468c443dde08b930c73de5465bc2bded61b8533c21529168da39

Observation 1d69bc32-d0ce-410f-bea8-2561a50cf2eb · outbound

This paper cites 2025 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2025 , eprint=

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.796752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.796752Z digest=sha256:a72529b90cb7f94594f981b736a4a0f3d7ec4bbe3a9aaf01bd8a0ed735588acd

Observation f7669d75-4a38-44d6-b00f-fa8d48395e8c · outbound

This paper cites 2025 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2025 , eprint=

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.802821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.802821Z digest=sha256:3fb75257bc238f344b83d389f944db17a29412216943abe439145d8d2b96c4b9

Observation afb8e98f-fbc7-4cd6-bd49-cc06cdef7786 · outbound

This paper cites 2023 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2023 , eprint=

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.812967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.812967Z digest=sha256:e0055e5e1e60505ccee495f4144e232cdc7787052479755b8e2ff8b359753197

Observation ad7425d2-83b8-4f67-86a0-1c4309d05b77 · outbound

This paper cites 2026 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2026 , eprint=

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:35.147811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:33.826957Z digest=sha256:af36c0c5034ee9c799115e9ed054fe6f2ec17055a5e6eee9dd965c813dbf0d73

Observation 434ab39b-75c8-4f47-98cf-9971e41f8b79 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning OpenVLA: An Open-Source Vision-Language-Action Model

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.843856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.843856Z digest=sha256:7d7a42fe443658c7d3f458ede117a3d78b7306828466c0cc8a24d10fa9c4ef54

Observation 899c45df-f990-45da-8988-4ce91fb96a9d · outbound

This paper cites ReinboT: Amplifying Robot Visual-Language Manipulation with Reinforcement Learning.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning ReinboT: Amplifying Robot Visual-Language Manipulation with Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.850650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.850650Z digest=sha256:5b58ac35360b8df9d33f02af0451ed6ac167d5efc9e686771f4952f1fa251239

Observation ddecd92c-4332-4e6e-96a4-82b697484ac0 · outbound

This paper cites SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.858488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.858488Z digest=sha256:fe821811823adf83cfd575c3a1d7bbc3904d383f7e27a69e35f6f6549071dafd

Observation 7ab6a76f-501a-484c-9f1f-5cd8834cc9ca · outbound

This paper cites arXiv preprint arXiv:2511.09515 , year=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning arXiv preprint arXiv:2511.09515 , year=

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.867217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.867217Z digest=sha256:61b3460fbf41b3abbbbb43cba4499663f50212fa1584b105d60eac0d0f82a967

Observation ab3ebb63-bbb8-4840-8d38-a33ff3daeb17 · outbound

This paper cites arXiv preprint arXiv:2510.00406 , year=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning arXiv preprint arXiv:2510.00406 , year=

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.877681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.877681Z digest=sha256:a1b22cc60363aa012a3f85f1139a6aab45b3def08fd63570a3e265daa7559bfb

Observation ff355e59-c5f4-4dee-a610-8c75c9d1ce71 · outbound

This paper cites arXiv preprint arXiv:2511.00091 , year=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning arXiv preprint arXiv:2511.00091 , year=

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.889487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.889487Z digest=sha256:d790419fb32bf2c6de4e917b83118b031ae000f2446c396fb7cbdbcadc8c506e

Observation 463ffd35-cc7a-4d8a-9694-0c890eb33c4d · outbound

This paper cites RobustVLA: On Robustness of Vision-Language-Action Model against Multi-Modal Perturbations.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning RobustVLA: On Robustness of Vision-Language-Action Model against Multi-Modal Perturbations

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.899884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.899884Z digest=sha256:a9901b5367fc4fbbde04f9189f42d8d6478e15d948e241631f7b14dd8fd8b7cf

Observation f9f1cae7-c5cd-4a06-ab10-9b978c160cc7 · outbound

This paper cites Interactive Post-Training for Vision-Language-Action Models.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Interactive Post-Training for Vision-Language-Action Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.909439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.909439Z digest=sha256:608f60b18910d1aabfddbfa641e8fe897ff1ea2cf5bfa2b0d211ce0981796991

Observation 7c113e5e-74d7-4871-9a54-1e942ab25433 · outbound

This paper cites Advances in neural information processing systems , volume=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Advances in neural information processing systems , volume=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.927727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.927727Z digest=sha256:3176fce7288d835e8951629907a22249e584baeae451b3b83f8a4ed72491cd63

Observation a3b986ed-abf6-4da0-bb13-9bb623b6c8ef · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Advances in Neural Information Processing Systems , volume=

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:35.079596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:33.947351Z digest=sha256:54aeec75a3cb4207446f28cb660aa0f53588446a6027e470faf4ec3d39feb194

Observation d60282f9-1343-48c8-9ff0-4e97a6bd045a · outbound

This paper cites Align-RUDDER: Learning From Few Demonstrations by Reward Redistribution.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Align-RUDDER: Learning From Few Demonstrations by Reward Redistribution

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.953684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.953684Z digest=sha256:2218d89ccb0fcf83e2bade1b5155457276ba64622bf858cb9707e12fd3fbd72d

Observation cad20952-1a16-4d52-b11c-c675fbc8d1ff · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Advances in Neural Information Processing Systems , volume=

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:35.053161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:33.964464Z digest=sha256:af0c6e4f3b15cd1e3f358e7adaded57130206582500de738e011cfe33aa8ff30

Observation e5a35949-7b26-4bf3-b165-64b8ebb71333 · outbound

This paper cites Reinforcement Learning with Segment Feedback.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Reinforcement Learning with Segment Feedback

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.975739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.975739Z digest=sha256:01f9a6f3c5337c63822c1f468be70abb98c0e652693991ba20354ddf260aa1ec

Observation 79542b24-c7f5-4ec2-b3f4-de899a5c6fe8 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Advances in Neural Information Processing Systems , volume=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.989588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.989588Z digest=sha256:98a57c41307ae7ba53d1bd34b9693640a0aea5be765ba361cf9de66befff4e13

Observation 0395fe38-777a-41f1-8c98-d990af966289 · outbound

This paper cites SARM: Stage-Aware Reward Modeling for Long Horizon Robot Manipulation.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning SARM: Stage-Aware Reward Modeling for Long Horizon Robot Manipulation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:34.003593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:34.003593Z digest=sha256:37092a6cfb2753e3f9e8af9945c61ecc6d8e12b039054b2c4ebbb30b0a887197

Observation 189c229c-13e8-4f0b-a5f7-f05ecb87fa2e · outbound

This paper cites International Conference on Machine Learning , pages=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning International Conference on Machine Learning , pages=

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:34.971063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:34.015049Z digest=sha256:f960017ea9c38c88ab319e492f4d3610679ee9a548fa93a674f8f4bdbddfbf05

Observation 6b6b3e3d-3f48-404e-9d3f-d6cb41d27881 · outbound

This paper cites International Conference on Machine Learning , pages=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning International Conference on Machine Learning , pages=

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:34.930931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:34.023324Z digest=sha256:8f582dc1885c42b442358448462566492ade500a845c9cf036bb6e815aa7bbb9

Observation 05e21ac1-ce4c-4a78-a8e7-e3e48570d794 · outbound

This paper cites Conference on Robot learning , pages=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Conference on Robot learning , pages=

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:34.891557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:34.029391Z digest=sha256:9687111acd54079789a222240f0ff1f8254b2499c35a64fa2b1e3c0316490137

Observation 86dd38a3-2008-47c8-ac77-b67ebdc99271 · outbound

This paper cites Conference on Robot Learning , pages=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Conference on Robot Learning , pages=

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:34.855099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:34.040918Z digest=sha256:363a6c738bd7b634b943b08438ff2f8fc4e2bab5ade62952f21c6aa250da44a0

Observation 1b351643-2240-4a64-a884-054ec0cd7cd7 · outbound

This paper cites Conference on Robot Learning , pages=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Conference on Robot Learning , pages=

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:34.816203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:34.048526Z digest=sha256:07836e22dd73bb59e751c257154a34a19b861aacdd78dcf245e9fc355b8a7e0e

Observation b4367387-7bd3-45b6-988d-1b16a760ed11 · outbound

This paper cites 8th Annual Conference on Robot Learning , year=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 8th Annual Conference on Robot Learning , year=

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:34.780841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T18:21:34.060079Z digest=sha256:0a234d725593ac66ad2d18b2dd3310b73cdb4ef2b49dd2bf5811364d03821d6f

Pith citing papers

No inbound Pith citation observations are available.