Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T04:17:00.065856Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 1 inbound Pith citation observation for arXiv:2505.01713.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T04:17:00.065856Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T17:55:44.834170Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-12T17:55:45.159103Z
40 of 40 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b23a21c9-dacb-4034-ad02-8fe9e7aa55fd · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45cdc215-036c-47dd-8de3-c187714b5a6e · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation The epic-kitchens dataset: Collection, challenges and baselines
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c903edfa-5d06-45ad-b073-80e193ab1f0a · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation A technique for computer detection and correction of spelling errors
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation bf4fbe5e-102b-4948-b6cf-40adccfb6ceb · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Rolling-unrolling lstms for action anticipation from first-person video
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 966d78ff-85f6-41c5-b0a4-f39ae2fa16f0 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Future transformer for long-term action anticipation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 134fc880-21de-436e-a5c3-d78ba8ae891a · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation LoRA: Low-Rank Adaptation of Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1b701bd-79f4-4c23-a6b0-d0bc66b07881 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Technical report for ego4d long term action anticipation challenge
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b838ab20-a63b-4e46-81f7-e4fa068c7359 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Technical Report for Ego4D Long Term Action Anticipation Challenge 2023
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 346cb39d-1f4a-47f4-856e-9d8bfb4f1ee6 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Anticipating the start of user interaction for service robot in the wild
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 493db213-0aa1-4eb4-867e-9e081f79fb03 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Palm: Predicting actions through language models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c29b4694-c060-4178-a363-3202275621d1 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Anticipating human activities using object affor- dances for reactive robotic response
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e76c131d-b624-4bc3-9f5a-612d6ab8fbf1 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Llama-vid: An image is worth 2 tokens in large language models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 54a044e8-3d9d-46ed-9577-4eee2b2bc33d · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Can’t make an omelette without breaking some eggs: Plausible action anticipation using large video-language models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ac04f3ad-eaf4-4502-96d5-190c0a1511db · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Ego-topo: Environment affordances from egocentric video
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4e59dfc2-f5c3-40e6-a55a-a4b39842175d · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Rethinking learning approaches for long- term action anticipation
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2cc08456-6da6-47d8-82a4-11d7a56e6d06 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Summarize the past to predict the future: Natural lan- guage descriptions of context boost multimodal object interaction anticipation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 089ac1e6-e416-4e2d-ba63-2d0d867454b6 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation EgoVideo: Exploring Egocentric Foundation Model and Downstream Adaptation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4702ef4c-8203-408b-9923-3c6b4558219d · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Learning transferable visual models from nat- ural language supervision
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09e33535-a780-44d2-b0b7-5cc90c7db377 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Pre- dicting the future from first person (egocentric) vision: A survey
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d3bb4032-a725-4d2c-8655-b13187b891e1 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation En- couraging lstms to anticipate actions very early
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 16c508a7-dc07-4503-8466-14239cd0594a · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation LLaMA: Open and Efficient Foundation Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cbdc48c-decb-4d83-b840-7caa03f667ff · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Attention is all you need
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb7b1e05-d3c1-452c-b982-e326b95fba8c · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Memory-and-anticipation transformer for online action understanding
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation bd421d6a-6661-470d-a005-533be7b31b9e · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Visionllm: Large language model is also an open-ended decoder for vision- centric tasks
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a4cae2e8-dbf4-46f2-8399-6c957c1fc5fb · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Black-box prompt tuning for vision-language model as a service
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0aacd786-8ec1-4f77-87aa-06406453995e · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation AntGPT: Can Large Language Models Help Long-term Action Anticipation from Videos?
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b5082f4-f0f6-471b-a946-3db1f8ecd69b · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Anticipative feature fusion transformer for multi-modal action anticipation
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 380a38c4-06a5-40fc-9484-30e59dd93977 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 9f27aa03-8246-463c-ba20-1948ea37e648 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation QueryMamba: A Mamba-Based Encoder-Decoder Architecture with a Statistical Verb-Noun Interaction Module for Video Action Forecasting @ Ego4D Long-Term Action Anticipation Challenge 2024
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47abcfae-6f50-438f-9b37-0320acd41712 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation The Llama 3 Herd of Models
Reference 1964
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 949ff927-2b61-4868-908f-17012bb77f5d · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation In the eye of beholder: Joint learning of gaze and actions in first person video
Reference 2015
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation aacadb1e-17b2-429d-98c5-224a8eeea937 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Temporal aggregate representations for long- range video understanding
Reference 2017
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ac976ea0-fd66-4dfc-a8d4-9caf30491c62 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Blip-2: Bootstrapping language-image pre- training with frozen image encoders and large language models
Reference 2018
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c9658f80-6d31-4319-83dd-b9d99b0cb6ed · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation VideoGraph: Recognizing Minutes-Long Human Activities in Videos
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bec848c-3850-4325-bb76-ae7419bda8a5 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Sgdcl: Semantic-guided dynamic correlation learning for explain- able autonomous driving
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 29be6859-98f4-457e-bdce-8813f3155c1e · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Using gaze patterns to pre- dict task intent in collaboration
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 6ccbc618-a512-4251-bba8-0a34906fb4d4 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Ego4d: Around the world in 3,000 hours of egocentric video
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 84f25e48-fe3f-435f-b837-06d6e38768f7 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Language models are few-shot learners
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79d26287-c79d-4db5-9add-acee4bed298e · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation A survey on multi- modal large language models for autonomous driving
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0b1b8fc5-8d06-4414-b6b2-03dd8cd766a8 · outbound
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation Ego- centric video-language pretraining
Reference 2025
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c57a096f-e4b6-494f-b811-0b43be3e0627 · inbound
Compositional Benchmark Synthesis for Hierarchical Human Action Recognition Vision and Intention Boost Large Language Model in Long-Term Action Anticipation
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.