Pith. sign in

Paper Citation Record · LEDGER

JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2311.05997.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.05997 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:18:28.360265Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

6
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7f5ae896-6dc5-4d83-a399-e905b05d9976 · inbound

AppAgent: Multimodal Agents as Smartphone Users cites this paper.

AppAgent: Multimodal Agents as Smartphone Users JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-17T10:16:43.859010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-17T10:16:43.364787Z digest=sha256:afd5a5455bda0849e8b23e12d6279c3f9455057753609cb33f756416c34b4525

Observation 15526ff7-f592-46b4-97f6-d0f44ac0f41c · inbound

A Survey on the Memory Mechanism of Large Language Model based Agents cites this paper.

A Survey on the Memory Mechanism of Large Language Model based Agents JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 159

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:21:39.639525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T07:21:39.440092Z digest=sha256:6a91d99801b40a451ff8dc5dfe2b3a7917a05f1e5f2666431605eb534f385e58

Observation 824b9835-d3f1-4f75-a8c2-726ef302f0a7 · inbound

Preference Goal Tuning: Post-Training as Latent Control for Frozen Policies cites this paper.

Preference Goal Tuning: Post-Training as Latent Control for Frozen Policies JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:22:44.348038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-23T08:20:05.898025Z digest=sha256:c0ee50647b448d78b720548f4035b64fb1d2fa6f2b74ed00e1f03c2d53aebced

Observation 7d1ff9b9-edd1-4000-a3c4-76583f6e71b5 · inbound

Conditional Multi-Stage Failure Recovery for Embodied Agents cites this paper.

Conditional Multi-Stage Failure Recovery for Embodied Agents JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:28.360265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:18:28.360265Z digest=sha256:6f3b83321dac441742ba40e1d419e7c1710e4f629c8038189a3c4e6e96fd1c9d

Observation 3a0e4777-3c75-4ded-818a-f0f47d392161 · inbound

SoK: Agentic Skills -- Beyond Tool Use in LLM Agents cites this paper.

SoK: Agentic Skills -- Beyond Tool Use in LLM Agents JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:19:31.125901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T23:19:31.024268Z digest=sha256:ad44efae08b3d623b404e7741e66796c681a627df0a22262761ba87e9e0cd092

Observation 739e99f3-24a0-4af3-9f5c-d42b8a4c71a9 · inbound

GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents cites this paper.

GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:46:07.697516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T17:57:36.038091Z digest=sha256:abf43e5411a8b22639e60b7e3da6d09061edfce59ebf0e6c94e4d2cf855bab28

Observation 52b44462-e38f-410a-be05-86e6d43ddfeb · inbound

Dream-Cubed: Controllable Generative Modeling in Minecraft by Training on Billions of Cubes cites this paper.

Dream-Cubed: Controllable Generative Modeling in Minecraft by Training on Billions of Cubes JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:31:02.641583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T01:36:23.517745Z digest=sha256:b33dc7d204c0993a16ffdae23ae45f633fda55cdbe680616d32d85af05e4880b

Observation 7c497296-ea37-4c23-811e-5fa4a4a2e53d · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:20:57.243431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T01:47:39.926540Z digest=sha256:ab1f93883a75d32f569697e95447312a4315ca4ba034824fbe9b1ff7526a0e49

Observation 3693cabc-54e7-4662-ad09-4ad65c97457a · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:19:15.144913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T23:15:44.550045Z digest=sha256:72e1aa4efbe747802c32852f97e6bf23927fa08b70d208d5081cbc3063e0dc68

Observation fe68687f-7002-4018-be4d-a9fef70d153a · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:25:07.322279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T23:23:42.883286Z digest=sha256:4a38367501fdc1cf3f5c727f90c394c4e1658d8a34598a5c2d9373de85ac74ff

Observation 9a54de7f-832a-45f2-a6a5-5807ba3488a9 · inbound

Do Vision-Language-Models show human-like logical problem-solving capability in point and click puzzle games? cites this paper.

Do Vision-Language-Models show human-like logical problem-solving capability in point and click puzzle games? JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T02:02:05.547051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T01:58:39.476408Z digest=sha256:a3ed0f222bb4e90ec31f2b638a973da35f908da526d02074bf5503bcbfc58397

Observation 38091461-5cb6-4fdf-9737-c8aeff9360c9 · inbound

SPIKE: An Adaptive Dual Controller Framework for Cost-Efficient Long-Horizon Game Agents cites this paper.

SPIKE: An Adaptive Dual Controller Framework for Cost-Efficient Long-Horizon Game Agents JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:33:12.266648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T10:32:26.668583Z digest=sha256:fa42603dcb6b13e7dc27a3250da7bd9ba592dee3de8b34dab02fc54138834cc9

Observation 0ce76651-5009-4389-bde0-20a88a50eefc · inbound

HiMe: Hierarchical Embodied Memory for Long-Horizon Vision-Language-Action Control cites this paper.

HiMe: Hierarchical Embodied Memory for Long-Horizon Vision-Language-Action Control JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-12T02:24:21.020383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T02:24:21.020383Z digest=sha256:0648bea3ae4b48d1000c196bcbc79c6fda2f8b695e6b9077171d3125106f8587

Observation 1a894190-d2c8-4030-ae46-267294b59629 · inbound

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills cites this paper.

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills JARVIS-1: Open-World Multi-task Agents with Memory-Augmented Multimodal Language Models

Reference 259

Resolution
unresolved
no resolver link, observed 2026-08-04T19:45:35.319727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:45:35.319727Z digest=sha256:cccd317e4944d69d5af13f586807aafa88ee30fd7c49f22be769d9e987ccd6a2