Pith. sign in

Paper Citation Record · LEDGER

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations

As of 19 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 8 inbound Pith citation observations for arXiv:2412.01441.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.01441 v3

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T04:25:49.930483Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:16:29.737125Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T07:14:02.611067Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy15
  • unresolved18
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4f61bd1f-ad53-41f9-96b2-f37d4ab0f81f · outbound

This paper cites The claude 3 model family: Opus, sonnet, haiku.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations The claude 3 model family: Opus, sonnet, haiku

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:25:50.802401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T04:25:49.738084Z digest=sha256:517f95cccd01140b948c9553df4433feb930a5f63b33811a87fc4aea44e41d71

Observation c972dd1b-830c-4305-8ec5-b3f383934511 · outbound

This paper cites Chen, L., Lu, K., Rajeswaran, A., Lee, K., Grover, A., Laskin, M., Abbeel, P., Srinivas, A., and Mordatch, I.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations Chen, L., Lu, K., Rajeswaran, A., Lee, K., Grover, A., Laskin, M., Abbeel, P., Srinivas, A., and Mordatch, I

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:25:50.786495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T04:25:49.745530Z digest=sha256:f6d899571bea00c880c8509c991ea1d041a73a7d8b0e1d4ac2f64f0270cfc608

Observation 9019d9a9-aea7-4e60-85e4-7007c4733807 · outbound

This paper cites Jack of All Trades, Master of Some, a Multi-Purpose Transformer Agent.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations Jack of All Trades, Master of Some, a Multi-Purpose Transformer Agent

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T04:25:49.771582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:25:49.771582Z digest=sha256:677c3fd6cb6222fe21309b710d76f3bc2a905f533d96bef4a76bb09a0f82a717

Observation 8655f358-a42e-46be-ad5c-c745ac6fac42 · outbound

This paper cites W., Grau-Moya, J., Ruoss, A., Orseau, L., and Hutter, M.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations W., Grau-Moya, J., Ruoss, A., Orseau, L., and Hutter, M

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T04:25:49.778486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:25:49.778486Z digest=sha256:ac50cb80ef4db61bc8d3435bbf242fd7922a4a5f47a4d960cce981fa0e29546a

Observation 6bc53b5d-ec16-4ccd-9b37-85846dff2984 · outbound

This paper cites Accessed: 2024-11-26.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations Accessed: 2024-11-26

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:25:50.769300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T04:25:49.791488Z digest=sha256:94faba974285afc703ec1ab85f28653d4e40c9b6a14ab464b0f102f27f99cbcc

Observation 5e2c2bec-83fc-4127-ad9e-2994b8239687 · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations Playing Atari with Deep Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T04:25:49.803221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:25:49.803221Z digest=sha256:f18a0f502c638d3081b66d0edff9a59548467be10fc71db0a78038c8ed491562

Observation 5599ae45-43e0-4b91-bc7b-f6bc928a0338 · outbound

This paper cites Hello gpt-4o, 2024a.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations Hello gpt-4o, 2024a

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:25:50.743265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T04:25:49.813528Z digest=sha256:3510293eba303ea8254f0861c8deab739288b5d46b9fd317b248a9a65e7c56f8

Observation a8736a3a-d482-48c5-9bb7-9f1cc3d25f66 · outbound

This paper cites GPT-4o System Card.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations GPT-4o System Card

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T04:25:49.826543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:25:49.826543Z digest=sha256:983728276a46225007b2eb25de68368f01bc45d14bbfb0bf1374d09636453af3

Observation 5b1d4099-aad8-486b-8ac0-268cf01704f5 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T04:25:49.841132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:25:49.841132Z digest=sha256:1c67fdbdcf1b836218cf893047a1d9900ff4bb3b0eb550d32772b12570da1f79

Observation 0cb6b272-845c-4156-8a36-153a9182a024 · outbound

This paper cites Scaling Instructable Agents Across Many Simulated Worlds.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations Scaling Instructable Agents Across Many Simulated Worlds

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T04:25:49.845816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:25:49.845816Z digest=sha256:e71b7bcbe86980dc61570755e0e19160c8905df02c4e601e5a000d41200475a2

Observation b7c0c764-1c8e-4b5d-873f-c5948c89303d · outbound

This paper cites REGENT: A retrieval-augmented generalist agent that can act in- context in new environments.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations REGENT: A retrieval-augmented generalist agent that can act in- context in new environments

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:25:50.709091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T04:25:49.852305Z digest=sha256:5393ea917035cc22bbc8bc9446131dc76e2af102252a627bd0c34ad1ffa8d4ad

Observation 1da024ce-5f1e-4160-baa2-d27fe26bfac3 · outbound

This paper cites DeepMind Control Suite.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations DeepMind Control Suite

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T04:25:49.858050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:25:49.858050Z digest=sha256:ecfd09b3d214b2ce7b739a3d0385a70ae278e48b29103b97c2ca580c30bd3548

Observation 1f4e1671-b4d2-4d94-b5e1-11d765fda8e1 · outbound

This paper cites Atari-GPT: Benchmarking Multimodal Large Language Models as Low-Level Policies in Atari Games.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations Atari-GPT: Benchmarking Multimodal Large Language Models as Low-Level Policies in Atari Games

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T04:25:49.863261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:25:49.863261Z digest=sha256:b128068321db0eb4292078fcbddd775a86ee980888facc7322b7ab043f50ea2e

Observation d0c58b0c-5663-4a1c-8322-e72e73e1e34f · outbound

This paper cites Why is prompting hard? Understanding prompts on binary sequence predictors.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations Why is prompting hard? Understanding prompts on binary sequence predictors

Reference 23

Resolution
metadata mismatch
local_arxiv, observed 2026-08-12T04:25:50.006761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T04:25:49.867988Z digest=sha256:d5b1ebb0f6cc8b9bc70e24a918820072ea63ab066b9e1cfb2708a1b95ef80810

Observation d837c2d6-547b-4f1d-b216-4caff4549492 · outbound

This paper cites In-Context Learning Enables Robot Action Prediction in LLMs.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations In-Context Learning Enables Robot Action Prediction in LLMs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T04:25:49.873674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:25:49.873674Z digest=sha256:4879ba7c2708f7a7f658517d296d7b1a5c34b1927a62a2cb664fe4d5097fde58

Observation 27123ea8-b373-4ab1-8fb3-d83d52b2e668 · outbound

This paper cites adaptive agent.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations adaptive agent

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:25:50.691641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T04:25:49.879119Z digest=sha256:a548cb0210b2d08ea6274130652cd1ae667ab40ad927ef29a8770cfeec23919c

Observation cabd335b-15b9-48c6-b44c-25bfb097863a · outbound

This paper cites Models, 2024b.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations Models, 2024b

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:25:50.724938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T04:25:49.819865Z digest=sha256:8cbadf900ab390097395e2494180f6954c87d2139f6ff9ba8f98375a656546fc

Observation bce1da48-2685-4895-b1fa-651c5b48b828 · outbound

This paper cites general pattern machines.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations general pattern machines

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:25:50.673815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T04:25:49.884724Z digest=sha256:7cf648b9c03e6b81ac5dcd8ef9703c8d3a8591ebc9ea498cbfc3f78a14191627

Observation 9c43d9df-290f-4aa9-aef0-87d9c3ae5b83 · outbound

This paper cites Take a deep breath and work on this problem step-by-step.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations Take a deep breath and work on this problem step-by-step

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:25:50.656669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T04:25:49.890027Z digest=sha256:ff73eac6022ae9f0d33572beb25ed58bf5e6ea3f0355f6387a95543e98c1fcb3

Observation cd11a01b-05eb-431e-8282-1a1322630f38 · outbound

This paper cites An alternative could be to virtually extend the context via retrieval based methods, e.g., REGENT (Sridhar et al., 2024), which trains a retrieval based agent.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations An alternative could be to virtually extend the context via retrieval based methods, e.g., REGENT (Sridhar et al., 2024), which trains a retrieval based agent

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:25:50.636976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T04:25:49.895673Z digest=sha256:ed6e03c3a18d083deaa2ff4d9cfd5409c7d11bfd4bc8578360e314dada86d271

Observation 138eb082-53fc-4511-84b2-5e4407e9703a · outbound

This paper cites knowing-doing.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations knowing-doing

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:25:50.618955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T04:25:49.900596Z digest=sha256:1fc6df2840b5be14ff7fd1f3e011a02b099ed59a126a038cb81063f87a5af79e

Observation d14d9c31-7195-4636-abbb-440a5d288974 · outbound

This paper cites an unresolved cited work.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-12T04:25:50.579808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T04:25:49.910816Z digest=sha256:31dcd59352ead7e266253da5547b8d0d80d2f7a77f8a2e114997ea11740af1ea

Observation fe074856-3feb-45e8-8f7a-8f5ba5c23743 · outbound

This paper cites sanity check.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations sanity check

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:25:50.556921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T04:25:49.915379Z digest=sha256:900cf010639758488f564b0d01caa2d4009bf5e847cc59f0920148f597a2b38c

Observation 7d9e2f77-2ee8-4581-866f-a9a74921de94 · outbound

This paper cites For completeness, we present all the ablation results in Tables A1 to A10.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations For completeness, we present all the ablation results in Tables A1 to A10

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:25:50.534528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T04:25:49.920178Z digest=sha256:1ebe82fff9ada1761dcfc395eaa536d5c175a8ac8062431a41c9080f91116847

Observation 4c701c19-b17f-4680-acb2-cafc73ef8c2a · outbound

This paper cites Note that o1-mini and o1-preview are text-only models and, therefore, cannot process RGB image observations.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations Note that o1-mini and o1-preview are text-only models and, therefore, cannot process RGB image observations

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:25:50.517992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T04:25:49.925459Z digest=sha256:0b81e7f4f2f68c9fd874b2007dd4ef636c50501581b1beb5f965ddd3753128b2

Observation eb2f2c57-2af6-4856-93b7-855676063cac · outbound

This paper cites Note that o1-mini and o1-preview are text-only and, thus, cannot process RGB observations.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations Note that o1-mini and o1-preview are text-only and, thus, cannot process RGB observations

Reference 35

Resolution
malformed identifier
raw_fallback, observed 2026-08-12T04:25:50.500345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T04:25:49.930483Z digest=sha256:4d612b5f355cf74201d3f42509c27c5c58b6ac09c2df258db633de39d86a317e

Observation 23f837ab-237e-4c3b-aa5f-bc9cee06ade0 · outbound

This paper cites In-Context Imitation Learning via Next-Token Prediction.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations In-Context Imitation Learning via Next-Token Prediction

Reference 2012

Resolution
unresolved
no resolver link, observed 2026-08-12T04:25:49.765864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:25:49.765864Z digest=sha256:3163ebeffc89b0640b313904705ddab4d785917b468f2e27a081f9b83b327fe5

Observation b73a40ec-83e6-4379-817f-dd0508e80288 · outbound

This paper cites GPT-4 Technical Report.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations GPT-4 Technical Report

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-12T04:25:49.808632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:25:49.808632Z digest=sha256:551980f158f6c0229c4943c8c339a3a8bcde3f8c69ea1dbb11cb0d5a95dcdeb9

Observation b08283de-c09f-4cc1-9ee3-0ff735c69d59 · outbound

This paper cites Shaking the foundations: delusions in sequence models for interaction and control.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations Shaking the foundations: delusions in sequence models for interaction and control

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-12T04:25:49.833514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:25:49.833514Z digest=sha256:a467d9119aff3448ef219cbfe177a9b2e1c46b5fbfce82cc6187f21dbc4f9419

Observation e6a6ccd1-e2cc-495c-b0e9-328d266b3a38 · outbound

This paper cites K., and Panov, A.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations K., and Panov, A

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-12T04:25:49.751934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:25:49.751934Z digest=sha256:f8e5a650b3505dad17e7a6bcb22e4e89630caa0447b816ab54dd3745fb4807db

Observation bda496a3-7261-4297-a81d-5ad25e3707ea · outbound

This paper cites Many-Shot In-Context Learning in Multimodal Foundation Models.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations Many-Shot In-Context Learning in Multimodal Foundation Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-12T04:25:49.796437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:25:49.796437Z digest=sha256:cabbc608a95907c687ef43b49f978399e85c1d4c813534549b4124bd8677eb6c

Observation c75bc4dc-c733-4b01-8752-fb58feb5b1c0 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations Gemini: A Family of Highly Capable Multimodal Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-12T04:25:49.731576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:25:49.731576Z digest=sha256:312f3cc694026e8e3b8bd5173c62a17ff923dcc878c8c10f8a99ea28c3548df1

Observation cbe054a6-9fd5-495c-af01-a765058d9605 · outbound

This paper cites GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-12T04:25:49.759564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:25:49.759564Z digest=sha256:b40a11be1f99b800dd2d8f423285cf5fc031c23c7b595d5c48c23f0d34339180

Observation 1f7f29f8-9d01-4dff-8441-24baffdc9061 · outbound

This paper cites FrontierMath: A Benchmark for Evaluating Advanced Mathematical Reasoning in AI.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations FrontierMath: A Benchmark for Evaluating Advanced Mathematical Reasoning in AI

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-12T04:25:49.785818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:25:49.785818Z digest=sha256:9e43ab30c7626ccd5b9b0acfda4c5364499000facfb1e545840a333afc1236a0

Observation 47f14cad-dd49-4fec-88bd-3e7a3f3c2ec6 · outbound

This paper cites Sample observations for each state representation format from our tic-tac- toe environment.

LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations Sample observations for each state representation format from our tic-tac- toe environment

Reference 2600

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:25:50.599979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T04:25:49.905721Z digest=sha256:040170a0c184449deae6ca38b885ec581e259a4a7fdbe2b83a0a0d1d5f5c43c7

Pith citing papers

Observation 8e1248b0-a985-460f-a564-2585acc1bb8c · inbound

BALROG: Benchmarking Agentic LLM and VLM Reasoning On Games cites this paper.

BALROG: Benchmarking Agentic LLM and VLM Reasoning On Games LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-12T16:21:08.336824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T16:21:08.336824Z digest=sha256:2188d684a796f1b22bec84ba7bf5050846f66e705646ce3b70ce59d6ac6244ea

Observation a2c90cd5-e3bc-40e4-adfd-d413b1486859 · inbound

Harnessing Language for Coordination: A Framework and Benchmark for LLM-Driven Multi-Agent Control cites this paper.

Harnessing Language for Coordination: A Framework and Benchmark for LLM-Driven Multi-Agent Control LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T14:42:18.143488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:42:18.143488Z digest=sha256:5762919cb642539e647d69bb022950415f80c4a272a682160e993f51c1390acf

Observation 13e8c5e2-ae42-4c78-92bb-ccdfaf60b886 · inbound

Mastering Board Games by External and Internal Planning with Language Models cites this paper.

Mastering Board Games by External and Internal Planning with Language Models LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T00:59:58.985560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:59:58.985560Z digest=sha256:c9c3df6a95e3fa6923f2ba190e882aa4809b0b444974ebb5e8180f7c8ae5c823

Observation e6abf4a5-82f5-48b3-b248-ead0e00d8691 · inbound

Disentangling Exploration of Large Language Models by Optimal Exploitation cites this paper.

Disentangling Exploration of Large Language Models by Optimal Exploitation LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T20:18:58.431750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:18:58.431750Z digest=sha256:deed8edea0f65e50ee90e2c23c96058e881f96fa6b9bf98c23b949e70bfd93e5

Observation 28463934-56eb-473d-b7ed-8149deb97f6b · inbound

LLMs are Greedy Agents: Effects of RL Fine-tuning on Decision-Making Abilities cites this paper.

LLMs are Greedy Agents: Effects of RL Fine-tuning on Decision-Making Abilities LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T11:16:29.737125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:16:29.737125Z digest=sha256:6ec05f141dec380d452c9b2f1ec81a9e6abce3d7a70dc7f138f0b8b660f75fac

Observation 4ac9bc53-09db-4ea1-a4e6-a698a1049401 · inbound

The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity cites this paper.

The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:10:31.567261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T16:10:31.440921Z digest=sha256:180e0028a65ff72cbe82cd88082acf3828a04b5e77125470f73806becddde2f9

Observation 9fa36af7-2daf-453a-8031-c0bfb98d7f1b · inbound

Instruction Agent: Enhancing Agent with Expert Demonstration cites this paper.

Instruction Agent: Enhancing Agent with Expert Demonstration LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:20.912878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T22:55:20.912878Z digest=sha256:7d6f2fd003a41e005577c9dcac57d738a182662e8a8eaf94cb1cb13bc2ff4665

Observation fb825063-655f-4ad6-98ce-cc7398056407 · inbound

Training Language Agents to Learn from Experience cites this paper.

Training Language Agents to Learn from Experience LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:14:02.612567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T07:11:09.642275Z digest=sha256:f0bae0b867e776634eb6cc167cb744289d7384d83efeaef8f348160c9452925b