Pith. sign in

Paper Citation Record · LEDGER

InSTA: Towards Internet-Scale Training For Agents

As of 15 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 11 inbound Pith citation observations for arXiv:2502.06776.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.06776 v2

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T14:24:50.484160Z

measured 69 of 69 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:19:15.802502Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

58 of 58 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved53
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation ee2d6038-0004-403e-9fdb-8f90a9517aaf · outbound

This paper cites write newline.

InSTA: Towards Internet-Scale Training For Agents write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.283326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.283326Z digest=sha256:580ee77816b2bfb40b368874482beb77a7f645cf125e82f4a0a0af24d1f4e13e

Observation 073770c8-88c3-430a-9ebf-4f4815ffca7a · outbound

This paper cites @esa (Ref.

InSTA: Towards Internet-Scale Training For Agents @esa (Ref

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.288008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.288008Z digest=sha256:20e505c9cfa24e11f43c37600dff9f5b16e95b1008a7df4d89a3a8ac2c1ad92d

Observation d92b4c8b-a731-4d17-a5b4-2f3b2d85bdbd · outbound

This paper cites an unresolved cited work.

InSTA: Towards Internet-Scale Training For Agents Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.291843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.291843Z digest=sha256:af6d9d6461d19e227cd5ae97d43b647563ebc8bb55d1e4232fbbe37cecb47e0f

Observation 92af65ee-1f28-4220-95d4-5948aafd16dc · outbound

This paper cites an unresolved cited work.

InSTA: Towards Internet-Scale Training For Agents Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:24:51.337301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.295342Z digest=sha256:e8431856ddc9d7cdbad821ffae96bd60b5f4bb160cf837f890dc9d2138880193

Observation f3921423-24e3-4645-b9ce-40b36b05b58a · outbound

This paper cites Language models as agent models.

InSTA: Towards Internet-Scale Training For Agents Language models as agent models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.298819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.298819Z digest=sha256:132c8aac51910f57e4fd0a412121ab2b64ad8d9330d01d3b424a323f25e8083f

Observation fbe027e7-122d-42f4-b87a-3849bd3aa532 · outbound

This paper cites Graph of thoughts: Solving elaborate problems with large language models.

InSTA: Towards Internet-Scale Training For Agents Graph of thoughts: Solving elaborate problems with large language models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.302297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.302297Z digest=sha256:5b9400a41b59883736ee118557a3045b66d970f446dec10a3a90f19038c99bb0

Observation cfde26e3-8a15-42c2-bd97-d056ffa81aa4 · outbound

This paper cites Language Models are Few-Shot Learners.

InSTA: Towards Internet-Scale Training For Agents Language Models are Few-Shot Learners

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.305890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.305890Z digest=sha256:849705d517e31ad634f52141093df01862c2ea405f251f21f96db872ebb00352

Observation 0305d162-35a3-4e8b-871f-307ab708f6e6 · outbound

This paper cites Sparks of Artificial General Intelligence: Early experiments with GPT-4.

InSTA: Towards Internet-Scale Training For Agents Sparks of Artificial General Intelligence: Early experiments with GPT-4

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.309685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.309685Z digest=sha256:ef091dfd250fd1f637f8c782de2b7b8686d1e1ee866aa4ace6cdd192e15a9dc7

Observation f718996a-ad8b-406c-8ffd-65ecf8b4eb47 · outbound

This paper cites FireAct: Toward Language Agent Fine-tuning.

InSTA: Towards Internet-Scale Training For Agents FireAct: Toward Language Agent Fine-tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.313595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.313595Z digest=sha256:33aaa0b651f90b4082b6d3c9d340aac61effd07a68ca113fe20ecf5bc2725e84

Observation c110ffad-3e6e-48ed-8d41-623f533a19cd · outbound

This paper cites The BrowserGym Ecosystem for Web Agent Research.

InSTA: Towards Internet-Scale Training For Agents The BrowserGym Ecosystem for Web Agent Research

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.317546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.317546Z digest=sha256:8b04b74c7f0ceb05e786ba2a77c1e39c407e6425ed62dcc50a46bb8b8618f5c5

Observation 7ac9c8c2-f6e7-4864-9f17-db73edb8daa0 · outbound

This paper cites Mind2Web: Towards a Generalist Agent for the Web.

InSTA: Towards Internet-Scale Training For Agents Mind2Web: Towards a Generalist Agent for the Web

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.321613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.321613Z digest=sha256:77df0deec2a12c7743b7c52e3901291d82984d310557c49fb0eca80f5e379001

Observation 3e458e21-9821-4ed6-abe6-1515b6d431c2 · outbound

This paper cites Better Synthetic Data by Retrieving and Transforming Existing Datasets.

InSTA: Towards Internet-Scale Training For Agents Better Synthetic Data by Retrieving and Transforming Existing Datasets

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.325236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.325236Z digest=sha256:70c1988f4b65087680a6992f54e0b09758ce60cb7b82fb5a35557723a46ad99e

Observation b76ffa0c-df13-42a3-8686-41db08222d2d · outbound

This paper cites The Llama 3 Herd of Models.

InSTA: Towards Internet-Scale Training For Agents The Llama 3 Herd of Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.328784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.328784Z digest=sha256:becf9d3ad6ef03a25ae74efa4eb5d6d39573c1e8c04b39fd595a93d02730c5ee

Observation 6a45155b-e0ad-4d65-a8ba-9af526c8c928 · outbound

This paper cites W eb V oyager: Building an end-to-end web agent with large multimodal models.

InSTA: Towards Internet-Scale Training For Agents W eb V oyager: Building an end-to-end web agent with large multimodal models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.332092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.332092Z digest=sha256:c364357fa2378467a0c94602e3f043891ade9f43fbcecd59c7fbeb953e1a5e2f

Observation c4a373ce-8807-43ad-8fc8-84f984a0d932 · outbound

This paper cites CogAgent: A Visual Language Model for GUI Agents.

InSTA: Towards Internet-Scale Training For Agents CogAgent: A Visual Language Model for GUI Agents

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.335297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.335297Z digest=sha256:c5ba445e32da9f5735a40ebaa637caa4edd8990ac9abeca10aa41694d5905c50

Observation 0ec5b2f6-a62a-446a-b3fa-fcddc1de3ada · outbound

This paper cites Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations.

InSTA: Towards Internet-Scale Training For Agents Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.338817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.338817Z digest=sha256:76ada451ab7a25bdc1b581f0888872ace44d52263bee7eefc33e0654fbb030d0

Observation cb1cf1c8-3ec2-44c5-8a62-5c49dd361be3 · outbound

This paper cites VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks.

InSTA: Towards Internet-Scale Training For Agents VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.342376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.342376Z digest=sha256:0e234abdc0d90fc16aa5490c5b7d2ce96cc233209ffca8e2c9d63cf6aea818c3

Observation e0ff1980-b081-4620-a325-683c051163b1 · outbound

This paper cites Tree search for language model agents, 2024 b.

InSTA: Towards Internet-Scale Training For Agents Tree search for language model agents, 2024 b

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.345982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.345982Z digest=sha256:dd2d74f7cf56ee6115640fed10f94aca53ece39b2b370e8443ec92cd3e8b1484

Observation 7749ecf6-5cff-41c9-a5e2-29246ca3984c · outbound

This paper cites Efficient Memory Management for Large Language Model Serving with PagedAttention.

InSTA: Towards Internet-Scale Training For Agents Efficient Memory Management for Large Language Model Serving with PagedAttention

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.349153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.349153Z digest=sha256:a9bdfdd908fa4a7e87da7ec2f6c06a4f9c22b547657f0f39140a6e79e2b3eaaa

Observation 9dade60d-eb87-49f7-b28b-630132912a4c · outbound

This paper cites RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback.

InSTA: Towards Internet-Scale Training For Agents RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.352547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.352547Z digest=sha256:cc8b6fa7dadabc8741c4f50a1470aabd460e86406760788df4ca489d372545be

Observation 958df537-a72c-4b84-9e31-a83b5e9c13de · outbound

This paper cites From generation to judgment: Opportunities and challenges of llm-as-a-judge.

InSTA: Towards Internet-Scale Training For Agents From generation to judgment: Opportunities and challenges of llm-as-a-judge

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.355887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.355887Z digest=sha256:d6ec4626000fc82b227de505c0f3c82c461c511b5ba212017e55fa2aff8b2e04

Observation cbf8049c-5d5b-4b04-ac59-83bd09176a9c · outbound

This paper cites VisualWebBench: How Far Have Multimodal LLMs Evolved in Web Page Understanding and Grounding?.

InSTA: Towards Internet-Scale Training For Agents VisualWebBench: How Far Have Multimodal LLMs Evolved in Web Page Understanding and Grounding?

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.359070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.359070Z digest=sha256:d7ddd8c4b479acfee35c3fb1954e0d1b27e065333e046f2c6376134e7fdc5fb5

Observation 987e6e94-0ef5-4376-b55a-a3221f83c769 · outbound

This paper cites Weblinx: Real-world website navigation with multi-turn dialogue, 2024.

InSTA: Towards Internet-Scale Training For Agents Weblinx: Real-world website navigation with multi-turn dialogue, 2024

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.362660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.362660Z digest=sha256:f4cf06827481fad6cb355b5f2b0c7b848e9fd90d5098870aed18f4ba81cdcd87

Observation 8188a3af-9600-4974-92aa-d068b7e3e320 · outbound

This paper cites Self-refine: Iterative refinement with self-feedback.

InSTA: Towards Internet-Scale Training For Agents Self-refine: Iterative refinement with self-feedback

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.366344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.366344Z digest=sha256:ee692a63b9c75c8a5cc6040442f82146479d60c72b4ac45910ca2fb26cdd8d14

Observation d2886aef-e7e6-4380-a9f6-3128d5b9fb03 · outbound

This paper cites Playwright.

InSTA: Towards Internet-Scale Training For Agents Playwright

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:24:51.319624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.369944Z digest=sha256:56451a250b6200f3f0538cec64d4e936178a660f8069c10d6dbbcd4e09b0827d

Observation c6b20d75-769c-4c46-a53b-29305c4a39d9 · outbound

This paper cites AgentInstruct: Toward Generative Teaching with Agentic Flows.

InSTA: Towards Internet-Scale Training For Agents AgentInstruct: Toward Generative Teaching with Agentic Flows

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.373086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.373086Z digest=sha256:0e1566b1fbf09ebaa89415ffa4f94eb98e08bc8e34ae2e2e0ffe5eba3021c64d

Observation 0f3e2a0f-c667-495f-8232-719fab758aa1 · outbound

This paper cites NNetNav: Unsupervised Learning of Browser Agents Through Environment Interaction in the Wild.

InSTA: Towards Internet-Scale Training For Agents NNetNav: Unsupervised Learning of Browser Agents Through Environment Interaction in the Wild

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.376394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.376394Z digest=sha256:e5dc8a3e74946c8a1deeb22f40aef048da606c2df87ce5b8c90074d35ab114b8

Observation b0678889-2d91-4977-9eb7-b206f822a189 · outbound

This paper cites Synatra: Turning Indirect Knowledge into Direct Demonstrations for Digital Agents at Scale.

InSTA: Towards Internet-Scale Training For Agents Synatra: Turning Indirect Knowledge into Direct Demonstrations for Digital Agents at Scale

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.379591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.379591Z digest=sha256:c1df0b98bba0d77376ad54a9e06eb3baf81327855b5b4da770b4c613fe9aaba3

Observation c7bc3b97-9d02-4cb0-a086-90a438797921 · outbound

This paper cites an unresolved cited work.

InSTA: Towards Internet-Scale Training For Agents Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:24:51.308230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.382859Z digest=sha256:113ece4fb2ef7222c78d4b78559c432c275e1fdf34e43871902bf0097085a0fa

Observation 0a7fb9ba-df97-4fc1-9ade-367acd1a7d3a · outbound

This paper cites Large Language Models Can Self-Improve At Web Agent Tasks.

InSTA: Towards Internet-Scale Training For Agents Large Language Models Can Self-Improve At Web Agent Tasks

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.385849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.385849Z digest=sha256:395f04eec0740113b6cfe4e2df90b3ad5239d5b19ce850c4b07c1722bc599b78

Observation da1d925a-a9bc-4dc7-84a9-52a65ff6659e · outbound

This paper cites REFINER : Reasoning feedback on intermediate representations.

InSTA: Towards Internet-Scale Training For Agents REFINER : Reasoning feedback on intermediate representations

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:24:51.297090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.389101Z digest=sha256:dd396c61402319293ff11ffc20c6444dfc4e2080324c93ccb7b6338794e7ea5d

Observation 34ab8fa1-d648-4230-803b-51f16d13253b · outbound

This paper cites Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents.

InSTA: Towards Internet-Scale Training For Agents Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.392087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.392087Z digest=sha256:4e7c4811b3229b922ccac76a5f5d85396687cd642988929628e62f0d3c0b21ab

Observation 5aefd61b-c7af-4b15-8c25-9615a88aa8b0 · outbound

This paper cites Language models are unsupervised multitask learners, 2019.

InSTA: Towards Internet-Scale Training For Agents Language models are unsupervised multitask learners, 2019

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:24:51.285581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.396681Z digest=sha256:7dc33a61e2855446bbeb951fde7a34a6a2eaca085fbc9f501e4afbaa55264b6c

Observation 96067855-455f-4833-a61d-1ab51bf68fb0 · outbound

This paper cites Android in the Wild: A Large-Scale Dataset for Android Device Control.

InSTA: Towards Internet-Scale Training For Agents Android in the Wild: A Large-Scale Dataset for Android Device Control

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.399754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.399754Z digest=sha256:28bf6aa97abb46321cf7045d95cf0eada131f1f4c66c0db3fbbc113b6c8f5a83

Observation a3f733c7-5748-4101-aabc-2baffffc950e · outbound

This paper cites Toolformer: Language models can teach themselves to use tools.

InSTA: Towards Internet-Scale Training For Agents Toolformer: Language models can teach themselves to use tools

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.403052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.403052Z digest=sha256:3162f68f3aa044f52744dc7696c34fbb64df49d6c973e566b1495126b0134562

Observation 0206bbf1-b54a-4626-bcb5-a65cff803769 · outbound

This paper cites RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold.

InSTA: Towards Internet-Scale Training For Agents RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.406158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.406158Z digest=sha256:7db543c78515e9faa945c68348282260ab4499330368eb5fc028f4b645ff31d6

Observation 17a875e4-7a63-453b-a9c5-d12662c1baf5 · outbound

This paper cites ScribeAgent: Towards Specialized Web Agents Using Production-Scale Workflow Data.

InSTA: Towards Internet-Scale Training For Agents ScribeAgent: Towards Specialized Web Agents Using Production-Scale Workflow Data

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.409647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.409647Z digest=sha256:f529a49b514565641f3861a5eb8a26ba4362f0161b3fb83d0acd9beb0de1cee3

Observation 5a806af7-f919-47b1-bf23-6c8b369eeed3 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

InSTA: Towards Internet-Scale Training For Agents Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.413143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.413143Z digest=sha256:b5058790e88e245bc97a106d2fd74946a153d79d7d95d2982d46ece0a6088268

Observation 5c808863-c9dd-4d0d-b6e7-358c2c4dc874 · outbound

This paper cites Fast best-of-n decoding via speculative rejection.

InSTA: Towards Internet-Scale Training For Agents Fast best-of-n decoding via speculative rejection

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.416715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.416715Z digest=sha256:35972eff52f75c598f3f1a73d0819a6653c96decdaba9ddd62a548f449fdb77a

Observation 8338de25-381f-4e0e-9505-09ee70966e48 · outbound

This paper cites Preference Fine-Tuning of LLMs Should Leverage Suboptimal, On-Policy Data.

InSTA: Towards Internet-Scale Training For Agents Preference Fine-Tuning of LLMs Should Leverage Suboptimal, On-Policy Data

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.420177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.420177Z digest=sha256:53fab949991d5c1a58815be892fc1256ccd8d5ea5ca3ccb08f2d0752ec362738

Observation da041530-0d95-4601-be8a-97927d1ea3e8 · outbound

This paper cites Common crawl, 2025.

InSTA: Towards Internet-Scale Training For Agents Common crawl, 2025

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:24:51.262720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.424981Z digest=sha256:8c6ddc1ea91d4be1ccfb57f3f8cd8e86d0eddd5652cec454e9888745eddca0cd

Observation ea94054e-d6a4-4c35-a502-8de8f602f350 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

InSTA: Towards Internet-Scale Training For Agents LLaMA: Open and Efficient Foundation Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.428386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.428386Z digest=sha256:001ecc1f52b4538bdc5a3a35bab6cfd8b862dd4ca61a2649c442bac65da2815b

Observation 4a3f1095-6f52-4036-bae6-8e4e41e07a58 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

InSTA: Towards Internet-Scale Training For Agents Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.432100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.432100Z digest=sha256:6f993c4da77010333607a89f85e88d4186396cc2d07e5559565e715be02a114e

Observation d0d07af5-0d1b-4b56-9522-de72f1c5921c · outbound

This paper cites Effective data augmentation with diffusion models.

InSTA: Towards Internet-Scale Training For Agents Effective data augmentation with diffusion models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:24:51.251327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.435592Z digest=sha256:cba02ab1f4cd75b93432c571d57ae4e0d25e4ec14fd187b416f8241c7f29cebf

Observation e2581515-9656-4d94-8da8-fefae582a480 · outbound

This paper cites LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench.

InSTA: Towards Internet-Scale Training For Agents LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.439003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.439003Z digest=sha256:899a18116c028ee00ae9b35ab7be7facceed56cd5fbf2524fbe78d6bdf8ff460

Observation ea5f47ac-2bc3-41b6-b0a4-1ba017f59dd5 · outbound

This paper cites A survey on large language model based autonomous agents.

InSTA: Towards Internet-Scale Training For Agents A survey on large language model based autonomous agents

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.442840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.442840Z digest=sha256:8b5b6afb44d75e2ccf574972f178d1b603d66f3dd6b82dc243d7199dd4ce9130

Observation 6b8baf50-50fc-4489-ac81-7eccc8655f94 · outbound

This paper cites Large Multimodal Agents: A Survey.

InSTA: Towards Internet-Scale Training For Agents Large Multimodal Agents: A Survey

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.446116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.446116Z digest=sha256:eb38d6f71034b733270fd4fb80fad76dbe1c32fcfefff2c767c4ae905faa55b3

Observation b2c28686-933b-49a4-ad6b-4da85fece042 · outbound

This paper cites An illusion of progress? assessing the current state of web agents.

InSTA: Towards Internet-Scale Training For Agents An illusion of progress? assessing the current state of web agents

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.449690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.449690Z digest=sha256:1c22e1bbccf89d564da4029d17df38b61d2a08b62565b424d471e56623bc5073

Observation d0ef471f-fea3-455e-ad14-3d8b99bed1b8 · outbound

This paper cites WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents.

InSTA: Towards Internet-Scale Training For Agents WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.453058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.453058Z digest=sha256:c81932077fe6b6a9fd55f093fb107039dcf6ce85e6e11ccaff5ae98a2dd2d82c

Observation acbe7994-16b6-4212-8d0f-8f4c6a3e901b · outbound

This paper cites Tree of Thoughts: Deliberate Problem Solving with Large Language Models.

InSTA: Towards Internet-Scale Training For Agents Tree of Thoughts: Deliberate Problem Solving with Large Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.456513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.456513Z digest=sha256:81e6b876b88ca91b000f126de28da3356a4c02288ba81691c9ba07f2218217d5

Observation 2510f647-cff4-4bc5-b7cb-2071f36393f2 · outbound

This paper cites TextGrad: Automatic "Differentiation" via Text.

InSTA: Towards Internet-Scale Training For Agents TextGrad: Automatic "Differentiation" via Text

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.460002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.460002Z digest=sha256:7a43dbd5a3efd3f6da61f7d6e229b73796126934b48b506336c05dcb6f91952a

Observation 97aee2eb-ff82-439a-8d43-3c209c1b7716 · outbound

This paper cites AgentTuning: Enabling Generalized Agent Abilities for LLMs.

InSTA: Towards Internet-Scale Training For Agents AgentTuning: Enabling Generalized Agent Abilities for LLMs

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.463527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.463527Z digest=sha256:d015a4d70006a49c1547eb680dff755b5e28538924ff1d43474c3eb17325c8d1

Observation 3d2010d8-6618-476a-9996-5423180fbf08 · outbound

This paper cites AppAgent: Multimodal Agents as Smartphone Users.

InSTA: Towards Internet-Scale Training For Agents AppAgent: Multimodal Agents as Smartphone Users

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.466991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.466991Z digest=sha256:e5aafba4a133f47a0619ce129ac77b105a216f7deeb9359ff3c0121d67957b5e

Observation bcceb7ae-022a-48ce-9579-cae963cd1bde · outbound

This paper cites Generative Verifiers: Reward Modeling as Next-Token Prediction.

InSTA: Towards Internet-Scale Training For Agents Generative Verifiers: Reward Modeling as Next-Token Prediction

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.470275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.470275Z digest=sha256:f600f4dd5ba1bec54199731d26bdb6ac3d4578d6a5efee24ef1d5381fc580da0

Observation fb46c824-8ea4-407c-b44e-585b13b8234f · outbound

This paper cites Evaluation of openai o1: Opportunities and challenges of agi, 2024.

InSTA: Towards Internet-Scale Training For Agents Evaluation of openai o1: Opportunities and challenges of agi, 2024

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.473513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.473513Z digest=sha256:f7a7b6f1e8fc416bdb5932cb692c030aa16173743387fb12dd1e76635378aec6

Observation 07220155-d6cb-4580-83c9-87aaf38fee44 · outbound

This paper cites Language Agent Tree Search Unifies Reasoning Acting and Planning in Language Models.

InSTA: Towards Internet-Scale Training For Agents Language Agent Tree Search Unifies Reasoning Acting and Planning in Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.477162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.477162Z digest=sha256:564cf98633469d2f1ddcebbaa532e9a09be790926d87fbb25255862d346736ac

Observation 0cd18ca5-7ec3-4e6a-b14c-98211eedf41f · outbound

This paper cites WebArena: A Realistic Web Environment for Building Autonomous Agents.

InSTA: Towards Internet-Scale Training For Agents WebArena: A Realistic Web Environment for Building Autonomous Agents

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.480712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.480712Z digest=sha256:b6706f689c32af4957b6c9b81d4ab35293c6f249ffb09363a2863dae3885e335

Observation 6a5421f1-b069-40e7-97eb-1429e6bd533f · outbound

This paper cites Proposer-Agent-Evaluator(PAE): Autonomous Skill Discovery For Foundation Model Internet Agents.

InSTA: Towards Internet-Scale Training For Agents Proposer-Agent-Evaluator(PAE): Autonomous Skill Discovery For Foundation Model Internet Agents

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.484160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.484160Z digest=sha256:fd7ef88337afbfd168df04757a146065ba43f8382d8e975ffc5ea1f8e6bac3fb

Pith citing papers

Observation 1d96973f-8934-484e-96ed-85db71ef9fe8 · inbound

Thinking vs. Doing: Agents that Reason by Scaling Test-Time Interaction cites this paper.

Thinking vs. Doing: Agents that Reason by Scaling Test-Time Interaction InSTA: Towards Internet-Scale Training For Agents

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.165917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.165917Z digest=sha256:db0c5b657425bc5bf03af2034b41ca0c894785e586f910988a732af04b56a438

Observation fddaf0a7-6b40-4a6c-b57c-0dda691a378f · inbound

DynaWeb: Model-Based Reinforcement Learning of Web Agents cites this paper.

DynaWeb: Model-Based Reinforcement Learning of Web Agents InSTA: Towards Internet-Scale Training For Agents

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:37:41.519723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-16T09:33:30.444057Z digest=sha256:557f89f7870747633e96371dd9e7b0da2f49ee0815992fd058c922503287e782

Observation 5db362cc-0231-4528-b70f-51b39367709a · inbound

WebFactory: Automated Compression of Foundational Language Intelligence into Grounded Web Agents cites this paper.

WebFactory: Automated Compression of Foundational Language Intelligence into Grounded Web Agents InSTA: Towards Internet-Scale Training For Agents

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:50:10.672001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-15T16:50:02.527821Z digest=sha256:e395087acd53cfd43532eea39e67127a06fb57f5b3eae58bb2474b7ca6491a72

Observation bc2c3cb3-33df-4191-9876-ed6ec68585d6 · inbound

ContractSkill: Repairable Contract-Based Skills for Multimodal Web Agents cites this paper.

ContractSkill: Repairable Contract-Based Skills for Multimodal Web Agents InSTA: Towards Internet-Scale Training For Agents

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-15T08:59:52.986286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-15T08:58:50.947580Z digest=sha256:319e65418a8511f24baae624bab74b0991f7d660f92d757a1dd4b0645e7ae528

Observation 3bd86329-3690-40a8-a34f-0320088f0b4b · inbound

Weblica: Scalable and Reproducible Training Environments for Visual Web Agents cites this paper.

Weblica: Scalable and Reproducible Training Environments for Visual Web Agents InSTA: Towards Internet-Scale Training For Agents

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:25:56.929155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-11T01:25:44.578007Z digest=sha256:9a398a13ab0a83a71869b37f9798e9bfa152c29be2ce2c109f0fca3653cdb03e

Observation 262c0d22-9ae3-4eaf-a79f-acb7ffbdded2 · inbound

OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents cites this paper.

OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents InSTA: Towards Internet-Scale Training For Agents

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:06:16.466449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T15:46:50.587684Z digest=sha256:98a33000edf2bb77af881e90ed8cb99880fb63c5db62fd4536dcb3f50361b33f

Observation c0103017-b7e6-426a-ae5a-27cdef8b4150 · inbound

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application cites this paper.

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application InSTA: Towards Internet-Scale Training For Agents

Reference 267

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:58:02.991562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-27T09:46:30.702256Z digest=sha256:af117b2198745b5f5e05ffe4e11e46aeae78c24a51e21ef9df00b57753cd7128

Observation 0ab252e1-a260-40eb-ac04-73dc2918661a · inbound

ProCUA-SFT Technical Report cites this paper.

ProCUA-SFT Technical Report InSTA: Towards Internet-Scale Training For Agents

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T18:08:46.699111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-27T03:18:04.149281Z digest=sha256:b98c6812a454fe31a2069968fa96f55077e5118e9ed9687d775c9b1eae3688c1

Observation 47e1d398-e6d2-4006-912a-9021323ea636 · inbound

Empowering GUI Agents via Autonomous Experience Exploration and Hindsight Experience Utilization for Task Planning cites this paper.

Empowering GUI Agents via Autonomous Experience Exploration and Hindsight Experience Utilization for Task Planning InSTA: Towards Internet-Scale Training For Agents

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-04T14:29:53.385736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-26T03:51:51.827622Z digest=sha256:3552f7b0578940d0ceff0f451ff82987b2910ef136d9c987376adfb214ea9550

Observation ae9c1a53-c75c-4d84-82f0-0a7ebaf1b595 · inbound

OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models cites this paper.

OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models InSTA: Towards Internet-Scale Training For Agents

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-31T02:18:01.822251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T02:18:01.822251Z digest=sha256:254d2f17bef88533475dc4613240fefbf82496096b3a4f716b2baa87fd5e4e8d

Observation a9b3ea1a-9dbf-4539-95c0-ddaaf0a6dd4d · inbound

Software Engineering for and with GUI Agent cites this paper.

Software Engineering for and with GUI Agent InSTA: Towards Internet-Scale Training For Agents

Reference 223

Resolution
unresolved
no resolver link, observed 2026-08-11T20:19:15.802502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:19:15.802502Z digest=sha256:01955c2340db4af4fd373d7ba77b8dbe084177f88f0a554d12bd3c766e7bd9f3