Pith. sign in

Paper Citation Record · LEDGER

InSTA: Towards Internet-Scale Training For Agents

As of 9 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 10 inbound Pith citation observations for arXiv:2502.06776.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.06776 v2

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T14:24:50.484160Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:27:50.165917Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

58 of 58 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved53
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation ee2d6038-0004-403e-9fdb-8f90a9517aaf · outbound

This paper cites write newline.

InSTA: Towards Internet-Scale Training For Agents write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.283326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.283326Z digest=sha256:65d2eb038f5a189de264fd3d302e3ef52ba6def5c0388bdf2f0e1f96b3b0703d

Observation 073770c8-88c3-430a-9ebf-4f4815ffca7a · outbound

This paper cites @esa (Ref.

InSTA: Towards Internet-Scale Training For Agents @esa (Ref

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.288008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.288008Z digest=sha256:5fd537a1b744b0ac833154fd6584f5792ef4f20ed378941f81493e30f2b19930

Observation d92b4c8b-a731-4d17-a5b4-2f3b2d85bdbd · outbound

This paper cites an unresolved cited work.

InSTA: Towards Internet-Scale Training For Agents Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.291843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.291843Z digest=sha256:d4697ddd10479ae849c2cba7c8308161b5c180ed809f70f6e170eda18e07030c

Observation 92af65ee-1f28-4220-95d4-5948aafd16dc · outbound

This paper cites an unresolved cited work.

InSTA: Towards Internet-Scale Training For Agents Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:24:51.337301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.295342Z digest=sha256:2f00c062034b1fb3e41532e09c872491d5be65614c83d5b3f2908b71dec0b154

Observation f3921423-24e3-4645-b9ce-40b36b05b58a · outbound

This paper cites Language models as agent models.

InSTA: Towards Internet-Scale Training For Agents Language models as agent models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.298819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.298819Z digest=sha256:5c04d0622a8a896b15d3d48f52cca1aff95ce50c7a3298378a2881aa5dd5799e

Observation fbe027e7-122d-42f4-b87a-3849bd3aa532 · outbound

This paper cites Graph of thoughts: Solving elaborate problems with large language models.

InSTA: Towards Internet-Scale Training For Agents Graph of thoughts: Solving elaborate problems with large language models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.302297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.302297Z digest=sha256:d2d4f5e69645815fabdfed0cea19c2e1a0420df76f4c255a9a46368d2a58fc1c

Observation cfde26e3-8a15-42c2-bd97-d056ffa81aa4 · outbound

This paper cites Language Models are Few-Shot Learners.

InSTA: Towards Internet-Scale Training For Agents Language Models are Few-Shot Learners

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.305890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.305890Z digest=sha256:2b2c03df0d4894f5bb214334493b568ae8602a1f08dcc1416ec6c06aa40d3d52

Observation 0305d162-35a3-4e8b-871f-307ab708f6e6 · outbound

This paper cites Sparks of Artificial General Intelligence: Early experiments with GPT-4.

InSTA: Towards Internet-Scale Training For Agents Sparks of Artificial General Intelligence: Early experiments with GPT-4

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.309685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.309685Z digest=sha256:dee6d70fb4c017a764ea43d2f6825717831537b61a66d368415a862acecc4a77

Observation f718996a-ad8b-406c-8ffd-65ecf8b4eb47 · outbound

This paper cites FireAct: Toward Language Agent Fine-tuning.

InSTA: Towards Internet-Scale Training For Agents FireAct: Toward Language Agent Fine-tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.313595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.313595Z digest=sha256:a9fea9f0a803ebf9b63cb2cbce6ba5b471c49dd764575f259bd9b0463bbd9ce5

Observation c110ffad-3e6e-48ed-8d41-623f533a19cd · outbound

This paper cites The BrowserGym Ecosystem for Web Agent Research.

InSTA: Towards Internet-Scale Training For Agents The BrowserGym Ecosystem for Web Agent Research

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.317546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.317546Z digest=sha256:cd40cef6be22f1a3e267f9220faf2f6b0b977ba26d66e0391e234ac0a59bae31

Observation 7ac9c8c2-f6e7-4864-9f17-db73edb8daa0 · outbound

This paper cites Mind2Web: Towards a Generalist Agent for the Web.

InSTA: Towards Internet-Scale Training For Agents Mind2Web: Towards a Generalist Agent for the Web

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.321613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.321613Z digest=sha256:737bf57dfbb62904052abd5998533117914a6b4daa45caeac7dc1e305a034c85

Observation 3e458e21-9821-4ed6-abe6-1515b6d431c2 · outbound

This paper cites Better Synthetic Data by Retrieving and Transforming Existing Datasets.

InSTA: Towards Internet-Scale Training For Agents Better Synthetic Data by Retrieving and Transforming Existing Datasets

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.325236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.325236Z digest=sha256:e5dbedcb1a9e6bb34bdb837ade2d7bca99ca9ea50faccf81d43a67d08ec03c0f

Observation b76ffa0c-df13-42a3-8686-41db08222d2d · outbound

This paper cites The Llama 3 Herd of Models.

InSTA: Towards Internet-Scale Training For Agents The Llama 3 Herd of Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.328784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.328784Z digest=sha256:07a455940ecfb07b838156025e09f28066a2b32c99e2c9822a5559ac15af929a

Observation 6a45155b-e0ad-4d65-a8ba-9af526c8c928 · outbound

This paper cites W eb V oyager: Building an end-to-end web agent with large multimodal models.

InSTA: Towards Internet-Scale Training For Agents W eb V oyager: Building an end-to-end web agent with large multimodal models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.332092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.332092Z digest=sha256:4140679667665ef0d2a7b2ec1eec6b9f223627cd647dac1a033661ee73943ad7

Observation c4a373ce-8807-43ad-8fc8-84f984a0d932 · outbound

This paper cites CogAgent: A Visual Language Model for GUI Agents.

InSTA: Towards Internet-Scale Training For Agents CogAgent: A Visual Language Model for GUI Agents

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.335297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.335297Z digest=sha256:0d0c84e8ab73f1d17815675bf77172285df0b608b1f503413fa65834949710ba

Observation 0ec5b2f6-a62a-446a-b3fa-fcddc1de3ada · outbound

This paper cites Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations.

InSTA: Towards Internet-Scale Training For Agents Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.338817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.338817Z digest=sha256:7d6bb2e6776cbb13e634c18eeb013140e561ac046409b4bce9cb1413888670ca

Observation cb1cf1c8-3ec2-44c5-8a62-5c49dd361be3 · outbound

This paper cites VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks.

InSTA: Towards Internet-Scale Training For Agents VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.342376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.342376Z digest=sha256:a6c649d1fb9dbfb15ff2b877d4b229d1d6e53abd6589af02ae4c765c08cbaf70

Observation e0ff1980-b081-4620-a325-683c051163b1 · outbound

This paper cites Tree search for language model agents, 2024 b.

InSTA: Towards Internet-Scale Training For Agents Tree search for language model agents, 2024 b

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.345982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.345982Z digest=sha256:430593d5bf053c03924c903e10ae87c46b109fe5b7bdc3e2fee0968a26c00cfa

Observation 7749ecf6-5cff-41c9-a5e2-29246ca3984c · outbound

This paper cites Efficient Memory Management for Large Language Model Serving with PagedAttention.

InSTA: Towards Internet-Scale Training For Agents Efficient Memory Management for Large Language Model Serving with PagedAttention

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.349153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.349153Z digest=sha256:0e9ef64c8979a06e29c523e27251d083f2013e33d18531cca4f1b937d15767d0

Observation 9dade60d-eb87-49f7-b28b-630132912a4c · outbound

This paper cites RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback.

InSTA: Towards Internet-Scale Training For Agents RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.352547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.352547Z digest=sha256:b5c042e5f218cc6783452e4291766e46b4baf79851d01fd2efd4591b62036f36

Observation 958df537-a72c-4b84-9e31-a83b5e9c13de · outbound

This paper cites From generation to judgment: Opportunities and challenges of llm-as-a-judge.

InSTA: Towards Internet-Scale Training For Agents From generation to judgment: Opportunities and challenges of llm-as-a-judge

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.355887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.355887Z digest=sha256:609cd89aade5279241bfc5b5e46276476d3e39861153d587090c76fa2107cafb

Observation cbf8049c-5d5b-4b04-ac59-83bd09176a9c · outbound

This paper cites VisualWebBench: How Far Have Multimodal LLMs Evolved in Web Page Understanding and Grounding?.

InSTA: Towards Internet-Scale Training For Agents VisualWebBench: How Far Have Multimodal LLMs Evolved in Web Page Understanding and Grounding?

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.359070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.359070Z digest=sha256:e388ea03abc65dcfe01d0a17df346e53b4f858e346f4997bbcae5f8b5463a46c

Observation 987e6e94-0ef5-4376-b55a-a3221f83c769 · outbound

This paper cites Weblinx: Real-world website navigation with multi-turn dialogue, 2024.

InSTA: Towards Internet-Scale Training For Agents Weblinx: Real-world website navigation with multi-turn dialogue, 2024

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.362660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.362660Z digest=sha256:eb2b39b4f884527e4a7819370205ae95420071b3140bd2b37c1218e9b7b377b0

Observation 8188a3af-9600-4974-92aa-d068b7e3e320 · outbound

This paper cites Self-refine: Iterative refinement with self-feedback.

InSTA: Towards Internet-Scale Training For Agents Self-refine: Iterative refinement with self-feedback

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.366344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.366344Z digest=sha256:3c58f98e2d318e15cf9b527a3813c5bd086b7d06962300f8faeec1d4ce8cdc84

Observation d2886aef-e7e6-4380-a9f6-3128d5b9fb03 · outbound

This paper cites Playwright.

InSTA: Towards Internet-Scale Training For Agents Playwright

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:24:51.319624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.369944Z digest=sha256:9a82c13446628b2a5d53486695b643cafb8ffb595bce050b8334f558cb9cba48

Observation c6b20d75-769c-4c46-a53b-29305c4a39d9 · outbound

This paper cites AgentInstruct: Toward Generative Teaching with Agentic Flows.

InSTA: Towards Internet-Scale Training For Agents AgentInstruct: Toward Generative Teaching with Agentic Flows

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.373086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.373086Z digest=sha256:fd4d83da4dd63de88032a9977e9fb68baddc2f57703480fb8fb6bb3a876946df

Observation 0f3e2a0f-c667-495f-8232-719fab758aa1 · outbound

This paper cites NNetNav: Unsupervised Learning of Browser Agents Through Environment Interaction in the Wild.

InSTA: Towards Internet-Scale Training For Agents NNetNav: Unsupervised Learning of Browser Agents Through Environment Interaction in the Wild

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.376394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.376394Z digest=sha256:18993d622e34cdb99bd16c4743d9f4219a6383122af72d145db5b9c10314f741

Observation b0678889-2d91-4977-9eb7-b206f822a189 · outbound

This paper cites Synatra: Turning Indirect Knowledge into Direct Demonstrations for Digital Agents at Scale.

InSTA: Towards Internet-Scale Training For Agents Synatra: Turning Indirect Knowledge into Direct Demonstrations for Digital Agents at Scale

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.379591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.379591Z digest=sha256:764273b751da01ccf8749cf09642061f6a9d427e3206417b7670b42c91e7c485

Observation c7bc3b97-9d02-4cb0-a086-90a438797921 · outbound

This paper cites an unresolved cited work.

InSTA: Towards Internet-Scale Training For Agents Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:24:51.308230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.382859Z digest=sha256:5857bf76e977ebd3e4412719fc312a24e04949b911ecb474e1c2b6a195f16265

Observation 0a7fb9ba-df97-4fc1-9ade-367acd1a7d3a · outbound

This paper cites Large Language Models Can Self-Improve At Web Agent Tasks.

InSTA: Towards Internet-Scale Training For Agents Large Language Models Can Self-Improve At Web Agent Tasks

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.385849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.385849Z digest=sha256:556488a5712ae798b036e9bf8342aa150163941d18970678b3d8e6a7154ff213

Observation da1d925a-a9bc-4dc7-84a9-52a65ff6659e · outbound

This paper cites REFINER : Reasoning feedback on intermediate representations.

InSTA: Towards Internet-Scale Training For Agents REFINER : Reasoning feedback on intermediate representations

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:24:51.297090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.389101Z digest=sha256:ef1180d58fb41d0b80ce9739f0aba9171a49c8dfc6e2258af40a787dcbff9a5f

Observation 34ab8fa1-d648-4230-803b-51f16d13253b · outbound

This paper cites Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents.

InSTA: Towards Internet-Scale Training For Agents Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.392087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.392087Z digest=sha256:2ecefc8ab926dc88f643ea2f96518bf7804b876b12e16d3b7332808076761f9f

Observation 5aefd61b-c7af-4b15-8c25-9615a88aa8b0 · outbound

This paper cites Language models are unsupervised multitask learners, 2019.

InSTA: Towards Internet-Scale Training For Agents Language models are unsupervised multitask learners, 2019

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:24:51.285581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.396681Z digest=sha256:77aedcefe0107f8c111be79efe6f9b191daf8344ac42c78d8f9475ee54cb3a78

Observation 96067855-455f-4833-a61d-1ab51bf68fb0 · outbound

This paper cites Android in the Wild: A Large-Scale Dataset for Android Device Control.

InSTA: Towards Internet-Scale Training For Agents Android in the Wild: A Large-Scale Dataset for Android Device Control

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.399754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.399754Z digest=sha256:45c2a01bcac99508a19032cb597e7fba0be3c0e0026b3758ca5772b0771c736a

Observation a3f733c7-5748-4101-aabc-2baffffc950e · outbound

This paper cites Toolformer: Language models can teach themselves to use tools.

InSTA: Towards Internet-Scale Training For Agents Toolformer: Language models can teach themselves to use tools

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.403052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.403052Z digest=sha256:75b6e2b3440654e8d6dea0925bb2b7fdf8bfcd14fa1a31e584e9d1d1e12abfff

Observation 0206bbf1-b54a-4626-bcb5-a65cff803769 · outbound

This paper cites RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold.

InSTA: Towards Internet-Scale Training For Agents RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.406158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.406158Z digest=sha256:98d49d34dc306d0280a9308eeb791426564595b66319d4a7ca8946ef2e9053bd

Observation 17a875e4-7a63-453b-a9c5-d12662c1baf5 · outbound

This paper cites ScribeAgent: Towards Specialized Web Agents Using Production-Scale Workflow Data.

InSTA: Towards Internet-Scale Training For Agents ScribeAgent: Towards Specialized Web Agents Using Production-Scale Workflow Data

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.409647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.409647Z digest=sha256:a86ecd38f760664e812624b54d5c568a372922ea4f8de063bd76ab99eef8f5df

Observation 5a806af7-f919-47b1-bf23-6c8b369eeed3 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

InSTA: Towards Internet-Scale Training For Agents Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.413143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.413143Z digest=sha256:ef085fca77a1add8b99e29872305cebb7f8d4635655d0963e3ea7ad0f88ad802

Observation 5c808863-c9dd-4d0d-b6e7-358c2c4dc874 · outbound

This paper cites Fast best-of-n decoding via speculative rejection.

InSTA: Towards Internet-Scale Training For Agents Fast best-of-n decoding via speculative rejection

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.416715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.416715Z digest=sha256:c45cc6901a4777870f86d7b77575e510d3debbd1ebfd81ce02b6b82d74ae4e4e

Observation 8338de25-381f-4e0e-9505-09ee70966e48 · outbound

This paper cites Preference Fine-Tuning of LLMs Should Leverage Suboptimal, On-Policy Data.

InSTA: Towards Internet-Scale Training For Agents Preference Fine-Tuning of LLMs Should Leverage Suboptimal, On-Policy Data

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.420177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.420177Z digest=sha256:f17a8e694060f370a9c3f6cf7d9228c0aaa2015d6d57e9b31829fb1ed0c9b01b

Observation da041530-0d95-4601-be8a-97927d1ea3e8 · outbound

This paper cites Common crawl, 2025.

InSTA: Towards Internet-Scale Training For Agents Common crawl, 2025

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:24:51.262720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.424981Z digest=sha256:3528b5c06b0195d12d89d4c5dae1009195f33c50c765c318c00d643a14bc4150

Observation ea94054e-d6a4-4c35-a502-8de8f602f350 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

InSTA: Towards Internet-Scale Training For Agents LLaMA: Open and Efficient Foundation Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.428386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.428386Z digest=sha256:5558dc78a8cd2dbdf8f2511952b8d1bf24f1e978e99dfd3710af0f18aab2d383

Observation 4a3f1095-6f52-4036-bae6-8e4e41e07a58 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

InSTA: Towards Internet-Scale Training For Agents Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.432100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.432100Z digest=sha256:79e74347d49a86fce04fd5e83b16e02670f7246b0d95c6f8c23bb6e7c550244f

Observation d0d07af5-0d1b-4b56-9522-de72f1c5921c · outbound

This paper cites Effective data augmentation with diffusion models.

InSTA: Towards Internet-Scale Training For Agents Effective data augmentation with diffusion models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:24:51.251327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.435592Z digest=sha256:aa4a3ae93aeb7c56120f444925bfcc60cd841c0410853ea58ed31488075f812a

Observation e2581515-9656-4d94-8da8-fefae582a480 · outbound

This paper cites LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench.

InSTA: Towards Internet-Scale Training For Agents LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.439003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.439003Z digest=sha256:89c85a4f394355a174b04d09189476535cd68314f72365bcbe3ed0f585c77325

Observation ea5f47ac-2bc3-41b6-b0a4-1ba017f59dd5 · outbound

This paper cites A survey on large language model based autonomous agents.

InSTA: Towards Internet-Scale Training For Agents A survey on large language model based autonomous agents

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.442840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.442840Z digest=sha256:793190ef2f96a3100a065b2aa42251f993464f4f7bd2784d94b1d3410816a94b

Observation 6b8baf50-50fc-4489-ac81-7eccc8655f94 · outbound

This paper cites Large Multimodal Agents: A Survey.

InSTA: Towards Internet-Scale Training For Agents Large Multimodal Agents: A Survey

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.446116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.446116Z digest=sha256:a0b08ff4460f70a2ecdcf4f42599cc8e79f543c89399c32da991b03d500240fe

Observation b2c28686-933b-49a4-ad6b-4da85fece042 · outbound

This paper cites An illusion of progress? assessing the current state of web agents.

InSTA: Towards Internet-Scale Training For Agents An illusion of progress? assessing the current state of web agents

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.449690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.449690Z digest=sha256:dc4ddbcab4c7e56e9463fa9b54f43636fb4474430205b763fabe51b847e6457b

Observation d0ef471f-fea3-455e-ad14-3d8b99bed1b8 · outbound

This paper cites WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents.

InSTA: Towards Internet-Scale Training For Agents WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.453058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.453058Z digest=sha256:3b856621fe54911f3b7fafab42467ba827834168f22cdbe5e8cb3ef3884b3ea5

Observation acbe7994-16b6-4212-8d0f-8f4c6a3e901b · outbound

This paper cites Tree of Thoughts: Deliberate Problem Solving with Large Language Models.

InSTA: Towards Internet-Scale Training For Agents Tree of Thoughts: Deliberate Problem Solving with Large Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.456513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.456513Z digest=sha256:60bcb3160b7fdca621899f2400e4314ba153c0b3e61f89cf3038477d6565836a

Observation 2510f647-cff4-4bc5-b7cb-2071f36393f2 · outbound

This paper cites TextGrad: Automatic "Differentiation" via Text.

InSTA: Towards Internet-Scale Training For Agents TextGrad: Automatic "Differentiation" via Text

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.460002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.460002Z digest=sha256:abd654170f8509f115cd676beda7b54422f30d2553c87bee16b5ab769242546d

Observation 97aee2eb-ff82-439a-8d43-3c209c1b7716 · outbound

This paper cites AgentTuning: Enabling Generalized Agent Abilities for LLMs.

InSTA: Towards Internet-Scale Training For Agents AgentTuning: Enabling Generalized Agent Abilities for LLMs

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.463527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.463527Z digest=sha256:a208e8249c8693c2237d4afafe8b84d932f4834b50d05c9894bd877ab7726556

Observation 3d2010d8-6618-476a-9996-5423180fbf08 · outbound

This paper cites AppAgent: Multimodal Agents as Smartphone Users.

InSTA: Towards Internet-Scale Training For Agents AppAgent: Multimodal Agents as Smartphone Users

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.466991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.466991Z digest=sha256:37ae517b9561c88f4775da3db94167b905b097efed50af9af12cff46abb7369c

Observation bcceb7ae-022a-48ce-9579-cae963cd1bde · outbound

This paper cites Generative Verifiers: Reward Modeling as Next-Token Prediction.

InSTA: Towards Internet-Scale Training For Agents Generative Verifiers: Reward Modeling as Next-Token Prediction

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.470275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.470275Z digest=sha256:f569c1c8254a1c47364ae6ad6f57068cd307b7a30f4aa923bf773221291d74ff

Observation fb46c824-8ea4-407c-b44e-585b13b8234f · outbound

This paper cites Evaluation of openai o1: Opportunities and challenges of agi, 2024.

InSTA: Towards Internet-Scale Training For Agents Evaluation of openai o1: Opportunities and challenges of agi, 2024

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.473513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.473513Z digest=sha256:93d05e2b300fde29a3aaf96e0ce628fcac119d8882572d088a3af1ec4f1a69d6

Observation 07220155-d6cb-4580-83c9-87aaf38fee44 · outbound

This paper cites Language Agent Tree Search Unifies Reasoning Acting and Planning in Language Models.

InSTA: Towards Internet-Scale Training For Agents Language Agent Tree Search Unifies Reasoning Acting and Planning in Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.477162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.477162Z digest=sha256:6fd0e2f87be84af04fbcebf7b34732b273283570aa61ee5c72bd2b4b27849e6e

Observation 0cd18ca5-7ec3-4e6a-b14c-98211eedf41f · outbound

This paper cites WebArena: A Realistic Web Environment for Building Autonomous Agents.

InSTA: Towards Internet-Scale Training For Agents WebArena: A Realistic Web Environment for Building Autonomous Agents

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.480712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.480712Z digest=sha256:4f2be1bcff51f9b00cffde38c1d52f2b4cc9f0c29fdcf3f5c4876494a8825191

Observation 6a5421f1-b069-40e7-97eb-1429e6bd533f · outbound

This paper cites Proposer-Agent-Evaluator(PAE): Autonomous Skill Discovery For Foundation Model Internet Agents.

InSTA: Towards Internet-Scale Training For Agents Proposer-Agent-Evaluator(PAE): Autonomous Skill Discovery For Foundation Model Internet Agents

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.484160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.484160Z digest=sha256:22965bc1b212ce5a39356e359e347c327c4fdffa024e31a93eb55a07561c669d

Pith citing papers

Observation 1d96973f-8934-484e-96ed-85db71ef9fe8 · inbound

Thinking vs. Doing: Agents that Reason by Scaling Test-Time Interaction cites this paper.

Thinking vs. Doing: Agents that Reason by Scaling Test-Time Interaction InSTA: Towards Internet-Scale Training For Agents

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.165917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.165917Z digest=sha256:0b4a31416cb38f8aa7353a849d9d55d323572e2b4d9683dc7693da10b367f85c

Observation fddaf0a7-6b40-4a6c-b57c-0dda691a378f · inbound

DynaWeb: Model-Based Reinforcement Learning of Web Agents cites this paper.

DynaWeb: Model-Based Reinforcement Learning of Web Agents InSTA: Towards Internet-Scale Training For Agents

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:37:41.519723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T09:33:30.444057Z digest=sha256:c48890d56906b1940195cc2be8fcf93c44d26f396c82ef819b838d39b85a47be

Observation 5db362cc-0231-4528-b70f-51b39367709a · inbound

WebFactory: Automated Compression of Foundational Language Intelligence into Grounded Web Agents cites this paper.

WebFactory: Automated Compression of Foundational Language Intelligence into Grounded Web Agents InSTA: Towards Internet-Scale Training For Agents

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:50:10.672001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T16:50:02.527821Z digest=sha256:c941d93ee7993e5d083c25a1dfd58ca6e9138018cdefcc05ee8267994eb5928b

Observation bc2c3cb3-33df-4191-9876-ed6ec68585d6 · inbound

ContractSkill: Repairable Contract-Based Skills for Multimodal Web Agents cites this paper.

ContractSkill: Repairable Contract-Based Skills for Multimodal Web Agents InSTA: Towards Internet-Scale Training For Agents

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-15T08:59:52.986286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T08:58:50.947580Z digest=sha256:bcc2fc0d24baedc70729d4e602323b0b9d0f0ba0c9b83e9c98adec7f56caab06

Observation 3bd86329-3690-40a8-a34f-0320088f0b4b · inbound

Weblica: Scalable and Reproducible Training Environments for Visual Web Agents cites this paper.

Weblica: Scalable and Reproducible Training Environments for Visual Web Agents InSTA: Towards Internet-Scale Training For Agents

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:25:56.929155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T01:25:44.578007Z digest=sha256:51cfa36558e2d69c544b03461ca5afcc9209fff39d3285425a49230635b3dc3c

Observation 262c0d22-9ae3-4eaf-a79f-acb7ffbdded2 · inbound

OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents cites this paper.

OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents InSTA: Towards Internet-Scale Training For Agents

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:06:16.466449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T15:46:50.587684Z digest=sha256:63bbe1a5f8cf17f7d0cbbabc8fd35a9adbcefc52ca12f5736102961fcd1eeccd

Observation c0103017-b7e6-426a-ae5a-27cdef8b4150 · inbound

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application cites this paper.

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application InSTA: Towards Internet-Scale Training For Agents

Reference 267

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:58:02.991562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T09:46:30.702256Z digest=sha256:f7013f2d47400e02fca399ca0f6c8748dc0ae9898badde998e074a776b34328b

Observation 0ab252e1-a260-40eb-ac04-73dc2918661a · inbound

ProCUA-SFT Technical Report cites this paper.

ProCUA-SFT Technical Report InSTA: Towards Internet-Scale Training For Agents

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T18:08:46.699111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T03:18:04.149281Z digest=sha256:11f17a9e150d58ac212a4d653bac379e55d7226285c5bb7198e75c1c84eb1930

Observation 47e1d398-e6d2-4006-912a-9021323ea636 · inbound

Empowering GUI Agents via Autonomous Experience Exploration and Hindsight Experience Utilization for Task Planning cites this paper.

Empowering GUI Agents via Autonomous Experience Exploration and Hindsight Experience Utilization for Task Planning InSTA: Towards Internet-Scale Training For Agents

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-04T14:29:53.385736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T03:51:51.827622Z digest=sha256:a5121210041fafd397f86e539466bba7c6153becc235a92d8fc57eb949458b38

Observation ae9c1a53-c75c-4d84-82f0-0a7ebaf1b595 · inbound

OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models cites this paper.

OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models InSTA: Towards Internet-Scale Training For Agents

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-31T02:18:01.822251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T02:18:01.822251Z digest=sha256:fa8765177313c2769d179cda1ad2542292fc1b78b4bfc7a9ca9df59835edd255