Pith. sign in

Paper Citation Record · LEDGER

InSTA: Towards Internet-Scale Training For Agents

As of 9 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 10 inbound Pith citation observations for arXiv:2502.06776.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.06776 v2

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T14:24:50.484160Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:27:50.165917Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

58 of 58 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved53
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation ee2d6038-0004-403e-9fdb-8f90a9517aaf · outbound

This paper cites write newline.

InSTA: Towards Internet-Scale Training For Agents write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.283326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.283326Z digest=sha256:65d2eb038f5a189de264fd3d302e3ef52ba6def5c0388bdf2f0e1f96b3b0703d

Observation 073770c8-88c3-430a-9ebf-4f4815ffca7a · outbound

This paper cites @esa (Ref.

InSTA: Towards Internet-Scale Training For Agents @esa (Ref

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.288008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.288008Z digest=sha256:5fd537a1b744b0ac833154fd6584f5792ef4f20ed378941f81493e30f2b19930

Observation d92b4c8b-a731-4d17-a5b4-2f3b2d85bdbd · outbound

This paper cites an unresolved cited work.

InSTA: Towards Internet-Scale Training For Agents Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.291843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.291843Z digest=sha256:d4697ddd10479ae849c2cba7c8308161b5c180ed809f70f6e170eda18e07030c

Observation 92af65ee-1f28-4220-95d4-5948aafd16dc · outbound

This paper cites an unresolved cited work.

InSTA: Towards Internet-Scale Training For Agents Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:24:51.337301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.295342Z digest=sha256:225591e0c9ea6feddfd2f5745584494518d966e96471d37b2f044024ddeccc31

Observation f3921423-24e3-4645-b9ce-40b36b05b58a · outbound

This paper cites Language models as agent models.

InSTA: Towards Internet-Scale Training For Agents Language models as agent models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.298819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.298819Z digest=sha256:5c04d0622a8a896b15d3d48f52cca1aff95ce50c7a3298378a2881aa5dd5799e

Observation fbe027e7-122d-42f4-b87a-3849bd3aa532 · outbound

This paper cites Graph of thoughts: Solving elaborate problems with large language models.

InSTA: Towards Internet-Scale Training For Agents Graph of thoughts: Solving elaborate problems with large language models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.302297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.302297Z digest=sha256:d2d4f5e69645815fabdfed0cea19c2e1a0420df76f4c255a9a46368d2a58fc1c

Observation cfde26e3-8a15-42c2-bd97-d056ffa81aa4 · outbound

This paper cites Language Models are Few-Shot Learners.

InSTA: Towards Internet-Scale Training For Agents Language Models are Few-Shot Learners

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.305890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.305890Z digest=sha256:2b2c03df0d4894f5bb214334493b568ae8602a1f08dcc1416ec6c06aa40d3d52

Observation 0305d162-35a3-4e8b-871f-307ab708f6e6 · outbound

This paper cites Sparks of Artificial General Intelligence: Early experiments with GPT-4.

InSTA: Towards Internet-Scale Training For Agents Sparks of Artificial General Intelligence: Early experiments with GPT-4

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.309685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.309685Z digest=sha256:dee6d70fb4c017a764ea43d2f6825717831537b61a66d368415a862acecc4a77

Observation f718996a-ad8b-406c-8ffd-65ecf8b4eb47 · outbound

This paper cites FireAct: Toward Language Agent Fine-tuning.

InSTA: Towards Internet-Scale Training For Agents FireAct: Toward Language Agent Fine-tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.313595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.313595Z digest=sha256:a9fea9f0a803ebf9b63cb2cbce6ba5b471c49dd764575f259bd9b0463bbd9ce5

Observation c110ffad-3e6e-48ed-8d41-623f533a19cd · outbound

This paper cites The BrowserGym Ecosystem for Web Agent Research.

InSTA: Towards Internet-Scale Training For Agents The BrowserGym Ecosystem for Web Agent Research

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.317546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.317546Z digest=sha256:cd40cef6be22f1a3e267f9220faf2f6b0b977ba26d66e0391e234ac0a59bae31

Observation 7ac9c8c2-f6e7-4864-9f17-db73edb8daa0 · outbound

This paper cites Mind2Web: Towards a Generalist Agent for the Web.

InSTA: Towards Internet-Scale Training For Agents Mind2Web: Towards a Generalist Agent for the Web

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.321613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.321613Z digest=sha256:737bf57dfbb62904052abd5998533117914a6b4daa45caeac7dc1e305a034c85

Observation 3e458e21-9821-4ed6-abe6-1515b6d431c2 · outbound

This paper cites Better Synthetic Data by Retrieving and Transforming Existing Datasets.

InSTA: Towards Internet-Scale Training For Agents Better Synthetic Data by Retrieving and Transforming Existing Datasets

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.325236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.325236Z digest=sha256:e5dbedcb1a9e6bb34bdb837ade2d7bca99ca9ea50faccf81d43a67d08ec03c0f

Observation b76ffa0c-df13-42a3-8686-41db08222d2d · outbound

This paper cites The Llama 3 Herd of Models.

InSTA: Towards Internet-Scale Training For Agents The Llama 3 Herd of Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.328784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.328784Z digest=sha256:07a455940ecfb07b838156025e09f28066a2b32c99e2c9822a5559ac15af929a

Observation 6a45155b-e0ad-4d65-a8ba-9af526c8c928 · outbound

This paper cites W eb V oyager: Building an end-to-end web agent with large multimodal models.

InSTA: Towards Internet-Scale Training For Agents W eb V oyager: Building an end-to-end web agent with large multimodal models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.332092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.332092Z digest=sha256:4140679667665ef0d2a7b2ec1eec6b9f223627cd647dac1a033661ee73943ad7

Observation c4a373ce-8807-43ad-8fc8-84f984a0d932 · outbound

This paper cites CogAgent: A Visual Language Model for GUI Agents.

InSTA: Towards Internet-Scale Training For Agents CogAgent: A Visual Language Model for GUI Agents

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.335297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.335297Z digest=sha256:0d0c84e8ab73f1d17815675bf77172285df0b608b1f503413fa65834949710ba

Observation 0ec5b2f6-a62a-446a-b3fa-fcddc1de3ada · outbound

This paper cites Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations.

InSTA: Towards Internet-Scale Training For Agents Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.338817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.338817Z digest=sha256:7d6bb2e6776cbb13e634c18eeb013140e561ac046409b4bce9cb1413888670ca

Observation cb1cf1c8-3ec2-44c5-8a62-5c49dd361be3 · outbound

This paper cites VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks.

InSTA: Towards Internet-Scale Training For Agents VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.342376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.342376Z digest=sha256:a6c649d1fb9dbfb15ff2b877d4b229d1d6e53abd6589af02ae4c765c08cbaf70

Observation e0ff1980-b081-4620-a325-683c051163b1 · outbound

This paper cites Tree search for language model agents, 2024 b.

InSTA: Towards Internet-Scale Training For Agents Tree search for language model agents, 2024 b

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.345982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.345982Z digest=sha256:430593d5bf053c03924c903e10ae87c46b109fe5b7bdc3e2fee0968a26c00cfa

Observation 7749ecf6-5cff-41c9-a5e2-29246ca3984c · outbound

This paper cites Efficient Memory Management for Large Language Model Serving with PagedAttention.

InSTA: Towards Internet-Scale Training For Agents Efficient Memory Management for Large Language Model Serving with PagedAttention

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.349153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.349153Z digest=sha256:0e9ef64c8979a06e29c523e27251d083f2013e33d18531cca4f1b937d15767d0

Observation 9dade60d-eb87-49f7-b28b-630132912a4c · outbound

This paper cites RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback.

InSTA: Towards Internet-Scale Training For Agents RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.352547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.352547Z digest=sha256:b5c042e5f218cc6783452e4291766e46b4baf79851d01fd2efd4591b62036f36

Observation 958df537-a72c-4b84-9e31-a83b5e9c13de · outbound

This paper cites From generation to judgment: Opportunities and challenges of llm-as-a-judge.

InSTA: Towards Internet-Scale Training For Agents From generation to judgment: Opportunities and challenges of llm-as-a-judge

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.355887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.355887Z digest=sha256:609cd89aade5279241bfc5b5e46276476d3e39861153d587090c76fa2107cafb

Observation cbf8049c-5d5b-4b04-ac59-83bd09176a9c · outbound

This paper cites VisualWebBench: How Far Have Multimodal LLMs Evolved in Web Page Understanding and Grounding?.

InSTA: Towards Internet-Scale Training For Agents VisualWebBench: How Far Have Multimodal LLMs Evolved in Web Page Understanding and Grounding?

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.359070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.359070Z digest=sha256:e2c15edf913a69c52c4218019390eb9c5d83a617d68a70554ebb0afe6c9e8a9e

Observation 987e6e94-0ef5-4376-b55a-a3221f83c769 · outbound

This paper cites Weblinx: Real-world website navigation with multi-turn dialogue, 2024.

InSTA: Towards Internet-Scale Training For Agents Weblinx: Real-world website navigation with multi-turn dialogue, 2024

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.362660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.362660Z digest=sha256:eb2b39b4f884527e4a7819370205ae95420071b3140bd2b37c1218e9b7b377b0

Observation 8188a3af-9600-4974-92aa-d068b7e3e320 · outbound

This paper cites Self-refine: Iterative refinement with self-feedback.

InSTA: Towards Internet-Scale Training For Agents Self-refine: Iterative refinement with self-feedback

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.366344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.366344Z digest=sha256:3c58f98e2d318e15cf9b527a3813c5bd086b7d06962300f8faeec1d4ce8cdc84

Observation d2886aef-e7e6-4380-a9f6-3128d5b9fb03 · outbound

This paper cites Playwright.

InSTA: Towards Internet-Scale Training For Agents Playwright

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:24:51.319624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.369944Z digest=sha256:da27d355af22e308c00e73b91cfa128664095de075d540b14819be5de549dda6

Observation c6b20d75-769c-4c46-a53b-29305c4a39d9 · outbound

This paper cites AgentInstruct: Toward Generative Teaching with Agentic Flows.

InSTA: Towards Internet-Scale Training For Agents AgentInstruct: Toward Generative Teaching with Agentic Flows

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.373086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.373086Z digest=sha256:fd4d83da4dd63de88032a9977e9fb68baddc2f57703480fb8fb6bb3a876946df

Observation 0f3e2a0f-c667-495f-8232-719fab758aa1 · outbound

This paper cites NNetNav: Unsupervised Learning of Browser Agents Through Environment Interaction in the Wild.

InSTA: Towards Internet-Scale Training For Agents NNetNav: Unsupervised Learning of Browser Agents Through Environment Interaction in the Wild

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.376394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.376394Z digest=sha256:18993d622e34cdb99bd16c4743d9f4219a6383122af72d145db5b9c10314f741

Observation b0678889-2d91-4977-9eb7-b206f822a189 · outbound

This paper cites Synatra: Turning Indirect Knowledge into Direct Demonstrations for Digital Agents at Scale.

InSTA: Towards Internet-Scale Training For Agents Synatra: Turning Indirect Knowledge into Direct Demonstrations for Digital Agents at Scale

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.379591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.379591Z digest=sha256:764273b751da01ccf8749cf09642061f6a9d427e3206417b7670b42c91e7c485

Observation c7bc3b97-9d02-4cb0-a086-90a438797921 · outbound

This paper cites an unresolved cited work.

InSTA: Towards Internet-Scale Training For Agents Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:24:51.308230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.382859Z digest=sha256:7991333ddfc4802aa7dd94aaa6492a5a7c534205aea8a03de47efc9e088c21cd

Observation 0a7fb9ba-df97-4fc1-9ade-367acd1a7d3a · outbound

This paper cites Large Language Models Can Self-Improve At Web Agent Tasks.

InSTA: Towards Internet-Scale Training For Agents Large Language Models Can Self-Improve At Web Agent Tasks

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.385849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.385849Z digest=sha256:556488a5712ae798b036e9bf8342aa150163941d18970678b3d8e6a7154ff213

Observation da1d925a-a9bc-4dc7-84a9-52a65ff6659e · outbound

This paper cites REFINER : Reasoning feedback on intermediate representations.

InSTA: Towards Internet-Scale Training For Agents REFINER : Reasoning feedback on intermediate representations

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:24:51.297090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.389101Z digest=sha256:b1c2b080d87695c5a0c67451e2c5939e018c06e1f55fb72712df6950d1f2c8f1

Observation 34ab8fa1-d648-4230-803b-51f16d13253b · outbound

This paper cites Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents.

InSTA: Towards Internet-Scale Training For Agents Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.392087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.392087Z digest=sha256:2ecefc8ab926dc88f643ea2f96518bf7804b876b12e16d3b7332808076761f9f

Observation 5aefd61b-c7af-4b15-8c25-9615a88aa8b0 · outbound

This paper cites Language models are unsupervised multitask learners, 2019.

InSTA: Towards Internet-Scale Training For Agents Language models are unsupervised multitask learners, 2019

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:24:51.285581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.396681Z digest=sha256:12d369956320ac1dfb28248a7867e46b45b998b74e08c55cb18f06841a190167

Observation 96067855-455f-4833-a61d-1ab51bf68fb0 · outbound

This paper cites Android in the Wild: A Large-Scale Dataset for Android Device Control.

InSTA: Towards Internet-Scale Training For Agents Android in the Wild: A Large-Scale Dataset for Android Device Control

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.399754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.399754Z digest=sha256:45c2a01bcac99508a19032cb597e7fba0be3c0e0026b3758ca5772b0771c736a

Observation a3f733c7-5748-4101-aabc-2baffffc950e · outbound

This paper cites Toolformer: Language models can teach themselves to use tools.

InSTA: Towards Internet-Scale Training For Agents Toolformer: Language models can teach themselves to use tools

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.403052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.403052Z digest=sha256:75b6e2b3440654e8d6dea0925bb2b7fdf8bfcd14fa1a31e584e9d1d1e12abfff

Observation 0206bbf1-b54a-4626-bcb5-a65cff803769 · outbound

This paper cites RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold.

InSTA: Towards Internet-Scale Training For Agents RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.406158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.406158Z digest=sha256:98d49d34dc306d0280a9308eeb791426564595b66319d4a7ca8946ef2e9053bd

Observation 17a875e4-7a63-453b-a9c5-d12662c1baf5 · outbound

This paper cites ScribeAgent: Towards Specialized Web Agents Using Production-Scale Workflow Data.

InSTA: Towards Internet-Scale Training For Agents ScribeAgent: Towards Specialized Web Agents Using Production-Scale Workflow Data

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.409647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.409647Z digest=sha256:a86ecd38f760664e812624b54d5c568a372922ea4f8de063bd76ab99eef8f5df

Observation 5a806af7-f919-47b1-bf23-6c8b369eeed3 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

InSTA: Towards Internet-Scale Training For Agents Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.413143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.413143Z digest=sha256:ef085fca77a1add8b99e29872305cebb7f8d4635655d0963e3ea7ad0f88ad802

Observation 5c808863-c9dd-4d0d-b6e7-358c2c4dc874 · outbound

This paper cites Fast best-of-n decoding via speculative rejection.

InSTA: Towards Internet-Scale Training For Agents Fast best-of-n decoding via speculative rejection

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.416715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.416715Z digest=sha256:c45cc6901a4777870f86d7b77575e510d3debbd1ebfd81ce02b6b82d74ae4e4e

Observation 8338de25-381f-4e0e-9505-09ee70966e48 · outbound

This paper cites Preference Fine-Tuning of LLMs Should Leverage Suboptimal, On-Policy Data.

InSTA: Towards Internet-Scale Training For Agents Preference Fine-Tuning of LLMs Should Leverage Suboptimal, On-Policy Data

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.420177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.420177Z digest=sha256:f17a8e694060f370a9c3f6cf7d9228c0aaa2015d6d57e9b31829fb1ed0c9b01b

Observation da041530-0d95-4601-be8a-97927d1ea3e8 · outbound

This paper cites Common crawl, 2025.

InSTA: Towards Internet-Scale Training For Agents Common crawl, 2025

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:24:51.262720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.424981Z digest=sha256:947010ed08d502348b51dfe78bf51270144c2b19c7014d406df8a12c6f991dc7

Observation ea94054e-d6a4-4c35-a502-8de8f602f350 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

InSTA: Towards Internet-Scale Training For Agents LLaMA: Open and Efficient Foundation Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.428386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.428386Z digest=sha256:5558dc78a8cd2dbdf8f2511952b8d1bf24f1e978e99dfd3710af0f18aab2d383

Observation 4a3f1095-6f52-4036-bae6-8e4e41e07a58 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

InSTA: Towards Internet-Scale Training For Agents Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.432100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.432100Z digest=sha256:79e74347d49a86fce04fd5e83b16e02670f7246b0d95c6f8c23bb6e7c550244f

Observation d0d07af5-0d1b-4b56-9522-de72f1c5921c · outbound

This paper cites Effective data augmentation with diffusion models.

InSTA: Towards Internet-Scale Training For Agents Effective data augmentation with diffusion models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:24:51.251327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:24:50.435592Z digest=sha256:fbef74bc107f9f7043f5641f6ce56d6c2951e4dd754c96a9151b4af92c32b442

Observation e2581515-9656-4d94-8da8-fefae582a480 · outbound

This paper cites LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench.

InSTA: Towards Internet-Scale Training For Agents LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.439003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.439003Z digest=sha256:89c85a4f394355a174b04d09189476535cd68314f72365bcbe3ed0f585c77325

Observation ea5f47ac-2bc3-41b6-b0a4-1ba017f59dd5 · outbound

This paper cites A survey on large language model based autonomous agents.

InSTA: Towards Internet-Scale Training For Agents A survey on large language model based autonomous agents

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.442840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.442840Z digest=sha256:793190ef2f96a3100a065b2aa42251f993464f4f7bd2784d94b1d3410816a94b

Observation 6b8baf50-50fc-4489-ac81-7eccc8655f94 · outbound

This paper cites Large Multimodal Agents: A Survey.

InSTA: Towards Internet-Scale Training For Agents Large Multimodal Agents: A Survey

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.446116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.446116Z digest=sha256:a0b08ff4460f70a2ecdcf4f42599cc8e79f543c89399c32da991b03d500240fe

Observation b2c28686-933b-49a4-ad6b-4da85fece042 · outbound

This paper cites An illusion of progress? assessing the current state of web agents.

InSTA: Towards Internet-Scale Training For Agents An illusion of progress? assessing the current state of web agents

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.449690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.449690Z digest=sha256:dc4ddbcab4c7e56e9463fa9b54f43636fb4474430205b763fabe51b847e6457b

Observation d0ef471f-fea3-455e-ad14-3d8b99bed1b8 · outbound

This paper cites WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents.

InSTA: Towards Internet-Scale Training For Agents WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.453058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.453058Z digest=sha256:3b856621fe54911f3b7fafab42467ba827834168f22cdbe5e8cb3ef3884b3ea5

Observation acbe7994-16b6-4212-8d0f-8f4c6a3e901b · outbound

This paper cites Tree of Thoughts: Deliberate Problem Solving with Large Language Models.

InSTA: Towards Internet-Scale Training For Agents Tree of Thoughts: Deliberate Problem Solving with Large Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.456513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.456513Z digest=sha256:60bcb3160b7fdca621899f2400e4314ba153c0b3e61f89cf3038477d6565836a

Observation 2510f647-cff4-4bc5-b7cb-2071f36393f2 · outbound

This paper cites TextGrad: Automatic "Differentiation" via Text.

InSTA: Towards Internet-Scale Training For Agents TextGrad: Automatic "Differentiation" via Text

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.460002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.460002Z digest=sha256:abd654170f8509f115cd676beda7b54422f30d2553c87bee16b5ab769242546d

Observation 97aee2eb-ff82-439a-8d43-3c209c1b7716 · outbound

This paper cites AgentTuning: Enabling Generalized Agent Abilities for LLMs.

InSTA: Towards Internet-Scale Training For Agents AgentTuning: Enabling Generalized Agent Abilities for LLMs

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.463527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.463527Z digest=sha256:a208e8249c8693c2237d4afafe8b84d932f4834b50d05c9894bd877ab7726556

Observation 3d2010d8-6618-476a-9996-5423180fbf08 · outbound

This paper cites AppAgent: Multimodal Agents as Smartphone Users.

InSTA: Towards Internet-Scale Training For Agents AppAgent: Multimodal Agents as Smartphone Users

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.466991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.466991Z digest=sha256:37ae517b9561c88f4775da3db94167b905b097efed50af9af12cff46abb7369c

Observation bcceb7ae-022a-48ce-9579-cae963cd1bde · outbound

This paper cites Generative Verifiers: Reward Modeling as Next-Token Prediction.

InSTA: Towards Internet-Scale Training For Agents Generative Verifiers: Reward Modeling as Next-Token Prediction

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.470275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.470275Z digest=sha256:f569c1c8254a1c47364ae6ad6f57068cd307b7a30f4aa923bf773221291d74ff

Observation fb46c824-8ea4-407c-b44e-585b13b8234f · outbound

This paper cites Evaluation of openai o1: Opportunities and challenges of agi, 2024.

InSTA: Towards Internet-Scale Training For Agents Evaluation of openai o1: Opportunities and challenges of agi, 2024

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.473513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.473513Z digest=sha256:93d05e2b300fde29a3aaf96e0ce628fcac119d8882572d088a3af1ec4f1a69d6

Observation 07220155-d6cb-4580-83c9-87aaf38fee44 · outbound

This paper cites Language Agent Tree Search Unifies Reasoning Acting and Planning in Language Models.

InSTA: Towards Internet-Scale Training For Agents Language Agent Tree Search Unifies Reasoning Acting and Planning in Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.477162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.477162Z digest=sha256:6fd0e2f87be84af04fbcebf7b34732b273283570aa61ee5c72bd2b4b27849e6e

Observation 0cd18ca5-7ec3-4e6a-b14c-98211eedf41f · outbound

This paper cites WebArena: A Realistic Web Environment for Building Autonomous Agents.

InSTA: Towards Internet-Scale Training For Agents WebArena: A Realistic Web Environment for Building Autonomous Agents

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.480712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.480712Z digest=sha256:4f2be1bcff51f9b00cffde38c1d52f2b4cc9f0c29fdcf3f5c4876494a8825191

Observation 6a5421f1-b069-40e7-97eb-1429e6bd533f · outbound

This paper cites Proposer-Agent-Evaluator(PAE): Autonomous Skill Discovery For Foundation Model Internet Agents.

InSTA: Towards Internet-Scale Training For Agents Proposer-Agent-Evaluator(PAE): Autonomous Skill Discovery For Foundation Model Internet Agents

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-08T14:24:50.484160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:24:50.484160Z digest=sha256:22965bc1b212ce5a39356e359e347c327c4fdffa024e31a93eb55a07561c669d

Pith citing papers

Observation 1d96973f-8934-484e-96ed-85db71ef9fe8 · inbound

Thinking vs. Doing: Agents that Reason by Scaling Test-Time Interaction cites this paper.

Thinking vs. Doing: Agents that Reason by Scaling Test-Time Interaction InSTA: Towards Internet-Scale Training For Agents

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.165917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.165917Z digest=sha256:0b4a31416cb38f8aa7353a849d9d55d323572e2b4d9683dc7693da10b367f85c

Observation fddaf0a7-6b40-4a6c-b57c-0dda691a378f · inbound

DynaWeb: Model-Based Reinforcement Learning of Web Agents cites this paper.

DynaWeb: Model-Based Reinforcement Learning of Web Agents InSTA: Towards Internet-Scale Training For Agents

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:37:41.519723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T09:33:30.444057Z digest=sha256:1c783e4240d2090af3e36f2293fcb151723f367bc79df8cd1aa1fce1158d72c0

Observation 5db362cc-0231-4528-b70f-51b39367709a · inbound

WebFactory: Automated Compression of Foundational Language Intelligence into Grounded Web Agents cites this paper.

WebFactory: Automated Compression of Foundational Language Intelligence into Grounded Web Agents InSTA: Towards Internet-Scale Training For Agents

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:50:10.672001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T16:50:02.527821Z digest=sha256:bfe78139edc4398ad729ecc1b0715e471b40b2b8bf016e7fe060170c1c525a4e

Observation bc2c3cb3-33df-4191-9876-ed6ec68585d6 · inbound

ContractSkill: Repairable Contract-Based Skills for Multimodal Web Agents cites this paper.

ContractSkill: Repairable Contract-Based Skills for Multimodal Web Agents InSTA: Towards Internet-Scale Training For Agents

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-15T08:59:52.986286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T08:58:50.947580Z digest=sha256:efc1aa39d8abe11b63714fd411d01ed365a21b01d6c70f066fca052005502e40

Observation 3bd86329-3690-40a8-a34f-0320088f0b4b · inbound

Weblica: Scalable and Reproducible Training Environments for Visual Web Agents cites this paper.

Weblica: Scalable and Reproducible Training Environments for Visual Web Agents InSTA: Towards Internet-Scale Training For Agents

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:25:56.929155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T01:25:44.578007Z digest=sha256:ffbef744771511d77396f87e5e91647190cede2ad99a3d268dc08140932de2ce

Observation 262c0d22-9ae3-4eaf-a79f-acb7ffbdded2 · inbound

OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents cites this paper.

OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents InSTA: Towards Internet-Scale Training For Agents

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:06:16.466449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T15:46:50.587684Z digest=sha256:f3032000ce73988368d9c67799e2e9b69283f4f98630b188627be4a193bf3ff9

Observation c0103017-b7e6-426a-ae5a-27cdef8b4150 · inbound

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application cites this paper.

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application InSTA: Towards Internet-Scale Training For Agents

Reference 267

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:58:02.991562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T09:46:30.702256Z digest=sha256:4d7307a9983da7457766d907198cdbb7b485ced3edf894a1f1c4f45e91ab7b35

Observation 0ab252e1-a260-40eb-ac04-73dc2918661a · inbound

ProCUA-SFT Technical Report cites this paper.

ProCUA-SFT Technical Report InSTA: Towards Internet-Scale Training For Agents

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T18:08:46.699111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T03:18:04.149281Z digest=sha256:1ab71fda30e46a3399978c919a5ff5f3c4bda46f17213bedcd202a8872b48f97

Observation 47e1d398-e6d2-4006-912a-9021323ea636 · inbound

Empowering GUI Agents via Autonomous Experience Exploration and Hindsight Experience Utilization for Task Planning cites this paper.

Empowering GUI Agents via Autonomous Experience Exploration and Hindsight Experience Utilization for Task Planning InSTA: Towards Internet-Scale Training For Agents

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-04T14:29:53.385736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T03:51:51.827622Z digest=sha256:e6536c62f5b6058e43a19427a4209a37ccf3dc8f824a3262ae5daa0e5330d304

Observation ae9c1a53-c75c-4d84-82f0-0a7ebaf1b595 · inbound

OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models cites this paper.

OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models InSTA: Towards Internet-Scale Training For Agents

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-31T02:18:01.822251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T02:18:01.822251Z digest=sha256:3c05625ce64fcfc1836706f80e1a84a6b368e8f0d9259dc68e98ed1e622d39b3