Pith. sign in

Paper Citation Record · LEDGER

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations

As of 22 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 6 inbound Pith citation observations for arXiv:2411.13451.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.13451 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T16:27:47.950451Z

measured 69 of 69 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:39:33.284170Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T19:36:08.845862Z

Reference resolution

63 of 63 outbound references displayed

  • verified exact0
  • verified fuzzy21
  • unresolved41
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 36dc160f-1a51-4c9d-b95a-67b8bd07de61 · outbound

This paper cites Apprenticeship learning via inverse reinforcement learning.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Apprenticeship learning via inverse reinforcement learning

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.653399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.700852Z digest=sha256:0ca9db7172b3d7ae91398d71f9bea74bfbed9153b9c9a636586bd7e74b8ec476

Observation b0fa7292-210a-4f63-abd5-d8d2523f5af9 · outbound

This paper cites GPT-4 Technical Report.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations GPT-4 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.705207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.705207Z digest=sha256:48962a8d69e661fb09ec2a0ed1de11dbab356a8ea44c454cc35f6777be6ee29e

Observation 512697fa-9823-4bf3-9468-0c11d5ac94a3 · outbound

This paper cites A survey of robot learning from demonstration.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations A survey of robot learning from demonstration

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.641628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.710116Z digest=sha256:3ac8169dc64b24fbdefb1ff2c04816731dac5e2d8181ba56048317edfc1c14d3

Observation d094a3c2-ee60-4ade-9a53-5a077effa084 · outbound

This paper cites In-Context Learning with Long-Context Models: An In-Depth Exploration.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations In-Context Learning with Long-Context Models: An In-Depth Exploration

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.714301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.714301Z digest=sha256:1840a85de4c7f63f6a6d966f5f7591c149f29c3510e28156e56144e7620000ea

Observation 22048b7f-6e44-48f6-88c3-a0c4f462d600 · outbound

This paper cites WorkArena++: Towards Compositional Planning and Reasoning-based Common Knowledge Work Tasks.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations WorkArena++: Towards Compositional Planning and Reasoning-based Common Knowledge Work Tasks

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.718496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.718496Z digest=sha256:e838e1699329ff10fc51db99aaf4c89e336104ec3c696155ce84f09ab9a5d953

Observation bc4cf55c-5ee0-4089-9157-db24a5ba320a · outbound

This paper cites Robots that imitate humans.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Robots that imitate humans

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.630462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.723127Z digest=sha256:b7c7ff0b92b4f090c72ffc5b48ef2a8b1c236ade3063a856883b35ce37bb7049

Observation bbfb3c10-2aed-4e9b-b3f0-1affec31d539 · outbound

This paper cites Extrapolating beyond sub- optimal demonstrations via inverse reinforcement learning from observations.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Extrapolating beyond sub- optimal demonstrations via inverse reinforcement learning from observations

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.619461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.727562Z digest=sha256:0f50a3b15beed939e604584ffcc7e0e76295613874d6d6a604e38dcdc7daa54a

Observation fee9e695-b9f0-47a0-a74b-f4e9d8d28ea4 · outbound

This paper cites Language Models are Few-Shot Learners.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Language Models are Few-Shot Learners

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.731275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.731275Z digest=sha256:8c097d5768f30401d5387087096f20002a1d380c859bedbbccf738311354263a

Observation 8f42f47b-25d0-4b60-b9e3-1b9078764edd · outbound

This paper cites Learning and reproduction of gestures by imitation.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Learning and reproduction of gestures by imitation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.607701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.735082Z digest=sha256:40e131d7cdb3e524a3b21cc5b76509dbd9dc9ce9fd2dd431c779fd94aa096069

Observation 5ca0d8fa-5faf-4810-8133-a0efbaca725b · outbound

This paper cites On learning, representing, and generalizing a task in a humanoid robot.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations On learning, representing, and generalizing a task in a humanoid robot

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.596177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.738915Z digest=sha256:bddbd4b1a16828586c3dc8a3110a8da46a3bdc763dae97e23001cc990dd92f46

Observation 10ef0918-974e-4a62-8ca4-3d71230ee7ab · outbound

This paper cites Learning from suboptimal demonstration via self-supervised reward regression.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Learning from suboptimal demonstration via self-supervised reward regression

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.742770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.742770Z digest=sha256:9362e71e964cf61dc7555e87523d5bc286cce54df0e0c7e9525571e479e676a0

Observation 68567748-7b01-4b7d-a599-8d198d91e000 · outbound

This paper cites SeeClick: Harnessing GUI Grounding for Advanced Visual GUI Agents.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations SeeClick: Harnessing GUI Grounding for Advanced Visual GUI Agents

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.746661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.746661Z digest=sha256:1561788fa1a4c9f8a926c6b1a778778aaa24e1927bc88b05e67ff79775dcb2f4

Observation e6cfdd05-6927-4ae3-8f9a-8aa130279110 · outbound

This paper cites Model-based inverse reinforcement learning from visual demonstrations.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Model-based inverse reinforcement learning from visual demonstrations

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.577094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.750462Z digest=sha256:f6fb9a33d0513cd5f8829c8676f26a778a41ed682aea4ba532f35b4611ab8853

Observation 26be4b1a-8596-4368-b384-83128fcb38a5 · outbound

This paper cites Mind2web: Towards a generalist agent for the web, 2023.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Mind2web: Towards a generalist agent for the web, 2023

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.754205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.754205Z digest=sha256:10d541d1f4be05e82878e657e8c4794706ecf09cc4041debf4d9501a11492dde

Observation 645cd879-aeef-498c-bf16-62ed8b84eaa8 · outbound

This paper cites Inverse kkt: Learning cost functions of manipulation tasks from demonstrations.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Inverse kkt: Learning cost functions of manipulation tasks from demonstrations

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.559417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.757739Z digest=sha256:5e8d0f4d19eb67b937379e9a95fbee4066cb8b1199656d87f89ae6339529b2f2

Observation 1f4cb291-fa46-4370-863c-debdc7cd83af · outbound

This paper cites Model-agnostic meta-learning for fast adap- tation of deep networks.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Model-agnostic meta-learning for fast adap- tation of deep networks

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.761368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.761368Z digest=sha256:50a97dafa3851a04a2d70dd318301a680c195a60e595cdbc225b860f39dc5021

Observation 65dd1256-87f8-4b10-af70-77b851a19385 · outbound

This paper cites One-shot visual imitation learning via meta-learning.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations One-shot visual imitation learning via meta-learning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.542552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.765045Z digest=sha256:c9c1cec1cc8e477476117e952632ccbff9a152afc98c0848294dc24424243fe7

Observation da3549be-eb60-4837-8a12-47486dcd957a · outbound

This paper cites Explaining and Harnessing Adversarial Examples.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Explaining and Harnessing Adversarial Examples

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.768429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.768429Z digest=sha256:4981f10423d7d5a490e9f771a71f8d233601eecc7a140bd699118dc4e3fda6c1

Observation 503ce36d-62a1-49b0-be57-810cdaa47455 · outbound

This paper cites A real-world webagent with planning, long context understanding, and program synthesis, 2024.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations A real-world webagent with planning, long context understanding, and program synthesis, 2024

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.530480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.773192Z digest=sha256:d6c5e8719545b46a28a08e4ef8433c230b5f7d372d045b194b161c168c2ce243

Observation 6ade81d5-e693-4ba9-b627-95df74eee4fd · outbound

This paper cites WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.777143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.777143Z digest=sha256:a2dfe3bcc6902abbe229da02eedbb827f961d5b4ecf31f8034a2f622ae562e63

Observation 2a008cdd-c087-4a69-83af-a9655f8c39c7 · outbound

This paper cites Generative adversarial imitation learning.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Generative adversarial imitation learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.780855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.780855Z digest=sha256:268acfdc80036589567118c98e7ff31477801475d4d63a456c3f5b6163a516a1

Observation a60f0ea9-a0eb-4ca4-ab11-b3057dcfe166 · outbound

This paper cites Cogagent: A visual language model for gui agents, 2023.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Cogagent: A visual language model for gui agents, 2023

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.784919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.784919Z digest=sha256:c92022b3513e1b48719523173114079b963e2b4e54f58d2b7d2ad19af424f4d6

Observation e878e8d1-cc08-4a6b-8d3a-7a2ae31b1a10 · outbound

This paper cites A data-driven approach for learning to control computers.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations A data-driven approach for learning to control computers

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.788505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.788505Z digest=sha256:c4417670cf531bf4f69a60cc5f17c988c2fa85739e69f49b58c076fc75d46980

Observation 23d1814c-6a2e-4bac-9db8-6dbb0f293a6c · outbound

This paper cites Imitation learning: A survey of learning methods.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Imitation learning: A survey of learning methods

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.497339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.792368Z digest=sha256:146133f609edeb8f88a5c2ffae7dade6b6dbf75f64a6ca67a0f07e5601629fc1

Observation 4ae50598-436b-4acf-9b48-44c0a66cce34 · outbound

This paper cites Do BERTs Learn to Use Browser User Interface? Exploring Multi-Step Tasks with Unified Vision-and-Language BERTs.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Do BERTs Learn to Use Browser User Interface? Exploring Multi-Step Tasks with Unified Vision-and-Language BERTs

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.796305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.796305Z digest=sha256:4e91647521cf953cfe3c958aeb3fb570c581e4642febd9341aaa09738a4dfd42

Observation 7e95e77f-1bf4-4590-b939-3e377113422d · outbound

This paper cites Data-efficient alignment of large language models with human feedback through natural language, 2023.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Data-efficient alignment of large language models with human feedback through natural language, 2023

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.486085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.800650Z digest=sha256:b1b13978fbbf404de7f6d060979a3bb53ccb807fd24c1a3a24253ba9dd5b6ca3

Observation 8d5919ed-eeac-4762-bc9a-291e66e85c36 · outbound

This paper cites DSPy: Compiling Declarative Language Model Calls into Self-Improving Pipelines.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations DSPy: Compiling Declarative Language Model Calls into Self-Improving Pipelines

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.804545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.804545Z digest=sha256:7f3fc6f8d86998ecec613658124e7f013099f2689eb2a5795087ad27901e6b03

Observation e73b708a-d00e-4714-983a-66b3884d8548 · outbound

This paper cites VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.808938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.808938Z digest=sha256:d6a8b3595a7e0e397e933d0cb0ddb823f3d0ed9d76c15dbfeb0a6f43038c0f42

Observation 8ef48656-a7b5-48ea-920e-8bcd54894348 · outbound

This paper cites Learning driving styles for autonomous vehicles from demonstration.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Learning driving styles for autonomous vehicles from demonstration

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.474745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.813101Z digest=sha256:5ec871c5e78104586011b213e113cbee8f67559c5256c3c220a00059a9d3a4b3

Observation 02546779-c88a-400d-b572-98acc7483112 · outbound

This paper cites AutoWebGLM: A Large Language Model-based Web Navigating Agent.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations AutoWebGLM: A Large Language Model-based Web Navigating Agent

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.817076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.817076Z digest=sha256:f90f152b4292a64092a05fc17fd0b30f002b0b6b927d50aee0bec9a91bfe6c3b

Observation 462787f6-5cee-427a-b63a-6863a5e82316 · outbound

This paper cites Reinforcement Learning on Web Interfaces Using Workflow-Guided Exploration.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Reinforcement Learning on Web Interfaces Using Workflow-Guided Exploration

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.821155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.821155Z digest=sha256:30476c92a90afb1fcc6d73ce9181ef07618c9997ddecc0680b1bbb657c19dc3c

Observation 129c7208-ae37-4fd9-927c-0ba4be703854 · outbound

This paper cites What makes good data for alignment? a comprehensive study of automatic data selection in instruction tuning, 2024.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations What makes good data for alignment? a comprehensive study of automatic data selection in instruction tuning, 2024

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.825408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.825408Z digest=sha256:ec330fd7ffd8c9be3f52145babae8c3922db553f4efd3477316200e179aed7dc

Observation b49a2526-84ce-4e61-8a3d-3a6f89f52f03 · outbound

This paper cites Algorithms for inverse reinforcement learning.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Algorithms for inverse reinforcement learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.829552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.829552Z digest=sha256:b5c292400ea76c56041f1f4fc690d1cfef8bc9e981bba8dbac4beb1d31a075a7

Observation e0eb73c7-227f-40ba-9e0b-9302b2425bda · outbound

This paper cites On First-Order Meta-Learning Algorithms.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations On First-Order Meta-Learning Algorithms

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.833325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.833325Z digest=sha256:0f9c7334ec6336d48f077ad10b06dd698e12c36a5c64cc85b53397880af4c3fa

Observation eb4f907c-2645-4cfd-a803-26757d210177 · outbound

This paper cites Experimental evidence on the productivity effects of generative artificial intelligence.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Experimental evidence on the productivity effects of generative artificial intelligence

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.837650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.837650Z digest=sha256:89aebfa99a4d5c3fa70c680d9824c78283e2b508a79e42fa57ee0fb988fa3331

Observation 606cc77e-edc6-47c3-b1f1-83201b1e33bd · outbound

This paper cites Oracle AI agents help organizations achieve new levels of productivity, Sep 2024.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Oracle AI agents help organizations achieve new levels of productivity, Sep 2024

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.442381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.841443Z digest=sha256:94821e1018ddadf9594e56fa94248e33f3896fafe27aa7a453e65c459a59b828

Observation c013a0ec-e490-4b2c-ae90-00bfe91b6494 · outbound

This paper cites Synatra: Turning Indirect Knowledge into Direct Demonstrations for Digital Agents at Scale.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Synatra: Turning Indirect Knowledge into Direct Demonstrations for Digital Agents at Scale

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.845089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.845089Z digest=sha256:1219414f47c688828aec807b41b7e8db9652f58f4a6fa62df64196628b1e251b

Observation 2f1c75ab-32f5-4c34-a7d7-e2a5b06f985e · outbound

This paper cites Training language models to follow instructions with human feedback.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Training language models to follow instructions with human feedback

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.848946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.848946Z digest=sha256:4e62d1ef7ad2cd93720b9edf703c24bf33153ae4a4c304d39e4f0904bdab37df

Observation 59bdfd54-f6b7-44b1-b1fb-f21a63d80c70 · outbound

This paper cites Alvinn: An autonomous land vehicle in a neural network.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Alvinn: An autonomous land vehicle in a neural network

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.852620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.852620Z digest=sha256:6263c52b04b2baab55790f640a148f617f25cb0197b0d5d6c852a10f0d7aebe6

Observation 516cfbd9-f719-44ae-8474-c7d5752f0b1b · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Direct preference optimization: Your language model is secretly a reward model

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.856131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.856131Z digest=sha256:da95be4c8954701103c256b1e8a6d645e8ef9a7958e447f1a0bc79692f3e1ea6

Observation 69c55096-7090-46a1-990f-06241076ec99 · outbound

This paper cites Recent advances in robot learning from demonstration.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Recent advances in robot learning from demonstration

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.859906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.859906Z digest=sha256:490c611be95633b62da1f4c59930d9400568448afa0321ab4a038b00edd24000

Observation 76831047-b8b2-4af4-95d8-c6c1fd10be38 · outbound

This paper cites Your Transformer is Secretly Linear.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Your Transformer is Secretly Linear

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.863781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.863781Z digest=sha256:3a4d01b5170d936789ad07cfbbf2771ea2b7605cd298c4c2b63205c5d2b293aa

Observation a8fcb006-1f57-476a-89eb-2dfa2cdecad6 · outbound

This paper cites Generalization guarantees for imitation learning.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Generalization guarantees for imitation learning

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.399030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.868558Z digest=sha256:6ab81d0d2fe2b00353137f3a090789103e83637d5a3a0a07a98bb79990f5992f

Observation 51c975c7-bf67-4a18-a47f-c11d4dc01d68 · outbound

This paper cites A reduction of imitation learning and structured prediction to no-regret online learning.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations A reduction of imitation learning and structured prediction to no-regret online learning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.872174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.872174Z digest=sha256:267249b4f6429d58e3ef50ab88af949e0119779464bb6c1950f5eb4ae4af4a95

Observation 6b9138ac-1179-4983-bf04-4413868487db · outbound

This paper cites Interactive robot task training through dialog and demonstration.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Interactive robot task training through dialog and demonstration

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.379477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.876142Z digest=sha256:c54a18e893a4acfbd8181255f3a4bc5d5b4c4c7a7e5837d6d5d1a705a9f2b8ed

Observation 7962338e-347c-418f-8d8d-2c0fd27f89f4 · outbound

This paper cites Learning from demonstration.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Learning from demonstration

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.879840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.879840Z digest=sha256:681fe1f4dccc2d647309c5857207bcf5a24082db1b47267654d87c625f28369c

Observation 0384e9f0-4689-45f2-ba5b-8a1de1c4a10b · outbound

This paper cites Dynamic movement primitives-a framework for motor control in humans and humanoid robotics.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Dynamic movement primitives-a framework for motor control in humans and humanoid robotics

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.883387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.883387Z digest=sha256:84a12ee99e538f42f108319ef9fc96d77ea573d952c6c37cef23d0ed2928e996

Observation 9f54fd47-c999-4625-be80-0a53131672fd · outbound

This paper cites Evolutionary principles in self-referential learning.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Evolutionary principles in self-referential learning

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.353433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.887214Z digest=sha256:e7b9f0e3a91a9adc9599f13b06ff5e64651f9658860a0a41423b1781d60b2e14

Observation 0a4fd0ef-9e60-46d5-a808-f148bfbb1e3f · outbound

This paper cites Quantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formatting.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Quantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formatting

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.890701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.890701Z digest=sha256:90ecc3114275f5e26f8b6e5282399e7a02bdc1c5039babf83524bf8bc104ba36

Observation 9d7e901c-18fe-46b1-afe1-adae2505ec81 · outbound

This paper cites Aligning Language Models with Demonstrated Feedback.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Aligning Language Models with Demonstrated Feedback

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.894820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.894820Z digest=sha256:daf23a95fe96d8d67c51577fa199a5f8db03cbf06d25526248552b17a4e48ab6

Observation 0d016a49-1904-4c28-823d-18d7f638bacc · outbound

This paper cites From pixels to ui actions: Learning to follow instructions via graphical user interfaces, 2023.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations From pixels to ui actions: Learning to follow instructions via graphical user interfaces, 2023

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.342472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.899039Z digest=sha256:173caacc35f66a053f2d92d2ac15d7c525a7ba19b31bee53bf9fc5fb5ac0d9ed

Observation c57e44e5-0813-47b1-9641-28cdc8a39646 · outbound

This paper cites World of bits: An open-domain platform for web-based agents.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations World of bits: An open-domain platform for web-based agents

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.902959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.902959Z digest=sha256:5dc79bf908129297355bd255789915d64b980a6886fbeabcdbda0f9706c2ca4e

Observation 2ef92842-0c88-471d-a13d-6d7d8aa633b8 · outbound

This paper cites Mas- tering the game of go with deep neural networks and tree search.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Mas- tering the game of go with deep neural networks and tree search

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.906781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.906781Z digest=sha256:3479d75abaafa3245d7aa83fbe135bd12cd5366f6809ed98f98f81cc6c84203e

Observation 6c8b2aa1-93ed-451e-b316-991b8af3b423 · outbound

This paper cites Beyond Browsing: API-Based Web Agents.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Beyond Browsing: API-Based Web Agents

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.911411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.911411Z digest=sha256:81696535865eaed7a5fba8b3adb43b69eeae9994f4ee96a584049d68da4d17fe

Observation 2a439e34-46d5-400d-9a76-60ae2d227b13 · outbound

This paper cites Perception, cognition, and action in teams of robots.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Perception, cognition, and action in teams of robots

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.316680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.916436Z digest=sha256:0740b2a187b41dd95edf477171abc4ee0f568f1298f432b345f19334478e9043

Observation 1767b3c0-e5e6-4c2c-bfd6-305adca70dc8 · outbound

This paper cites CogVLM: Visual Expert for Pretrained Language Models.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations CogVLM: Visual Expert for Pretrained Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.920413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.920413Z digest=sha256:4dd929ef2401302c2b02cbed96481a2437bd9996972a0bd034568e1ebaf6e1e5

Observation eb410e8c-32bc-40f7-9df4-0412f7c5431a · outbound

This paper cites Transformers: State- of-the-art natural language processing.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Transformers: State- of-the-art natural language processing

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.304753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.924639Z digest=sha256:b7c73d8e80608407cbea61772629c62e72972ff08c9d646f597faade10d78651

Observation 1b2c5ae1-7234-4fb4-82b8-75dad0398fd1 · outbound

This paper cites Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.928829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.928829Z digest=sha256:fa389d4048051b2f3af58967b66b97b5f32324fd1a54a46b82f968bece9e2945

Observation 772a2a84-c345-4631-9b6e-71ce483761d2 · outbound

This paper cites ReAct: Synergizing Reasoning and Acting in Language Models.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations ReAct: Synergizing Reasoning and Acting in Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.933320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.933320Z digest=sha256:e6226d4810570257556d12430408a3a3e12218d6e79d345ba2c7a235b48fe066

Observation 8eb4bab8-bd01-4b1b-88f3-8fb586dbd0b3 · outbound

This paper cites Selfd: self-learning large-scale driving policies from the web.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Selfd: self-learning large-scale driving policies from the web

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:27:48.292158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.937628Z digest=sha256:611c2be3b74e7aaee5b69c7b45c466b3a86e68b46dcaa374ab4b571b84d376ea

Observation 677bf8aa-74db-4028-a681-dad0d6b8755c · outbound

This paper cites Group Preference Optimization: Few-Shot Alignment of Large Language Models.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Group Preference Optimization: Few-Shot Alignment of Large Language Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.941734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.941734Z digest=sha256:793cfe8122bc3bce7937915f0fd7642d2ea072570b96331aef0a5ff926a06f3a

Observation f5ddd3cf-0ac9-4a67-8ccc-b16bb853f23c · outbound

This paper cites Gpt-4v(ision) is a generalist web agent, if grounded.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations Gpt-4v(ision) is a generalist web agent, if grounded

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-12T16:27:47.946320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:27:47.946320Z digest=sha256:9d1c7a5e8eebe261c14644ddc4107d0d4083dbd40e0f99683be674e9b2adf339

Observation 33e426d8-13ea-4add-9539-d3711507e979 · outbound

This paper cites ideal answers.

AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations ideal answers

Reference 63

Resolution
malformed identifier
raw_fallback, observed 2026-08-12T16:27:48.270464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T16:27:47.950451Z digest=sha256:2f7cb23ac7c548b35554a0ebedb29832b651aed49f1eab9ec094f6191a9a9f7f

Pith citing papers

Observation 3c7bcfbb-7281-4ecf-adcc-28c37d490299 · inbound

Large Language Model-Brained GUI Agents: A Survey cites this paper.

Large Language Model-Brained GUI Agents: A Survey AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations

Reference 294

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:08:27.935795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-19T11:08:27.472508Z digest=sha256:07bb9414941e76e6c2b0054e9bc06e93f5218db27accb343b38153b161852a66

Observation 0ab2517b-68f6-48ab-8557-ccb0ef8e018c · inbound

How Far Are We from Generating Missing Modalities with Foundation Models? cites this paper.

How Far Are We from Generating Missing Modalities with Foundation Models? AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-25T08:15:33.568089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-25T08:15:12.947854Z digest=sha256:49fce673982a025e64b4b6c289ed880d59ec31f728aebb338513b8bbd615d48e

Observation 4cee231e-b9a5-4934-8753-36ce2714661d · inbound

A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence cites this paper.

A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations

Reference 141

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T22:23:14.807843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-14T22:23:14.621091Z digest=sha256:e4ebb671cb274e14185ad1d0ca720bf21b3c19983695a6671acc6ae221495a01

Observation d2b029cd-acf1-4010-b500-ee287d712c1d · inbound

Meta-Tool: Efficient Few-Shot Tool Adaptation for Small Language Models cites this paper.

Meta-Tool: Efficient Few-Shot Tool Adaptation for Small Language Models AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:54:48.826369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-10T00:50:27.257888Z digest=sha256:c7fc2600695c06d7730bdd6fa8e71e1155b6d3cfbde47ad8826af5c7df3116f7

Observation 98a0c1ce-927c-4561-819e-a0cfa7292d0f · inbound

Learning to Adapt: Self-Improving Web Agent via Cognitive-Aware Exploration cites this paper.

Learning to Adapt: Self-Improving Web Agent via Cognitive-Aware Exploration AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:36:08.847746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T22:18:15.244033Z digest=sha256:9b596a4d1c20a02d634c7ea6cc65fdda7fb3f161e2eebfa57a515685f402ede4

Observation bba51d67-c2ef-4b44-a4a9-0d2a042fd53e · inbound

CoAdapt-GUI: Joint Workflow Context and Policy Adaptation for Unseen GUI Applications cites this paper.

CoAdapt-GUI: Joint Workflow Context and Policy Adaptation for Unseen GUI Applications AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T00:39:33.284170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:39:33.284170Z digest=sha256:994cee193464b4c85687fd4d8635fa7f026c189c38dbd807d29e43ab52edc46a