Pith. sign in

Paper Citation Record · LEDGER

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding

As of 7 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 1 inbound Pith citation observation for arXiv:2509.04243.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.04243 v1

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:17:59.198656Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-19T14:24:48.938948Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T14:27:24.243476Z

Reference resolution

25 of 25 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a50b85eb-42ff-4fb8-b63d-7de7fd27ae09 · outbound

This paper cites arXiv preprint arXiv:2505.20272.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding arXiv preprint arXiv:2505.20272

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.042313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.042313Z digest=sha256:b88213362e02b138976c3a05dc7fc7ceb2dcf35527ab7070aa0f595ef0184450

Observation 6e027198-d1b9-4fef-88d9-d6adf7da15b3 · outbound

This paper cites Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.048480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.048480Z digest=sha256:c46689f834f95431c31226ffb755eb414d094046ff7200498bbd24150f51711d

Observation ec0add16-468f-477e-835c-69dd9f66da52 · outbound

This paper cites SeeClick: Harnessing GUI Grounding for Advanced Visual GUI Agents.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding SeeClick: Harnessing GUI Grounding for Advanced Visual GUI Agents

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.059404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.059404Z digest=sha256:f4830706a5812f767461a7e54f0d31a5051fbda569511154eee5dac764607b2a

Observation 4bd59021-7d4f-4189-a3ea-4b2b0e1c6928 · outbound

This paper cites Navigating the Digital World as Humans Do: Universal Visual Grounding for GUI Agents.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Navigating the Digital World as Humans Do: Universal Visual Grounding for GUI Agents

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.066630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.066630Z digest=sha256:b1356de6993552a4aaf14989c123a073c6cafd73835ed3bb655b17cbb2d2036b

Observation 72b709d3-f758-4b82-ba34-795e4ea39d38 · outbound

This paper cites Understanding the planning of LLM agents: A survey.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Understanding the planning of LLM agents: A survey

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.073173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.073173Z digest=sha256:e1ce632a9147cf88d9c423304493b2200b62c71e71689d79c019b68e4df142bb

Observation 95307b67-6919-4634-96ff-13a2d2c1422e · outbound

This paper cites GPT-4o System Card.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding GPT-4o System Card

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.084357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.084357Z digest=sha256:3fd144280a084d334f728cd4630f24ac2ec791d0169acfd1ce03a718c5c76a0a

Observation a2ea24a2-94db-4886-96be-fa7c27bb4302 · outbound

This paper cites ScreenSpot-Pro: GUI Grounding for Professional High-Resolution Computer Use.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding ScreenSpot-Pro: GUI Grounding for Professional High-Resolution Computer Use

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.093353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.093353Z digest=sha256:467bc0899790ac20304c19bbced577094c1cfe9991cb38713495ea961138e4ec

Observation 07dac15a-233d-4014-b99e-d37547d24629 · outbound

This paper cites UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.101162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.101162Z digest=sha256:f04db68a98d1d46ff2762f0189faa702a05bbe36f096332f2e411a725c1a9182

Observation 55c9e07d-5d37-4f98-b557-cd7e69f5f738 · outbound

This paper cites GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.107172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.107172Z digest=sha256:e552a3b63bfe77bdf3b92362b8a1444a51b3576c1e01a08a9e6b74cdb1e4667a

Observation 36c9096e-0085-428f-b270-6ec2c5373643 · outbound

This paper cites https: //openai.com/index/o3-o4-mini-system-card/.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding https: //openai.com/index/o3-o4-mini-system-card/

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:18:00.281563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:17:59.113099Z digest=sha256:fd0978aaca6172bca40b20e419e30f4ce352c34293938dd1549fcac85e1a5187

Observation 1deaa277-9d02-46b8-9e0b-489853b93817 · outbound

This paper cites UI-TARS: Pioneering Automated GUI Interaction with Native Agents.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding UI-TARS: Pioneering Automated GUI Interaction with Native Agents

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.119944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.119944Z digest=sha256:b1fd263cd3d69b3b730f1dddbda963ad0bfd4e19b94f6f52569b67cfab3778bd

Observation 5191626a-7adf-41c9-bc82-0d25b332f5d3 · outbound

This paper cites Kimi-VL Technical Report.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Kimi-VL Technical Report

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.126766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.126766Z digest=sha256:2d66065e804ed9532d9367b118a23cd164ab50bacd931286ebb9a4ba71d4a99d

Observation 832c4ab2-5750-4a4a-8deb-a2f2e34a6f25 · outbound

This paper cites OS-ATLAS: A Foundation Action Model for Generalist GUI Agents.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding OS-ATLAS: A Foundation Action Model for Generalist GUI Agents

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.141447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.141447Z digest=sha256:56b708751aa0c9285b4203019a7b392716dac679d371ae444f5f49b81307ed7c

Observation d2bacd15-d514-4503-ad5b-72b764f494c6 · outbound

This paper cites arXiv preprint arXiv:2505.13227.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding arXiv preprint arXiv:2505.13227

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.146788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.146788Z digest=sha256:88662412ad69d8a0035ad3a2d12baa85387190f0366cb3d7099eda0e7f6404aa

Observation 05b268ee-9e9a-4069-85da-ba16ba02202b · outbound

This paper cites Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.153421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.153421Z digest=sha256:eee3ac0ecb3f58564e68c91156d105c7f1ab930d02119837a3e997830301129a

Observation 4ad3a2cb-568a-4a04-93ae-b56b78fbf339 · outbound

This paper cites GTA1: GUI Test-time Scaling Agent.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding GTA1: GUI Test-time Scaling Agent

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.159913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.159913Z digest=sha256:af45a644ff3f1b39903ed90defb9b68b4c9d6987b120163c2f57b10bc24fd2b3

Observation 22d4f0a1-cba8-44cc-b13e-9f5fc856c6c9 · outbound

This paper cites Aria-UI: Visual Grounding for GUI Instructions.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Aria-UI: Visual Grounding for GUI Instructions

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.166418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.166418Z digest=sha256:60adddddd17fa2cce9212b16d4e10df21a3acc2b2955f939566020a97facfc8e

Observation 7abe08bc-6ef1-46e9-ae9c-82cf2434754b · outbound

This paper cites Enhancing Visual Grounding for GUI Agents via Self-Evolutionary Reinforcement Learning.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Enhancing Visual Grounding for GUI Agents via Self-Evolutionary Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.172567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.172567Z digest=sha256:1a473be3b1791a21220309e6ba6ed8483174a1f475bd228a415898162fb09ec6

Observation 19d5177d-1d48-41d2-bc68-bbb2d52e5a67 · outbound

This paper cites Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.179579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.179579Z digest=sha256:af3e00c09649e56b62169208ad3b3bb8c81baa31428d80b5bdc18e5b7837e70b

Observation 1290d5e6-8370-49c8-beff-0aa2fa43589b · outbound

This paper cites DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.185051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.185051Z digest=sha256:7a7c60893ca9df72a80beb2c4d424f731a9e42e7e83b7b35b0f06dae5e7a2368

Observation aaff90cb-1c79-4cff-8850-bcc83d0c235c · outbound

This paper cites WebArena: A Realistic Web Environment for Building Autonomous Agents.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding WebArena: A Realistic Web Environment for Building Autonomous Agents

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.192008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.192008Z digest=sha256:c966fbe0254e15dfef7e4aee4d448963a92ade1845c76477f711f103d675c12d

Observation ca9b32fc-55d6-44e6-a355-b616183806ab · outbound

This paper cites ACTIVE-o3: Empowering MLLMs with Active Perception via Pure Reinforcement Learning.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding ACTIVE-o3: Empowering MLLMs with Active Perception via Pure Reinforcement Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.198656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.198656Z digest=sha256:6c4d49ca23736060d6ab5dd6d649f0ba1076c718e3b8c3884353b6d2b327a518

Observation 25a6e7b9-77e1-4110-a5d3-99b23f3b1fd7 · outbound

This paper cites Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.134588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.134588Z digest=sha256:ee659d00734502588b49eea2aef19637ef63fbb775dfa023bdb9867ea570656f

Observation dc214ed5-a72b-4c17-b12c-ba17a143c73f · outbound

This paper cites https: //www.anthropic.com/news/developing-computer-use.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding https: //www.anthropic.com/news/developing-computer-use

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:18:00.312755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:17:59.028076Z digest=sha256:d4008eddf3345b83e3df24c0b1f43abed4889d0283e83b23e53fd4873e461c37

Observation cdaedacc-94d8-4e55-a78f-de3fa83cbb14 · outbound

This paper cites Qwen2.5-VL Technical Report.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Qwen2.5-VL Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.034735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.034735Z digest=sha256:0fdced2a5f2a722804c1dcda6f3e2e063da1e59342dc0ec25bb79857e35a352f

Pith citing papers

Observation f7e38fa0-b9c5-45cd-99db-71c21005896a · inbound

DRS-GUI: Dynamic Region Search for Training-Free GUI Grounding cites this paper.

DRS-GUI: Dynamic Region Search for Training-Free GUI Grounding Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-19T14:27:24.245090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T14:24:48.938948Z digest=sha256:911a30f40c28908f031a55e5af0a3e3bc4ffae643fab02eaa8190ae3fb73b343