Pith. sign in

Paper Citation Record · LEDGER

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization

As of 18 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 14 inbound Pith citation observations for arXiv:2502.04306.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.04306 v1

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T22:57:12.869989Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:33:02.869499Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T00:12:50.066065Z

Reference resolution

56 of 56 outbound references displayed

  • verified exact0
  • verified fuzzy16
  • unresolved39
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d3730524-bea0-40f6-9a2c-abdb6403fa3f · outbound

This paper cites GPT-4 Technical Report.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.618521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.618521Z digest=sha256:b94c2e67e9758cc6fcca30ccaecf89605ca84ef6a3c4f44594eb2c13b4cd984f

Observation 76f5f5bf-e1a0-46e3-a21e-6ac3c9782fb8 · outbound

This paper cites PaLM 2 Technical Report.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization PaLM 2 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.623748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.623748Z digest=sha256:4c8d9832c57c78bede00824ae388e480c9a9369b7a31d6b64d5906144e758581

Observation 4037b6a2-85c3-4b69-b849-17dc8a9ab4b2 · outbound

This paper cites Program Synthesis with Large Language Models.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Program Synthesis with Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.628595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.628595Z digest=sha256:51ab1584912fd64a9f8c22ac86f42fb6a4452031a7f951151f92518dae0bbcc7

Observation abad24c7-c5c3-4cd5-a387-e10f708402d2 · outbound

This paper cites A general theoretical paradigm to understand learning from human preferences.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization A general theoretical paradigm to understand learning from human preferences

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:57:13.869742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.633813Z digest=sha256:1d725c69502b7b099a57932039a47329df7cae39e1ba622954dd1b3bedce4760

Observation 6715730b-9f74-4afe-beba-3889b44ef644 · outbound

This paper cites Rank analysis of incomplete block designs: I.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Rank analysis of incomplete block designs: I

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.638465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.638465Z digest=sha256:554d0764c3ea7f8744edc198359c6f23adb94c98596ec3879a67bb33f9f1f830

Observation 783e1512-f46d-4dda-8cd1-732d3404c361 · outbound

This paper cites Preference Learning Algorithms Do Not Learn Preference Rankings.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Preference Learning Algorithms Do Not Learn Preference Rankings

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.643098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.643098Z digest=sha256:d6581490652a9cda11bccf92b38c5b851368b81f553e9c7d6ec99c592e0be165

Observation db93fe2f-df8b-4c0a-8d2f-fef6567543bb · outbound

This paper cites AutoAgents: A Framework for Automatic Agent Generation.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization AutoAgents: A Framework for Automatic Agent Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.648358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.648358Z digest=sha256:cd18ecca58fc13df3477e1a7455caf2a7f7812ad79fac71cd10df134db228998

Observation bbe9f0a0-813a-492a-8ed9-a5e07f1aeae5 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Evaluating Large Language Models Trained on Code

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.653008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.653008Z digest=sha256:95313d5af3ebbeb332c0b47b8c9e96da5aecd4346b98274d3041b64aa9bf5546

Observation 39132a60-e940-4c92-b3b1-5acf31439456 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Training Verifiers to Solve Math Word Problems

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.657855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.657855Z digest=sha256:f6f9409c97c0b8a4d6176459a7b9fcfccfeb82e59b220920d813e72a1b0ceac6

Observation 9ed8e6d1-280b-488f-a08c-4529977cc6b1 · outbound

This paper cites DROP: A reading comprehension benchmark requiring discrete reasoning over paragraphs.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization DROP: A reading comprehension benchmark requiring discrete reasoning over paragraphs

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.662885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.662885Z digest=sha256:217782c9ff329c3d589602eef2757d3689a605570180af3a1b089ea3bf7667da

Observation 22a688a3-97ee-4aec-b0a7-5bcd43a35937 · outbound

This paper cites Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.667166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.667166Z digest=sha256:aaa22ef94c9f71b9756c682a8f412b5c6c030007b7e926903bf67fd5f870394f

Observation 523419a9-71b1-49a0-bd12-65b3a6e44f69 · outbound

This paper cites Data Interpreter: An LLM Agent For Data Science.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Data Interpreter: An LLM Agent For Data Science

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.672145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.672145Z digest=sha256:e3002d05da421fe5e0f0e9df6a6580c0a2965111a96979fe110c90e56d3aea5e

Observation b17995c6-1247-46e8-8fd5-82887e161ca6 · outbound

This paper cites Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, and Weizhu Chen.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, and Weizhu Chen

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:57:13.836771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.676662Z digest=sha256:bb19a94057f106672ac29f7680deada864d9f50acf4f8ed2f35c86a56bf1a786

Observation 66f9f549-15d9-4508-9cce-51f1e93e5fb4 · outbound

This paper cites Automated design of agentic systems.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Automated design of agentic systems

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:57:13.822483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.680923Z digest=sha256:1b209212062dc95dafd8a04222018541902274dab26b2e481f8b8822722be170

Observation f9036684-3777-4203-a531-692cecb8b42b · outbound

This paper cites The N+ Implementation Details of RLHF with PPO: A Case Study on TL;DR Summarization.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization The N+ Implementation Details of RLHF with PPO: A Case Study on TL;DR Summarization

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.685241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.685241Z digest=sha256:91911ebc060b409facc9a00d0de7f59dc8fe4fd6090a878dbd6d11172b29b535

Observation 34217c21-d590-4e49-9398-d3b734f99674 · outbound

This paper cites SELF-[IN]CORRECT: LLMs Struggle with Discriminating Self-Generated Responses.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization SELF-[IN]CORRECT: LLMs Struggle with Discriminating Self-Generated Responses

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.690459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.690459Z digest=sha256:ee1dc70424d4ffd352944ec88c681b1e8bb7681c71a60e3d3456605e41b9de8f

Observation 0eacf067-512e-41a4-946a-f8b685c3d21f · outbound

This paper cites Dspy: Compil- ing declarative language model calls into state-of-the-art pipelines.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Dspy: Compil- ing declarative language model calls into state-of-the-art pipelines

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:57:13.806667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.694921Z digest=sha256:74cda0ec05d2af940749f4633e13ffb331c130a377b2f319729aef8d64392d36

Observation a57525fb-7c3a-44ad-8b1d-0b1ebe9974ad · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Gonzalez, Hao Zhang, and Ion Stoica

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.699289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.699289Z digest=sha256:c72adcb227642fdd8e5530e20ec5bdb43086b6cee1cd2d3eba3ec509d11efeda

Observation 5aec9b39-8049-4888-bd3c-0ba7e85892e1 · outbound

This paper cites AutoFlow: Automated Workflow Generation for Large Language Model Agents.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization AutoFlow: Automated Workflow Generation for Large Language Model Agents

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.703649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.703649Z digest=sha256:2edc05cc955a2383eee63a557ef40631ebb383e48297c1b76e10d585f64ee1fc

Observation 9182281b-7a15-4fd4-9b4f-7ffc17c828cd · outbound

This paper cites DeepSeek-V3 Technical Report.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization DeepSeek-V3 Technical Report

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.708501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.708501Z digest=sha256:d168db6a084cacefc506594422cacfab788c302494083320c1de4eeebf1c4285

Observation b6dcd6c3-db4d-424a-997a-ca66c168ca8c · outbound

This paper cites A dynamic llm-powered agent network for task-oriented agent collaboration.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization A dynamic llm-powered agent network for task-oriented agent collaboration

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.713084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.713084Z digest=sha256:1fd3b2e91ce1d8f1c2488fad4f4e32221c26aa22bd31aadd1391bd8a11b9de82

Observation 68acb43a-21cb-41fb-8287-a80311037be4 · outbound

This paper cites Self-refine: Iterative refinement with self-feedback.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Self-refine: Iterative refinement with self-feedback

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:57:13.772353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.717444Z digest=sha256:902f67265ca33c1b66d1ff8ffd6bfe71064f9b12c000db3759f4720212c32859

Observation afa58b2c-fcc2-4aad-b839-c1d5bb82114a · outbound

This paper cites SimPO: Simple Preference Optimization with a Reference-Free Reward.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization SimPO: Simple Preference Optimization with a Reference-Free Reward

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.721732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.721732Z digest=sha256:553d2ed042e28d3c366b907ee942a63d02781b38926cc1073683a9192897a4ed

Observation 3c524b6e-e454-4f60-8800-0d5e6d45c54e · outbound

This paper cites Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.726195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.726195Z digest=sha256:478cc3a856d43c96ba1d4550bb958857db174932ae7354e4ba701131af912c2c

Observation 995c473c-f3e3-40b5-b607-113fb98b3c4b · outbound

This paper cites Training language models to follow instructions with human feedback.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Training language models to follow instructions with human feedback

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:57:13.757784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.730665Z digest=sha256:cc16401225396ac7d6f3855edd31103c4ecd3d821226f58d3de88fb5f2000fcc

Observation e09a88e1-aa5e-4373-adf5-19920e30f149 · outbound

This paper cites Disentangling length from quality in direct preference optimization.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Disentangling length from quality in direct preference optimization

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:57:13.742715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.734619Z digest=sha256:1adeae3c73ae138da7a19434eb396aa09ee9ac39775be4df5254a1771386bfd8

Observation 96c6840c-58d5-4b4e-9a3d-c39a35a5bcaa · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Direct preference optimization: Your language model is secretly a reward model

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.738726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.738726Z digest=sha256:107e4dc1b51c1fa041741c1a9a64a5cca93a53a1ab09be17f1df09cc98395e06

Observation 730793c5-2d09-4834-bd38-8a53edb308ad · outbound

This paper cites Code Generation with AlphaCodium: From Prompt Engineering to Flow Engineering.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Code Generation with AlphaCodium: From Prompt Engineering to Flow Engineering

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.743463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.743463Z digest=sha256:a3433709b17ccb5f23ee3d43cfff01e5da6931543a262b5a6f22759c3e3f37a5

Observation c7637e98-1356-47e2-bc61-66235a1b08d6 · outbound

This paper cites Archon: An Architecture Search Framework for Inference-Time Techniques.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Archon: An Architecture Search Framework for Inference-Time Techniques

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.748007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.748007Z digest=sha256:421634cb0315b70eab7368738c3e09ec9ed3b06fc3f6a94039bf65ce7489b12c

Observation 590e455e-1797-4d30-b5a0-3605336b837d · outbound

This paper cites Proximal Policy Optimization Algorithms.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Proximal Policy Optimization Algorithms

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.752579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.752579Z digest=sha256:6a2e0e2ffdcf278dc9492a3791c3884bcc541fd59c188307a78d79f7472ddfee

Observation 1552c035-da6d-403d-9411-129bc9a05d11 · outbound

This paper cites Reflexion: Language agents with verbal reinforcement learning.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Reflexion: Language agents with verbal reinforcement learning

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:57:13.716896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.757173Z digest=sha256:1f9368d97bf101cac8fafaf170663f81ff9753d83f2423000e32f3a75fc1f0da

Observation cf0d7583-db8f-4d2a-a930-ba87247585b7 · outbound

This paper cites Adaptive In-conversation Team Building for Language Model Agents.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Adaptive In-conversation Team Building for Language Model Agents

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.761310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.761310Z digest=sha256:87b196ddade2d093ef5dd7f626f76b0fcbded0b5a2c6cebd4bbbf1557beec1be

Observation 957d2528-5de7-43b6-9827-a7763559cb4f · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.765587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.765587Z digest=sha256:c083fe09df8648acc4615e104ffc69903e488a0823e4d974621b5289665fcd34

Observation cfa4f3c3-9a2d-4b3b-a7b0-b030e6a684a6 · outbound

This paper cites Self-consistency improves chain of thought reasoning in language models.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Self-consistency improves chain of thought reasoning in language models

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:57:13.701555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.770342Z digest=sha256:67a1f5dba75f2282c03c703a5e486c43d481ea332d270e7bfd86ca91f4a69656

Observation 9a66ec3b-ca81-4111-8059-3af9c63ef40c · outbound

This paper cites Unleashing the emergent cognitive synergy in large language models: A task-solving agent through multi- persona self-collaboration.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Unleashing the emergent cognitive synergy in large language models: A task-solving agent through multi- persona self-collaboration

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:57:13.685561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.774619Z digest=sha256:69c208c8f0e883ea080f28202b2043627c114ba8b78ca3716be260d87fb8a7e9

Observation 0d8999c4-e272-4ada-9252-1e886f0e2e92 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Chain-of-thought prompting elicits reasoning in large language models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.779256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.779256Z digest=sha256:92eb58f47d917c5a2b90a89475d6ed3b5e003849f31439c3dfa9e4a0b699c439

Observation 75839ae2-a980-484e-ad82-47651ca6bd13 · outbound

This paper cites Is dpo superior to ppo for llm alignment? a comprehensive study.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Is dpo superior to ppo for llm alignment? a comprehensive study

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:57:13.661527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.783467Z digest=sha256:09ebdb9ffcea035a0506c8661901ee7fc1a841d4defe473f2b1acd4767109d5d

Observation 29de14fb-a760-4134-8026-859dbf0b7960 · outbound

This paper cites Lemur: Harmonizing natural language and code for language agents.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Lemur: Harmonizing natural language and code for language agents

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:57:13.646406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.787784Z digest=sha256:fba24a9707d0e6dc92184e0ccb95f584283ae0bf2d4f90fb0916495f5c58882f

Observation af447a64-30b5-454f-a568-7a1014933a12 · outbound

This paper cites Qwen2.5 Technical Report.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Qwen2.5 Technical Report

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.792051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.792051Z digest=sha256:a545fa728dd2542387bff1fe7e903496185dfb33ec6958a8521dfa7be34e7e11

Observation 872f6fb6-0f25-49dd-8a5c-51a2cfa4610e · outbound

This paper cites Large Language Models as Optimizers.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Large Language Models as Optimizers

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.796606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.796606Z digest=sha256:c6c90acea3392529648ecd06ce9fde55eccaaa14c81d1c53bd02fc569438f3b0

Observation 34433255-8367-41b9-b74a-996117b5ff42 · outbound

This paper cites Buffer of thoughts: Thought-augmented reasoning with large language models.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Buffer of thoughts: Thought-augmented reasoning with large language models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:57:13.631900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.801065Z digest=sha256:28e333236202ae3b6e5d0aee98932401b244aab0767568fc877033354478d49a

Observation a8523c9a-d6f9-495e-a4c8-9ba3a264929d · outbound

This paper cites SuperCorrect: Advancing Small LLM Reasoning with Thought Template Distillation and Self-Correction.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization SuperCorrect: Advancing Small LLM Reasoning with Thought Template Distillation and Self-Correction

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.805488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.805488Z digest=sha256:99d2e3f8604ee6f17542569e25e045c0ffb0149c7ff16512435fc3a4b6d11e62

Observation eebd04e1-989f-4d39-bc8d-b7abc32cb453 · outbound

This paper cites Hotpotqa: A dataset for diverse, explainable multi-hop question answering.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Hotpotqa: A dataset for diverse, explainable multi-hop question answering

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:57:13.616769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.809888Z digest=sha256:cf225c9aa8ff2e32fa027508977310cbdf9844b597b19faa892ababd3fc64fcc

Observation 056913e2-58e7-4a41-b99d-e8a92a24c65e · outbound

This paper cites TextGrad: Automatic "Differentiation" via Text.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization TextGrad: Automatic "Differentiation" via Text

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.814593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.814593Z digest=sha256:cf58da7fa892f12608b220ed04e777737a89e062a18869dd93c419c90de53f3e

Observation 28aae596-721c-49e4-9ebf-e193d166fa80 · outbound

This paper cites G-Designer: Architecting Multi-agent Communication Topologies via Graph Neural Networks.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization G-Designer: Architecting Multi-agent Communication Topologies via Graph Neural Networks

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.819438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.819438Z digest=sha256:57bb831a5686b3f4f0a2960d5a37a6f96f5dd4481a31f1f9a853bca588127c89

Observation c84e06ba-b204-4691-869a-70b5adf62783 · outbound

This paper cites AFlow: Automating Agentic Workflow Generation.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization AFlow: Automating Agentic Workflow Generation

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.824216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.824216Z digest=sha256:e21e303bd66642f3acd144c5def91df77a809c21bf2c9a566884b596e621c6e9

Observation 2aa3e9ba-73b0-476d-9f36-219b6f159f2b · outbound

This paper cites Achieving >97% on GSM8K: Deeply Understanding the Problems Makes LLMs Better Solvers for Math Word Problems.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Achieving >97% on GSM8K: Deeply Understanding the Problems Makes LLMs Better Solvers for Math Word Problems

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.828725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.828725Z digest=sha256:36cc2f877e89e7aa6cb3c4dce5036e5fe34a5a2e28b1dd244501da8263993bc3

Observation 7ae5b654-d6f7-4043-b10e-578f07115c06 · outbound

This paper cites Symbolic Learning Enables Self-Evolving Agents.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Symbolic Learning Enables Self-Evolving Agents

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.833466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.833466Z digest=sha256:ba4fc86905e6dbf2d0ba4f3f13b75634b6486af2a6abad955e2cf1f32d08b816

Observation 4a2e3f73-e73a-4640-99e3-4f25bb359481 · outbound

This paper cites "" This is a wor kf low graph.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization "" This is a wor kf low graph

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-08T22:57:12.838215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:57:12.838215Z digest=sha256:f4f55f1859a368180482ed0933b435b990d4eea5be28a090aa916eb89a198975

Observation 8a7763ed-99b6-40b2-b47c-9a5e5f0c87b0 · outbound

This paper cites Format MUST follow : custom ( i n s t r u c t i o n : str ) -> str You can modify the i n s t r u c t i o n prompt.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Format MUST follow : custom ( i n s t r u c t i o n : str ) -> str You can modify the i n s t r u c t i o n prompt

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:57:13.601950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.842713Z digest=sha256:1cbab2972c95aad4bb79c8e847999def20266d9742e80ae2a2b849cfa08551c7

Observation f7132786-fb17-40e0-877a-7e974988f781 · outbound

This paper cites an unresolved cited work.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:57:13.586250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.847115Z digest=sha256:acaa37a612f08bab10b6f4a4c18b7771c890ef4c06e685e6e0e9aeb6d1a308ba

Observation 6e5e3a19-ae31-4779-a77e-fa615ea6dd22 · outbound

This paper cites Format MUST follow : a n s w e r _ g e n e r a t e () -> str For example : so lu tio n = await self.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Format MUST follow : a n s w e r _ g e n e r a t e () -> str For example : so lu tio n = await self

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T22:57:13.571072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.852322Z digest=sha256:3bab7c31dd32e47da1ce8f902c6e3e7899f8c456c45886cc0a3287fc031b51b6

Observation 7e8f3ce8-6a70-498b-9671-0c4c8b1cde36 · outbound

This paper cites an unresolved cited work.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:57:13.555653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.856799Z digest=sha256:452c2d53a62a8d2124754e0c6c748fb1f44349fa45096dada6cd4fa98192aaaf

Observation 37c5180b-f107-4b6e-b2c7-fd36eaa2b531 · outbound

This paper cites an unresolved cited work.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:57:13.540101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.861442Z digest=sha256:5cf4411c86d004dd300abee15f7a5e4b23980590a7698e81afffac68939761e6

Observation b0b2d1a0-ab32-4f7e-a8de-7e1c27894b5f · outbound

This paper cites an unresolved cited work.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-08T22:57:13.524891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.865732Z digest=sha256:1bbfbac4294ea36435926786ac0173aaa89455afa87e8324e74dc48d74f3e509

Observation 8576080f-d72b-4221-b476-829f078754ea · outbound

This paper cites "" This is a wor kf low graph.

ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization "" This is a wor kf low graph

Reference 56

Resolution
malformed identifier
raw_fallback, observed 2026-08-08T22:57:13.509722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T22:57:12.869989Z digest=sha256:8db1a88b0dfc4f963856a0ed5b81baff8e211e82e053929fb0071040f05fdd8e

Pith citing papers

Observation 0a41fbf5-200c-4f0b-a399-205d9a46a1b9 · inbound

FlowReasoner: Reinforcing Query-Level Meta-Agents cites this paper.

FlowReasoner: Reinforcing Query-Level Meta-Agents ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T11:33:02.869499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:33:02.869499Z digest=sha256:2491eefbdcf868791c815a637c16fdc7d13d88cdaf789111f92eece99b4eb9bd

Observation c72814f9-1626-49eb-81e5-f5da80437d1b · inbound

HALO: Hierarchical Autonomous Logic-Oriented Orchestration for Multi-Agent LLM Systems cites this paper.

HALO: Hierarchical Autonomous Logic-Oriented Orchestration for Multi-Agent LLM Systems ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T20:51:13.601509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:51:13.601509Z digest=sha256:50f34e87c5981a9ed3d1d4fa32a94f1f0274da8612d16853fc6d81624047d2b9

Observation 8e15cffa-39ec-400a-9577-83390127f903 · inbound

MermaidFlow: Redefining Agentic Workflow Generation via Safety-Constrained Evolutionary Programming cites this paper.

MermaidFlow: Redefining Agentic Workflow Generation via Safety-Constrained Evolutionary Programming ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T13:00:08.840573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:00:08.840573Z digest=sha256:c9a36044acedcb6b052280f561f0983c9f8c71f231850e501633c9b33b0961e8

Observation a721dc97-9037-4479-b644-6c9584fbabac · inbound

GenoMAS: A Multi-Agent Framework for Scientific Discovery via Code-Driven Gene Expression Analysis cites this paper.

GenoMAS: A Multi-Agent Framework for Scientific Discovery via Code-Driven Gene Expression Analysis ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization

Reference 126

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:40:51.269575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-22T00:37:11.945418Z digest=sha256:35138e9a63c1b9c0950953a6bf9be7cbbc8de1e5bb491b0abfb4ac5e66396e99

Observation f3c0730e-c85b-4733-8c99-a39eff0612f1 · inbound

Rethinking Query Optimization for Multi-Agent Systems [Vision] cites this paper.

Rethinking Query Optimization for Multi-Agent Systems [Vision] ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T17:21:28.646860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T17:21:28.646860Z digest=sha256:0bdb3a4b53783eb9f12403778f8eac8fb0aa9e6c78e3282e44df314226fa974a

Observation ff54a86d-7777-42ff-84a2-9d05f1a66d12 · inbound

Autogenesis: A Self-Evolving Agent Protocol cites this paper.

Autogenesis: A Self-Evolving Agent Protocol ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:29:24.729524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T10:27:34.143483Z digest=sha256:ce5e4b6c79fae459959e1dc045a311590f176d6a9bad88b05436603a4a41496f

Observation b2c07b09-5da3-4bfd-8109-fc522d1443b7 · inbound

Autogenesis: A Self-Evolving Agent Protocol cites this paper.

Autogenesis: A Self-Evolving Agent Protocol ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-21T00:19:16.783581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-21T00:15:18.079458Z digest=sha256:ab17e406941cdf21854b33524f2cce2ed1185d860f3e024f54d26b2ef1f6d721

Observation 879697a1-c5a3-4c0c-ab0c-6d9f2a50c1e1 · inbound

EVOCHAMBER: Test-Time Co-evolution of Multi-Agent System at Individual, Team, and Population Scales cites this paper.

EVOCHAMBER: Test-Time Co-evolution of Multi-Agent System at Individual, Team, and Population Scales ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:47:08.474776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:38:33.870289Z digest=sha256:078f7fbe0ac097eaa70a09e8a6000f6c7fea0a6266e584ec0b755c1e61498999

Observation a8b83235-1d34-42e4-a096-2ae6ff7980ac · inbound

Towards Direct Evaluation of Harness Optimizers via Priority Ranking cites this paper.

Towards Direct Evaluation of Harness Optimizers via Priority Ranking ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:14:40.387427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-22T06:14:28.559147Z digest=sha256:43bf9c4eb8726c803ecf9681bc14483793f009d90bf6ec143ed9bb72cf33f639

Observation e287924a-aa6a-4bed-ba34-0e1a35ca3380 · inbound

Evolve as a Team: Collaborative Self-Evolution for LLM-based Multi-Agent Systems cites this paper.

Evolve as a Team: Collaborative Self-Evolution for LLM-based Multi-Agent Systems ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-06-29T00:12:50.067704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T00:05:31.780655Z digest=sha256:c1f31d06e50038359c4694eaaa38fd94f5eb3b78df91679e4fdf3b6ec240276a

Observation 75c452f2-a934-4964-995b-d6c0c695d327 · inbound

A Workflow-Aware Serving Layer for Agentic Applications cites this paper.

A Workflow-Aware Serving Layer for Agentic Applications ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-12T05:56:18.797766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T05:56:18.797766Z digest=sha256:db1af9c755976887cbcd85d1a477c2ab4373819506fb8086f8551bc06aed0742

Observation af735b10-0f92-44ae-9c45-29658aba67bf · inbound

CONTRA: Red-Teaming Configurations of Personalizable Agents cites this paper.

CONTRA: Red-Teaming Configurations of Personalizable Agents ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-12T04:03:40.067404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T04:03:40.067404Z digest=sha256:8dc20d501d01fcb144765123bf15d12b02e6d410c4feec71c0a6843912f0f2bd

Observation d32dd730-36fd-4726-b31b-54bf42d8bf75 · inbound

Reward-Free Evolving Agents via Pairwise Validator cites this paper.

Reward-Free Evolving Agents via Pairwise Validator ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T02:15:34.145140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T02:15:34.145140Z digest=sha256:8681634678f161e66b4e53b629045694cfeb27465739ccea326195628cb69a4f

Observation 689c9664-2d66-48fa-9e2b-80feda11857e · inbound

FlowEvo: Self-Evolving Agents through the Co-Evolution of Workflows and Executable Skills cites this paper.

FlowEvo: Self-Evolving Agents through the Co-Evolution of Workflows and Executable Skills ScoreFlow: Mastering LLM Agent Workflows via Score-based Preference Optimization

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T16:08:11.758646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T16:08:11.758646Z digest=sha256:238f8198c36a220dd1724917eec7e9e6968bb65bb580daff20017e0160f473a9