Pith. sign in

Paper Citation Record · LEDGER

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation

As of 7 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 3 inbound Pith citation observations for arXiv:2506.16024.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.16024 v1

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:49:49.180444Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T17:38:50.313305Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T20:46:14.342643Z

Reference resolution

65 of 65 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved63
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b9f414b3-ef0a-4c8a-b75f-41f8bef65d7d · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:54.408301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:43.699129Z digest=sha256:3ad76e64b0ce99c83ea56e7d43690cbae771ce44c9295e4800240e650eea4dd3

Observation 8cef63ee-899a-4c2a-a768-7c7e7237abcf · outbound

This paper cites LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:43.738491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:43.738491Z digest=sha256:2db026228feea873e8b0269661a2ac0891d75ae1790e92f43c1eb328b45ff428

Observation 2e556eaf-749a-4e32-bb5e-5633c28f2cb8 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:54.237578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:43.814403Z digest=sha256:3e455fed4f5b4ca8494d818f5ca78c49b58b84ee30e3f0351c1518e3f0ea2eae

Observation 2c265f62-e1ab-424a-ae77-ee52b5c3ca48 · outbound

This paper cites LongWriter: Unleashing 10,000+ Word Generation from Long Context LLMs.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation LongWriter: Unleashing 10,000+ Word Generation from Long Context LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:43.860165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:43.860165Z digest=sha256:7ccf1d1fb7db93f537c444fdd3bd76b3f52ea6df99f8fd7d379fbcb2aec28238

Observation 236c8de7-949f-42ca-afef-dbfbce19af36 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:54.052816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:43.907794Z digest=sha256:9d39e7f9ff8a75a2d3d42d32b89fd7c9c48b0c7d046e34c0e8a68939ce45df3b

Observation 4d36f76e-7a00-4171-a2b9-2da1305902de · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.022270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.022270Z digest=sha256:05f65af578aefd9aa5bd6820d8f0bbff99828bb3354e0fc600e8763db23e43b9

Observation 5f8acf18-99c1-4c75-817f-aa49b9edf5f5 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.120598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.120598Z digest=sha256:09254782d7d5992cc3d641e207b16e96cf92ce05ace5409a11f39145811dcf38

Observation 5f68243e-48e3-4e26-b3b3-ef1c2eeadc7c · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.272093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.272093Z digest=sha256:847dc71b45f727e3e2559ad6deaf01d69083c3f18bd33ec8bd87804ceb43e64a

Observation 5f09216f-3f4f-4346-8399-c217cdace7bf · outbound

This paper cites Extending Context Window of Large Language Models via Positional Interpolation.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Extending Context Window of Large Language Models via Positional Interpolation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.314588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.314588Z digest=sha256:6a0509287218ead8b627fb05b57718b66e0ee25c455e9240c431ea5bcacc9acf

Observation 4417de2c-c90c-4496-8e18-b2eb200b88d8 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:53.894572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:44.329946Z digest=sha256:00fe39b1006fd4534236b2c1d8bb5d55a2971888458c5b5c8c9d3d550a52a358

Observation 5974bde4-aca8-44c4-92c9-6f0a835ffd91 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:53.735912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:44.396554Z digest=sha256:9d6e69b134207346ecccfb8b5a44170e04a0fb74395b8b8f7b570c19f72a2c59

Observation 12cae417-0717-4ead-9967-196f7ce980de · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:53.562496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:44.479076Z digest=sha256:c47bc3c3f8eda9f67fb71cfb3a855f30db18ecddde9f8819b163a1fc5cc563dd

Observation 6d4cbf06-a589-4f36-8b55-fe83b1168291 · outbound

This paper cites Human-like Summarization Evaluation with ChatGPT.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Human-like Summarization Evaluation with ChatGPT

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.575596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.575596Z digest=sha256:c1e2ba685e908d7593b04bd5ab42aa417594e401e67161e2237adbbad299d8f4

Observation ee2d02a8-d372-4b79-83ea-5b07b22a068b · outbound

This paper cites The Llama 3 Herd of Models.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation The Llama 3 Herd of Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.657266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.657266Z digest=sha256:211c2c3384cf413d2ec747b8adc9c011e9c91d3da974f037759dedc225084f97

Observation d7fb7a6a-2efb-44d7-b6cd-fb0a866d80fa · outbound

This paper cites Efficiently Modeling Long Sequences with Structured State Spaces.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Efficiently Modeling Long Sequences with Structured State Spaces

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.770766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.770766Z digest=sha256:e52af62036246c5e174bd7f0a7d5935062258d800010aa2127b6b70e4f804ade

Observation 7477ff42-3194-405c-b262-185315e163df · outbound

This paper cites A Survey on LLM-as-a-Judge.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation A Survey on LLM-as-a-Judge

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.861710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.861710Z digest=sha256:cfec0febf0b7af60ade7e4ec226550e1635b29864cbb44eb35e17e9e0ffd2ad2

Observation 7042a2e3-5ddb-4d6c-8f2a-26fb77cc0a4a · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.944982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.944982Z digest=sha256:dfdde7e460e5c170c2eefe9d0d7c4a3952742e3d9361bf1b6fd1b0c0d4f3f239

Observation c35b17de-6ccc-4761-af96-722dbe43503c · outbound

This paper cites MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.021883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.021883Z digest=sha256:294b2bb1259a65eec9c77c1fa7e6d1edf794e87329280effb27be8dab2ca9e4b

Observation 529850c2-aedf-4374-af1a-c1af52d8c8fd · outbound

This paper cites Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.103956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.103956Z digest=sha256:7f323fa087a5f4a3fef2cb5da24d788786e90d44cc627b487475346ff1e09455

Observation 34d087a0-cd4a-4e21-b2f9-d566fe2f35c2 · outbound

This paper cites LongForm: Effective Instruction Tuning with Reverse Instructions.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation LongForm: Effective Instruction Tuning with Reverse Instructions

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.194684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.194684Z digest=sha256:6aad734398939d45f193e37b64d81b7eb7f650c92be710d7bd3970e844c1ccc3

Observation a5f31b34-4056-49a4-8c61-6cf39760e3e5 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:53.402762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:45.242809Z digest=sha256:c919fa3d8f459ff76798785400e7e70b76285f5e6e4c44a0b4ab68ec6aa24fe9

Observation 0d698436-d03b-4e13-a145-6c37421a7044 · outbound

This paper cites LongLaMP: A Benchmark for Personalized Long-form Text Generation.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation LongLaMP: A Benchmark for Personalized Long-form Text Generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.315741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.315741Z digest=sha256:39a00ccf884725a5e6ebf77cb024d3e71ee4a7e56bf9084b18a3df72aa4a8728

Observation c3c620a6-2120-4b1c-a9c2-a7dcadeb3e73 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:53.216011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:45.393122Z digest=sha256:968c4355ad8a1f8f497ca813669a9239887ad145d704a3ad7fd350cdefc5ebf5

Observation 0179e988-ab6e-43f0-92ac-1d417b699cf5 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:53.054656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:45.467796Z digest=sha256:4ae9cf3adfc26114b24f329847d65d6f09d02dfbdfa3c61c13576fa43622a994

Observation a4d8107f-edb6-4c45-b171-df8fd309b3d6 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:52.915515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:45.539371Z digest=sha256:2229ef439d783ff052ba09b286703dcaf4978b12fee89bf4fa9ae487d92d28c1

Observation dfa0033a-f9b9-48e1-b771-49d7b24b3522 · outbound

This paper cites Long-context LLMs Struggle with Long In-context Learning.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Long-context LLMs Struggle with Long In-context Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.617067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.617067Z digest=sha256:085f21b8887e94b9f9158e4531bba67c2dc6a0da2f747529b5ae505ba0a0545a

Observation 5c469c74-a5be-4097-ac07-bf4a99c668c9 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:52.753067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:45.687633Z digest=sha256:ae0172f218e770e0cd4cf362b4b7f7ea823c68f07baaa33bb19994e04162ce8a

Observation 88c43ae2-e0bf-4f22-a362-20ad6e67b8e9 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:52.520620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:45.739551Z digest=sha256:bc0e2a19116fb7f225e9ccfab3b81abac317f4932c550757de6688793b7152b6

Observation 13afb7af-db6b-4057-8a33-5dd80c6b2a6e · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.792936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.792936Z digest=sha256:ca6c2ee7b3470fb1aa9acd4392c6cb2e21a26d002e6dd4790edc85a617e5ee15

Observation a6fe5f2d-6df3-4d70-a985-0738b2b35da9 · outbound

This paper cites DeepSeek-V3 Technical Report.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation DeepSeek-V3 Technical Report

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.880835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.880835Z digest=sha256:dc0709dcae724c050e08cd04d731450d85d85feab736e931113068c36168cd27

Observation 830307c8-b165-4305-bbf1-ec2a4ef09ced · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.941061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.941061Z digest=sha256:5e3a3ad6c976e563676d3218c04cbedb068f171d6656c54331fb2c2f7d35d9b7

Observation 086c634b-f677-4d9c-8c9a-9cef0dbeb295 · outbound

This paper cites LongGenBench: Long-context Generation Benchmark.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation LongGenBench: Long-context Generation Benchmark

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.997697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.997697Z digest=sha256:4cec0c7ad22491f5365b928f070ed8710dd7082f572cb70b15e8b1a81c69f52c

Observation a592615a-3980-41cd-aced-ef607163ab2d · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:52.378599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.093389Z digest=sha256:e0c779ac7c019194bc2ad6da22cc68488e581e0af370fcf50be47fcceb03f933

Observation a32eb9f3-2544-4fb3-8e49-00b627fadd6e · outbound

This paper cites Exploration of Masked and Causal Language Modelling for Text Generation.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Exploration of Masked and Causal Language Modelling for Text Generation

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:49:49.759678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.175021Z digest=sha256:23eead5a86e7b883f63478c05539f08f13e16f159ac943a5961b2d45556bafcb

Observation 10dafc76-c9c9-4275-8617-cbff01502816 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:52.222337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.255452Z digest=sha256:a8ae6550fffbe18f853da1cbe405e745d1463485223809033ef59b26417b70d3

Observation f888ce7a-2b12-49b9-9897-b8e67fc92e8f · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:52.074672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.319450Z digest=sha256:ac7fadbf59f7fe6b15619f91f0d07f6fb864c2d4d3e2c172166e5097f393b295

Observation d63995fc-019c-46d9-9870-002206de8f2a · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:46.383781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:46.383781Z digest=sha256:a010fd5b44215065a52dff226f12a5bef7c2ce0abfb250ec2e1cc432b2d47b30

Observation d42f8e59-0122-4261-9780-078650f3b807 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:46.451664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:46.451664Z digest=sha256:4079ee2816a446303ca42754adfe50a094aaff5150fda82f4cfb07370d3771ad

Observation 92eb86ef-20e4-4d99-af77-27d2cbf87aca · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:51.955371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.512069Z digest=sha256:6d9f0eef117f00c13b1245e9e83765a650fec5a803f3f9ce8d602d42a0c40681

Observation 12605ee8-c6ba-4fd6-9164-6c7aac2cafc5 · outbound

This paper cites Suri: Multi-constraint Instruction Following for Long-form Text Generation.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Suri: Multi-constraint Instruction Following for Long-form Text Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:46.578703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:46.578703Z digest=sha256:2e74fe31711f03ad3b9c2293a6dc90ecef1dd0186bb33426046db20e6a669403

Observation 51abed4d-7cb7-4fb6-b823-3491463ad441 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:51.856072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.677908Z digest=sha256:ef340901fbd40f9056009db8eca64028403cd49d4fa7bd842e0da1b4a971d9fd

Observation 61142bbc-84ef-42df-8c1f-b9fea274b958 · outbound

This paper cites HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:46.777966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:46.777966Z digest=sha256:b24c67222f74f86b2cfa29a6abe0aeead110dee507b4fc810384bb6fa8d3ebec

Observation d6570b55-f2a4-4d3a-86ad-9cac1299f8aa · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:46.874265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:46.874265Z digest=sha256:278089ac3fe2925d8045ab875c5c9bf96f3bb53f379173de6ff943f3e488a9c5

Observation 9eed7ccb-e89a-4efb-9ae5-7a14134382ce · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:51.683561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.997928Z digest=sha256:5c1327d174ef1d5f961412ef85ddfed1cbd2a2dc511bd8e6c03d4f0c0b6fc690

Observation 40630c9e-4032-4943-8cbe-acfa5b8f5174 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:47.087005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:47.087005Z digest=sha256:fe8aa71b203d16d98f23566ee1220fa25072c252b04b5b08a5acaeb7c18d5aba

Observation 5879515a-38ee-40e5-bfec-08e0afebe51f · outbound

This paper cites Preference Ranking Optimization for Human Alignment.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Preference Ranking Optimization for Human Alignment

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:47.131971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:47.131971Z digest=sha256:395e1645212df1f53bcab19e6b2fe5d23dbbca3218b111db5525e81b793a00a4

Observation d61cef29-e34b-401a-896f-fe36f61d1391 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:51.412239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:47.192922Z digest=sha256:b5968db74564aa07e08845b266a154ce1868a8310623aa948fd28f65f8bcf933

Observation c63adc82-6224-4d5b-9678-88897cd54dfe · outbound

This paper cites PROXYQA: An Alternative Framework for Evaluating Long-Form Text Generation with Large Language Models.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation PROXYQA: An Alternative Framework for Evaluating Long-Form Text Generation with Large Language Models

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:49:49.444103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:47.284775Z digest=sha256:d7b940897576151a7710064ab4b97ecba28a1160079ec0e159eac1f183f030af

Observation 6272fee1-3493-4ae4-a8b7-0fe84e220959 · outbound

This paper cites Long Range Arena: A Benchmark for Efficient Transformers.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Long Range Arena: A Benchmark for Efficient Transformers

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:47.407269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:47.407269Z digest=sha256:020ea543bef39a0b9b10da20add87fc32043aae2e17f738e274ba925e267df00

Observation b233dbd9-8c9d-474e-a345-722c41b0d59b · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:51.205623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:47.467394Z digest=sha256:9e0b3caa28a0ba0d90496ded6b4a87f2597cd743b64906c6d53e3b1cc002f620

Observation 0883774c-4db3-4d91-a393-76b9d3224910 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:51.006117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:47.581044Z digest=sha256:774dfebadc925380e750c55e8cf528a1f77d9835084e0fa00d110664d58c0887

Observation 16eff6a0-1561-4ddd-9491-e91ea99aa61c · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:47.732683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:47.732683Z digest=sha256:7812df6bbfd5f4a40a61b8157c8ea9fe75acbe72be9ce0fb57a5432801a9463a

Observation 164df63e-321b-459d-a263-6131ebd555aa · outbound

This paper cites Self-Taught Evaluators.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Self-Taught Evaluators

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:47.837626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:47.837626Z digest=sha256:a41116bf76609529697642cd19b322aae1396b53eed8ab3ee09d6bd7aff0bb43

Observation 95be38fa-aece-4e88-8109-e4c96cdb5f47 · outbound

This paper cites Meta-Rewarding Language Models: Self-Improving Alignment with LLM-as-a-Meta-Judge.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Meta-Rewarding Language Models: Self-Improving Alignment with LLM-as-a-Meta-Judge

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:47.916012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:47.916012Z digest=sha256:644bbef1b08884ecc94ec97b18817fa7861d5d528e9b49d4b2fd2a098dae4d1f

Observation 1d64bf2a-1c1b-4886-bc56-f5b34b402675 · outbound

This paper cites Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:48.044179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:48.044179Z digest=sha256:9aa627474465311434543747ca54a8cbdb7675603711af58e6d192129ab66ca4

Observation 1af95218-cebe-4276-b85b-922499cfaf73 · outbound

This paper cites Effective Long-Context Scaling of Foundation Models.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Effective Long-Context Scaling of Foundation Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:48.167491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:48.167491Z digest=sha256:a1b456f6d2a43ef7c49f8d3b089399963b3d87ba6fc510eb1d8be451091aad55

Observation 94a2298e-0f4d-4ee8-ac87-a8499243abc1 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:50.776994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:48.262762Z digest=sha256:1603e60192621f5e4eaf62b4dcbfe2833ccfe48f2112759b0c5c436b36ae1bad

Observation 43cd9b3f-4865-4c7d-8ad1-16fd701443bd · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:50.613756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:48.378006Z digest=sha256:bfa7732fe0737aa366ac08d76d075cd35f313c482ea13967b3ad3b44a454da79

Observation a6f978f5-fece-457e-b9b1-eb2383e21ca4 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:50.390405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:48.477481Z digest=sha256:db6dff230d6aeff41b251ee6d90cdee77e4a1803c920df18a6b4fbc4720913de

Observation 835d83dc-d79c-49c9-9398-323753002d8e · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:50.209143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T23:49:48.564204Z digest=sha256:426d3fa6d7ff6da82522a5face084d88168d3c8b31af69f40b4cc4bceee60403

Observation a45831ae-3018-49f0-b5f9-0c5dc3b00652 · outbound

This paper cites RRHF: Rank Responses to Align Language Models with Human Feedback without tears.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation RRHF: Rank Responses to Align Language Models with Human Feedback without tears

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:48.654069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:48.654069Z digest=sha256:7d153d79c8cec4872cbe5fb364ff314f5f629b2e71e6ed4aef49c8c3ad6e09b9

Observation 5c289b82-76f9-485d-9729-a6e8826b1212 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:48.826479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:48.826479Z digest=sha256:f631df1c6f4c35a72ab40c9b812ebe1f5acaf65b3a4f1c7eb425c4b39a2dab5f

Observation 3fe594dc-d118-4ddf-9c68-3c4e21e6461c · outbound

This paper cites LongReward: Improving Long-context Large Language Models with AI Feedback.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation LongReward: Improving Long-context Large Language Models with AI Feedback

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:48.916548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:48.916548Z digest=sha256:20e850b137182e6afc72fd83ee7f8bb0af843268c486bcd5adcb9c3b52521f76

Observation 3e0be4d5-c228-4e61-a556-972f360edc0e · outbound

This paper cites online" 'onlinestring :=.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation online" 'onlinestring :=

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:49.024302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:49.024302Z digest=sha256:0fa65f6cd4c24f4816e4e406ffcd0588403db7dac01c27671d52dceff73d0f1b

Observation 3fd049d2-2af4-4dbb-8589-4f3b74950240 · outbound

This paper cites write newline.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation write newline

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:49.180444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:49.180444Z digest=sha256:6e7f83c82ac850c69fcc33c2132c25510939384d6ed22531c3de59ba363020c7

Pith citing papers

Observation 30c3a4b0-31ad-41ca-941c-9751ccf1f351 · inbound

A Survey of Reinforcement Learning for Large Reasoning Models cites this paper.

A Survey of Reinforcement Learning for Large Reasoning Models From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation

Reference 176

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:05:31.664012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-18T00:02:24.352947Z digest=sha256:35b2fb1eb785b43bfb6958c579553331a99275f5ac0e73c78bbd427224632a50

Observation 2130b475-363d-4d1c-8fd9-220de7dafc4d · inbound

SEIF: Self-Evolving Reinforcement Learning for Instruction Following cites this paper.

SEIF: Self-Evolving Reinforcement Learning for Instruction Following From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:20:55.949246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T01:51:09.514927Z digest=sha256:478b829a9d68accc1c6754b04548f5ec9ccfff3b39173279efcb2c879350575a

Observation 6c6922a7-4223-4a49-87bf-676531d8ee2e · inbound

Trust Region On-Policy Distillation cites this paper.

Trust Region On-Policy Distillation From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation

Reference 171

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T20:46:14.343964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-28T17:38:50.313305Z digest=sha256:8811412c3e75dfce880458fcfced3eb6d4d28d7638034640578ea69f73830fad