Pith. sign in

Paper Citation Record · LEDGER

Large Language Models Can Self-Improve in Long-context Reasoning

As of 22 August 2026, this Paper Citation Record lists 100 of 112 outbound references and 4 inbound Pith citation observations for arXiv:2411.08147.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.08147 v1

Coverage vector

measured 100 of 112 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T21:59:11.621135Z

measured 104 of 104 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:43:44.229927Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-10T18:16:17.926450Z

Reference resolution

100 of 112 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved99
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8a918d31-e3fc-4ea4-93b3-159d55fde0e2 · outbound

This paper cites GPT-4 Technical Report.

Large Language Models Can Self-Improve in Long-context Reasoning GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.306537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.306537Z digest=sha256:d0cd6691c72c9691774ab1c83a0ed1808ae21534affdeaef10ffe03f4086b8ef

Observation 3dabd5db-0c68-4d89-b0ce-a90f2a924cb7 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.311066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.311066Z digest=sha256:d48d4b2db6a57f56b1562b641b550f5f27d65690e4c13fcd1fc1bbbba390096b

Observation 509026b9-55e0-4fbc-a7a6-dd7391a07281 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.314878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.314878Z digest=sha256:e5ebf01e0672b03ef52f297a5d93cbadecff0af10e5d22834d5c9b79903f5fef

Observation e88427c0-687d-4f77-b162-da7af3e8fca8 · outbound

This paper cites Why Does the Effective Context Length of LLMs Fall Short?.

Large Language Models Can Self-Improve in Long-context Reasoning Why Does the Effective Context Length of LLMs Fall Short?

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.318494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.318494Z digest=sha256:58e05eaf6077b32d609eef770409ef7f7eabd0a913c0a29bba8238d9ea35907d

Observation 2e99b3c0-4e25-4e48-804d-41a66801d0a2 · outbound

This paper cites Make Your LLM Fully Utilize the Context.

Large Language Models Can Self-Improve in Long-context Reasoning Make Your LLM Fully Utilize the Context

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.322583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.322583Z digest=sha256:21d55856b9c69258f22328a2c9a532a6f0b1cbf322a7e8484f3db778567e062b

Observation 392585d0-a57b-4dce-af7f-0a2ec01a3c8b · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.326458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.326458Z digest=sha256:c1538a757fc2f364a2f54a6d4fcc20720f14e8bbeceaba9a7d313bee97fc3b05

Observation 354655eb-68fe-4eed-b326-cd72f108b74f · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.330307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.330307Z digest=sha256:3c5428a7f312a4cb578ed0f9de03a3e4c98952a83f233dda5e9411368e593086

Observation 7d990902-ad1d-454f-bba7-970403b1daaa · outbound

This paper cites LongAlign: A Recipe for Long Context Alignment of Large Language Models.

Large Language Models Can Self-Improve in Long-context Reasoning LongAlign: A Recipe for Long Context Alignment of Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.333514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.333514Z digest=sha256:fe508be9671eb0dd8b3e96c710f70ee98973659bbf1339fb4c11d418f0ac3e1d

Observation ef285698-d967-47a1-968f-11681ce2145c · outbound

This paper cites LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding.

Large Language Models Can Self-Improve in Long-context Reasoning LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.337185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.337185Z digest=sha256:05e7d2598ac88410fe354d3057a7887aaafb088e2d3c0d81489c0528e231a9bc

Observation 37a5a221-87c8-41d5-87ac-b8c68c9f7bf9 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.340595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.340595Z digest=sha256:959d7409cede975ee6fac814540908a9ce81c2199331fa762b6eba6c4a74bb0d

Observation ff9ab0f2-33a6-4696-829d-b45653353371 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.343731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.343731Z digest=sha256:5619f7beeeb122761d0e3971e9b084ce8c7a802bdf66c39625c42bc7e847d0de

Observation 9561d0c8-6a24-42f9-8dde-222482d790bf · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.346733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.346733Z digest=sha256:c20ee034f457e2d094f1c30f89440426675d80de1c0ff88d256467e2d2404e20

Observation 739b5958-0b73-4fcb-a02a-b9dec1801d39 · outbound

This paper cites Bickel and K.A.

Large Language Models Can Self-Improve in Long-context Reasoning Bickel and K.A

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.349631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.349631Z digest=sha256:217e508bafa71418d07dd5b28161240f04d86bebc78a2b9e1888fab77ca680dd

Observation 6ee08289-7d7c-4705-bc8b-60d9657628e9 · outbound

This paper cites Sparks of Artificial General Intelligence: Early experiments with GPT-4.

Large Language Models Can Self-Improve in Long-context Reasoning Sparks of Artificial General Intelligence: Early experiments with GPT-4

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.352448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.352448Z digest=sha256:956dac3b553120205832f73019fba3d10949a744defb80c4034675110dcd5a57

Observation 7e925acf-8f64-4f92-8f8b-6df75b5f8399 · outbound

This paper cites Extending Context Window of Large Language Models via Positional Interpolation.

Large Language Models Can Self-Improve in Long-context Reasoning Extending Context Window of Large Language Models via Positional Interpolation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.355767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.355767Z digest=sha256:526bf040cd6e80d178ed2c5f470e9a00102624e7e528c07237bcd7c42049b39e

Observation abf60e79-3821-40a9-a09c-04386207b308 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.359164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.359164Z digest=sha256:4ab2f623f2283807ca54fb6c3204a01a6ede91ee8cf60e1eab1ce47616b1aef2

Observation a4c9daad-e1b8-46be-bb93-4559074dfb04 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.362368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.362368Z digest=sha256:7a6ecfffa035f4ff87b25440b78560ace935f973f0d9497820200708f6393248

Observation e3f0d919-6313-43fe-a42c-985bdb14ffbe · outbound

This paper cites What are the Essential Factors in Crafting Effective Long Context Multi-Hop Instruction Datasets? Insights and Best Practices.

Large Language Models Can Self-Improve in Long-context Reasoning What are the Essential Factors in Crafting Effective Long Context Multi-Hop Instruction Datasets? Insights and Best Practices

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.365346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.365346Z digest=sha256:9a22b7948d3cc9cdd56f57d0ceec6d55fcbf7825d7ee839e10f5c03ea1c3af25

Observation c0f5226c-65a3-43d5-ac16-74175ea3ad2b · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

Large Language Models Can Self-Improve in Long-context Reasoning Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.368436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.368436Z digest=sha256:b96c85f3de5b88f06843221bd50db5dfc9b15f5ad3fd2feba419aeb6a4294935

Observation cb5bf869-a2fb-4b9e-ac8f-3e77d9e3fcba · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Large Language Models Can Self-Improve in Long-context Reasoning Training Verifiers to Solve Math Word Problems

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.371647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.371647Z digest=sha256:cce0f3b6ffeeeb982d574bb2dcf9af861d5c083aabbe665d7035a9a5cab3be04

Observation b19e1b83-3fc0-4dab-bc38-926d3911d61c · outbound

This paper cites UltraFeedback: Boosting Language Models with Scaled AI Feedback.

Large Language Models Can Self-Improve in Long-context Reasoning UltraFeedback: Boosting Language Models with Scaled AI Feedback

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.374842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.374842Z digest=sha256:0833fc3c504d26d6464256bf199dcb3a41d2c3f96ef1a13e45ca382e39cdddb4

Observation 5172e102-cacd-49f1-a271-3b1711b65aab · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.377998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.377998Z digest=sha256:a5eeb18c6f509a66a4f79742acaedf194e0bd68c5d7427b60899150656f62038

Observation 3bd30292-d155-457a-829f-10bf187509e1 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.381347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.381347Z digest=sha256:4c906bdf335c79d807e31f08b4fb454d5a2def328ef3a6c06f66f540f162162d

Observation e1d38f53-668b-45bd-986e-c9910b2847a6 · outbound

This paper cites LongNet: Scaling Transformers to 1,000,000,000 Tokens.

Large Language Models Can Self-Improve in Long-context Reasoning LongNet: Scaling Transformers to 1,000,000,000 Tokens

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.384267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.384267Z digest=sha256:b78faf66a146627c608857977aa4c010439c3b94f15896c6527dfaf2eb9e4088

Observation 6f31a60e-d1f3-4289-83f0-7e4877a4b4ff · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.387537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.387537Z digest=sha256:d9f921a8d0ae7cc845f3a7c93facc36d1a3cd28af67ea9d39cb4e5718610a98d

Observation 82d94504-c315-494a-a177-3a331dcb2e14 · outbound

This paper cites The Llama 3 Herd of Models.

Large Language Models Can Self-Improve in Long-context Reasoning The Llama 3 Herd of Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.390488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.390488Z digest=sha256:dfd686d0d5eb3ad8139a58bb42c9b14dc64134a047826fb3b1eafb03e2849a76

Observation 2f3bc8b2-97fd-4d87-a5ab-3f458e140287 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.393791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.393791Z digest=sha256:88dcc48e7a6542f9e811e16ef72eb2a42c014a7bc9163234dace0c8473054d22

Observation bb6a64d0-6233-4a37-9cab-f0f1ba3243d1 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.396861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.396861Z digest=sha256:4a7c6e1906f06b5739013fa16b6b901a477ae5988671dd8d367e43d4fcab7fc2

Observation 5283bd0b-4f60-478a-b28c-5fc97fc64b0a · outbound

This paper cites Data Engineering for Scaling Language Models to 128K Context.

Large Language Models Can Self-Improve in Long-context Reasoning Data Engineering for Scaling Language Models to 128K Context

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.399879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.399879Z digest=sha256:cae7bc5d82774aaf526f037f27327d50e96b80c470d59196249879cd7de88c11

Observation 6dedd3d2-c6c3-436a-84b4-af3b8e919535 · outbound

This paper cites The Capacity for Moral Self-Correction in Large Language Models.

Large Language Models Can Self-Improve in Long-context Reasoning The Capacity for Moral Self-Correction in Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.403148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.403148Z digest=sha256:4baaadd50883ca4331bca411bf6e91440d32eed3a59e7f2364581b0b6deefa44

Observation e4aa1090-0d5e-4a7e-9055-d0b56a25d104 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.406956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.406956Z digest=sha256:3700987247ecffb28a3a24bed5640e9326d9ba07a768d1d1f3cdd89045d9e7a7

Observation d36d4899-7809-4eb9-96de-33aa57067f44 · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

Large Language Models Can Self-Improve in Long-context Reasoning ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.409930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.409930Z digest=sha256:f1a853446faa179ff2d69dfd3c85688201beadd20bc3bf771baceedd79b55d26

Observation 7f06b897-ccd0-49d1-8925-f798f58d15c4 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.413071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.413071Z digest=sha256:bf6e787dae8b885f4ebf2ba10dc65f8cfda13361e09f687350872abac0c1b9b5

Observation cbe9ffd4-fc00-4b25-a7f0-dd4a1fff88b4 · outbound

This paper cites Reinforced Self-Training (ReST) for Language Modeling.

Large Language Models Can Self-Improve in Long-context Reasoning Reinforced Self-Training (ReST) for Language Modeling

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.415869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.415869Z digest=sha256:1f7baeb3982e3e6a784378d5e95549ac1bc8cdc4f15149e06c6cd4f0299a655b

Observation 1762ad22-dc37-428a-ae9f-d433fcae8340 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.418981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.418981Z digest=sha256:db00fdf77de8b1db3c0b57cdc049647adaf647ed1e21dfea2183a00b763f21db

Observation 8e6aa272-51e4-4a3e-bcc2-f6ad74fac1f2 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.421706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.421706Z digest=sha256:8dfc64becfdc9931c1ed92c5ccd2da8afbf5c14f4619d10d4cad5fea8a6cf46e

Observation a247b1b7-f16e-4b2a-8edb-674c8601e163 · outbound

This paper cites ORPO: Monolithic Preference Optimization without Reference Model.

Large Language Models Can Self-Improve in Long-context Reasoning ORPO: Monolithic Preference Optimization without Reference Model

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.424607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.424607Z digest=sha256:e2ed02726db19925f5063092ee0645e7340ee2d6eb18800f37c9ada762ddb753

Observation a6d6cbf8-1023-4901-8447-69da31179a18 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.427457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.427457Z digest=sha256:489d8401f37d82b7b8ee5685dc61193f855a46865d5ac142497d1dbc3a3b188a

Observation e2a5de38-6888-493c-a419-5fcd61ecccc4 · outbound

This paper cites RULER: What's the Real Context Size of Your Long-Context Language Models?.

Large Language Models Can Self-Improve in Long-context Reasoning RULER: What's the Real Context Size of Your Long-Context Language Models?

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.430548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.430548Z digest=sha256:0203a7c8a73e7f07f0c7332df1551de3380ac219f876054b908b945d79313923

Observation f9624050-6c9d-4344-8497-db53506adce2 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.433714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.433714Z digest=sha256:d9b444dbb0eb6921c0c40942f15b62cdfb52381f4cfb7ea5dc269759a92298d9

Observation 6e4d9a5f-68c4-4308-a50d-c58376af60d8 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.436778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.436778Z digest=sha256:8d9c35bf4dbacbf57483bfe094ef70bfa3d2e05e1e8b3aed08ea512f65217917

Observation 8ec6703e-cebf-470d-8c86-fded145b4a81 · outbound

This paper cites GPT-4o System Card.

Large Language Models Can Self-Improve in Long-context Reasoning GPT-4o System Card

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.439591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.439591Z digest=sha256:0e105e1491d7397c672596485aa14e722f038f06e147d4152793fc8ea0df1023

Observation d1e7c5d9-ad45-4730-bc51-95efb318188d · outbound

This paper cites Camels in a Changing Climate: Enhancing LM Adaptation with Tulu 2.

Large Language Models Can Self-Improve in Long-context Reasoning Camels in a Changing Climate: Enhancing LM Adaptation with Tulu 2

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.442695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.442695Z digest=sha256:f151ba2c44fe442f518c86dd7c47429ed9a9798615650c6942d3207837edf638

Observation d9273d0f-97c5-4a90-8012-f42e1105f403 · outbound

This paper cites DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models.

Large Language Models Can Self-Improve in Long-context Reasoning DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.445695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.445695Z digest=sha256:4494c1e6e5d40758587891a6d37f78f7bcc8b289375ff7b4aa6f8d20d3316bdd

Observation 61a0b5a9-cc02-46c6-8735-7f2cf0f1ad99 · outbound

This paper cites SELF-[IN]CORRECT: LLMs Struggle with Discriminating Self-Generated Responses.

Large Language Models Can Self-Improve in Long-context Reasoning SELF-[IN]CORRECT: LLMs Struggle with Discriminating Self-Generated Responses

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.449089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.449089Z digest=sha256:45f9595ff4577517747b767abe0038e2137f705da3fcf55b97a2ad2487a17efb

Observation 9681af41-2f08-4e33-9465-f9cdd956a3ad · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.452296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.452296Z digest=sha256:b9174bd7c7899ccedf7a08de05ccdfa1f2ae2a00bf7c80820e2a9d78ee71afc1

Observation 6e422970-bc84-4f76-b750-a0a64c7b1626 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.455500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.455500Z digest=sha256:a889a3691fad45d5c3caa21e97cfc83fbfb2e77b044d81bf9e5f1dc1fbba5bdc

Observation f7e91e14-a153-4ca6-ae4d-0258bd882691 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.458498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.458498Z digest=sha256:b0eddd49e8867d19f26ecfd53ef2ebd80cc881e18d56349c52d15305330bc2a3

Observation e1156245-d11f-4974-8b38-d32615c8ab4c · outbound

This paper cites One Thousand and One Pairs: A "novel" challenge for long-context language models.

Large Language Models Can Self-Improve in Long-context Reasoning One Thousand and One Pairs: A "novel" challenge for long-context language models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.461492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.461492Z digest=sha256:29dfd7c6f6470fbce187d145970949737cb70fa4bc419c0dcdf62532b9825743

Observation dee2e371-e6d9-44ea-b8c6-79b428defc9d · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.464840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.464840Z digest=sha256:70da9f6e4c644bae44805681f68cc5c314c8fd2e525522d0a6c3f3e4826a2e08

Observation 974ba318-2539-4414-b5e1-112525dfc358 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.468356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.468356Z digest=sha256:2bac86c36043ff81451e3b28d05d8bd36b31d570a72a0c1a569b51a04b634b29

Observation 80d4dc1a-5b77-48b6-aaac-edb217397620 · outbound

This paper cites Training Language Models to Critique With Multi-agent Feedback.

Large Language Models Can Self-Improve in Long-context Reasoning Training Language Models to Critique With Multi-agent Feedback

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-08-12T21:59:12.006873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-12T21:59:11.471411Z digest=sha256:46cc767105ae9cf61d5d40ee4bdfc2ad402140577434bbcae488d3147ce2236b

Observation 3d6373db-0923-4f5f-9efe-8e4f21f093b4 · outbound

This paper cites CriticEval: Evaluating Large Language Model as Critic.

Large Language Models Can Self-Improve in Long-context Reasoning CriticEval: Evaluating Large Language Model as Critic

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.474773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.474773Z digest=sha256:4587678bb5dae3a000daaa18da1f69fb0999c8083f9f98e0045cf9a8a1ef00cc

Observation 4b3ca501-514d-45b2-83b9-9f6d138490f3 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.478016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.478016Z digest=sha256:14834019479ae8b0b0b20f58349c4b797a444da77a6c0668acf8c224f7bee58f

Observation 7a299b91-40f2-4427-8bc4-791eb3b56e0f · outbound

This paper cites ALR$^2$: A Retrieve-then-Reason Framework for Long-context Question Answering.

Large Language Models Can Self-Improve in Long-context Reasoning ALR$^2$: A Retrieve-then-Reason Framework for Long-context Question Answering

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.481767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.481767Z digest=sha256:d7c303fac5d863cbb6c33eefac0b3f8e5d7231599dc9d3b63f91f20619d6b9a4

Observation c0cf0d74-be64-45e6-8750-e41c703122c1 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.485205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.485205Z digest=sha256:a6d9559529112e518dfb87d764637de696386cfbbba73dccd6f059ef9e8a3b3f

Observation 6523ec48-8509-49ad-9b5a-729520146a8e · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.488339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.488339Z digest=sha256:77c241f3d7d34c955cacb8cddb525124d93baf4110a018a1a71d55caf62c88cd

Observation 419a9974-5a29-4d11-9bb9-7c1c9b52e551 · outbound

This paper cites Jamba: A Hybrid Transformer-Mamba Language Model.

Large Language Models Can Self-Improve in Long-context Reasoning Jamba: A Hybrid Transformer-Mamba Language Model

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.491299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.491299Z digest=sha256:1fb814c16544b383d518fae1cabde3c6f8104b4639284b3dca389da106373645

Observation c9bac916-d2b2-4e8a-b2e6-7e0ae6134fe3 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.494523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.494523Z digest=sha256:fef0e8094553f50b4088caa71c2d19f2ceb7ba510b354dc07c268628998fff96

Observation 9c5982ca-cfa1-44a6-a9bc-4b0342e10c9a · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.497467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.497467Z digest=sha256:4c69531c91813cf1dd9a0e61c1155bbc4a84dde5e9f2145e7a9ded5510563589

Observation 55bd8efa-d946-469d-b6eb-e6601d1c8245 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.500312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.500312Z digest=sha256:7a98383998576091af5737fd95da7adb8f65ef99771b3aaed7d095a78e51b9f7

Observation f50ed608-6eb8-44e8-ac3a-1a6efb078d9f · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

Large Language Models Can Self-Improve in Long-context Reasoning RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.503157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.503157Z digest=sha256:16ecadbb0a69380bc18b8edfb06297b27dac0447c8d6c61a45cd20a34709917f

Observation 23363eee-02df-4ebb-b7e9-31e27ad75bbb · outbound

This paper cites AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents.

Large Language Models Can Self-Improve in Long-context Reasoning AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.506456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.506456Z digest=sha256:e5a9b5360ea6986cdbd64fc9a46cd0b0e498ff1f30204e517c27cbd8453ba82c

Observation fd1b14dc-e64d-4505-ae7b-375c3626dd73 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.509653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.509653Z digest=sha256:4b5d8b7643a7f5fe4fde140a5473b0e17b0612b1b49572a7e0d4645593bfdf99

Observation 39ce6e49-77ea-4eb8-b2a0-70f42683350f · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.512655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.512655Z digest=sha256:aed25a3a241b0c214f5ced51a25a441e2d57db747f26022d5c070d29ba41d8af

Observation 2032335d-8d27-4947-9dad-9c74f00b81c1 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.515509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.515509Z digest=sha256:048f059186c044ac8e253f0dbc2c703c9e78358d9408fc122ad3de92923ccd6e

Observation 18c794d3-e2b7-46e9-ae57-43d3f4987e06 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:59:12.564988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-12T21:59:11.518421Z digest=sha256:9355bf60d96b66074f80653f7207ce1e7111de95e54f79526940339f5a2c7424

Observation 4860d0f3-6582-4e0c-ba29-b429538df09a · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.521265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.521265Z digest=sha256:dc86c779a89ec81879750366e711af5c133382f648e40fdf84ff02b3ffe69553

Observation 7a31241d-3999-4b3d-b11a-c5b999828ab7 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:59:12.548112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-12T21:59:11.524202Z digest=sha256:51ccdb987af5b53df42bffa8f3acfccfa1275df332ffe93593b2badb4e032214

Observation c155a58d-f7ab-4080-a6b3-6b509dc9c0e9 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:59:12.538319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-12T21:59:11.527124Z digest=sha256:c7cbfb05341bb6ca7fda91c2967f5b0959b7966f3ccbcebb0e8b4a147417023d

Observation 7349d838-9ab5-4026-8f64-8924a1a2a455 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:59:12.528543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-12T21:59:11.530354Z digest=sha256:1ae60ce1545edb008d8b4b7ebfe4849d795e370df8a71587db0d3b5261a26cd1

Observation a4cb839f-9b51-4a5b-b4e7-7eda387fc184 · outbound

This paper cites Self-Consistency Preference Optimization.

Large Language Models Can Self-Improve in Long-context Reasoning Self-Consistency Preference Optimization

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.533287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.533287Z digest=sha256:fd283ef29fc2e490d860b1169c8269beab1e5d3c5e74ca439f84e9ff20f87bd6

Observation 37a00a3a-e394-442b-99fe-8d510734ead7 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.536801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.536801Z digest=sha256:20cfea6d71b086ff4e7de10c2b6f17260415ea7010329fab9882b67cb828e59f

Observation de6760af-633a-4963-bc01-bb9d72098497 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.539758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.539758Z digest=sha256:b4f8ef939911260d45334bb527a76cedb546eea76e9623032f6cf71577a668ce

Observation dfb24677-d9b0-4b09-ab2d-5f53213dec6f · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.542752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.542752Z digest=sha256:05588c62d31afe4da110598d04885e76dab970f72a56cb7867b698d36d0744f7

Observation f3d1d7b9-e80b-4e5b-8aec-746b9671f09d · outbound

This paper cites jina-embeddings-v3: Multilingual Embeddings With Task LoRA.

Large Language Models Can Self-Improve in Long-context Reasoning jina-embeddings-v3: Multilingual Embeddings With Task LoRA

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.545616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.545616Z digest=sha256:6c20b9d7ac86985f6d77b5d50fd538c88702d08fff799496336b8b314a670fd5

Observation 03d9c865-d3ed-4065-b639-a14223410df7 · outbound

This paper cites You Only Cache Once: Decoder-Decoder Architectures for Language Models.

Large Language Models Can Self-Improve in Long-context Reasoning You Only Cache Once: Decoder-Decoder Architectures for Language Models

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.548817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.548817Z digest=sha256:72980618519d27b71f407bead8ce12503785355d8696290fe98caa6186af76a3

Observation bfe10510-beba-4749-9495-cdcf59f3ea56 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.552150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.552150Z digest=sha256:978a0cff4f97a7f579652d6da40aa03f73d035ecab232468d2ae26669759ce2b

Observation 8695dbf8-44fe-4b26-9dbd-d773b3462f48 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 79

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:59:12.493056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-12T21:59:11.555573Z digest=sha256:ba853e9fc38279c79b6d02791e63491d0fd7ca0036462a059e607f1a3128b94c

Observation 134e0421-c1f6-4e5e-ba9a-305ad9b52663 · outbound

This paper cites Michelangelo: Long Context Evaluations Beyond Haystacks via Latent Structure Queries.

Large Language Models Can Self-Improve in Long-context Reasoning Michelangelo: Long Context Evaluations Beyond Haystacks via Latent Structure Queries

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.558906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.558906Z digest=sha256:8571e62facee41e7a010e55f4d3b1686d227cf2a3052da3b0774d9c71533292b

Observation 64c024f7-d61d-4c4c-8722-3fa84c0defe3 · outbound

This paper cites Don't Throw Away Data: Better Sequence Knowledge Distillation.

Large Language Models Can Self-Improve in Long-context Reasoning Don't Throw Away Data: Better Sequence Knowledge Distillation

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.562032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.562032Z digest=sha256:473bb757090078ff6b90783a7c82a30fdd28824fe56602be21da446782217a70

Observation dc8741a5-ada9-4872-bbdc-76dbf92c8c9d · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.565165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.565165Z digest=sha256:a3d296fa2d86c41e9cbcd9e2af6c7543479f7246a1aadd41aba6db9633d3032d

Observation de30b7c5-5679-486c-a051-ddd988b49aac · outbound

This paper cites Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA.

Large Language Models Can Self-Improve in Long-context Reasoning Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.568384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.568384Z digest=sha256:3d85e984da8cd4140ff5b3242b7fccb4c26bcd65ada6d2f369a0febd949c7e01

Observation 99c7ef06-162a-4d9a-b0d1-46cea58f70bd · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 84

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:59:12.476279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-12T21:59:11.572681Z digest=sha256:a55c3a02a9446f328f8677349552c8b3fa764c0f2bdac90e9362428d1fcae36d

Observation 753c194b-dda0-4d9b-91d0-754ca68d1748 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 85

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:59:12.466599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-12T21:59:11.575884Z digest=sha256:23c9d156aea6d49768b167ad08bdd98b6393a429a2b9d956f26b727f7ed34238

Observation 3a4fa418-b8e5-415b-b1e6-4b23aa46a164 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.578863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.578863Z digest=sha256:b1d0bd346d7f604809a0593227cc75bc3e81b85a5a023e7ceebf1c27fff51c9d

Observation 22abef6c-9cfc-44f0-b731-2d1cc6444caa · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.581857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.581857Z digest=sha256:09e8cb35ef16e90338c5c6a5607c94c3e90460cf857920bc50a6456bb388213d

Observation 551fa1df-4c13-47e4-9416-7b02612eb989 · outbound

This paper cites Better Instruction-Following Through Minimum Bayes Risk.

Large Language Models Can Self-Improve in Long-context Reasoning Better Instruction-Following Through Minimum Bayes Risk

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.584965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.584965Z digest=sha256:44c268f2bdd133679d86e2794afe9de3098094ad83df74f603adf071408e4839

Observation 3019c49e-9785-4b8d-9c26-14670eca582c · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.588160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.588160Z digest=sha256:67d59edd5018a52989f689342ecda0652cf36c94ae10a352967cd894960005e3

Observation 05541181-61bf-495c-838c-87fe1e1a4d1b · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.590962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.590962Z digest=sha256:fbb317a7946cf71e47807c71c267adde6c06888ae1dec52bda57e5b241a802e0

Observation 044160ba-4ce3-40b1-827c-ce5b0a79c113 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 91

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:59:12.430587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-12T21:59:11.593746Z digest=sha256:a3b3c71e8abbe50902d7ae46e4dff3de44443c5591b07aabc82673114a03b139

Observation 89cb1edf-e351-4791-b2b6-19e5ed6efd87 · outbound

This paper cites Qwen2 Technical Report.

Large Language Models Can Self-Improve in Long-context Reasoning Qwen2 Technical Report

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.596551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.596551Z digest=sha256:5870cb90298acc5b204de0df4e3ab802e2b17a648e72aee50338166430ebfb79

Observation dfc4cb05-ce2c-40c7-86a4-e125bbc03f1b · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 93

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:59:12.420929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-12T21:59:11.599542Z digest=sha256:5e8d32c0795f692b3b09cd7f54b7e4ab375cb60ed933a587ce6cc9aa6d30a2e4

Observation c691c278-1685-4eeb-944b-0d6daf410905 · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.602396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.602396Z digest=sha256:86a0e3549f8e5ca73be292198270d324634ddc33a103e69f13839a73d186bf6c

Observation 9d542e08-3b47-4884-bfbc-f8c883a8cbe7 · outbound

This paper cites Differential Transformer.

Large Language Models Can Self-Improve in Long-context Reasoning Differential Transformer

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.605565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.605565Z digest=sha256:fde2cd4ad39606d1fa0fd45519592da76406d5c067a44cbf8e927a1c97145be5

Observation dddbd108-437b-4dcf-b9d2-f8b4ed5c62e5 · outbound

This paper cites Long-Context Language Modeling with Parallel Context Encoding.

Large Language Models Can Self-Improve in Long-context Reasoning Long-Context Language Modeling with Parallel Context Encoding

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.608721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.608721Z digest=sha256:9861256a29de50dc41c0428303bd5d56f1eaf05df0efa64c7de7dbfd33d55d74

Observation 06ea400d-30db-4d2d-8789-57870e1c5112 · outbound

This paper cites HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly.

Large Language Models Can Self-Improve in Long-context Reasoning HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.611829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.611829Z digest=sha256:16d4c787d7fa010960b76ca7167b5baab394ebd70e881145ca74790a0c892083

Observation 4c2dc3f5-e634-48a4-af55-843f3f5ea1fa · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.615042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.615042Z digest=sha256:6923ea8f3a71fc5848d808dda98796660a375e2000a36f3b4d817fb47ded7c1c

Observation 105b7cf0-29dd-43b9-b7be-8532f1d9909c · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 99

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:59:12.396679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-12T21:59:11.618143Z digest=sha256:5437449964fe58cd40bda51244dfdfea35ad697da901ca9addd186b51851a75b

Observation 9b78972c-dfa9-4f06-8527-be2c5c428ded · outbound

This paper cites an unresolved cited work.

Large Language Models Can Self-Improve in Long-context Reasoning Unresolved cited work

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.621135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.621135Z digest=sha256:8d3961fc49022de478e9de2fbe9ddf5235b8eb52169c6c39bd622009ba878807

Pith citing papers

Observation b4ee988a-3f88-438a-885e-7930504f9ce1 · inbound

Does Table Source Matter? Benchmarking and Improving Multimodal Scientific Table Understanding and Reasoning cites this paper.

Does Table Source Matter? Benchmarking and Improving Multimodal Scientific Table Understanding and Reasoning Large Language Models Can Self-Improve in Long-context Reasoning

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-08-10T16:34:40.617192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-10T16:34:40.543242Z digest=sha256:fc2594b346b43ec45ae56313f0b69649c82f2e06d4ef32a0218736e77650efaa

Observation 14a587b3-fe87-4db8-9ebb-748a1039dd17 · inbound

100 Days After DeepSeek-R1: A Survey on Replication Studies and More Directions for Reasoning Language Models cites this paper.

100 Days After DeepSeek-R1: A Survey on Replication Studies and More Directions for Reasoning Language Models Large Language Models Can Self-Improve in Long-context Reasoning

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:44.229927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:44.229927Z digest=sha256:96cdecd93e1a7ea3bd5f0952586a16306aec0cc1b340c4c22bce18d449c486e2

Observation 27c72fba-56b9-40bb-9cb5-9c63d6cf06fd · inbound

CAFE: Retrieval Head-based Coarse-to-Fine Information Seeking to Enhance Multi-Document QA Capability cites this paper.

CAFE: Retrieval Head-based Coarse-to-Fine Information Seeking to Enhance Multi-Document QA Capability Large Language Models Can Self-Improve in Long-context Reasoning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T21:23:46.376793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:23:46.376793Z digest=sha256:1438bcb221067df3112fa3c1ed2efbcf5fa931ccd80cad95ba8125dbe83602b4

Observation 1725a611-9fe4-4818-aec8-d5fc573dd06a · inbound

UNIBROWSE: A Data-to-Agent Framework for Multimodal BrowseComp cites this paper.

UNIBROWSE: A Data-to-Agent Framework for Multimodal BrowseComp Large Language Models Can Self-Improve in Long-context Reasoning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-14T10:51:16.019022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:51:16.019022Z digest=sha256:865ee24fbd67f176423fae581706e9e48f347a109561ec4c450f85699a910e29