Pith. sign in

Paper Citation Record · LEDGER

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization

As of 20 August 2026, this Paper Citation Record lists 94 of 94 outbound references and 1 inbound Pith citation observation for arXiv:2505.11166.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.11166 v3

Coverage vector

measured 94 of 94 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:06:21.682241Z

measured 95 of 95 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T06:07:36.830550Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T06:11:20.403487Z

Reference resolution

94 of 94 outbound references displayed

  • verified exact4
  • verified fuzzy19
  • unresolved69
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c4f53665-b4c7-46ac-85a3-0e29ffabc915 · outbound

This paper cites A general theoretical paradigm to understand learning from human preferences.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization A general theoretical paradigm to understand learning from human preferences

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.287572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.287572Z digest=sha256:05960337f02dbbb576b09b53fa94a6cf7634b84c4590683003c7b510a8e9b1ae

Observation 50f25e68-8a31-40ed-88f7-f69ee8f07547 · outbound

This paper cites Unifying cross-lingual summarization and machine translation with compression rate.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Unifying cross-lingual summarization and machine translation with compression rate

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.297575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.297575Z digest=sha256:5ead8871ecfb41565765c48e70fa0590d9bb20a5041ed1911f7f88ffb66d0840

Observation 321dcabc-c338-412e-9745-145c6390b12d · outbound

This paper cites CItruS: Chunked instruction-aware state eviction for long sequence modeling.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization CItruS: Chunked instruction-aware state eviction for long sequence modeling

Reference 3

Resolution
verified exact
doi, observed 2026-08-15T21:06:21.931963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.308569Z digest=sha256:6edb5bc5703b9495e94cc94d753585f7bc2c5eba7853aafe244a53f43f3c4fb5

Observation 11500c01-8f75-4510-b022-87f70d5f7911 · outbound

This paper cites LongAlign: A recipe for long context alignment of large language models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LongAlign: A recipe for long context alignment of large language models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.313134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.313134Z digest=sha256:364f08f3c37dc2c871da3ebbbacf9b5e9e331748360b073c485ff6d348ebe2a8

Observation 66a26a42-e044-4bf9-9533-5cd68f07d188 · outbound

This paper cites LongBench: A bilingual, multitask benchmark for long context understanding.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LongBench: A bilingual, multitask benchmark for long context understanding

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.317522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.317522Z digest=sha256:91af6cf06606d6e32d2411be494d0ab60869019c4e127ac648dace0c82377835

Observation ea7e55a8-00ce-4a44-96c6-3cda48e65be6 · outbound

This paper cites LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.321988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.321988Z digest=sha256:e020a3530f5f23e695769a542429dffc62e8f3c520c95b636e26bb0e9184926a

Observation 56d46e58-0eb1-462e-99bd-7a7d72042d52 · outbound

This paper cites Longwriter: Unleashing 10,000+ word generation from long context LLMs.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Longwriter: Unleashing 10,000+ word generation from long context LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.326027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.326027Z digest=sha256:42a6ae959a385be00e05cbc86863372c150a764326c3266a8b52a63a54b3dfcc

Observation 51b86b48-fcf2-4f56-835e-0b333994341a · outbound

This paper cites Luna: A lightweight evaluation model to catch language model hallucinations with high accuracy and low cost.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Luna: A lightweight evaluation model to catch language model hallucinations with high accuracy and low cost

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.329421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.329421Z digest=sha256:cab09f09edbac4b4a1497e3c6fcc716792830b4736cd23a82c31cba65777199c

Observation 3ab93836-3b06-485c-b764-c8e80f7dc7b5 · outbound

This paper cites Context-DPO: Aligning Language Models for Context-Faithfulness.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Context-DPO: Aligning Language Models for Context-Faithfulness

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.333355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.333355Z digest=sha256:7b061df25aa13cda2b325c9e2ddf77e448ef37bef53b450d7311861ca8eded8e

Observation 59aa538a-0ed0-4da0-8ece-95846396d7f5 · outbound

This paper cites an unresolved cited work.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.337415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.337415Z digest=sha256:e855a949d1107a127440464b80f97abc310391a9a1020aa90e0c39f576a74bf1

Observation b8869fc9-9c52-484c-9693-7fbcd9f45470 · outbound

This paper cites LongPO: Long context self-evolution of large language models through short-to-long preference optimization.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LongPO: Long context self-evolution of large language models through short-to-long preference optimization

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.340734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.340734Z digest=sha256:9563de3ef3ff0e490b9c179e50b06f23604e0e4106fd4262b63c28a3d4ac5739

Observation a4ca1c02-8e36-429a-9e58-8d740d7c2e31 · outbound

This paper cites LongloRA: Efficient fine-tuning of long-context large language models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LongloRA: Efficient fine-tuning of long-context large language models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.344025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.344025Z digest=sha256:2eb6f77c1c5b337c609aaa5f6896f7ed87480d63f1de6eacd72b51f9874be1f7

Observation 0d5d4a65-3291-41c8-ab84-6cc638913e68 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.347826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.347826Z digest=sha256:2b955b9f2075742ece14f847f320764b9679b3d9b2f8ff5d3135d8e47e4ad156

Observation 2481da52-7be7-484e-b2ff-a47dec067b2f · outbound

This paper cites Smith, and Matt Gardner.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Smith, and Matt Gardner

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.351982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.351982Z digest=sha256:4daeb2a1b0dff307359bd4829f63104379ee898a77070bcaedd59da3692c98be

Observation 9e4a7a43-0673-4a1a-8247-121571271f8a · outbound

This paper cites A Survey on Long Text Modeling with Transformers.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization A Survey on Long Text Modeling with Transformers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.355585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.355585Z digest=sha256:18cbda47627486d2228b64fa6db96f834627833c4aee381b078e023cbdd9e5c0

Observation fade82b0-368b-4e18-af54-4af488960c82 · outbound

This paper cites LongReD: Mitigating Short-Text Degradation of Long-Context Large Language Models via Restoration Distillation.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LongReD: Mitigating Short-Text Degradation of Long-Context Large Language Models via Restoration Distillation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.360895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.360895Z digest=sha256:6904c6261d3926881ba97ee7999e9635784fe885965feeea2a10df160d426c23

Observation 1feec261-6e05-4325-b823-57e9383eb807 · outbound

This paper cites The Llama 3 Herd of Models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization The Llama 3 Herd of Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.365652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.365652Z digest=sha256:2ede483fbe70e03a0ff2b914159df543ea16f547d864541130eeb9e4285ba1d3

Observation e0edc79e-c9b6-4a07-b186-bcba7152b21c · outbound

This paper cites What is wrong with perplexity for long-context language modeling? InThe Thirteenth International Conference on Learning Representations, 2025.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization What is wrong with perplexity for long-context language modeling? InThe Thirteenth International Conference on Learning Representations, 2025

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.369420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.369420Z digest=sha256:2673cf080cdd8019e906acf43426f1f5fed82d0adca87c558a8b962360adab4b

Observation a27871c9-4476-414f-8f1b-05dfe229e00c · outbound

This paper cites Open llm leader- board v2.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Open llm leader- board v2

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.373009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.373009Z digest=sha256:b7e94946b6e559799c724b2e8a89e415eb1dad15889ef5a10dbc37351c36b13b

Observation 35d1c77e-a666-48e7-b37e-0c3d58aa3571 · outbound

This paper cites Data engineering for scaling language models to 128k context.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Data engineering for scaling language models to 128k context

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.376538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.376538Z digest=sha256:74708fe23e5063a88cc46b3600e74425e0b8c638a846a9ac2e1a0069e293683e

Observation b7a683b8-e35a-4cd3-9387-fd7db4173c3b · outbound

This paper cites How to train long-context language models (effectively).

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization How to train long-context language models (effectively)

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.380547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.380547Z digest=sha256:c756f89295266f74664329d83fb4d358e4da30b16102f09a09f6c5cfe4901633

Observation 52c67a86-e965-4a58-abf8-9c6271cdc260 · outbound

This paper cites How to train long-context language models (effectively), 2025.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization How to train long-context language models (effectively), 2025

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.384297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.384297Z digest=sha256:32858ca1d48b06f9159a55c9cc83193ecc79310e75c3b8e98b6820894f00287f

Observation 065e54ff-13f2-4898-ae24-09197a77bbea · outbound

This paper cites Mamba: Linear-time sequence modeling with selective state spaces.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Mamba: Linear-time sequence modeling with selective state spaces

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.387800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.387800Z digest=sha256:d324e344deb851b979eb0dac788b9f57fdabc99608489324172a6e7d351f5768

Observation a26a5ab3-843c-425d-acc5-73433f8835de · outbound

This paper cites Efficiently modeling long sequences with structured state spaces.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Efficiently modeling long sequences with structured state spaces

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.817391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.391260Z digest=sha256:6440ce258dcbac0d56185ac84c0d782cb798ec83c24aa33ca518fa3798432441

Observation e9a5b1c8-669d-4cec-93d2-f760c180d162 · outbound

This paper cites Two stones hit one bird: bilevel positional encoding for better length extrapolation.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Two stones hit one bird: bilevel positional encoding for better length extrapolation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.804718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.394973Z digest=sha256:c9c3a3649d977c5fc5162ff24cbb243e0bc5eefd90bcf5a576af7c8e6dc2cd0d

Observation 3997cc7e-3d58-4c9e-bf2f-796c8a1d61fb · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Measuring Mathematical Problem Solving With the MATH Dataset

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.398494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.398494Z digest=sha256:03fb832b3944186dc78ea37a37ff99eee374a933ee6018dfd6f8317b9b778f93

Observation fadf7f49-889f-474b-ba40-67b0cb1fac51 · outbound

This paper cites Improving long context document-level machine translation.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Improving long context document-level machine translation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.402718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.402718Z digest=sha256:c942710fb818f2f21c49d9fa7140298123f69227ba96c4e8b5d11f9cef18bb8e

Observation 2efbadcd-b722-42da-a661-31e8210c58c5 · outbound

This paper cites Constructing a multi-hop QA dataset for comprehensive evaluation of reasoning steps.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Constructing a multi-hop QA dataset for comprehensive evaluation of reasoning steps

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.406422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.406422Z digest=sha256:04dfbbb677aa8f36e91119588af942380668a6499a449442894f7c378b20e5c7

Observation bab910bd-d6d3-403a-a56c-6e6cbdab86a0 · outbound

This paper cites ORPO: Monolithic preference optimization without reference model.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization ORPO: Monolithic preference optimization without reference model

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.409833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.409833Z digest=sha256:3489e51582607f3a495fcb657824626deb003f0bff714cbf9311c7b85d5f0581

Observation 0d0a6897-d98d-4b99-b3fc-ca98eedb7816 · outbound

This paper cites RULER: What’s the real context size of your long-context language models? InFirst Conference on Language Modeling, 2024.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization RULER: What’s the real context size of your long-context language models? InFirst Conference on Language Modeling, 2024

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.413167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.413167Z digest=sha256:71f7c33d61a0415aac0140dfcf30352d2b6b890ba087eb841f0ebb0fe7a00ea1

Observation 79ac3bf2-a78e-46ff-8aaa-f542247aebf4 · outbound

This paper cites Fewer is more: Boosting math reasoning with reinforced context pruning.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Fewer is more: Boosting math reasoning with reinforced context pruning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.416410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.416410Z digest=sha256:9128270f8993a7cc4e7bd01678cd17edb9ca2cb6d4f73e574652789fa39b7276

Observation 919931ce-cf80-4df9-be3c-625b4596fcba · outbound

This paper cites DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.419979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.419979Z digest=sha256:d08481450a9bb1c37e4e572f5ed918d8f31999a82453be80fedb0860aee25b1e

Observation 74afce17-8d91-4ff3-9c03-1477b9532a83 · outbound

This paper cites Mistral 7B.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Mistral 7B

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.423777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.423777Z digest=sha256:1d5d2809fd45d748de7f402562a1eb3edfbf379d8bd219810410984edff41449

Observation 10423ee0-da8a-469d-9caa-44979560b076 · outbound

This paper cites The NarrativeQA reading comprehension challenge.Transactions of the Association for Computational Linguistics, 6:317–328, 2018.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization The NarrativeQA reading comprehension challenge.Transactions of the Association for Computational Linguistics, 6:317–328, 2018

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.427797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.427797Z digest=sha256:3019eed749e8950cdc14049153fac07afd57117bf19932dac62514cb2fbb6293

Observation 251ae323-563d-4083-9c3a-e87e54e9b168 · outbound

This paper cites Babilong: Testing the limits of llms with long context reasoning-in-a-haystack.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Babilong: Testing the limits of llms with long context reasoning-in-a-haystack

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.786196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.432249Z digest=sha256:23a7ee6a7d6e4191b414984d0589c0305d35e0e4be05a70e296e6e0d27886c28

Observation 09f00ea0-30d3-401b-ac34-e5cc7d3f1e70 · outbound

This paper cites Bart: Denoising sequence-to-sequence pre-training for natural lan- guage generation, translation, and comprehension.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Bart: Denoising sequence-to-sequence pre-training for natural lan- guage generation, translation, and comprehension

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.773554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.436524Z digest=sha256:fd68a06d17c2fd0b9965a4d24988446fb7022906bccd16f82147737892c21b14

Observation 1b5592b0-42f7-42a1-8e89-6fde66ce2232 · outbound

This paper cites Fundamental capabilities and applications of large language models: A survey.ACM Comput.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Fundamental capabilities and applications of large language models: A survey.ACM Comput

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.440458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.440458Z digest=sha256:1162df74261e70d3529127343e8ecd1f1b89f9b737f08dd9ec5514f98e3af49f

Observation 625b7171-e13e-4332-9566-8c078399fdcb · outbound

This paper cites PSPO*: An Effective Process-supervised Policy Optimization for Reasoning Alignment.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization PSPO*: An Effective Process-supervised Policy Optimization for Reasoning Alignment

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-15T21:06:22.181532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.444253Z digest=sha256:5695954e5d7973b788f224e36e299de11ab0b8f89b44eb28fcd4df47c4811fa6

Observation 563bd622-0fbf-4fe0-9c15-92456885cb30 · outbound

This paper cites Making long-context language models better multi-hop reasoners.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Making long-context language models better multi-hop reasoners

Reference 39

Resolution
verified exact
doi, observed 2026-08-15T21:06:21.861204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.447870Z digest=sha256:91395ac500564a755b278dc9a39f1e15fdeb555c0021030a4fc87bf2768d024e

Observation 0bfd3b9f-87cb-4c75-84fe-eaff1d4905ca · outbound

This paper cites Compressing context to enhance inference efficiency of large language models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Compressing context to enhance inference efficiency of large language models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.451694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.451694Z digest=sha256:7f2cf0ea20d2606a2f793c5fae7b9fe07b9acaecdd9c3167ad3dbcc66ff5ac2c

Observation ff924a2d-3838-4d0e-8fbd-5e4238a6f418 · outbound

This paper cites SnapKV: LLM knows what you are looking for before generation.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization SnapKV: LLM knows what you are looking for before generation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.455879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.455879Z digest=sha256:06aec0211e694e2cb8fbe4295ea82a393731c5c7c692b6140982591425193d0a

Observation ea3b987e-acdf-4b55-ac8e-e83dbb7f2fc8 · outbound

This paper cites MDCure: A Scalable Pipeline for Multi-Document Instruction-Following.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization MDCure: A Scalable Pipeline for Multi-Document Instruction-Following

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-15T21:06:21.842018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.459722Z digest=sha256:09b6fd6e0bd5938d7d756b15724158c749f7fcb1127621b83902adc52bc652e1

Observation b63d8b18-e504-41a8-ac88-5dde435dd145 · outbound

This paper cites A comprehensive survey on long context language modeling.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization A comprehensive survey on long context language modeling

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.751632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.463912Z digest=sha256:fa896f3f68d544f21a42123975e88de29d1a892bcd06d4b8ec6459d51dae0f63

Observation 97ef50cd-29cf-4fc6-9e73-e8069ca1d66c · outbound

This paper cites Decoupled weight decay regularization.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Decoupled weight decay regularization

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.467936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.467936Z digest=sha256:72e1e8f67c1208636a518b309b2912f747c0b4734f79f8308a596778e991cb6c

Observation 6ff09e3e-5d2c-4e24-a198-efcd65399a6a · outbound

This paper cites MoBA: Mixture of Block Attention for Long-Context LLMs.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization MoBA: Mixture of Block Attention for Long-Context LLMs

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.471878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.471878Z digest=sha256:ea423067e1ba0238ad44bccd51edbcab4509db62c8365b7c73b81b044aa8a72d

Observation bb756b43-6729-4e29-a0ac-6c0d4e32f8a2 · outbound

This paper cites When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.476220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.476220Z digest=sha256:449fddc7afb6ad4ebd0a37cc4066b9db3cefffbc9e5dadd8432fbd413339e130

Observation 3beb0357-2152-4fb4-91ac-edfd83b20d49 · outbound

This paper cites SimPO: Simple preference optimization with a reference-free reward.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization SimPO: Simple preference optimization with a reference-free reward

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.731953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.480122Z digest=sha256:b8aca7d480fb818efced1864eff2b78601ee73362908e1f9eb62e95a36b238e4

Observation 67558fb0-8f55-41cf-b6ce-77d4553746c4 · outbound

This paper cites Training language models to follow instructions with human feedback.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Training language models to follow instructions with human feedback

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.718328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.483717Z digest=sha256:3c0983b0961c13a5d16fd1ac2ccd656d4f0bfcd7954c7ee2e6629824a2d9d66a

Observation d8ca97d2-d082-4236-9627-0fa09848f9f9 · outbound

This paper cites Vicky Zhao, Lili Qiu, and Dongmei Zhang.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Vicky Zhao, Lili Qiu, and Dongmei Zhang

Reference 49

Resolution
malformed identifier
no resolver link, observed 2026-08-15T21:06:21.488823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.488823Z digest=sha256:eb03ada1a986f789580734d75838f85990eeef0fe9328ee922727c614ff26057

Observation 4a933ae7-cec9-4b42-a662-a2010e2ce6f0 · outbound

This paper cites YaRN: Efficient context window extension of large language models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization YaRN: Efficient context window extension of large language models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.493201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.493201Z digest=sha256:c04b0ea3a35c6f1fc814f51105b78b1b90d48cc0809d89fc18b39894dd834992

Observation a8154c0d-8865-4710-9146-983fe574b3c5 · outbound

This paper cites Handling Very Long Contexts in Neural Machine Translation: a Survey.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Handling Very Long Contexts in Neural Machine Translation: a Survey

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.688508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.502109Z digest=sha256:f1e870d323c17f928b8ef143c43547a7862bc51839fc055f5af0210d135428d8

Observation de4410f3-6d63-4499-853f-f7b6bf1b1f63 · outbound

This paper cites Infobatch: Lossless training speed up by unbiased dynamic data pruning.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Infobatch: Lossless training speed up by unbiased dynamic data pruning

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.676606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.506022Z digest=sha256:f17834d1299ddf6fc3986e5ce74118cef376e8e472b27f4a1bb1ee54fa54b383

Observation 28999005-27d5-4b0a-a3aa-3dc4b198dfa5 · outbound

This paper cites Direct Preference Optimization: Your Language Model is Secretly a Reward Model.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.510415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.510415Z digest=sha256:0b64c07e6bc3ddb389e08b2939eff70e2554854e854851e16e8e273152db22da

Observation ef0fe2ee-8f47-47e9-b086-1b3a71c86f96 · outbound

This paper cites ZeRO: Memory Optimizations Toward Training Trillion Parameter Models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization ZeRO: Memory Optimizations Toward Training Trillion Parameter Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.522867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.522867Z digest=sha256:01ca1de10681848dcc51dcc1e2ee3c11695236662ac725974b50ea7e29d2da86

Observation ae96f388-7aa7-476b-bf48-e23d1b7b8143 · outbound

This paper cites SQuAD: 100,000+ questions for machine comprehension of text.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization SQuAD: 100,000+ questions for machine comprehension of text

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.527301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.527301Z digest=sha256:2b270feb9e9d7b1b75e4c5ba9750b17e204b19bb05417e715307d9a2e7316b6c

Observation ae3ae815-e323-4bbb-86b5-782aca30583f · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.532177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.532177Z digest=sha256:c2d9c20c1fa4cf441aaa41178b41b308636f6ecbf2debedc0d1997ad684ed338

Observation fecc4298-f24e-4cee-9d61-3d1146fd47a0 · outbound

This paper cites an unresolved cited work.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:06:22.658195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.537001Z digest=sha256:27bd889943bb40d4efb803de3b7dfe44acf127ba823d86bd53b3da140a86058b

Observation bf4f98c5-4af6-4ebd-ad01-b5072f34c0e3 · outbound

This paper cites Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.541137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.541137Z digest=sha256:e8c8aa50b71e6340fa30037d77b815ea1422133ff34248e332558a9dfaf0e0f0

Observation 8bc4674d-699f-4116-936b-361c00f14c1d · outbound

This paper cites Generalized preference optimization: A unified approach to offline alignment.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Generalized preference optimization: A unified approach to offline alignment

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.647333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.545856Z digest=sha256:756e27167c67b0c7a6b84fbc5d62e9579505c0936060682c2ddd826c95ed0a46

Observation 371e2cfc-b42a-4ba8-96a3-e5a0ee3f2b0b · outbound

This paper cites LOGO -- Long cOntext aliGnment via efficient preference Optimization.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LOGO -- Long cOntext aliGnment via efficient preference Optimization

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.549930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.549930Z digest=sha256:0973c8a78da01dc9c748f43444f8a1661262e4804a11039b86f6f04ccce1643c

Observation b2108d74-9dfe-4810-823b-e5b26ba5db3b · outbound

This paper cites Musique: Multihop questions via single-hop question composition.Transactions of the Association for Computational Linguistics, 10:539–554, 2022.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Musique: Multihop questions via single-hop question composition.Transactions of the Association for Computational Linguistics, 10:539–554, 2022

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.553894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.553894Z digest=sha256:c05d64dc73cdb113edf482944ef52ccc9a65c136a6dccc03713d241b38c26ba9

Observation 43fc8ca2-f7c8-442f-825a-1c42138e4f9e · outbound

This paper cites Leave no document behind: Benchmarking long-context LLMs with extended multi-doc QA.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Leave no document behind: Benchmarking long-context LLMs with extended multi-doc QA

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.557707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.557707Z digest=sha256:7fd50f23093cf026a8ece7eb4093db00ca3457bfea91c1892915b5063d302044

Observation 443c3306-3e0a-4667-acf9-f9919d2bc10b · outbound

This paper cites MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.561955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.561955Z digest=sha256:85dc434d7ed2d0026cc3ab92285fe7de86b319d1ce6676ca7c14fa3cd512cbd7

Observation 8399af25-98bb-4eb4-ae1d-1f4eca0b9a84 · outbound

This paper cites Chi, Quoc V Le, and Denny Zhou.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Chi, Quoc V Le, and Denny Zhou

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.628552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.565886Z digest=sha256:e87be055bfedfd1d548b95381a59cd48cb8e1466eec2d1e24dbf5b33d1f09c31

Observation 6059e6b7-97e0-403e-a987-fe68453342a2 · outbound

This paper cites Kullback–Leibler divergence.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Kullback–Leibler divergence

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.615703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.569471Z digest=sha256:f3774a2cd2705ebdac46b6eb55ebf6addabc3d34adc42365e8bf3973652e6c91

Observation 49b05a65-9a5d-454d-8bd5-7e924200befe · outbound

This paper cites Wit and Marie Gillette.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Wit and Marie Gillette

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.604784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.573105Z digest=sha256:8fe78d0f425aae7827c52365011494323069278ceeabc7dc34b8e91338cdd45e

Observation 505c72a8-e97f-420f-84dc-ec67e3eae96f · outbound

This paper cites An efficient recipe for long context extension via middle- focused positional encoding.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization An efficient recipe for long context extension via middle- focused positional encoding

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.594373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.576890Z digest=sha256:9e5bcbeb7c61ba9d03f7a85027d71dba78e53d1fc7a2aff280b701eb2f8b7602

Observation 414a2203-ee51-4946-b2e3-9f5bb1264a36 · outbound

This paper cites Effective long-context scaling of foundation models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Effective long-context scaling of foundation models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.580733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.580733Z digest=sha256:b201dda025cf16b77a58d162fe2e790cf595f00602f2c36674d1309d64592047

Observation 8cce9127-4172-4fea-83e7-ff025a3d6794 · outbound

This paper cites Large language models for generative information extraction: A survey.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Large language models for generative information extraction: A survey

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.584357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.584357Z digest=sha256:a3232ba79ee0889a00bf7eb88206c3c2f4ee942078539f1c26555c8edd807197

Observation ef386735-cb16-414a-9306-7b121d41d857 · outbound

This paper cites RECOMP: Improving retrieval-augmented LMs with con- text compression and selective augmentation.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization RECOMP: Improving retrieval-augmented LMs with con- text compression and selective augmentation

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.568096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.587957Z digest=sha256:a0f2081e8b1764bd7bae68b491dba5a348e0086f8474b7b57ce86615507763d8

Observation 0832a35e-eece-4803-b931-2251f5ef98d8 · outbound

This paper cites LongFaith: Enhancing Long-Context Reasoning in LLMs with Faithful Synthetic Data.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LongFaith: Enhancing Long-Context Reasoning in LLMs with Faithful Synthetic Data

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.591936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.591936Z digest=sha256:a9ab527eae656d2e11fdb4f325f91de11c4256e9047e5d66d5da13e8ba82e970

Observation 1a79735a-77f7-49a6-9ef7-68c09f2c17a0 · outbound

This paper cites Qwen2.5 Technical Report.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Qwen2.5 Technical Report

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.600241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.600241Z digest=sha256:2382a59f9605a67764f8682d74dd1ffd0e1cf7769a509045e0690982ce7b2740

Observation ccc194f9-4edd-4fe2-9940-e9dd6ed15fcb · outbound

This paper cites Mindllm: Lightweight large language model pre-training, evaluation and domain application.AI Open, 5: 1–26, 2024.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Mindllm: Lightweight large language model pre-training, evaluation and domain application.AI Open, 5: 1–26, 2024

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.556789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.604178Z digest=sha256:100f3e94fcb9bfb71aae79c1e23092ddf7d66bcde7dfaac8a7c269b865162c00

Observation 35f54b63-142f-4928-b67a-47f4b80787f3 · outbound

This paper cites an unresolved cited work.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Unresolved cited work

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.608200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.608200Z digest=sha256:feffca4ccd433c1cc5a9f41dab60d6b9f76cded7cc4fee945ae0d72c64a5a1b4

Observation 5384a0cb-441d-4438-94b5-26ccc5be58c8 · outbound

This paper cites Longcite: Enabling LLMs to generate fine-grained citations in long-context QA,.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Longcite: Enabling LLMs to generate fine-grained citations in long-context QA,

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.546000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.611974Z digest=sha256:2d3acfd0a3d1ad7d6fe0d1a1f08ebef5b1919a922b020e64300c16be8cdb7b6a

Observation 44de6843-0d4c-4884-b833-33b6f39c6000 · outbound

This paper cites LongReward: Improving Long-context Large Language Models with AI Feedback.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LongReward: Improving Long-context Large Language Models with AI Feedback

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.620099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.620099Z digest=sha256:db38fd1108a45a3e87ccb5622feea6376f339233bc3704ebe1f15689d03186d6

Observation 75363c18-7a12-48e7-97fc-8323d8c71b60 · outbound

This paper cites Extending Llama-3's Context Ten-Fold Overnight.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Extending Llama-3's Context Ten-Fold Overnight

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.625023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.625023Z digest=sha256:80036658c05e582cb062d1ec67f08f2063a8b898ff922d35b6b8a08e140f6f5e

Observation 23f5ddc1-134a-4958-8589-54d81ba8e1a5 · outbound

This paper cites IOPO: Empowering LLMs with Complex Instruction Following via Input-Output Preference Optimization.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization IOPO: Empowering LLMs with Complex Instruction Following via Input-Output Preference Optimization

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.629057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.629057Z digest=sha256:76b8f14f160b47963181c07555b67c5c02f194018b7322a098aa7fe5d8e05b7b

Observation 6b168475-3b2c-4c08-9593-503be193790f · outbound

This paper cites LONGA- GENT: Achieving question answering for 128k-token-long documents through multi-agent collaboration.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LONGA- GENT: Achieving question answering for 128k-token-long documents through multi-agent collaboration

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.633192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.633192Z digest=sha256:63fd96632d24eca32c6b887367311ea322d0482e471aaf308a1b5d4f3c242109

Observation 98b2e85e-9240-4041-a61b-dc25b16a9a96 · outbound

This paper cites an unresolved cited work.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Unresolved cited work

Reference 80

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:06:22.534771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.616025Z digest=sha256:6c5651ec1554f1b7a58c717dc536334c4c6aeb0ad1202c87a7ac545ae0a8ed0c

Observation 817db5e0-9104-480d-8ff4-f2ed4f5e0d81 · outbound

This paper cites LlamaFactory: Unified effi- cient fine-tuning of 100+ language models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LlamaFactory: Unified effi- cient fine-tuning of 100+ language models

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.641958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.641958Z digest=sha256:5caeef33509c44f6ec13a91a0374f5134f97452c5335975bc0b1c764cf82b999

Observation 568bcd0e-55f2-4f65-9c02-a17f075660b8 · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Instruction-Following Evaluation for Large Language Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.646243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.646243Z digest=sha256:bf5e7177a7857c781d66ba18ab8043c5eb963621147012ff741651bb8a90e50f

Observation bf6342b9-9d05-45ab-901d-b49bee060473 · outbound

This paper cites PoSE: Efficient con- text window extension of LLMs via positional skip-wise training.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization PoSE: Efficient con- text window extension of LLMs via positional skip-wise training

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.523288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.650411Z digest=sha256:cae5810638d4f822707b1340b604e4edd69e21ffc4d2d8376a8f85919cbc66ec

Observation 08d217a7-7671-4793-ac30-d7708bc2385b · outbound

This paper cites Chain-of-Thought Matters: Improving Long-Context Language Models with Reasoning Path Supervision.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Chain-of-Thought Matters: Improving Long-Context Language Models with Reasoning Path Supervision

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.653955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.653955Z digest=sha256:7a2c5f51e04cbd8ccbbe0c348bc38a16ce0b9d2151f56e0f018317aba24430bf

Observation 9f678ab7-fc80-4070-93ff-4800f562e6a4 · outbound

This paper cites SLiC-HF: Sequence Likelihood Calibration with Human Feedback.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization SLiC-HF: Sequence Likelihood Calibration with Human Feedback

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.637749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.637749Z digest=sha256:51a1b002c7c0e29ca1d8dacb863c0330e684d96e6e7918104be0542279e913e3

Observation 4150fec0-92d4-4e75-aa66-4825d206d802 · outbound

This paper cites Generalizing From Short to Long: Effective Data Synthesis for Long-Context Instruction Tuning.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Generalizing From Short to Long: Effective Data Synthesis for Long-Context Instruction Tuning

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.668121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.668121Z digest=sha256:6cb1a88f9b80091b9cb75660712f1f9587aa2b4a39c98c096f92dc66dece8d5e

Observation 38dae01b-6055-48ce-8958-1fe75512ae93 · outbound

This paper cites Generalizing From Short to Long: Effective Data Synthesis for Long-Context Instruction Tuning.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Generalizing From Short to Long: Effective Data Synthesis for Long-Context Instruction Tuning

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.658781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.658781Z digest=sha256:efcc0789394870ce28eb510e8a1c723934b9f3cc459493ba429acff63991cd76

Observation 23949f5e-f084-4d8b-a964-69886c64b526 · outbound

This paper cites an unresolved cited work.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Unresolved cited work

Reference 93

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:06:22.512210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.672749Z digest=sha256:e61acae0d619ccdd5a748b404b6343da2cb2e04844563a046db75dfbcdcf1c74

Observation b188141b-00ba-4189-99af-e8a3fb7bd7e4 · outbound

This paper cites The Collegian.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization The Collegian

Reference 94

Resolution
malformed identifier
raw_fallback, observed 2026-08-15T21:06:22.019293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.676766Z digest=sha256:3d6de3ed225f9a2b567809b23a22ba6fef07766744d724716ab73c2c472cc64d

Observation 2b25102c-228a-481c-b64a-cecaef3a30c6 · outbound

This paper cites In addition, we further examine the theoretical validity of the first motivation from adata-sampling perspective.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization In addition, we further examine the theoretical validity of the first motivation from adata-sampling perspective

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.499465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:06:21.682241Z digest=sha256:f3d465aac1b997e802bba371a7ca5b6b5ac49b1f5e217e38ffedda248879fce6

Observation ee0517ce-950d-4a4b-8084-6f15a5b410b5 · outbound

This paper cites an unresolved cited work.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Unresolved cited work

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.304223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.304223Z digest=sha256:298588e9e0dc1a4ceb349440de9d5055b3252037dcd6b2e1b4a40b6be52479ae

Observation 64d86e4e-55c4-40f4-946d-c82fdcce0c1d · outbound

This paper cites an unresolved cited work.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Unresolved cited work

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.514589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.514589Z digest=sha256:af8630206d725d31b4deaf65e68898900832e657129a36b89bd4a52e85686623

Observation 15c472a3-c4d3-4cdb-b69e-181134c4aa87 · outbound

This paper cites an unresolved cited work.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Unresolved cited work

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.497385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.497385Z digest=sha256:e500d378919b4dd99ef2c9dadf84ea076e2fd991ece793eb6bc21d539b6ec175

Observation 23a3064a-7101-46e7-a0c4-d54932ef993f · outbound

This paper cites LongFaith: Enhancing Long-Context Reasoning in LLMs with Faithful Synthetic Data.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LongFaith: Enhancing Long-Context Reasoning in LLMs with Faithful Synthetic Data

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.595959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.595959Z digest=sha256:d20546d28b7385b7307bb71ff2f44e0a2f1750e8156bfc0d4a0f0267fcd4ae58

Pith citing papers

Observation 401592e4-5322-4dd6-81c0-5d0a9dcb878f · inbound

OPSDL: On-Policy Self-Distillation for Long-Context Language Models cites this paper.

OPSDL: On-Policy Self-Distillation for Long-Context Language Models SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-04T02:07:03.503226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T06:07:36.830550Z digest=sha256:6a18294a799d58ead09fd93f0e26ec3c272d79cc01f8850d4b80629d6ec6e860