Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T16:52:49.881585Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 100 of 117 outbound references and 5 inbound Pith citation observations for arXiv:2501.12766.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T16:52:49.881585Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:40:09.996269Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-15T11:17:24.563972Z
100 of 117 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b8ae1682-3c40-4687-a630-b011c5e2c9a5 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Many-Shot In-Context Learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef92d5aa-ab39-4f31-8643-62951e46d8bf · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Yi: Open foundation models by 01.ai, 2024
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec82821e-59b9-4268-a165-a7757384fe65 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents M., Neyshabur, B., and Zhai, X
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3783d258-7d66-4799-b8cb-242a19435443 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Training-Free Long-Context Scaling of Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b417ecd5-918b-45ef-a074-782e32c56f70 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Many-shot jailbreaking
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07b44178-b8f0-465e-b75f-c2976d39b72d · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3904a4a-e84d-4761-9527-a968441afa4c · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents and Lempitsky, V
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a766e8d3-1417-404c-b18c-dc00bde38027 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68ebb1a4-f872-41d9-a115-c915e5d26d6a · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b76c459-c5ed-4732-afba-1b8ff64cbaa7 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Codeplan: Repository-level coding using llms and planning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a44e6632-6b02-4616-b9d7-6a5dcfe9324d · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Smollm-corpus, 2024
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b70d8eb-f91c-48c7-a662-edec16b9cd5e · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents and LeCun, Y
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2da1e6e3-832c-4efe-8489-cb87fe6f1226 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents In-Context Learning with Long-Context Models: An In-Depth Exploration
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbcd98db-3ab5-4344-9f77-7e4fc31a0133 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3f9f1bc-6061-4d09-b14b-0967a2b00c40 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents G., Bradley, H., O’Brien, K., Hallahan, E., Khan, M
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 598f7e81-aec7-4508-83b7-ee3e75952e1a · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Piqa: Reasoning about physical commonsense in natural language
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ad28309-b0af-4b6a-a967-9a8f025aa474 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Peek across: Improving multi-document modeling via cross-document question-answering
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e88be9cc-dcf6-4e41-b2bd-2648398b65c6 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Resolving the imbalance issue in hierarchical disciplinary topic inference via llm-based data augmentation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fdc6626-4477-4d76-b0dc-b27a5289afda · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents CLEX: Continuous Length Extrapolation for Large Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7304842d-9c25-478e-8b0f-2d9e08e7aba0 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Extending Context Window of Large Language Models via Positional Interpolation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e13da2df-cfe2-4191-ae01-62b044db974f · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Incremental False Negative Detection for Contrastive Learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d71abe2e-6f35-4f4b-8360-3188ae0c8af1 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents LongLoRA: Efficient Fine-tuning of Long-Context Large Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aac55ebb-5eed-4fcc-b41b-b55ae99c9ecd · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Language Models as Science Tutors
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64fb3737-ecca-4717-8d7c-77917eebca21 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb4f1c2c-c51a-4b11-9265-d25e99141dda · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 511ee546-b07c-4fd5-82c8-e0679dd3465e · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Deepseek-v2: A strong, economical, and efficient mixture-of-experts language model, 2024
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8afd3d6d-faa3-4021-8457-4d6d2ac0588b · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Fewer truncations improve language modeling
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 108b24ad-0d6a-4d4e-90e1-15314906cff0 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Enhancing Chat Language Models by Scaling High-quality Instructional Conversations
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 762099b0-450b-4b28-a9ca-52d373fc15ba · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents The Llama 3 Herd of Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2964de3-5c80-4fef-843a-4412e1b95952 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents O., Hart, P
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88e10fbb-9019-43e8-962a-ff2634938633 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Data Engineering for Scaling Language Models to 128K Context
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fa4a2af-350f-4345-ae4f-8cece78bd100 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Quest: Query-centric Data Synthesis Approach for Long-context Scaling of Large Language Model
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1f16e7b-eaed-4cdc-a00c-9ab2404259b7 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents The Pile: An 800GB Dataset of Diverse Text for Language Modeling
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecc9ea48-3fab-483a-9850-07d03379ff00 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents How to train long-context language models (effectively)
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b9bfb5a-1d8f-48ae-aff6-c7f609c6f012 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Deep learning, volume 1
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 380909dd-b631-41ed-ac72-bc7edd35b795 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Retrieval augmented language model pre-training
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f90e41c-179f-4239-9296-8d1b5e1491c4 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents LM-Infinite: Zero-Shot Extreme Length Generalization for Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f09505fc-3b04-49c7-8867-4413fb92b92d · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Measuring Massive Multitask Language Understanding
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d4597f3-7373-4045-a5cc-3db81c221f26 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Scaling Laws for Autoregressive Generative Modeling
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a2382c3-c54c-4c74-a721-12e032519888 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents E., Osindero, S., and Teh, Y
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation def041a3-b0ad-4389-b601-a299c4a07578 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Training Compute-Optimal Large Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba128b55-5878-480b-ac9e-cbf96b1fec11 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents RULER: What's the Real Context Size of Your Long-Context Language Models?
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7edebc2f-17b2-44bb-b871-e6875b46cd1a · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Camels in a Changing Climate: Enhancing LM Adaptation with Tulu 2
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce720dd9-6067-4225-9cf1-349f28ad3a1c · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Long-Context LLMs Meet RAG: Overcoming Challenges for Long Inputs in RAG
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb940a89-586e-4315-b614-922cfaec16d4 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Llm maybe longlm: Self-extend llm context window without tuning, 2024 b
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0fd8472-b10c-436f-901a-587dcfe507f6 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Understanding the effects of language-specific class imbalance in multilingual fine-tuning
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48b86255-52ec-4763-b0cd-8ff13f1cd786 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents B., Pion, N., Weinzaepfel, P., and Larlus, D
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aaea3f9b-36d6-49b7-92dd-a7c9faf3cf45 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Scaling Laws for Neural Language Models
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1f00387-35bd-45ac-a33e-387b2d27b890 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents F., and Ramakrishnan, R
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75cbb18a-ae66-4dc5-9e8a-30e0546ebeaf · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Unresolved cited work
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0315e77-57ff-496d-afa3-af0eb05494de · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents The Stack: 3 TB of permissively licensed source code
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33455ace-2951-4073-9d23-21d5da625b67 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Crafting papers on machine learning
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67259b43-c288-4cd9-88f8-db1048e9fd6e · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents S., Yvon, F., Gall \'e , M., et al
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60eacaa4-c011-4097-bc00-9c2edf7c902f · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents The Inductive Bias of In-Context Learning: Rethinking Pretraining Example Design
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4efbd360-56d0-449b-9ed2-5ef69d0d0a80 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents u ttler, H., Lewis, M., Yih, W.-t., Rockt \
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a902c6a-73af-46a8-9f56-d24a860efe74 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Functional Interpolation for Relative Positions Improves Long Context Transformers
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d83144d2-1bf8-41ee-84a0-225e65017feb · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents From System 1 to System 2: A Survey of Reasoning Large Language Models
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fdb8ded2-0c1d-4086-aca4-63dadbac0a58 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents World Model on Million-Length Video And Language With Blockwise RingAttention
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b49441af-a9a2-41f1-aa27-c946f7615cd5 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents LogiQA: A Challenge Dataset for Machine Reading Comprehension with Logical Reasoning
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d814e29-aac4-4e04-8f28-f752da1f4228 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Decoupled Weight Decay Regularization
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3236e0b-e274-4afb-8fc9-d7a3205fac85 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Fineweb-edu, May 2024
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe6d6e35-b7e6-473a-89bb-72eb8eb2fa15 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Learning passage impacts for inverted indexes
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 873dbf4f-fad0-4c84-aed0-d1538d325e6d · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents and Liu, B
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d09e2b34-1770-448a-acd2-735c740246aa · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Introducing meta llama 3: The most capable openly available llm to date, 2024
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d19d97c5-34d2-4fca-bcce-b21f9202bb57 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents S., Carbonell, J
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7768ffa1-ad0c-49eb-a762-975947208f82 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Unresolved cited work
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46d83723-8b22-4624-baee-3717e079833a · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents and Rosenbloom, P
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad86d884-2cb2-4678-8f49-0eab00173228 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Document Expansion by Query Prediction
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2792de7c-82df-4926-a951-62f47a88274e · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents GPT-4 Technical Report
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20e8d39d-7f00-4235-87a0-b8935d354b5e · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Training language models to follow instructions with human feedback
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd3fd6e0-61eb-4b64-86ab-c2ca03e46a07 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents The LAMBADA dataset: Word prediction requiring a broad discourse context
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 111ca4d8-84d9-4467-b4a8-7058be393f07 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents OpenWebMath: An Open Dataset of High-Quality Mathematical Web Text
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4658dd33-4243-4d1e-92d6-03e97b3974ad · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents YaRN: Efficient Context Window Extension of Large Language Models
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76c77de0-f723-4676-969d-872242b9c1df · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Improving language understanding by generative pre-training
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd943460-5dce-42b8-bb44-fe9dd660b12f · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Unresolved cited work
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d48ddb90-ea9a-4d9a-9578-30827b4a637b · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Zero: Memory optimizations toward training trillion parameter models
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bedb4a9-9893-4519-ad16-5a5f5bf00d63 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Contrastive Learning with Hard Negative Samples
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a210bede-9715-4da9-92de-6fdb2ef66552 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Code Llama: Open Foundation Models for Code
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7820db58-1300-4ff6-b1aa-261ef793e1d2 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Distributionally Robust Neural Networks for Group Shifts: On the Importance of Regularization for Worst-Case Generalization
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6919cf99-db81-404c-a474-334b9235cffd · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents L., Bhagavatula, C., and Choi, Y
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4ebea3b-00c9-43ae-a172-28c1c62a66c5 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Unresolved cited work
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f36bbba1-6f0c-45f2-81be-d497e46ffee8 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents H., Sch \"a rli, N., and Zhou, D
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d8672f83-f6bf-4502-baf1-c61c0f9622df · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents In-context Pretraining: Language Modeling Beyond Document Boundaries
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d00f2de-e8e2-440a-ba4c-2f944ad14301 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents R., Hestness, J., and Dey, N
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c658b6fc-b89e-4a89-b928-bdbd76b383ac · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95e98ca3-b90c-4735-919f-e68b48da28cd · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Unraveling the Mystery of Scaling Laws: Part I
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6973dcef-bb11-43f6-a1f5-da94b8ecfd97 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Rectified rotary position embeddings
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a0576dd2-76e7-4439-a759-1d46675e5bd1 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Roformer: Enhanced transformer with rotary position embedding, 2021
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0ccd1e5f-90a2-4157-8659-4c307700c6de · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Qwen2.5: A party of foundation models, September 2024
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70cc8a88-2e52-4715-9330-6bdf65767ab9 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Untie the Knots: An Efficient Data Augmentation Strategy for Long-Context Pre-Training in Language Models
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 35e519c5-c2f7-4a48-9c4a-6a3dd00a3fad · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents LLaMA: Open and Efficient Foundation Language Models
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4617071d-0497-4297-8e80-f3535433745a · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cfe92fa-969c-4db3-a504-c19cd265f158 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Focused transformer: Contrastive training for context scaling
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e9016387-c687-4aa8-a474-8750aec6a482 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents and Hinton, G
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18cdf622-fb3e-4cee-8bbe-7ce20167c6de · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents A survey on large language model based autonomous agents
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 57957cc3-6561-4abb-b3a6-b72cac500fed · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Query-as-context Pre-training for Dense Passage Retrieval
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a58715eb-3b72-417d-8e33-021e29c0753c · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Contextual masked auto-encoder for dense passage retrieval
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5133502f-36cc-4dc9-96d0-a39b4837f685 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Less is More for Long Document Summary Evaluation by LLMs
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1960fb80-c0c4-426d-abb0-9d2023308396 · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents Efficient Streaming Language Models with Attention Sinks
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcfb9f50-dc73-4812-aaad-74c207584e1f · outbound
NExtLong: Toward Effective Long-Context Training without Long Documents M., Pham, H., Dong, X., Du, N., Liu, H., Lu, Y., Liang, P
Reference 101
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 48b4e371-f0c5-4084-92ef-69e093ce72b9 · inbound
From System 1 to System 2: A Survey of Reasoning Large Language Models NExtLong: Toward Effective Long-Context Training without Long Documents
Reference 253
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b3a1a969-801d-4be4-91d7-b2e5e67ff30b · inbound
MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent NExtLong: Toward Effective Long-Context Training without Long Documents
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b5722cbe-bd39-4a3e-ae62-3e775887f637 · inbound
MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent NExtLong: Toward Effective Long-Context Training without Long Documents
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 136bcd82-410a-4977-9243-58df5a5f1f00 · inbound
Libra: Large Chinese-based Safeguard for AI Content NExtLong: Toward Effective Long-Context Training without Long Documents
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae8b22b1-d0f9-41cb-8a37-1b0031b7fe89 · inbound
Masked Diffusion Language Models are Strong and Steerable Text-Based World Models for Agentic RL NExtLong: Toward Effective Long-Context Training without Long Documents
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.