Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-11T11:50:26.030339Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 83 of 83 outbound references and 0 inbound Pith citation observations for arXiv:2605.22567.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-11T11:50:26.030339Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
83 of 83 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 5bae1058-3cc9-4b12-94b0-af4fa8223685 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Advances in neural information processing systems , volume=
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3bf6452e-08eb-467b-b33a-0cd7d666021a · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6ee4a7c3-c25b-4d02-9cd6-d64d2ba24e7d · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Linjuan Wu, Hao-Ran Wei, Jialong Tang, Shuang Luo, Baosong Yang, Fei Huang, Yongliang Shen, Weiming Lu
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b8dbf2a1-6ea5-439a-be72-65acad0a72cb · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance author Feng, Y
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d1e0b85b-49dd-4146-91c6-acddfbf6890b · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance arXiv preprint arXiv:2603.04597 , year=
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2b6cd2a7-4e11-41c1-bcfc-6ec30de4e053 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Language-Specific Layer Matters: Efficient Multilingual Enhancement for Large Vision-Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9335aab5-fd03-402d-87e2-93709a5939e1 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance SLAM : Towards Efficient Multilingual Reasoning via Selective Language Alignment
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7503c952-5389-4ac6-8628-beba2da63df0 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance arXiv preprint arXiv:2603.19097 , year=
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ec387173-3331-4db6-be7a-7e341ab79c9e · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance The State and Fate of Linguistic Diversity and Inclusion in the NLP World
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9840533d-2457-4aa3-a52d-24c48e4aa7bf · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance First Conference on Language Modeling , year=
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 4a0c7831-f9bd-48cf-8e11-4ec8bc526bf0 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Proceedings of the AAAI Conference on Artificial Intelligence , author=
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 30689f68-2d9f-4fe7-981c-5de38d68dd74 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Crosslingual Reasoning through Test-Time Scaling
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3535989b-63c8-41f7-9c24-184949ddf50b · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2ef391bd-b367-4118-9c1d-a0d3a37edb50 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Language Mixing in Reasoning Language Models: Patterns, Impact, and Internal Causes
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 436603ab-cb2a-4bfa-af7d-ad303dec3d7e · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Is LLM an Overconfident Judge? Unveiling the Capabilities of LLM s in Detecting Offensive Language with Annotation Disagreement
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation cc2b17b6-010c-44c9-bdc5-147dc4987fb9 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9953abc3-9f37-45cc-ae1a-ad8445afdabe · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance arXiv preprint arXiv:2504.18428 , year=
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1a3dde83-27e3-4fc0-a356-cef691412fc4 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance MMATH : A Multilingual Benchmark for Mathematical Reasoning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 33ec0069-9e7e-4184-b0a1-941996780236 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance When Models Reason in Your Language: Controlling Thinking Language Comes at the Cost of Accuracy
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f37eed6b-c595-4646-9f36-97379e886015 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Multilingual Test-Time Scaling via Initial Thought Transfer
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f39d86af-1277-411f-bba2-62e51deb5309 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance arXiv preprint arXiv:2507.05418 , year=
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 5ec0a21a-04a2-4143-9f15-a7cfd43cd343 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance arXiv preprint arXiv:2506.05850 , year=
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9ab76d68-14be-492e-9f8a-fb775ea418be · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance arXiv preprint arXiv:2510.07300 , year=
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation eef01423-0a90-4078-ae3c-95505687db18 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance arXiv preprint arXiv:2510.02272 , year=
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 33edb356-7d2c-4636-ab08-b5cc52b0963d · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Efficient Reinforcement Finetuning via Adaptive Curriculum Learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9ab7fb25-0ce4-449e-b178-1a2e6f78d139 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Don't Tell the Answer, Truly Guide the Reasoning During RL Rollouts
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 23f61bfc-1d93-4359-bfc2-29a17a5582dd · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance F ast C u RL : Curriculum Reinforcement Learning with Stage-wise Context Scaling for Efficient Training R1-like Reasoning Models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a78a41be-ffcf-42fa-a9a9-8721c63cc08b · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance arXiv preprint arXiv:2507.13266 , year=
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 191902fe-f9b7-457f-a8ab-53c5cf24bd31 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Understanding the Repeat Curse in Large Language Models from a Feature Perspective
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 03328549-8306-432e-adc6-00a66af510bd · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance ArXiv , year=
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 8bb3ca51-c57b-4648-895e-466727ea8ea2 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d4cdce2b-625d-4c17-bdfb-ace2231e4964 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 85c87d47-dcf2-477d-bfef-63f09b9b67a1 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6d10f557-afb1-4da6-84ed-02ee772bdb8d · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Notion Blog , year=
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3e96a67c-9130-412f-8d9a-fe85ecd039a3 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 15dd5c13-112e-4a4c-9f75-4b4d98e09169 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance The Eleventh International Conference on Learning Representations
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 38d0d04d-794f-4238-b296-0219924fbefb · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance AIME 2025 , url =
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 262b4b21-8ef4-4131-9c4d-a50118df3874 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance AIME 2024 , url =
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 163a6bcd-ac58-4ea1-b4fc-cf4fa01bddf8 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance The Twelfth International Conference on Learning Representations
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 5c255bab-ba08-4569-97da-1ae470e1de98 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance GPT-4 Technical Report
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f1254178-3e0a-4a06-a02c-7111d0ab9c6f · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance 2025 , url =
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 604f0662-faf6-49e7-87a6-41b31def1ce5 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance StepHint: Multi-level Stepwise Hints Enhance Reinforcement Learning to Reason
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0147729d-f7ca-43ce-9a93-407c4950d8c0 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance 2025 , eprint=
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f8877d16-bc46-4f28-92fc-c83228390b44 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance arXiv preprint arXiv:2508.11408 , year=
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b407e988-56e8-41a3-a359-0386d21965ba · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance GHPO: Adaptive Guidance for Stable and Efficient LLM Reinforcement Learning
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 24c0fa72-ffde-406c-a295-f2e6dd829cc8 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance arXiv preprint arXiv:2505.16984 , year =
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 96a80a11-f43d-46f4-a21d-eb95fc43ec92 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance SRFT: A Single-Stage Method with Supervised and Reinforcement Fine-Tuning for Reasoning
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1f647d1a-7b48-4272-b0b0-88039a0600c4 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Qwen3 Technical Report
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e241ea70-e93b-4277-9fe6-c43d77933497 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance XCOPA : A Multilingual Dataset for Causal Commonsense Reasoning
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0b3cf418-91b4-4d6c-8da8-c0fb19df7a76 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Few-shot Learning with Multilingual Language Models
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 79d7532d-6fc9-4146-ad84-e8bf14541905 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Unresolved cited work
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f28d80af-1271-49f0-a6f9-581440dd6fb9 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Generalized Slow Roll for Tensors
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c67d6216-ad29-4182-bc2c-ddc9650365db · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance MMLU-ProX: A Multilingual Benchmark for Advanced Large Language Model Evaluation
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation eac3a3e9-3c43-4588-a489-71857d31bfa1 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance 2022 , eprint=
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 80918205-7b94-4e92-8d60-4cf7297f8a1f · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance 2021 , eprint=
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9949bdc8-7595-4e9d-81b4-1bd9f298d1ad · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance 2020 , url =
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ef4c0240-c9e9-4a1d-a140-45b15ff05cac · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Is Bigger and Deeper Always Better? Probing LLaMA Across Scales and Layers
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b2b82dbb-5c6e-4f3d-967d-bae71c2c05fd · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance How do Large Language Models Handle Multilingualism?
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation fc7f0780-3e14-4efd-9804-e6de663bdc8c · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance P - MME val: A Parallel Multilingual Multitask Benchmark for Consistent Evaluation of LLM s
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 401e6c79-938b-4fab-b5d2-a292dfd0eec1 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Humanity's Last Exam
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b6338a0e-fb06-4393-b47d-af13677cf1db · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Levesque and Ernest Davis and Leora Morgenstern , editor =
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d642e2eb-c366-4b69-84dd-ce065cbe0ec9 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Gordon , title =
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c554c4b0-c53d-4307-b081-29027461c322 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Few-shot Learning with Multilingual Generative Language Models
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9e7f99fb-98a7-4555-939b-313ef401c19d · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance A Corpus and Cloze Evaluation for Deeper Understanding of Commonsense Stories
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation dc8a9e6f-825a-4f86-b01a-8c46163da85d · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance The Llama 3 Herd of Models
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d4a5b6d9-9cbb-40e4-9b29-12e08523d53f · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Light-R1: Curriculum SFT, DPO and RL for Long COT from Scratch and Beyond
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b3185cc4-cbfb-4e96-bb74-bd1616d44d53 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a222b94a-6ac9-499c-9e86-780c9595b19f · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance NumGLUE
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7b9e04f8-a09a-4c74-9532-263ab512dbd7 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Unresolved cited work
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 248f9198-6fbb-494b-afb2-e1d062060765 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance No Language Left Behind: Scaling Human-Centered Machine Translation
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation fee103aa-f327-4ab2-9981-5cbe9ded6431 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance 2025 , url=
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation aabf15fa-57fc-424d-82bb-8d6e8d0d5114 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 24bc853d-44fe-43e1-beb3-2c3fd367a9b6 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ed58e4bc-598c-434f-a64a-003bc0dca810 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance A survey on hallucination in large language models: Principles, taxonomy, challenges, and open questions
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3161205a-7f90-493b-848a-d4c213efa3aa · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Learning Fine-Grained Grounded Citations for Attributed Large Language Models , booktitle =
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9cbed096-a97c-4ca0-8c4f-2ffd1fae5da4 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Improving Contextual Faithfulness of Large Language Models via Retrieval Heads-Induced Optimization , booktitle =
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation dea12292-f86d-4ecd-806e-b39a128e56c6 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Alleviating Hallucinations from Knowledge Misalignment in Large Language Models via Selective Abstention Learning , booktitle =
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ab426810-bc68-4d26-bc79-ae328f402fe0 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Advancing Large Language Model Attribution through Self-Improving , booktitle =
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7c644739-e1ac-4c3a-be99-f4a8a5a4d1a5 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance C lue A nchor: Clue-Anchored Knowledge Reasoning Exploration and Optimization for Retrieval-Augmented Generation
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d4c905e8-5870-4fb9-bbc3-0291593e3da5 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Know More, Know Clearer: A Meta-Cognitive Framework for Knowledge Augmentation in Large Language Models
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e64c8e8c-f3be-4664-a8ca-a110edcb6ab1 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance CM -Align: Consistency-based Multilingual Alignment for Large Language Models
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d481b79a-ca56-4797-aa7a-68b596109b6e · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Multilingual Knowledge Editing with Language-Agnostic Factual Neurons
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 20d2ccfc-d17f-43a9-ac68-3682258b3bd3 · outbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Less, but Better: Efficient Multilingual Expansion for LLM s via Layer-wise Mixture-of-Experts
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
No inbound Pith citation observations are available.