Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-18T07:39:58.366423Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 48 inbound Pith citation observations for arXiv:2011.01060.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-18T07:39:58.366423Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T21:38:56.042227Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-11T03:17:50.841893Z
15 of 15 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b36a629a-3b0f-4c30-bbd9-ae4fa763a8e7 · outbound
Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps In Proceedings of the 1993 ACM SIGMOD International Conference on Management of Data, SIGMOD ’93, page 207–216, New York, NY , USA
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7432c0d9-f441-4f71-bf49-e932fe90b9f0 · outbound
Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps In Proceedings of the 2013 Conference on Empirical Methods in Natural Language Processing , pages 1533–1544, Seattle, Washington, USA, October
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 99b0b215-057c-451f-94e2-d0cca7acc3e1 · outbound
Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps Large-scale Simple Question Answering with Memory Networks
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9df89687-67e3-4778-8798-3decd40b9cc9 · outbound
Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8b5b711c-46c3-41b4-946f-a7fd4c879c23 · outbound
Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps HybridQA: A Dataset of Multi-Hop Question Answering over Tabular and Textual Data
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 552106ad-af2d-481e-8349-27bae1bc1adb · outbound
Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7a87f1b8-3a3e-4485-86e7-8d24c8a0541a · outbound
Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps In Proceedings of COLING 2016, the 26th International Conference on Computational Linguistics: Technical Papers , pages 2956–2965, Osaka, Japan, December
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c076e7a4-fb88-4667-a4fe-3aeaeca42fdf · outbound
Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps In Proceedings of the 2017 Conference on Empirical Methods in Natural Language Processing, pages 2021–2031, Copenhagen, Denmark, September
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2329a960-8cb7-4f8a-8101-57f7deb6abee · outbound
Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps RoBERTa: A Robustly Optimized BERT Pretraining Approach
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4e3d2b22-2778-40bc-961b-99321182a240 · outbound
Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps In Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing, pages 2383–2392, Austin, Texas, November
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 45f0188f-60ca-4dbc-8da8-c4d64668043e · outbound
Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps In Proceedings of the 2010 Conference on Empirical Methods in Natural Language Processing, pages 1088–1098, Cambridge, MA, October
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 70991d22-2845-4549-9c7c-e4999d799c4c · outbound
Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps Association for Computational Linguistics
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4da38f82-70cc-4859-9f23-aa44603f3a31 · outbound
Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6bcf5f79-439d-490a-9bf2-a5433923ad3c · outbound
Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps In Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing , pages 2369–2380, Brussels, Belgium, October-November
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ad1860ed-25b2-4403-89d6-0d2c59d374c8 · outbound
Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ae5eec23-4273-41eb-9d42-1e5ef63d67eb · inbound
Retrieval-Augmented Generation for Large Language Models: A Survey Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 119
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dd0020fb-070e-4471-8847-de50cade18db · inbound
ZeroSearch: Incentivize the Search Capability of LLMs without Searching Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d28a5430-c1ac-4c5e-8573-a16d1fc5e446 · inbound
ZeroSearch: Incentivize the Search Capability of LLMs without Searching Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3d86b65a-ddb7-49e1-992b-ec0509af1c5d · inbound
Group-in-Group Policy Optimization for LLM Agent Training Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 80f108ca-f080-49c1-9fb5-deb496e9de7a · inbound
Mixture-of-Retrieval Experts for Reasoning-Guided Multimodal Knowledge Exploitation Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation da03794f-a6c1-433d-a9b3-e7bf9a975659 · inbound
From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 43168c99-88ac-4736-9b1b-7433bd3e196b · inbound
Erase to Improve: Erasable Reinforcement Learning for Search-Augmented LLMs Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b003c741-9e5f-46c9-88f4-8ef24f77c6b4 · inbound
Question-Adaptive Graph Learning for Multi-hop Retrieval Augmented Generation Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a98880f0-e466-4bcf-984d-e1e9c9050839 · inbound
EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2cabc30a-8a71-4219-a5ee-ae8465f18a02 · inbound
EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9e879b36-c0bc-45f1-9f7b-7eccd5082b8e · inbound
Sharpness-Guided Group Relative Policy Optimization via Probability Shaping Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 461e1e52-7266-457c-a79f-6bfbbb36bcac · inbound
MemSearcher: Training LLMs to Reason, Search and Manage Memory via End-to-End Reinforcement Learning Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0e924b27-a3d9-4b1d-961a-a6bb325df8f7 · inbound
MiroThinker: Pushing the Performance Boundaries of Open-Source Research Agents via Model, Context, and Interactive Scaling Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9be4fa95-cb86-4366-bc53-2f5f809e3e78 · inbound
Agent-R1: A Unified and Modular Framework for Agentic Reinforcement Learning Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bee7e3bc-b35a-4479-97aa-965fc489d581 · inbound
LocalSearchBench: Benchmarking Agentic Search in Real-World Local Life Services Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c98c48de-7d5b-41cb-a0df-86cef427f83a · inbound
AutoTool: Dynamic Tool Selection and Integration for Agentic Reasoning Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f322f208-5048-4895-9dba-2754ff805845 · inbound
Leveraging Spreading Activation for Improved Document Retrieval in Knowledge-Graph-Based RAG Systems Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e9f187f-9e0d-4882-8e7a-cff50d3c057c · inbound
Attention Sink Forges Native MoE in Attention Layers: Sink-Aware Training to Address Head Collapse Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 44943469-5037-495a-9788-3dac570e3991 · inbound
Attention Sink Forges Native MoE in Attention Layers: Sink-Aware Training to Address Head Collapse Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40674f8f-a68f-4f4f-8132-81250b5f2863 · inbound
Adaptive Information Control for Search-Augmented LLM Reasoning Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccecdd0d-125a-4631-a6ce-5cd4740c5371 · inbound
DeepResearch-9K: A Challenging Benchmark Dataset of Deep-Research Agent Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 956d9de2-e2f6-47ee-bd88-f3450e643a6b · inbound
MSA: Memory Sparse Attention for Efficient End-to-End Memory Model Scaling to 100M Tokens Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8fc202cb-85db-4836-96c9-7cbe4707b7b6 · inbound
Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b25c39b-d3e0-4bbb-9739-9083f855146b · inbound
OASES: Outcome-Aligned Search-Evaluation Co-Training for Agentic Search Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3eacbc76-875d-4df7-ac24-333baa2b351a · inbound
Do We Still Need GraphRAG? Benchmarking RAG and GraphRAG for Agentic Search Systems Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a446c81c-b306-4b3f-9e82-29d563975943 · inbound
Transforming External Knowledge into Triplets for Enhanced Retrieval in RAG of LLMs Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0c1ea5b9-cf19-424e-86ff-6ea1ba61678e · inbound
HeadRank: Decoding-Free Passage Reranking via Preference-Aligned Attention Heads Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b437cd33-2d0b-4b5e-b03a-e836aeba4b31 · inbound
MemSearch-o1: Empowering Large Language Models with Reasoning-Aligned Memory Growth in Agentic Search Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 46c4525e-fb01-4b87-9bc9-ff19219d2ba7 · inbound
EHRAG: Bridging Semantic Gaps in Lightweight GraphRAG via Hybrid Hypergraph Construction and Retrieval Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5a18daf1-b189-4fb4-9b8c-256aee05d83b · inbound
AtomicRAG: Atom-Entity Graphs for Retrieval-Augmented Generation Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 71898f25-1322-4658-9064-4eab9dabec81 · inbound
Reformulating KV Cache Eviction Problem for Long-Context LLM Inference Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 34be19f9-2c42-48e1-bc6d-cf50ad4c0d76 · inbound
Query Symbolically or Retrieve Semantically? A Dataset and Method for Semi-Structured Question Answering Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 840b8620-2233-4ff9-b976-b6626e99a153 · inbound
MoG: Mixture of Experts for Graph-based Retrieval-Augmented Generation Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 922a8a2a-aefe-4b32-a77a-6bb10bdbfe9a · inbound
MemGraphRAG: Memory-based Multi-Agent System for Graph Retrieval-Augmented Generation Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e3366a89-d66a-4fe9-8fb0-8e132299b4a6 · inbound
Efficient RAG with Intent-Aware Retrieval and Semantics-Preserving Chunking Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 76007fe2-3744-45f9-827d-38802c53fbe5 · inbound
Policy and World Modeling Co-Training for Language Agents Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 155c6d3f-b9e7-43e3-a831-2cd7182b4323 · inbound
QCFuse: Query-Aware Cache Fusion via Compressed View for Efficient RAG Serving Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 42bded46-782b-4d79-9915-8adf1c24620f · inbound
Selection Integrity for LLM Graph Memory: An Accumulability Criterion for Information-Flow-Blind Retrieval Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1fb04bc6-912e-4048-9266-453a5115a3d8 · inbound
Agents-K1: Towards Agent-native Knowledge Orchestration Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c6885ddd-dfcd-44ff-895b-b47c24e904d8 · inbound
Agents-K1: Towards Agent-native Knowledge Orchestration Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f1300452-95d2-46e2-a638-26489191cced · inbound
Agents-K1: Towards Agent-native Knowledge Orchestration Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d3a9853-1de3-4856-88e9-53f07e45e6cb · inbound
FlowRAG: Synergizing Explicit Reasoning via Frequency-Aware Multi-Granularity Graph Flow Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 21d41ae4-e7c1-4b7a-90ad-7ded4e39e8af · inbound
R$^2$-Searcher: Calibrating Retrieval and Reasoning Boundaries for Agentic Search Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ede3a1b8-1754-4d9b-aea1-d9301445b575 · inbound
STAPO: Selective Trajectory-Aware Policy Optimization for LLM Agent Training Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 472c4520-63b8-4094-9577-a4ce9154dbf7 · inbound
Retrieving a Set, Not Independent Passages: Set-Level Compatibility Learning for Efficient Set Exploration Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 04829529-6ee2-487f-bc6f-849892b168f8 · inbound
UNIBROWSE: A Data-to-Agent Framework for Multimodal BrowseComp Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 435da1a2-d03b-4139-a24d-6d6fb00a3ecf · inbound
Shapley Context Pruning: A Cooperative Game Perspective for Context Reranking and Pruning Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a74b787b-e60c-4ba7-88e0-0eee59648c1b · inbound
From Outcomes to Actions: Leveraging Hindsight for Long-Horizon Language Agent Training Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.