Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 100 inbound Pith citation observations for arXiv:1712.01815.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T21:17:01.988107Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
1084
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 0eaca5ad-6510-4b06-932c-d60eb3d17d03 · inbound
AI safety via debate Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7ee7e5c5-693f-44f8-8edc-2d9c24b21975 · inbound
Near-optimal Bayesian Solution For Unknown Discrete Markov Decision Process Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation fc83e076-d3cd-4d7e-9e6d-0e9a58688e09 · inbound
Inductive general game playing Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f0a6c56a-c69e-47ea-9569-3676dba578e4 · inbound
On Multi-Agent Learning in Team Sports Games Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5d96b3c8-3011-408c-8fc9-7a61619caf22 · inbound
Growing Action Spaces Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a640cb50-f99a-495d-8840-4f16cc9e329e · inbound
General Board Game Playing for Education and Research in Generic AI Game Learning Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0abe9487-8397-4217-af11-baa2fe004dee · inbound
Solving Rubik's Cube with a Robot Hand Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 102
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0c02283a-22cb-4f6a-933e-a72bbff4bd6e · inbound
On the Measure of Intelligence Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 8381a50b-b274-4684-9ff2-d1a1133fc4d7 · inbound
Generative Language Modeling for Automated Theorem Proving Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 73302cc3-b4e0-40d2-86f1-59127324da1f · inbound
Language Models (Mostly) Know What They Know Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 208
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5b094429-2807-4abd-bfe9-6fa891c7c3ef · inbound
Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 171
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5d1d5c01-630d-491d-9367-0b0ca6d33ad0 · inbound
Reasoning with Language Model is Planning with World Model Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 130
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation bc59fdd5-f186-4d79-8d15-01b35dd647b3 · inbound
Reinforcement Learning with Foundation Priors: Let the Embodied Agent Efficiently Learn on Its Own Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 82bf8c01-107d-4a33-bae1-912a5bf63842 · inbound
Language Agent Tree Search Unifies Reasoning Acting and Planning in Language Models Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 54a06b1a-4101-4aec-89aa-8058b0f629eb · inbound
Learning Interactive Real-World Simulators Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 93501dce-6bb5-4f92-a496-7ee3a7ad24a9 · inbound
MTSpark: Enabling Multi-Task Learning with Spiking Neural Networks for Generalist Agents Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02838376-9eab-42fe-a73f-7553b9b6589b · inbound
Exponential Speedups by Rerooting Levin Tree Search Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6e5d3b2-1623-4529-a6c0-c9278cffbdb6 · inbound
Augmenting the action space with conventions to improve multi-agent cooperation in Hanabi Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb8e9ed8-20a7-4907-a60d-8687c5fc7c36 · inbound
Monte Carlo Tree Search based Space Transfer for Black-box Optimization Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dcbdc2d-a295-438d-ad05-1702f5c610a1 · inbound
Seed-CTS: Unleashing the Power of Tree Search for Superior Performance in Competitive Coding Tasks Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33e22423-7f97-4549-aba4-c0eb5d12a761 · inbound
RAG-Star: Enhancing Deliberative Reasoning with Retrieval Augmented Verification and Refinement Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ca98572-cd7b-402e-832b-2556fcb4081e · inbound
Ensembling Large Language Models with Process Reward-Guided Tree Search for Better Complex Reasoning Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe3f1d92-ba3d-40d6-be20-3812083b914e · inbound
Towards Intrinsic Self-Correction Enhancement in Monte Carlo Tree Search Boosted Reasoning via Iterative Preference Learning Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a231f84-78a8-49c5-bc27-805f2653b871 · inbound
Automating the Search for Artificial Life with Foundation Models Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ba3a3f3-7c77-478a-9912-b58697dd6e95 · inbound
NS-Gym: Open-Source Simulation Environments and Benchmarks for Non-Stationary Markov Decision Processes Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99a84f7f-9ae3-499d-913b-b7166dcb7dd7 · inbound
A Survey on Multi-Turn Interaction Capabilities of Large Language Models Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92a1232c-ee55-4e65-b270-8433a89c431c · inbound
AirRAG: Autonomous Strategic Planning and Reasoning Steer Retrieval Augmented Generation Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c01dde67-806f-4de8-81c3-a6d3d5c47842 · inbound
Beyond the Sum: Unlocking AI Agents Potential Through Market Forces Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eeb9d69a-03ce-4f16-8622-951c2ef8f387 · inbound
Revisiting Rogers' Paradox in the Context of Human-AI Interaction Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6ad1393-60cf-4f1d-bc29-dc0df14e6709 · inbound
HEPPO-GAE: Hardware-Efficient Proximal Policy Optimization with Generalized Advantage Estimation Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d8b1132-0942-45fb-bbf0-f08977538e45 · inbound
What if Eye...? Computationally Recreating Vision Evolution Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6bc3594-ea50-41aa-b4b8-b1c5686cd307 · inbound
Lipschitz Lifelong Monte Carlo Tree Search for Mastering Non-Stationary Tasks Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f82252d-cc4f-4fde-8f0b-458d0bec97b8 · inbound
A Variational Inequality Approach to Independent Learning in Static Mean-Field Games Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71f9a99c-4ded-44b2-bf3f-f543e487623b · inbound
Develop AI Agents for System Engineering in Factorio Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42f2bac0-71e6-4c31-847b-d3c222489a9a · inbound
CH-MARL: Constrained Hierarchical Multiagent Reinforcement Learning for Sustainable Maritime Logistics Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2531ecd-f9c3-43ff-b4dd-243733311256 · inbound
Synthesis of Model Predictive Control and Reinforcement Learning: Survey and Classification Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4b80ebf-5a1e-400e-8a2b-dc6c33c4488a · inbound
LIMO: Less is More for Reasoning Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e3f38b74-f831-47e1-a2a6-91f7c19c4fca · inbound
Beyond Interpolation: Extrapolative Reasoning with Reinforcement Learning and Graph Neural Networks Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13770f35-3892-417f-b92e-e4aff51cea6e · inbound
Holistically Guided Monte Carlo Tree Search for Intricate Information Seeking Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a36ca27a-c381-46e4-a85b-a0960fc6b2fa · inbound
Probabilistic Artificial Intelligence Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ac05b3c-09a7-4cc2-8855-cce582f6580b · inbound
LLMs Can Teach Themselves to Better Predict the Future Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ffa25c5-17da-454a-801f-0ae3455dcae6 · inbound
Two-Player Zero-Sum Differential Games with One-Sided Information Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51fffe79-a280-413d-b851-57e036a9be1b · inbound
Policy Guided Tree Search for Enhanced LLM Reasoning Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 058bfed7-c6c2-4d82-ac31-7778c2ce6739 · inbound
We Can't Understand AI Using our Existing Vocabulary Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d75d4da-4c03-46e7-a36e-48077c96761a · inbound
$\texttt{SEM-CTRL}$: Semantically Controlled Decoding Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation bec547c2-e15e-4994-9ade-b935f7dcc8b1 · inbound
Time-R1: Post-Training Large Vision Language Model for Temporal Video Grounding Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9d420f29-727e-4b67-8ded-c93a1f5a20f8 · inbound
Scalable Multi-Task Learning through Spiking Neural Networks with Adaptive Task-Switching Policy for Intelligent Autonomous Agents Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 143ce6d6-de68-44ed-84b8-5b90c3a4a49a · inbound
DiagnosisArena: Benchmarking Diagnostic Reasoning for Large Language Models Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb50a7b3-b697-4853-8960-64934c712f57 · inbound
Large Language Models for Planning: A Comprehensive and Systematic Survey Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 221
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce2edbaa-5f35-45f9-8e1e-649dd05b8711 · inbound
A Framework for Adversarial Analysis of Decision Support Systems Prior to Deployment Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9703d2b-f199-4dd6-8544-44b9b91e86e5 · inbound
Decomposing Elements of Problem Solving: What "Math" Does RL Teach? Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd313a9f-5794-4dbb-9f22-8b1be12cc3de · inbound
LlamaRL: A Distributed Asynchronous Reinforcement Learning Framework for Efficient Large-scale LLM Training Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c343496-eb31-43df-98bc-4424873161c5 · inbound
Bregman Centroid Guided Cross-Entropy Method Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 419ef1f1-a177-49e1-8a10-5c862632ae18 · inbound
Rewarding the Unlikely: Lifting GRPO Beyond Distribution Sharpening Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb642055-069f-49a9-a58c-65798539348c · inbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96b5d1a6-42ac-479d-9522-b113d999f338 · inbound
Learning to Plan via Supervised Contrastive Learning and Strategic Interpolation: A Chess Case Study Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c712f5b-949f-4280-a6ee-3b068b91deb3 · inbound
Boosting LLM Reasoning via Spontaneous Self-Correction Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 2008
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8f2eee2-1bb1-44aa-bead-6ad9de2b6459 · inbound
TreeRL: LLM Reinforcement Learning with On-Policy Tree Search Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99368f62-42ff-4941-b9c3-b2d8ea26e4b2 · inbound
Data-Driven Policy Mapping for Safe RL-based Energy Management Systems Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd2ea97c-aa52-4054-999b-1fa2e8176f5e · inbound
The Decrypto Benchmark for Multi-Agent Reasoning and Theory of Mind Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d9dc32f-db41-44a7-8c75-c83962b8fee6 · inbound
A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5e27596-3d6c-4b64-9aa1-59f9f7e7627a · inbound
Gym4ReaL: A Suite for Benchmarking Real-World Reinforcement Learning Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e9e4522-fbce-4ae5-ac6d-afa829209472 · inbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation deed1063-266c-4cc0-a404-41faaf89ca6f · inbound
Reinforcement Learning for Automated Cybersecurity Penetration Testing Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edcbd904-06a2-410c-809f-67344d3ac2e1 · inbound
Partial Label Learning for Automated Theorem Proving Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7ffe987-480b-4136-8cc5-1641a34a07b1 · inbound
Critique of World Model Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22e431cc-f1da-48bd-92ef-a9c77a23ed9a · inbound
Artificial Generals Intelligence: Mastering Generals.io with Reinforcement Learning Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c36e4d87-d47b-44e8-8317-feb967e0fa88 · inbound
Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d31759d9-4456-4e96-82a5-6807f12a6c34 · inbound
What Does it Mean for a Neural Network to Learn a "World Model"? Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 774b86a0-9dd7-4312-9234-e4225be897b5 · inbound
General Agentic Planning Through Simulative Reasoning with World Models Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f6ecdc91-0157-44c1-a375-e8a38252c7e3 · inbound
Tail-Risk-Safe Monte Carlo Tree Search under PAC-Level Guarantees Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 562758e3-95f4-4845-be47-56ffaeeb86ab · inbound
Hybrid Physics-Machine Learning Models for Quantitative Electron Diffraction Refinements Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 504cf053-eea8-41b3-8015-1c90d66a0ba9 · inbound
The Fair Game: Auditing & Debiasing AI Algorithms Over Time Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 117
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 783b6d4e-df78-4688-b664-bf61f0902eba · inbound
Evolutionary Optimization of Deep Learning Agents for Sparrow Mahjong Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14e5daec-6c9c-4667-8f5b-90fb8fe6d078 · inbound
Edge General Intelligence Through World Models and Agentic AI: Fundamentals, Solutions, and Challenges Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 101
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cde19674-1179-4713-ad61-86d99b27d526 · inbound
Mirage or Method? How Model-Task Alignment Induces Divergent RL Conclusions Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b65bf5b5-13a2-49c3-8c8f-d92500571d16 · inbound
Scalable Option Learning in High-Throughput Environments Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e3d5166e-e192-4b50-9afd-3bbb249ff6f5 · inbound
TransZero: Parallel Tree Expansion in MuZero using Transformer Networks Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 525fa1de-ba25-4a75-9403-e531e4b1fd88 · inbound
Student-Centered Distillation Narrows the Agentic Gap Between Small and Large LLMs Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72b104ca-c66f-4952-b4b4-a77ff7925207 · inbound
Use the Online Network If You Can: Towards Fast and Stable Reinforcement Learning Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 353cde8c-65a4-435e-b4be-b5577287cd77 · inbound
People use fast and flat simulation to reason about new games Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e257f57-fbcd-40c5-839b-f9aeb7887c89 · inbound
Optimal control of the future via prospective learning with control Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7410afe3-f7c1-4cd2-b1c9-81ad21f0280e · inbound
Olmo 3 Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1a2b69ba-8b4f-4e69-b893-e00b80a8a1c1 · inbound
Learning to Plan, Planning to Learn: Adaptive Hierarchical RL-MPC for Sample-Efficient Decision Making Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9cc4fbe1-2280-4386-8fd8-d51d67c38b86 · inbound
Toward Training Superintelligent Software Agents through Self-Play SWE-RL Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 718bc116-399b-4cb8-8ebc-c01cdc475e14 · inbound
Toward Training Superintelligent Software Agents through Self-Play SWE-RL Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e6e7293-f947-406f-81aa-c91cd60daeb6 · inbound
Safety Alignment of LMs via Non-cooperative Games Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c278cb97-5e22-497d-89f4-23db7ec1a477 · inbound
Multi-agent DRL-based Lane Change Decision Model for Cooperative Platooning in Mixed Traffic Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fe87430-de22-42e4-abfb-0a787823317f · inbound
Output-Space Search: Targeting LLM Generations in a Frozen Encoder-Defined Output Space Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32557d67-3451-44be-8156-77395016f9c0 · inbound
CPMobius: Iterative Coach-Player Reasoning for Data-Free Reinforcement Learning Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e56cbe2-06ef-401e-ac4c-1ef5b24d998a · inbound
Self-Supervised Bootstrapping of Action-Predictive Embodied Reasoning Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 349ee6f0-86de-41a1-8875-d019a4a94633 · inbound
Multi-agent imitation learning with function approximation: Linear Markov games and beyond Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06ebd88d-ee9c-40f6-a8db-7f3cc2982c46 · inbound
Quantum entanglement provides a competitive advantage in adversarial games Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f58c790d-9f7d-4d1b-acad-e30429f8a847 · inbound
Computer Architecture's AlphaZero Moment: Automated Discovery in an Encircled World Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0c3e1e3d-00f1-477e-bd1b-c30f092d37ea · inbound
Probabilistic Language Tries: A Unified Framework for Compression, Decision Policies, and Execution Reuse Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7e16cd7d-a8a1-487e-91e0-5dd59846ac29 · inbound
Advantage-Guided Diffusion for Model-Based Reinforcement Learning Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0b6162e5-857b-4a59-a128-e27154cf945a · inbound
AdverMCTS: Combating Pseudo-Correctness in Code Generation via Adversarial Monte Carlo Tree Search Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2324b110-0069-4883-ba49-ca862400a268 · inbound
AlphaCNOT: Learning CNOT Minimization with Model-Based Planning Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5d0a9622-d16a-48a0-bd6e-9fc7119fd27b · inbound
PAWN: Piece Value Analysis with Neural Networks Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 632febf0-5ba1-4a6b-ac89-72ce804b4ea6 · inbound
Causal inference for social network formation Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.