Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T15:00:08.031991Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 1 inbound Pith citation observation for arXiv:2412.11417.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T15:00:08.031991Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T23:05:56.982882Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T23:05:57.178299Z
33 of 33 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d00e597f-15e3-4487-99b4-bfa81ceff794 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Emulating human play in a leading mobile card game,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6f55e57d-deef-4103-97b2-3a6274f2cd30 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Strategy generation for multiunit real-time games via voting,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 87c4d2a5-0288-4ebf-9d9e-30afd255b7d3 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Mastering the game of go with deep neural networks and tree search,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b6b15f2-8625-4cfd-b729-feda5e1c1556 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Mastering the game of go without human knowledge,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation be9a6928-f320-4ee0-8bad-b796dd5bd260 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Dota 2 with Large Scale Deep Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 770441ca-cea8-4476-86fd-f79a168e5dff · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Grand- master level in starcraft ii using multi-agent reinforcement learning,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff7780fc-ad59-479b-b86d-4fa626738e96 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Full douzero+: Improving doudizhu ai by opponent modeling, coach-guided training and bidding learning,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 64b263ad-2f62-4386-97b7-3821930a4ef8 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Danzero+: Dominating the guandan game through reinforcement learning,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation df76c031-04ab-433d-be60-0f896c32bdb1 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Mastering curling with rl-revised decision tree,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6ea2a705-d8d3-4556-b82a-63bbefc77ab9 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Emergent Abilities of Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07566e5c-beea-47a0-b7e4-4189e3ac2f7b · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement ReAct: Synergizing Reasoning and Acting in Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2b7a8cf-4cae-4b0d-85b9-efaf1665636b · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Do As I Can, Not As I Say: Grounding Language in Robotic Affordances
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16bed98f-5c20-4161-ae53-5f6b7836fd32 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Toolformer: Language models can teach themselves to use tools,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c2cc585e-8cea-4c99-994d-da740061e6a3 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Hugginggpt: Solving ai tasks with chatgpt and its friends in hugging face,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ff0bb081-c039-44d1-afc8-da712c7c5971 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement WebGPT: Browser-assisted question-answering with human feedback
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7b8a1a4-dbd3-47d4-8bd5-81c9654ba6ed · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 859a69ad-8b17-47d3-b864-e660a24d1f91 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Qwen2.5-Coder Technical Report
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 472aa2cf-26a8-49fd-a4b1-c82d535c664f · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Google research football: A novel reinforcement learning environment,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e9cee1c4-38e1-4fe1-a281-c4f7f4ba7bd5 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Fever Basketball: A Complex, Flexible, and Asynchronized Sports Game Environment for Multi-agent Reinforcement Learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c92bffee-6d9a-4986-b24b-c007d951a374 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Suphx: Mastering Mahjong with Deep Reinforcement Learning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec4cdb99-e01f-491c-ade4-7e571b6f9a31 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Douzero: mastering doudizhu with self-play deep reinforcement learning,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1c62660c-41c8-4ee4-9c49-9d2b2b42bbba · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Chessgpt: Bridging policy learning and language mod- eling,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 707485f5-32a9-4022-bd5e-63f75c031243 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Voyager: An Open-Ended Embodied Agent with Large Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f367805-8182-4f17-9b07-4a01c2234c06 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Minedojo: Building open- ended embodied agents with internet-scale knowledge,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 20691713-3637-45f5-a115-b1beba7a27b9 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Large Language Models Play StarCraft II: Benchmarks and A Chain of Summarization Approach
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff03e40f-694b-46b5-b091-258372deea50 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement PokeLLMon: A Human-Parity Agent for Pokemon Battles with Large Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63d3722e-674d-47f5-9302-560bdf60efd5 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Towards general computer control: A multi- modal agent for red dead redemption ii as a case study,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1983a87b-3a89-4a8f-a862-cf6c2ba03a60 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Executable Code Actions Elicit Better LLM Agents
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80fca482-4d62-44f2-b6c1-9deee9c155c9 · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Eureka: Human-Level Reward Design via Coding Large Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72106eca-28e0-44b2-a468-1638ec2aa15d · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c30c1fa6-e9bb-4ffd-a3fa-0bc0f7a68f8b · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Proximal Policy Optimization Algorithms
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed43bd58-15d1-48ef-a345-ca377a0bf7bc · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement Trust region policy optimization,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7c7265be-94ef-4eeb-b933-6a2a144dd2dc · outbound
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement High- dimensional continuous control using generalized advantage estimation,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 706dfce6-7892-42bf-b350-c2ebf1cbde03 · inbound
Multi-Armed Bandits-Based Optimization of Decision Trees RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.