Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-13T13:05:01.450957Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 0 inbound Pith citation observations for arXiv:2604.03562.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-13T13:05:01.450957Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
22 of 22 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 8778e897-b65e-4c41-b1cc-e6563e6e9da0 · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling SpaceX Starlink,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e0ccf51-d8e2-4b09-b645-143595936ab4 · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling Beam hopping for multi-beam GEO satellite communication systems,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d614cea-32e2-415c-b96c-0708011ce9b4 · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling Deep reinforcement learning for dynamic spectrum access in satellite communications,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edea8cad-01d6-431c-92cb-3ed4f736261f · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling Eureka: Human-level reward design via coding large language models,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1cd6ae7-12d7-4c8e-95d1-32ab17174fc6 · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling Deep reinforcement learning for resource management in network slicing,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da16fb74-2dff-4773-97f6-73fe454f95d3 · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling Deep Reinforcement Learning Architecture for Continuous Power Allocation in High Throughput Satellites
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22c11e2f-ecfd-439a-ad3b-b358a0bfd956 · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling Multi-objective optimization for cognitive satellite communications using deep reinforcement learning,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36dd3ab4-dd4a-4312-b72c-94a7dd0c7651 · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling Deep reinforcement learning for satellite communication: A survey,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fdef6d9b-0790-4b2d-8b4c-29a80c808a92 · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling Policy invariance under reward transformations: Theory and application to reward shaping,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a02ac5e8-931f-4039-b021-ca2e421b24d6 · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling A practical guide to multi-objective rein- forcement learning and planning,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31a6e1dc-8a81-4995-86bd-218aeb620d77 · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling What Can Learned Intrinsic Rewards Capture?
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3735d2af-d4b2-48a3-aac5-eacecfc2de4e · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling Large language models for telecom: Opportunities and challenges,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da95fcba-51c7-493d-955b-588eb84223bc · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling Networking with large language models,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 392c7006-c0b6-4315-92cd-d93ffa57cc57 · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling Large Language Models for Telecom: Forthcoming Impact on the Industry
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e2ecb67-045b-4c29-b263-50b17bc91190 · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling WirelessLLM: Empowering Large Language Models Towards Wireless Intelligence
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b97aedd-78bc-4c1f-928a-4c7361c03e17 · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling The Evolving Landscape of LLM- and VLM-Integrated Reinforcement Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19818408-4229-48fc-88a5-d38a02cafde1 · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling Continuous inspection schemes,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc6084af-2ea0-4f23-8a1a-243e09fa68f0 · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling Proximal Policy Optimization Algorithms
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f22e198-fa1e-4ad6-9e89-ead6537b9627 · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling A definition of continual reinforcement learning,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7ff86de-5294-4358-a5a1-363b6f949572 · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling Gradient surgery for multi-task learning,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b640f302-29e5-411a-9048-de4f8e15c978 · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling Retrieval- augmented generation for knowledge-intensive NLP tasks,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa0ae670-76ed-47e1-9c13-6d8bbf177cd0 · outbound
When Adaptive Rewards Hurt: Causal Probing and the Switching-Stability Dilemma in LLM-Guided LEO Satellite Scheduling Grounding by Trying: LLMs with Reinforcement Learning-Enhanced Retrieval
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.