Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2307.16039.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T23:23:32.490637Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-23T04:55:24.835950Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 9adc86c2-8131-4a2c-ace1-61ac0989f332 · inbound
RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation dd7540e7-cb29-4bee-8b02-c108a7c09538 · inbound
Training Language Models to Self-Correct via Reinforcement Learning Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 060b73f1-9848-4154-a328-ac2b062ff445 · inbound
AdaMCoT: Rethinking Cross-Lingual Factual Reasoning through Adaptive Multilingual Chain-of-Thought Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bd3debf8-77b6-4a9f-a89c-07734bb1bd2a · inbound
Whose Truth? Pluralistic Geo-Alignment for (Agentic) AI Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62ee54a1-3688-45e8-97f0-c9fe9286fb83 · inbound
TASE: Token Awareness and Structured Evaluation for Multilingual Language Models Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.