Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T04:40:33.122717Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 0 inbound Pith citation observations for arXiv:2607.28545.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T04:40:33.122717Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
15 of 15 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a67d0831-c5a3-4c43-8f87-13ac6660e72f · outbound
ORCA-bench: How Ready Are Language Model Agents for Oncall? Long Code Arena: a Set of Benchmarks for Long-Context Code Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6c465db-c1d9-4df1-92c7-714a303b96f3 · outbound
ORCA-bench: How Ready Are Language Model Agents for Oncall? SREGym: A Live Benchmark for AI SRE Agents with High-Fidelity Failure Scenarios
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c631941-4708-4774-8bf1-f51af23ab502 · outbound
ORCA-bench: How Ready Are Language Model Agents for Oncall? Itbench: Evaluating ai agents across diverse real-world it automation tasks
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d4f1571b-17c5-4f18-9ce8-1fb6f9658f37 · outbound
ORCA-bench: How Ready Are Language Model Agents for Oncall? Swe-bench: Can language models resolve real-world github issues? In International Conference on Learning Representations, volume 2024, pages 54107--54157, 2024
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c1f0203f-d55f-46db-a1e5-0f68d07d964b · outbound
ORCA-bench: How Ready Are Language Model Agents for Oncall? Terminal-Bench: Benchmarking Agents on Hard, Realistic Tasks in Command Line Interfaces
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3caebaa7-a09c-4176-83a1-95f96065a7ce · outbound
ORCA-bench: How Ready Are Language Model Agents for Oncall? Rcaeval: A benchmark for root cause analysis of microservice systems with telemetry data
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 703cdbae-f913-497b-85eb-0dc67a3bfe5e · outbound
ORCA-bench: How Ready Are Language Model Agents for Oncall? Building ai agents for autonomous clouds: Challenges and design principles
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5b613cb5-f74a-422f-abcf-d769fe4da5d5 · outbound
ORCA-bench: How Ready Are Language Model Agents for Oncall? Openrca: Can large language models locate the root cause of software failures? In The Thirteenth International Conference on Learning Representations, 2025
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 29c43866-cf73-42dc-badd-965ef8c563a2 · outbound
ORCA-bench: How Ready Are Language Model Agents for Oncall? Swe-bench multimodal: Do ai systems generalize to visual software domains? In The Thirteenth International Conference on Learning Representations, 2025
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 47f386de-8d47-4459-919b-1bd9a2799903 · outbound
ORCA-bench: How Ready Are Language Model Agents for Oncall? Swe-smith: Scaling data for software engineering agents
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0bdee2f2-9d83-4c48-81fd-52effa5b812e · outbound
ORCA-bench: How Ready Are Language Model Agents for Oncall? Judging llm-as-a-judge with mt-bench and chatbot arena
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1b018bd-c6e6-4698-8063-b7a62fff3da5 · outbound
ORCA-bench: How Ready Are Language Model Agents for Oncall? Graders should cheat: privileged information enables expert-level automated evaluations
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f595175f-81f7-4df1-b643-eb90207a3f55 · outbound
ORCA-bench: How Ready Are Language Model Agents for Oncall? @esa (Ref
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 136b50f1-8939-4830-8fd4-305350631b9a · outbound
ORCA-bench: How Ready Are Language Model Agents for Oncall? Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2397f58b-10e4-4d29-b909-fa110d9a7746 · outbound
ORCA-bench: How Ready Are Language Model Agents for Oncall? Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.