Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:26:36.964619Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2608.11888.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:26:36.964619Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
36 of 36 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 880a41a7-b057-477a-a641-ac2b56d97433 · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ccb5c91-4931-405b-acf9-b5167a0c7fba · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Swe-skills- bench: Do agent skills actually help in real-world software engineering?
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d304092-63ab-4396-911d-6e9fbd363b04 · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Lost in the middle: How language models use long contexts,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6b2d0fdb-26a6-4c9d-aaa6-7252deb63e06 · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Large language models can be easily distracted by irrelevant context,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc9253a3-202f-41f4-8214-72a7ac68772c · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Quantifying language models’ sensitivity to spurious features in prompt design or: How i learned to start worrying about prompt formatting,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 39a5e884-301e-48c9-896c-acd4d461b098 · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents More doc- uments, same length: Isolating the challenge of multiple documents in rag,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fe4a77f-ffcf-4967-b0da-39d15fdacaaf · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Differential testing for software,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 899f2ebc-47de-4604-a7d0-98e4f6010311 · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Equipping agents for the real world with agent skills,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 35bca5fa-fd82-448d-892c-47827473563d · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents OpenCode: The open source coding agent,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6346949e-831b-4403-a265-b1612ef26aae · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Introducing claude opus 4.6,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ef7072a1-38dc-4a28-a1c0-7d7d76fe5171 · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Introducing gpt-5.5,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c8db957b-9181-465c-8170-6f4d4bba60b7 · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Ensemble methods in machine learning,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 804f795f-296f-4e4b-a27e-dba9b01f30f7 · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Self-Consistency Improves Chain of Thought Reasoning in Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c92d4d58-f465-432e-9e65-a6430f05a91f · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Voyager: An Open-Ended Embodied Agent with Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b92d28f2-2adb-47fb-8579-e06b17052d83 · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Executable code actions elicit better llm agents,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fbc1d13-ac23-42e8-8b33-a4929bb13464 · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Gpt4tools: Teaching large language model to use tools via self-instruction,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 92d17c18-a0c1-4eec-aa6a-5f2a880bc79b · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents ReAct: Synergizing Reasoning and Acting in Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1310748-b14b-4210-ac1d-08f76de121ce · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Same task, more tokens: the impact of input length on the reasoning performance of large language models,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 03e8beaa-7204-4700-96e7-570a8d7f02f3 · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Why do multi-agent llm systems fail?
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf1d2fc8-ca0a-41b7-b68d-53691326170b · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Agentless: Demystifying LLM-based Software Engineering Agents
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6463b77a-1cf8-4f4a-b391-c5dfc1859dd0 · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Swe-agent: Agent-computer interfaces enable automated software engineering,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5bc29c56-70f1-48ab-84e1-f501c9cfda0b · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Swe-bench: Can language models resolve real-world github issues?
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ab050f45-5cea-44e7-af9b-674eae1fa51f · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents OpenHands: An Open Platform for AI Software Developers as Generalist Agents
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e805c5d-84d0-4549-bce8-35fbe75cbd37 · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Multi-swe-bench: A multilingual benchmark for issue resolving,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9166c087-402a-4c56-99f9-2ea25e418323 · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents SWE-Lancer: Can Frontier LLMs Earn $1 Million from Real-World Freelance Software Engineering?
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cb7da78-2257-48de-a27f-b6a68ab0ca69 · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Measuring ai ability to complete long software tasks,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b96ea37b-c03d-469c-8da3-06c81b357a64 · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Understanding and detecting real-world performance bugs,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation aab1f957-4ef4-497c-aa30-a1be108ea07f · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Discovering, reporting, and fixing perfor- mance bugs,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 37fceef0-e62d-4b25-8ad1-cf135d4532d1 · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Performance issues and optimizations in javascript: an empirical study,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d7045bb6-25f1-4f10-965b-4f946cdc590a · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Understanding performance problems in deep learning systems,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f2735e11-491d-42db-905b-cb677e01f4dd · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents A comprehensive study on deep learning bug characteristics,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0c106420-0050-4149-ac60-09e30d9750cf · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Taxonomy of real faults in deep learning systems,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cad5573d-0fcc-4644-bad8-cec289311a6a · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents An empirical study on tensorflow program bugs,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6132151-456e-4736-b5b3-6abaced6b2b2 · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents An empirical study on configuration errors in commercial and open source systems,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8eb61ccd-7549-48f4-bcc6-eb991ce5e6df · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents Hey, you have given me too many knobs!: Understanding and dealing with over-designed configuration in system software,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3cb0f87a-665b-4040-a1ae-60f9b1a61530 · outbound
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents An empirical study on performance bugs for highly config- urable software systems,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
No inbound Pith citation observations are available.