Pith. sign in

Paper Citation Record · LEDGER

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses?

As of 18 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 0 inbound Pith citation observations for arXiv:2608.04828.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04828 v1

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:38:43.670203Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

26 of 26 outbound references displayed

  • verified exact1
  • verified fuzzy6
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a4266c2f-6900-4907-be5e-e97ba459722a · outbound

This paper cites Large Language Model Agent: A Survey on Methodology, Applications and Challenges.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Large Language Model Agent: A Survey on Methodology, Applications and Challenges

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:41.370321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:41.370321Z digest=sha256:e48a71dc529012e262732d6e92ad5ef549ef3209a441b0a819883039b20ba428

Observation 7589ae9d-d7f5-4bd1-90c8-da0b6ae5b8af · outbound

This paper cites Theagentcompany: benchmarking llm agents on consequen- tial real world tasks.Advances in Neural Information Processing Systems, 38, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Theagentcompany: benchmarking llm agents on consequen- tial real world tasks.Advances in Neural Information Processing Systems, 38, 2026

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:38:46.467109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T15:38:41.416490Z digest=sha256:dbd962ba309132f166f8732cecdcd363a30abc062dd6a27297e3757ce5851ded

Observation 80bb1fca-ef73-4483-a08a-a46343d5056e · outbound

This paper cites Skill-R1: Agent Skill Evolution via Reinforcement Learning.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Skill-R1: Agent Skill Evolution via Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:41.569596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:41.569596Z digest=sha256:193fcb3b0b33c2b671cc1c1f34de90c6d44cb5db8a6a67a5f7d632100318391f

Observation 5f1ac2f4-bfda-4d8d-9d64-fb38090d1ccb · outbound

This paper cites Agent skills: A data-driven analysis of claude skills for extending large language model functionality.arXiv preprint arXiv:2602.08004, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Agent skills: A data-driven analysis of claude skills for extending large language model functionality.arXiv preprint arXiv:2602.08004, 2026

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:41.769513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:41.769513Z digest=sha256:b9303d2d2d3cfc65b8c20bb0e5c28c6d34f257eb1cfc5ec8a660a6561deb94bb

Observation 8a87d7b9-7bd6-48a7-96b4-fb039b38cb34 · outbound

This paper cites SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:41.966146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:41.966146Z digest=sha256:3707691b358b1199ad87561e15cca66aaeec2c353ea3928dfceb64c248d91fee

Observation 8cf06d0b-ed6a-4d69-8d5c-d4b138ca2a04 · outbound

This paper cites SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.040035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.040035Z digest=sha256:177f245dd0fe88ca8c26bde469591c671afbaaf379607716a2a9655aa1193829

Observation 90ac718d-d84d-4824-acef-9aeb41f4943a · outbound

This paper cites Orga- nizing, orchestrating, and benchmarking agent skills at ecosystem scale.arXiv preprint arXiv:2603.02176, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Orga- nizing, orchestrating, and benchmarking agent skills at ecosystem scale.arXiv preprint arXiv:2603.02176, 2026

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.079299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.079299Z digest=sha256:39b3aa8f4f077c321d283a9ae8e244531e9501e7ff57a82b41f559985942eb98

Observation f0ee1b70-516b-47a8-9423-c4facdd1a8ff · outbound

This paper cites Skillnet: Create, evaluate, and connect ai skills.arXiv preprint arXiv:2603.04448, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Skillnet: Create, evaluate, and connect ai skills.arXiv preprint arXiv:2603.04448, 2026

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.169353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.169353Z digest=sha256:d5cca5b38e68fc192b3a37e3a4d23ae0c70f69a667a2387f953fe47f1cc2fc13

Observation a58fb1c1-d4fd-4fbf-9923-db935930c6b9 · outbound

This paper cites Autoskill: Experience-driven lifelong learning via skill self-evolution.arXiv preprint arXiv:2603.01145, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Autoskill: Experience-driven lifelong learning via skill self-evolution.arXiv preprint arXiv:2603.01145, 2026

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.251238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.251238Z digest=sha256:a5db468302a0146d39ae5385a19aaf7476ae7dbab1b7e6759eef029459fb1707

Observation 075b6390-617e-4a08-86a9-68a5ad758c82 · outbound

This paper cites Memento-skills: Let agents design agents.arXiv preprint arXiv:2603.18743, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Memento-skills: Let agents design agents.arXiv preprint arXiv:2603.18743, 2026

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.343681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.343681Z digest=sha256:1b3a334e8bf12a3004e381571b5abfa0ce768b1d0d8f9ead629f7e200d01a166

Observation 8bc841d7-9ea7-4f04-a959-4cd49f36f544 · outbound

This paper cites SLBench: Evaluating How LLM Agents Follow Logical Relations in Skills.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? SLBench: Evaluating How LLM Agents Follow Logical Relations in Skills

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:38:43.949998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T15:38:42.420308Z digest=sha256:9f131d62c3ae32d42a4df868a61ecbaef16a1b1d9fe6cd333723cd06b0fa6775

Observation 36817ce0-c3c4-4a23-871f-a55c30c2e68d · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Instruction-Following Evaluation for Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.552670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.552670Z digest=sha256:5578d5eb7355eebe1e3ae7ab7928baa9ae0e4cf8b43265640e9a45f6de9d7909

Observation 606233ca-dde7-41df-8d6f-d25fc75a3a00 · outbound

This paper cites Followbench: A multi-level fine-grained constraints following benchmark for large language models.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Followbench: A multi-level fine-grained constraints following benchmark for large language models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:38:46.246312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T15:38:42.639788Z digest=sha256:9372a253c1b6eb1dbb03d9688f32e37c19aa2bd577833070074998e939f9f299

Observation d2ce020d-3290-40f0-86f7-c18c856f7429 · outbound

This paper cites Infobench: Evaluating instruction following ability in large language models.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Infobench: Evaluating instruction following ability in large language models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.724332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.724332Z digest=sha256:9874b14a2af51430d0d5005fc186ac05bca21d3802026e5c9fb569202be9c183

Observation cf969f85-d06d-4fc5-b477-7d19481f3089 · outbound

This paper cites Benchmarking complex instruction-following with multiple constraints composition.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Benchmarking complex instruction-following with multiple constraints composition

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:38:46.007953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T15:38:42.800877Z digest=sha256:3e9462ef204674af55f372e43c8822fcd5d6cdb850f34afdf5b263018ec0b26c

Observation 07c56e30-fdd0-4fbd-bf28-edf6934c4a1c · outbound

This paper cites Agentif: Benchmarking large language models instruction following ability in agentic scenarios.Advances in Neural Information Processing Systems, 38, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Agentif: Benchmarking large language models instruction following ability in agentic scenarios.Advances in Neural Information Processing Systems, 38, 2026

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:38:45.742772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T15:38:42.897685Z digest=sha256:fa7396e2dccb22016c054f07909d458106a67dfe2960a223e0e09053d6c53fff

Observation c2026e00-7ded-4a48-b435-ab6447975160 · outbound

This paper cites Stop Comparing LLM Agents Without Disclosing the Harness.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Stop Comparing LLM Agents Without Disclosing the Harness

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.965973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.965973Z digest=sha256:bc1713c2a47b32a5c3c68a043aebb7bf50d164860878b85fd4e427f8fe6ce419

Observation 395af4c8-0ea7-40dd-9ffb-33525113c553 · outbound

This paper cites Harness-Bench: Measuring Harness Effects across Models in Realistic Agent Workflows.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Harness-Bench: Measuring Harness Effects across Models in Realistic Agent Workflows

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:43.056799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:43.056799Z digest=sha256:57ca7ec6f1ac65cb8c0d60112f17f911e1e057b8e21cde074d0afdd4826884ce

Observation 3470191b-dca1-4e18-8fc5-7d464a915668 · outbound

This paper cites Sysbench: Can llms follow system message? InThe Thirteenth International Conference on Learning Representations, 2024.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Sysbench: Can llms follow system message? InThe Thirteenth International Conference on Learning Representations, 2024

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:38:45.568923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T15:38:43.120628Z digest=sha256:d77349440492b55b4b5b636b209facdc0c0d9e3615665d7b412376412f4fea47

Observation a5877146-a186-4571-8180-498013d79e94 · outbound

This paper cites Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:43.205719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:43.205719Z digest=sha256:2d75d07d5814e1248718f26086dbfffe4f72b17cb4fa93ce308953df33d21235

Observation aa05d250-1d29-4d0f-89e4-cd52dd829559 · outbound

This paper cites Judging llm-as-a-judge with mt-bench and chatbot arena.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Judging llm-as-a-judge with mt-bench and chatbot arena

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:43.274595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:43.274595Z digest=sha256:d25a664a66dbb05d7858a103d049ad8a189d669ca3f88dc19e1a0046e3683a37

Observation 581b78c4-ada0-4025-b4bf-4daf605c7e84 · outbound

This paper cites SOPBench: Evaluating Language Agents at Following Standard Operating Procedures and Constraints.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? SOPBench: Evaluating Language Agents at Following Standard Operating Procedures and Constraints

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:43.353203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:43.353203Z digest=sha256:4855056d7566b11cd3fdeaf82a1913d4018d9f59698f680331411bb00c98bf3e

Observation 4200e5b6-42e8-4fd7-90d2-5fa0a37f6bd6 · outbound

This paper cites an unresolved cited work.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:38:45.416713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T15:38:43.436668Z digest=sha256:6e922bea8ac07845f5d4eafb90aa979513d5f4ce9f23e716f94be7da005b859e

Observation 486e0e27-5156-4bdc-a184-fe574b52ac5a · outbound

This paper cites an unresolved cited work.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:38:45.104227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T15:38:43.511861Z digest=sha256:ef074d225db493d443774185e5bf6674a03055482769b0794e02ab9a9a9ffc2e

Observation b6bc6ce9-2d16-4f92-a12c-2763c6f3590f · outbound

This paper cites an unresolved cited work.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:38:44.786035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T15:38:43.586596Z digest=sha256:bb1361ba1f7f98ce56bd60ffeb51e79fd01e50dc380696e4c6ba1e83fa81edd8

Observation 115f6e3a-ae6c-4593-bd56-6b6c61ec52f1 · outbound

This paper cites pdftk-server.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? pdftk-server

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:38:44.514436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T15:38:43.670203Z digest=sha256:7523fe9173305337d1ddb331477b1170326dacf8519c5c25663431261726a045

Pith citing papers

No inbound Pith citation observations are available.