Pith. sign in

Paper Citation Record · LEDGER

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses?

As of 10 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 0 inbound Pith citation observations for arXiv:2608.04828.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04828 v1

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:38:43.670203Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

26 of 26 outbound references displayed

  • verified exact1
  • verified fuzzy6
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a4266c2f-6900-4907-be5e-e97ba459722a · outbound

This paper cites Large Language Model Agent: A Survey on Methodology, Applications and Challenges.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Large Language Model Agent: A Survey on Methodology, Applications and Challenges

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:41.370321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:41.370321Z digest=sha256:51c72bffcdf94b1c549ce937f2fb4b878c142f8807d1d8d404391b8db44ab216

Observation 7589ae9d-d7f5-4bd1-90c8-da0b6ae5b8af · outbound

This paper cites Theagentcompany: benchmarking llm agents on consequen- tial real world tasks.Advances in Neural Information Processing Systems, 38, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Theagentcompany: benchmarking llm agents on consequen- tial real world tasks.Advances in Neural Information Processing Systems, 38, 2026

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:38:46.467109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:38:41.416490Z digest=sha256:a93da1f489324ae66d1d49d56122fe432fea2190319cd2fb3462e581292cdda0

Observation 80bb1fca-ef73-4483-a08a-a46343d5056e · outbound

This paper cites Skill-R1: Agent Skill Evolution via Reinforcement Learning.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Skill-R1: Agent Skill Evolution via Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:41.569596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:41.569596Z digest=sha256:e1176e86b6f676f87f0148bfaa7aff4ccf75d41dcdadc38fa7efc886ffe01597

Observation 5f1ac2f4-bfda-4d8d-9d64-fb38090d1ccb · outbound

This paper cites Agent skills: A data-driven analysis of claude skills for extending large language model functionality.arXiv preprint arXiv:2602.08004, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Agent skills: A data-driven analysis of claude skills for extending large language model functionality.arXiv preprint arXiv:2602.08004, 2026

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:41.769513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:41.769513Z digest=sha256:6eff9590e4539cde554862a9f1dea4f89c9af2f3f509e25d659b8251db0cab33

Observation 8a87d7b9-7bd6-48a7-96b4-fb039b38cb34 · outbound

This paper cites SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:41.966146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:41.966146Z digest=sha256:a11521788b4ff8361bbec59bde734409577e5e739a32db4ead3b9f97f61d0bcd

Observation 8cf06d0b-ed6a-4d69-8d5c-d4b138ca2a04 · outbound

This paper cites SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.040035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.040035Z digest=sha256:bd8f4aa908ba8f1d66ede9dcedb194fc2b6835626faa7b743a8bc689e5cfb0db

Observation 90ac718d-d84d-4824-acef-9aeb41f4943a · outbound

This paper cites Orga- nizing, orchestrating, and benchmarking agent skills at ecosystem scale.arXiv preprint arXiv:2603.02176, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Orga- nizing, orchestrating, and benchmarking agent skills at ecosystem scale.arXiv preprint arXiv:2603.02176, 2026

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.079299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.079299Z digest=sha256:87dd170d4984054c1ad7e105fe9e84ad0ae08b470027a4867f03527f5f7aaf29

Observation f0ee1b70-516b-47a8-9423-c4facdd1a8ff · outbound

This paper cites Skillnet: Create, evaluate, and connect ai skills.arXiv preprint arXiv:2603.04448, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Skillnet: Create, evaluate, and connect ai skills.arXiv preprint arXiv:2603.04448, 2026

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.169353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.169353Z digest=sha256:d56fdbf83b1aec3356b4b8dc4cd6f0e602c267d6649fba1721e2bd3780b7b015

Observation a58fb1c1-d4fd-4fbf-9923-db935930c6b9 · outbound

This paper cites Autoskill: Experience-driven lifelong learning via skill self-evolution.arXiv preprint arXiv:2603.01145, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Autoskill: Experience-driven lifelong learning via skill self-evolution.arXiv preprint arXiv:2603.01145, 2026

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.251238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.251238Z digest=sha256:a025980da5d6b2901914fbae34331ff83c7167277a7a70644696b95b548bf49c

Observation 075b6390-617e-4a08-86a9-68a5ad758c82 · outbound

This paper cites Memento-skills: Let agents design agents.arXiv preprint arXiv:2603.18743, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Memento-skills: Let agents design agents.arXiv preprint arXiv:2603.18743, 2026

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.343681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.343681Z digest=sha256:e50fbeaaeaf4f9ebdb5606b4ad6fad0758b774ea12dd3d06480a96f9c6aca3b5

Observation 8bc841d7-9ea7-4f04-a959-4cd49f36f544 · outbound

This paper cites SLBench: Evaluating How LLM Agents Follow Logical Relations in Skills.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? SLBench: Evaluating How LLM Agents Follow Logical Relations in Skills

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-06T15:38:43.949998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:38:42.420308Z digest=sha256:93dd7d8f9e446c2dadaa6ada5ee9f6993757554335b561e2731e511a468628f0

Observation 36817ce0-c3c4-4a23-871f-a55c30c2e68d · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Instruction-Following Evaluation for Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.552670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.552670Z digest=sha256:d567e8a16855b7a737ee62bb857ba8906100cddb3edb9b2f0085de8428f716be

Observation 606233ca-dde7-41df-8d6f-d25fc75a3a00 · outbound

This paper cites Followbench: A multi-level fine-grained constraints following benchmark for large language models.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Followbench: A multi-level fine-grained constraints following benchmark for large language models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:38:46.246312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:38:42.639788Z digest=sha256:3fc72a413a4027109ad4e8958e67a5cb95238dcef08d9b1a4b02b52a49f4e9b2

Observation d2ce020d-3290-40f0-86f7-c18c856f7429 · outbound

This paper cites Infobench: Evaluating instruction following ability in large language models.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Infobench: Evaluating instruction following ability in large language models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.724332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.724332Z digest=sha256:c47187cad85e6a6df2e254683c8bdcd51f1cd9c26c82a4082a809696b9c3e430

Observation cf969f85-d06d-4fc5-b477-7d19481f3089 · outbound

This paper cites Benchmarking complex instruction-following with multiple constraints composition.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Benchmarking complex instruction-following with multiple constraints composition

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:38:46.007953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:38:42.800877Z digest=sha256:5d8e0189e371a2cf16d54f70337b026b067d8a21452fb5a0223e8f41faa0a4af

Observation 07c56e30-fdd0-4fbd-bf28-edf6934c4a1c · outbound

This paper cites Agentif: Benchmarking large language models instruction following ability in agentic scenarios.Advances in Neural Information Processing Systems, 38, 2026.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Agentif: Benchmarking large language models instruction following ability in agentic scenarios.Advances in Neural Information Processing Systems, 38, 2026

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:38:45.742772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:38:42.897685Z digest=sha256:73bda81d6f11fd485ce78f373d04cf26a5a6270fc9df835cd9ff5399761c27fd

Observation c2026e00-7ded-4a48-b435-ab6447975160 · outbound

This paper cites Stop Comparing LLM Agents Without Disclosing the Harness.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Stop Comparing LLM Agents Without Disclosing the Harness

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:42.965973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:42.965973Z digest=sha256:83be3283b1376d490984c2daaf359f615edf6a09de64af6dc71cd3a88c338f55

Observation 395af4c8-0ea7-40dd-9ffb-33525113c553 · outbound

This paper cites Harness-Bench: Measuring Harness Effects across Models in Realistic Agent Workflows.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Harness-Bench: Measuring Harness Effects across Models in Realistic Agent Workflows

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:43.056799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:43.056799Z digest=sha256:c0e7af85d8da977b17c9ceaa05ff47e74afdc9c9cc5b7db2505c7c211a8de231

Observation 3470191b-dca1-4e18-8fc5-7d464a915668 · outbound

This paper cites Sysbench: Can llms follow system message? InThe Thirteenth International Conference on Learning Representations, 2024.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Sysbench: Can llms follow system message? InThe Thirteenth International Conference on Learning Representations, 2024

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:38:45.568923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:38:43.120628Z digest=sha256:cd6e1c1899f2094595450fa67ece94a9ca86f1055b9288b65afc8aef4dcb40ce

Observation a5877146-a186-4571-8180-498013d79e94 · outbound

This paper cites Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:43.205719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:43.205719Z digest=sha256:63026fa55a33bd31aabc9296fbcd1fa41c2722b330f0d820a9223d760a5bfdf8

Observation aa05d250-1d29-4d0f-89e4-cd52dd829559 · outbound

This paper cites Judging llm-as-a-judge with mt-bench and chatbot arena.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Judging llm-as-a-judge with mt-bench and chatbot arena

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:43.274595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:43.274595Z digest=sha256:1ada9b371a47006b2563f30ba18a147a7c446123f09952a35fb1b6ea4252050f

Observation 581b78c4-ada0-4025-b4bf-4daf605c7e84 · outbound

This paper cites SOPBench: Evaluating Language Agents at Following Standard Operating Procedures and Constraints.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? SOPBench: Evaluating Language Agents at Following Standard Operating Procedures and Constraints

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:43.353203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:43.353203Z digest=sha256:744eba7203aeaf54a7ac5a4b2350dfeda71557c76baa74fa0a8bfd3ebe189078

Observation 4200e5b6-42e8-4fd7-90d2-5fa0a37f6bd6 · outbound

This paper cites an unresolved cited work.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:38:45.416713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:38:43.436668Z digest=sha256:2e216630a4fc35181b5f651d17c5417f663d081d422ea219243356fce9fd93c2

Observation 486e0e27-5156-4bdc-a184-fe574b52ac5a · outbound

This paper cites an unresolved cited work.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:38:45.104227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:38:43.511861Z digest=sha256:81d632d6269ac0d8209563ec4be8e2eac0e7003e1089db209f436b37fb504e15

Observation b6bc6ce9-2d16-4f92-a12c-2763c6f3590f · outbound

This paper cites an unresolved cited work.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:38:44.786035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:38:43.586596Z digest=sha256:20f7b0c29fdbbd59176cb55d20251b7d967c8b5e7d33eed4d0b9060de01c36ee

Observation 115f6e3a-ae6c-4593-bd56-6b6c61ec52f1 · outbound

This paper cites pdftk-server.

Skill-Use: Can LLMs Actually Use Skills in Agentic Harnesses? pdftk-server

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:38:44.514436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:38:43.670203Z digest=sha256:d4fbd18a7ff997b9708b28f93dd20f281b5d5282ddf3b3ef8dec28d79ff95fe8

Pith citing papers

No inbound Pith citation observations are available.