Pith. sign in

Paper Citation Record · LEDGER

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models

As of 16 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 0 inbound Pith citation observations for arXiv:2607.19608.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.19608 v1

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T12:21:27.902347Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8a5cf67e-ce82-42fc-bd4b-95d8ae5124cc · outbound

This paper cites Scaling Laws for Neural Language Models.

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models Scaling Laws for Neural Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T12:21:27.187689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:21:27.187689Z digest=sha256:6c39b7e49d2824fb1f01775815bd9dfd284248683c9c21c53a2b2a19af873307

Observation 9905b508-1e12-42b5-a6f7-c3507413d7a0 · outbound

This paper cites InProceedings of the 2018 Conference on Empirical Methods in Natural Language Processing, pages 2381–2391, Brussels, Belgium.

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models InProceedings of the 2018 Conference on Empirical Methods in Natural Language Processing, pages 2381–2391, Brussels, Belgium

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T12:21:27.434213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:21:27.434213Z digest=sha256:1841bcc86077abbeafd8081d8addce00b1bfbe03b6cfe0b81604999f3180e24e

Observation 7b233d8a-7bab-46f6-ae6c-c3af734f5e82 · outbound

This paper cites Jason Wei, Maarten Bosma, Vincent Y.

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models Jason Wei, Maarten Bosma, Vincent Y

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T12:21:27.732290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:21:27.732290Z digest=sha256:e2cbffe5c337172f651711aaa1810d6f509b0949b40582e1e5e7a4d6471dc4ed

Observation 1454e511-acdc-4761-85ae-7fcfc4019a0a · outbound

This paper cites In International Conference on Learning Representa- tions, volume 2024, pages 40193–40219.

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models In International Conference on Learning Representa- tions, volume 2024, pages 40193–40219

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T12:21:27.817717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:21:27.817717Z digest=sha256:8e0f437f564aca3cf444ffefe9787e2418b6e08b38a76918fa617598a143145f

Observation 51222f51-d4e5-4ec0-8ff6-a7b2d03ac1f5 · outbound

This paper cites InProceedings of the 2015 Conference on Empirical Methods in Natural Lan- guage Processing, pages 1743–1752, Lisbon, Portu- gal.

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models InProceedings of the 2015 Conference on Empirical Methods in Natural Lan- guage Processing, pages 1743–1752, Lisbon, Portu- gal

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-01T12:21:27.522508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:21:27.522508Z digest=sha256:12bbb1a1d2896d50f621b74b4b54c8df50725b3105495d495842c96c9c6f7a9c

Observation 3aa5ae0f-8446-43d3-9d19-f18c9d9cbba5 · outbound

This paper cites an unresolved cited work.

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models Unresolved cited work

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-01T12:21:27.263986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:21:27.263986Z digest=sha256:d19283444b3f1d8ca504f31da9b4a92e4c120505465223aa32f3516a28041ecb

Observation eaab194c-25fc-4e3c-a2f0-9b801b099ed5 · outbound

This paper cites In Proceedings of the 2017 Conference on Empirical Methods in Natural Language Processing, pages 785– 794, Copenhagen, Denmark.

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models In Proceedings of the 2017 Conference on Empirical Methods in Natural Language Processing, pages 785– 794, Copenhagen, Denmark

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-01T12:21:27.344201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:21:27.344201Z digest=sha256:1440c1b7c6d9ee3c71846fdea68d64a9d374261e46cf4d9cdf9ca225f4902a56

Observation 8fbfd539-fd3a-4116-b824-a70cce5c63d8 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-01T12:21:27.069244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:21:27.069244Z digest=sha256:04fd62558d0a1c5e68bcf8452db5a52cf569835d69c0e7b41f97c2be1ba69c86

Observation 35ad0d64-4d48-4513-8a40-e99799afc4f9 · outbound

This paper cites FinBERT: Financial Sentiment Analysis with Pre-trained Language Models.

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models FinBERT: Financial Sentiment Analysis with Pre-trained Language Models

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-01T12:21:26.881155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:21:26.881155Z digest=sha256:60e527aa5386e2e695390c6c930e359529fa21863a7e13fabd62f1a86a73cb9e

Observation 3ba57595-a891-4b7b-98ed-1f7bc48748ce · outbound

This paper cites InAdvances in Neural Information Processing Systems, volume 33, pages 1877–1901.

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models InAdvances in Neural Information Processing Systems, volume 33, pages 1877–1901

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-01T12:21:26.949299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:21:26.949299Z digest=sha256:19804d43fd9c5c4cc306aba629a5d732e62fd8c843b4e74f343b0a599501a789

Observation 83e97f75-e045-41b8-bc55-953a4db40ff0 · outbound

This paper cites Multitask Prompted Training Enables Zero-Shot Task Generalization.

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models Multitask Prompted Training Enables Zero-Shot Task Generalization

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-01T12:21:27.618990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:21:27.618990Z digest=sha256:79949dc3c1b25d1ec0d0d0a2350fd3418e10f6544dbbc113335d25da5b6e4215

Observation bc5af543-c740-4fdf-ad7d-be4b7cf8fb3c · outbound

This paper cites InProceedings of the 2022 conference on empirical methods in natu- ral language processing, pages 5085–5109.

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models InProceedings of the 2022 conference on empirical methods in natu- ral language processing, pages 5085–5109

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-01T12:21:27.683212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:21:27.683212Z digest=sha256:31a62c68a64f2a0fb8e9d6b4da29197041be48171547719c094f86196ed52873

Observation cfbc7349-0859-404c-b1c9-28b8cfe6629c · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models Instruction-Following Evaluation for Large Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-01T12:21:27.902347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:21:27.902347Z digest=sha256:90a20e8a9965e95bab322ece6446eceb0d35b48ee14ce6ffbeb3e7aac2dfeee4

Observation 42b8d5a2-167a-4323-ae69-8184e535f646 · outbound

This paper cites InFindings of the Association for Computational Linguistics: EMNLP 2024, pages 1691–1706.

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models InFindings of the Association for Computational Linguistics: EMNLP 2024, pages 1691–1706

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T12:21:27.011124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:21:27.011124Z digest=sha256:1b98ee0d73c7aca987525d1820a9e81e99bf75655de8693d6f44bf632c637ec7

Observation 6b8c3433-9075-43b9-84cc-c0f0a60d5cc9 · outbound

This paper cites Qwen-Scope: Turning Sparse Features into Development Tools for Large Language Models.

Task Competence Is Not Instruction Following: Evaluating Instruction-Conflicting Behavior in Small Language Models Qwen-Scope: Turning Sparse Features into Development Tools for Large Language Models

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-01T12:21:27.112037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:21:27.112037Z digest=sha256:3ce89141ef29f743e61fbc9e72283cd25ffeb535c719fb1c3fc5f0b507504800

Pith citing papers

No inbound Pith citation observations are available.