Pith. sign in

Paper Citation Record · LEDGER

Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2306.11270.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.11270 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T22:54:48.347005Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:58:57.677920Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7efef2b2-85b0-4390-957a-33ec2e23c547 · inbound

Beyond Prompt Content: Enhancing LLM Performance via Content-Format Integrated Prompt Optimization cites this paper.

Beyond Prompt Content: Enhancing LLM Performance via Content-Format Integrated Prompt Optimization Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-08T22:54:48.347005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:54:48.347005Z digest=sha256:456f6f2fb6c7051b38e942360d08aa47faa52eb8631fba9063166ab39e08e778

Observation 99af6741-ddb7-48db-8fce-5ec9a5da0b7e · inbound

Investigating the Robustness of Retrieval-Augmented Generation at the Query Level cites this paper.

Investigating the Robustness of Retrieval-Augmented Generation at the Query Level Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T18:55:54.909052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:55:54.909052Z digest=sha256:78f89e9a079fa09bb1fa276d50e264a7c38e307347d4b94820593ab156bc20a9

Observation 65904cfc-6982-4dd2-b3c3-78d6f22486b2 · inbound

PlotTwist: A Creative Plot Generation Framework with Small Language Models cites this paper.

PlotTwist: A Creative Plot Generation Framework with Small Language Models Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T18:06:40.507466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:06:40.507466Z digest=sha256:d73731a445b44b7240532a956883132d0f2d01cc0f731c26a4279fcb49359a26

Observation 7443e39b-e8e7-4138-9734-159fa2c6c416 · inbound

Compared to What? Baselines and Metrics for Counterfactual Prompting cites this paper.

Compared to What? Baselines and Metrics for Counterfactual Prompting Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

Reference 164

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:56:27.679391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-09T19:02:46.991897Z digest=sha256:4a1004a8e0e737c5983d668bd79840ba0b59cd11b6dbd5ee3153b9338fde1e52

Observation bb2d9f26-1830-4dd4-af72-4cb87a97d3f6 · inbound

Towards Context-Invariant Safety Alignment for Large Language Models cites this paper.

Towards Context-Invariant Safety Alignment for Large Language Models Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:39:40.890191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-21T05:36:12.562807Z digest=sha256:2a8dd36fd909524bec812cd7fb601377cdaef86921067332b223d5e5004e804a

Observation e58e76aa-9eb0-4a23-ad73-5cbda748458f · inbound

Same Patient, Different Words, Different Diagnosis? Evaluating Semantic Stability in Clinical LLMs cites this paper.

Same Patient, Different Words, Different Diagnosis? Evaluating Semantic Stability in Clinical LLMs Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T07:23:13.274114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T07:14:22.444020Z digest=sha256:05ff78913b1f3f07a99af71e3ff00623fd2856d9bf348b33fbe7145fdf3a1821

Observation c103f8f5-6aaf-4a19-8e22-f6956c057a90 · inbound

Prompt Perturbation for Reliable LLM Evaluation over Comparison Graphs cites this paper.

Prompt Perturbation for Reliable LLM Evaluation over Comparison Graphs Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:58:57.679681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T01:03:20.261268Z digest=sha256:cf44e51b9c725877f57961ca59204d4eab111b959fbdc78df1cb6ced7bde6a86

Observation 6d28b749-88ac-4093-a4af-38eab487840a · inbound

The Powerless Noise: How Experimental Settings Shape the Reported Power of Noise cites this paper.

The Powerless Noise: How Experimental Settings Shape the Reported Power of Noise Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-12T01:10:52.578408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:10:52.578408Z digest=sha256:6485ffd12f9bc75b5044adf914df396bba4dc7762025f554adeb93b95c279c5a

Observation 688187b5-de1e-443f-b59d-069303e3f4ca · inbound

The Powerless Noise: How Experimental Settings Shape the Reported Power of Noise cites this paper.

The Powerless Noise: How Experimental Settings Shape the Reported Power of Noise Evaluating the Zero-shot Robustness of Instruction-tuned Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-13T07:04:18.645464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T07:04:18.645464Z digest=sha256:8b6b6a25915ee1453bcdc2d9835d0e4c46fe55afeb7025713cfee5c90fcd2086