Pith. sign in

Paper Citation Record · LEDGER

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models

As of 11 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2509.10744.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.10744 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T17:41:26.749626Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved25
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8486a350-9521-4f10-84f8-264e20c14bf8 · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.324388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.324388Z digest=sha256:d6ec559dbe14e7224b645ccc926e97cb3718b3b44d4896b9fe60263607350b48

Observation 42f280f6-3db1-4a6f-b8b0-d8da1ce95671 · outbound

This paper cites 2023.RADIATION AND CAN- CER BIOLOGY STUDY GUIDE.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models 2023.RADIATION AND CAN- CER BIOLOGY STUDY GUIDE

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.392622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.392622Z digest=sha256:f070d2a0208a3f47f5f7d31f49bcf1d08dec25eed753dfa983e0186b0dc881c6

Observation 2723ce5c-c6d4-474e-854d-4209523dabdd · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.466641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.466641Z digest=sha256:c8efda35988b67cf814fe50e697271aadbb00c1995137c251e7b662de541511c

Observation 9ed580f0-67bd-44f9-ab7b-ad35badd4874 · outbound

This paper cites Qwen Technical Report.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Qwen Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.542294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.542294Z digest=sha256:8128ca71143f0cda8384691734a1ba755678a384efbd1190ccf2edbcc3216bbc

Observation 38455af8-88ef-45e3-b4cc-7a583862dd47 · outbound

This paper cites Beattie, S.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Beattie, S

Reference 5

Resolution
verified exact
doi, observed 2026-08-04T17:44:06.657194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-04T17:41:24.599332Z digest=sha256:484859ba3420fa7b32afc57dba66a26d899cfe5dc371ef9603bb754eb1575872

Observation 4361f049-0e8b-407b-a518-08ab7e77b44d · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.693090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.693090Z digest=sha256:4585224cf7e5ecda294655f386f98f3c5a782a1c66e1e17d39064932a3a396c8

Observation 3180879f-8562-4e9e-aceb-d2f3dc9be986 · outbound

This paper cites 2024.Science and Engineering Indicators 2024: The State of U.S.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models 2024.Science and Engineering Indicators 2024: The State of U.S

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.730975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.730975Z digest=sha256:f0784df6bd08c19c9df9d6855ccb7bfccd6c54386363b18d46f30023fa7f79f2

Observation 989a7007-af5a-41f5-b050-a7c96927f59e · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.794393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.794393Z digest=sha256:56e6f83463540022525951db02a0543ae721769b0dd8870e069cfbb6bf8bfc37

Observation a6c64ead-25fc-4b2d-9c18-e5671b6153c4 · outbound

This paper cites The Faiss library.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models The Faiss library

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.828179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.828179Z digest=sha256:30169353281384831fe8f4c2d3def163c07347948d90c22a681b9af746fab778

Observation 4048d7d2-cec4-4399-aee4-bdcb8c762105 · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.900523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.900523Z digest=sha256:1663f90f984069d459313d14ce60882173a4f7bdbfbc15e1de712c65063cbe36

Observation c675e996-8c89-482c-aac3-1b4f477fddc5 · outbound

This paper cites The Llama 3 Herd of Models.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models The Llama 3 Herd of Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.955443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.955443Z digest=sha256:b02b28eab36f512ef2c2cf9126103c6514deb9688e08e7306224ca72624b25db

Observation 62aebe7e-b7c5-4516-ad99-d2d10bb53810 · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:25.057261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.057261Z digest=sha256:0a7722e96be388e50d1a52c251d3e402ce922e8e5962759a4470e4a2035b5b68

Observation e6a80fa6-8f8a-4100-b825-882ace8c95dc · outbound

This paper cites Measuring Massive Multitask Language Understanding.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Measuring Massive Multitask Language Understanding

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:25.150748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.150748Z digest=sha256:bf74c1248f124e19d5a907f66abf2c3b92618d0cc354a53aba505075b8014eb4

Observation 12cefcf7-075d-48fd-b8d9-25def0368f62 · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:25.253320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.253320Z digest=sha256:fa66c79b4d5445b0b9613ccba828ea6d2e67bd9981332683ec3daa7bed791a02

Observation f1e518fe-2069-408d-915c-6c20be29c67d · outbound

This paper cites Mistral 7B.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Mistral 7B

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:25.356677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.356677Z digest=sha256:353bf7a74d98b0ff7ee54fb4f6b69248b70d0f86dd7208d520393f0dede0048b

Observation d27d0ddc-3885-4cf0-9321-4380fbebe63a · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 16

Resolution
malformed identifier
no resolver link, observed 2026-08-04T17:41:25.470163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.470163Z digest=sha256:b221497ca20917ffe60c29f974605ed790da7c3545ef13f33779ff9e237f503f

Observation 3d437974-6ad9-4d14-bf0f-ffd8a2cb1c7f · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:25.555929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.555929Z digest=sha256:900153ac97a49557974071c92247efbf2446bc40e2aebfbe044506551404321c

Observation 84c28f20-78fa-4c01-b41e-8b6ed0414fdd · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:25.608152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.608152Z digest=sha256:72663de6c2fc572b84ad3cb85a24d5534c0d244d2c43e01551e3267e76617d7e

Observation 8a08fa86-f3d1-48a8-b5bc-162b2240f87e · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:25.685924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.685924Z digest=sha256:dbb3737ddafde4491524d5f48eb050ce9c6fbb66d1c085bd169f192810bed1b6

Observation 9c914f5a-12a2-4d9f-9177-e6e9d074ff7c · outbound

This paper cites OLMo: Accelerating the Science of Language Models.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models OLMo: Accelerating the Science of Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:25.771827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.771827Z digest=sha256:e406f7ee4d76c917e524b1b0623470084f6d9b039e37b849fa40962337f7413b

Observation fe18cd74-ba37-409d-8e09-481a06146177 · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:25.884961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.884961Z digest=sha256:ce375ff50d6ff68044b37ac923f308ede001e9bde570e68261101595a87ba3f7

Observation 7439b183-a787-42a9-8c11-efce408a164a · outbound

This paper cites AdaParse: An Adaptive Parallel PDF Parsing and Resource Scaling Engine.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models AdaParse: An Adaptive Parallel PDF Parsing and Resource Scaling Engine

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:26.013320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:26.013320Z digest=sha256:4fb6c5656acd4b62e2e0e6a1cb00c379f023f38ddf373ffe4cd0e72abadf32a6

Observation 090749f7-9de1-42a0-8109-3e0e569522a0 · outbound

This paper cites Gemma 3 Technical Report.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Gemma 3 Technical Report

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:26.134486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:26.134486Z digest=sha256:29a00d3d97f70dffccb9b0d7c8a0243f1440e19169dc9db5c044bc23a14c72fa

Observation dcd6ef8f-2025-4bdb-bd5d-88e958fd75f3 · outbound

This paper cites AstroMLab 1: Who Wins Astronomy Jeopardy!?.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models AstroMLab 1: Who Wins Astronomy Jeopardy!?

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-04T17:44:06.150232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-04T17:41:26.265139Z digest=sha256:56e81cf46560d006d51ad83861a48851cbd0d234349a51ed07d8b21ab2d9a90f

Observation cc4e81db-6201-47b3-8dab-57b9685da8b5 · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:26.408422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:26.408422Z digest=sha256:229305980022afdaf3d491a34fb66fd9992088f80b932479f69e470b76f72e93

Observation d1ec3a66-d79d-44e4-8d7e-513ba62d7d8e · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:26.633666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:26.633666Z digest=sha256:4972f871c636690bbbd2bb7fe9c36e47eb517c94f71a000da61a2d6e449412c1

Observation c24f4045-7edf-465f-b4bf-a04bb640e1fb · outbound

This paper cites TinyLlama: An Open-Source Small Language Model.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models TinyLlama: An Open-Source Small Language Model

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:26.749626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:26.749626Z digest=sha256:a8241410f8c142274cf75cffe6292f650e7f0433676102726619a1636892dc59

Observation 7c317ed6-eaf3-4c24-9754-50e0d01248c6 · outbound

This paper cites InExtended Semantic Web Conference.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models InExtended Semantic Web Conference

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:26.516157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:26.516157Z digest=sha256:510c7972a18f58f5e0323c5b77cde820a7fb0cdfe5d88dea8c06584196264643

Pith citing papers

No inbound Pith citation observations are available.