Pith. sign in

Paper Citation Record · LEDGER

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning

As of 7 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2607.22996.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.22996 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T04:00:54.997773Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

19 of 19 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a28832b8-786c-4f28-835e-fa85a29f33d3 · outbound

This paper cites GPT-4 Technical Report.

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T04:00:54.939835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:00:54.939835Z digest=sha256:b0985d8da0e7b04ca2c661146634205833b4ff39bbf99f94157ae2214cb48d79

Observation 81b2dffd-c3e2-489f-ae3f-d0d9bacef4ca · outbound

This paper cites Addison Wesley Longman, Inc.

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning Addison Wesley Longman, Inc

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T04:00:54.944418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:00:54.944418Z digest=sha256:8c240fbf4386b13fca9e571624366d1df63416a9d967d91a54eeaa074fde58ba

Observation d0d031f8-9a0a-4af7-b8bb-431945942992 · outbound

This paper cites Handbook 1: Cognitive domain.

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning Handbook 1: Cognitive domain

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T04:00:54.947754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:00:54.947754Z digest=sha256:aaf7d1ed88ba330d609a48bc1b5a0b711e5105e9e1adcdb04a3fd20509090158

Observation 0514c1df-1f5e-4680-8a83-3c25ae562808 · outbound

This paper cites IEEE Transactions on Education 48(4), 612–618 (2005).

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning IEEE Transactions on Education 48(4), 612–618 (2005)

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T04:00:54.951100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:00:54.951100Z digest=sha256:7a20c1b70d217969e1295ebb1ff106dc401a138dbde71420cd082a8324dd6f34

Observation dc6b65db-6aa4-44fe-842b-811eeaf6f6a1 · outbound

This paper cites In: HGAIS@ ISWC (2024).

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning In: HGAIS@ ISWC (2024)

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T04:00:54.954431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:00:54.954431Z digest=sha256:eac7b2f1b43d8c2fb66cdf335a41b3e07bdf5c6a54518aeb4adac6a035f5d70d

Observation 2557b7f1-2458-41cc-bfed-92ea4d90e849 · outbound

This paper cites In: International Conference on Learning Representations.

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning In: International Conference on Learning Representations

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T04:00:54.958006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:00:54.958006Z digest=sha256:ecc1dc6dd264208fb6f22c98519fb8bcc9225bd6493b076981e69dcfb1b62072

Observation d59d26da-0620-49bc-ba2d-e93034b74acf · outbound

This paper cites In: Text summarization branches out.

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning In: Text summarization branches out

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T04:00:54.961399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:00:54.961399Z digest=sha256:ed8d430e7440afc2da5f5a6f3c3bcc38044eb57ec0e033fecb379db1948f48f1

Observation d842d21a-1454-473c-ac49-e2a926e24c4b · outbound

This paper cites arXiv preprint arXiv:2508.06583 (2025).

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning arXiv preprint arXiv:2508.06583 (2025)

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T04:00:54.964373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:00:54.964373Z digest=sha256:0ad1d415843875b30475f6130915c7b1b4bcb0063f5b8ec86132170f67dcfdbf

Observation 5999ecfe-c60d-4b5b-94bd-846cfd0e9548 · outbound

This paper cites In: Findings of the Association for Computational Linguistics: EMNLP 2023.

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning In: Findings of the Association for Computational Linguistics: EMNLP 2023

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T04:00:54.967429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:00:54.967429Z digest=sha256:d8a22c45a935378641bd8d6ab25f198e1980fbf621dbee165d335fe0313b4d73

Observation 2b2fc866-584a-48d4-89fd-b214e73a1ae0 · outbound

This paper cites In: Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics.

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning In: Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T04:00:54.970153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:00:54.970153Z digest=sha256:0c31eb2511fed27e477163fe9389a070f80f3ef04214521261fc13664f49b1ca

Observation 26b3d22a-b454-4ce6-ac57-ef4d7a4219d0 · outbound

This paper cites Advances in neural information processing systems35, 27730–27744 (2022).

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning Advances in neural information processing systems35, 27730–27744 (2022)

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T04:00:54.972904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:00:54.972904Z digest=sha256:bf2876b2d3f86beab7e5a95ad9fd91bc741811f43004d853db723a38f0e42ad4

Observation 1848c2a5-4421-4783-9980-d891baf5b47e · outbound

This paper cites In: Proceedings of the 40th annual meeting of the Association for Computational Linguistics.

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning In: Proceedings of the 40th annual meeting of the Association for Computational Linguistics

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T04:00:54.975744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:00:54.975744Z digest=sha256:705f1f37c0cd1ccc7d33b0fa4fc5ba41a30ea8d53c4fff8970ff65a5c5543cde

Observation 1eb3251b-39a2-44fc-81cb-d2645cb7628d · outbound

This paper cites In: Proceedings of the 2019 conference on empirical methods in natural language processing and the 9th international joint conference on natural language processing (EMNLP-IJCNLP).

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning In: Proceedings of the 2019 conference on empirical methods in natural language processing and the 9th international joint conference on natural language processing (EMNLP-IJCNLP)

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T04:00:54.979022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:00:54.979022Z digest=sha256:026f4e2e0216b6a8a58d2ecb908e886921273f34927c30a08557f5875206294c

Observation 9c0f9505-bd6c-4103-9acf-3d2b73cf73e9 · outbound

This paper cites In: Proceedings of the 2022 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies.

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning In: Proceedings of the 2022 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T04:00:54.982121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:00:54.982121Z digest=sha256:c995c481e2442b5511e5f88b7d26344a9f4c1eeecff626aa42da52cdac714454

Observation f52e00ed-97c1-43f9-a3f0-b4a27a4d85d9 · outbound

This paper cites Simulated Students in Tutoring Dialogues: Substance or Illusion?.

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning Simulated Students in Tutoring Dialogues: Substance or Illusion?

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T04:00:54.984768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:00:54.984768Z digest=sha256:1e1e1700918dd4d9c9ad7c406f9b20640677369e3d6b7343f77a1df2ae4185da

Observation ba13ca8b-8c8b-47ba-b7ed-316bcf040b86 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T04:00:54.988084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:00:54.988084Z digest=sha256:23ac843a6fbe2255f834fea2b2fe625e7b44ee9aeab8f64985e0345b1ff9ff45

Observation ce6b6d91-1879-4855-962e-645e52ebe9b4 · outbound

This paper cites The AI Teacher Test: Measuring the Pedagogical Ability of Blender and GPT-3 in Educational Dialogues.

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning The AI Teacher Test: Measuring the Pedagogical Ability of Blender and GPT-3 in Educational Dialogues

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T04:00:54.991558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:00:54.991558Z digest=sha256:710ab0332977d11d40f3de5f90bb777f8e0e5639313c3dd186441a42f42c2d67

Observation 6cc5f055-9f18-4687-ac23-aec4a34e87b5 · outbound

This paper cites an unresolved cited work.

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T04:00:54.994742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:00:54.994742Z digest=sha256:1b3c95d01f080bba15dfe072d315836b56a584463cce9481235c61ed2a0dfdd2

Observation 7dfdbb62-80a9-4e68-8f97-759ead65877b · outbound

This paper cites Emergent Abilities of Large Language Models.

Beyond Direct Answering: Aligning Educational LLMs as Socratic Guides via Heuristic Reinforcement Learning Emergent Abilities of Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T04:00:54.997773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:00:54.997773Z digest=sha256:9ce7fcac657dd1ce52c5466e3fd81291044a754c538aa2f60872d9046fb4108c

Pith citing papers

No inbound Pith citation observations are available.