Pith. sign in

Paper Citation Record · LEDGER

Exploring LLM Autoscoring Reliability in Large-Scale Writing Assessments Using Generalizability Theory

As of 18 August 2026, this Paper Citation Record lists 8 of 8 outbound references and 1 inbound Pith citation observation for arXiv:2507.19980.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.19980 v2

Coverage vector

measured 8 of 8 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:55:24.034052Z

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T14:19:01.351948Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-15T14:19:01.474990Z

Reference resolution

8 of 8 outbound references displayed

  • verified exact2
  • verified fuzzy2
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fbb965c4-d1f7-49f1-83b2-77e60388529b · outbound

This paper cites an unresolved cited work.

Exploring LLM Autoscoring Reliability in Large-Scale Writing Assessments Using Generalizability Theory Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-06T13:55:24.390766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:55:24.006017Z digest=sha256:3a2df0c7a07bc667962e7037a24382d95502ef05e4cd5b44c5808a3655687693

Observation 4a9686e6-eba0-4cff-8dd0-b62fd262fb19 · outbound

This paper cites an unresolved cited work.

Exploring LLM Autoscoring Reliability in Large-Scale Writing Assessments Using Generalizability Theory Unresolved cited work

Reference 3

Resolution
verified exact
raw_fallback, observed 2026-08-06T13:55:24.277218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:55:24.023357Z digest=sha256:005ef37437dfdebfcb685259fd40e229512ef0cea08dd6d39348200ef9e2a63f

Observation 4a359c86-654b-4540-adb6-4d57f8f9bdb0 · outbound

This paper cites an unresolved cited work.

Exploring LLM Autoscoring Reliability in Large-Scale Writing Assessments Using Generalizability Theory Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T13:55:24.372501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:55:24.014410Z digest=sha256:410d4835586f83cec32c4179cb164206f2dcf952f392652a98a8331d1aee0d12

Observation 35037e15-774f-474b-b170-dacdad40baaf · outbound

This paper cites G study variance and covariance components were estimated for the seven score effects associated with the 𝑝•×𝑡∘×𝑟• design.

Exploring LLM Autoscoring Reliability in Large-Scale Writing Assessments Using Generalizability Theory G study variance and covariance components were estimated for the seven score effects associated with the 𝑝•×𝑡∘×𝑟• design

Reference 5

Resolution
verified exact
raw_fallback, observed 2026-08-06T13:55:24.350485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:55:24.018359Z digest=sha256:1dc0cceb5cea0d371ce23a93dcfe08f52deb6a40d2107d86833bd216f5be60d8

Observation 204749dd-5bc4-4fb4-980c-f49378a09a54 · outbound

This paper cites an unresolved cited work.

Exploring LLM Autoscoring Reliability in Large-Scale Writing Assessments Using Generalizability Theory Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T13:55:24.026919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:55:24.026919Z digest=sha256:378b709168a47f21928b4b5fd5fe71f9411e1d867c209ce2bb5ab84a1ed03a58

Observation 13c161e8-c8c6-4b5b-a8b2-98b2cc8d0b1b · outbound

This paper cites https://doi.org/10.1016/j.jsp.2017.12.005 31 Wilson, J., Chen, D., Sandbank, M.

Exploring LLM Autoscoring Reliability in Large-Scale Writing Assessments Using Generalizability Theory https://doi.org/10.1016/j.jsp.2017.12.005 31 Wilson, J., Chen, D., Sandbank, M

Reference 7

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T13:55:24.152575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:55:24.031006Z digest=sha256:4f6c8fa6ff3446c3218254f145aed1bdf462c8ef5efada8fc1c4dd054bd5183d

Observation 0b2c7cc9-1437-4fd4-aea4-34aeaf2623d7 · outbound

This paper cites Similarly, Yancey et al.

Exploring LLM Autoscoring Reliability in Large-Scale Writing Assessments Using Generalizability Theory Similarly, Yancey et al

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:55:24.381854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:55:24.010684Z digest=sha256:f9e3e8bf36603e8c88ef41d51c9a0272ed4489ffaa42f3e07a38e4ff03ffb8d1

Observation 5a192637-10bd-4b4e-ad93-e9db15f36e75 · outbound

This paper cites 576-584).

Exploring LLM Autoscoring Reliability in Large-Scale Writing Assessments Using Generalizability Theory 576-584)

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:55:24.363814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:55:24.034052Z digest=sha256:1efc1afba2a817742d2393f70ac62b613fa5a5eb0498999cb8a0b91f75297d86

Pith citing papers

Observation db44260c-5bec-440f-9530-32e592f4064e · inbound

Deployment Decision Reliability: A Generalizability-Theory Framework for Sizing Long-Horizon Agent Evaluations cites this paper.

Deployment Decision Reliability: A Generalizability-Theory Framework for Sizing Long-Horizon Agent Evaluations Exploring LLM Autoscoring Reliability in Large-Scale Writing Assessments Using Generalizability Theory

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-15T14:19:01.482213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T14:19:01.351948Z digest=sha256:bdd1e008b925d886b8f5f984f8e46c1f5853bda8f166be6ae84b54e06391cce7