Pith. sign in

Paper Citation Record · LEDGER

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories

As of 8 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 0 inbound Pith citation observations for arXiv:2507.22086.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.22086 v1

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:12:27.279857Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

20 of 20 outbound references displayed

  • verified exact1
  • verified fuzzy5
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8522e1f7-83b8-4382-8ada-484bc2d8e6fa · outbound

This paper cites GPT-4 Technical Report.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T13:12:25.144642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:12:25.144642Z digest=sha256:a21316cdca1224bdd8ee15ca7916ab5dc419a7d5354d658af42aa7b7cde3d3e7

Observation ec4ecc15-ae91-4602-96e3-e0e07251f99c · outbound

This paper cites The Llama 3 Herd of Models.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories The Llama 3 Herd of Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T13:12:25.587220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:12:25.587220Z digest=sha256:1e60c6d54a9cf4c93765161e006e480f0f0784feb89a6eed7e5e7c1b41bed212

Observation a2053efd-16bd-41ca-9a2e-829bd07f44ba · outbound

This paper cites J., Bird, C., Barr, E.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories J., Bird, C., Barr, E

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:12:29.306879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:12:25.670459Z digest=sha256:9cae5849fa7b1cc9d504b90eb5ff66ebcc0292feadf225a1f88984301882ac80

Observation 313ae3af-7d69-4a9a-991c-b8d67e535382 · outbound

This paper cites DeepSeek-V3 Technical Report.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories DeepSeek-V3 Technical Report

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T13:12:25.896581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:12:25.896581Z digest=sha256:604f8077aeb0ecc9f066e23942474cebebba88d61e3636bbdaf5f789d8371d3a

Observation 3d0ffcb1-2cfd-4476-a5c5-b2a92334c48d · outbound

This paper cites an unresolved cited work.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T13:12:28.816108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:12:26.149430Z digest=sha256:ebc2e145d815e9da1adc09ab5b2b4cd6dffc8476fbd0e5a2c5279fd7a3f1dcac

Observation 62677a37-6445-43a2-a151-546f381c4141 · outbound

This paper cites P., Sabu, S., Wang, J., M.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories P., Sabu, S., Wang, J., M

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:12:28.579917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:12:26.369220Z digest=sha256:fd191b38db18533765f6b257ac060cf37245485f8dbb482126e98677cbe270ee

Observation b9adfcaf-32a1-4c59-8e7d-4503dcc2fa11 · outbound

This paper cites Repotrans- bench: A real-world benchmark for repository-level code translation.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories Repotrans- bench: A real-world benchmark for repository-level code translation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T13:12:26.467618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:12:26.467618Z digest=sha256:bb429cba8c82a329b3197e138caed5113c10c944818f035289b7de8fccf1d61c

Observation 8addf16b-77ef-4ab1-8250-d51afb7b2ac5 · outbound

This paper cites Qwen2.5 Technical Report.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories Qwen2.5 Technical Report

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T13:12:26.605674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:12:26.605674Z digest=sha256:1aee805323f88c7a6ef0e0b2a1c3e463f885cc2dbceef4c6777b070a41d36766

Observation d619baaa-5ab7-4ab6-94c7-115876ba40bb · outbound

This paper cites BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T13:12:26.731265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:12:26.731265Z digest=sha256:0f5cdbf73e741b954880b7fbeec44f14c491984800c7a32dad46edf9c2ba50d6

Observation 0f4ab164-89c7-4161-bd98-8b34d370b2b0 · outbound

This paper cites an unresolved cited work.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T13:12:28.329334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:12:26.835649Z digest=sha256:837205954df084d5518661a2cd940b9440206fc397667288feb5aba4ef394bbc

Observation c42e22cd-197f-4ada-9a83-c266a5426851 · outbound

This paper cites The test sets are further split into two test1 and test2 based on the date created to test the data contamination issue.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories The test sets are further split into two test1 and test2 based on the date created to test the data contamination issue

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:12:28.216862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:12:26.928560Z digest=sha256:6c5b4c41cd36479e6d94aaadcfb06fd40aca40359283a9c7be5a5375d38b701a

Observation 8ceb5ea3-078b-43e0-b82f-5fee4bfd689f · outbound

This paper cites an unresolved cited work.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-06T13:12:28.100105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:12:27.069811Z digest=sha256:90bd9932b627c7f5d876cbe4235398b77a12e092727813be19edc87d3d0ce4a3

Observation ff0cbb13-768b-4a6b-aba0-9004afde136c · outbound

This paper cites an unresolved cited work.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-06T13:12:28.020057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:12:27.204923Z digest=sha256:88de2fbfccd281bc725ce0580caf78dd7484387fd76abd0be12bc2103b052d0b

Observation 02b0cae8-38ac-489e-be96-ef4b0d8276e4 · outbound

This paper cites an unresolved cited work.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-06T13:12:27.881031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:12:27.279857Z digest=sha256:86155cd8ce3317774f5f50b7bc3f748319f3a4bd6abb0734e3f84c13dccc638f

Observation f51def23-a51e-4190-9756-42c279a66d77 · outbound

This paper cites Data Contamination Through the Lens of Time.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories Data Contamination Through the Lens of Time

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-06T13:12:26.242662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:12:26.242662Z digest=sha256:c34a037a3d6108db5721d4a7122fe0e1adf6fee034f914c2e34b776aa92ded3a

Observation b3b83ac6-b9a9-4554-aea3-49dd60cbc2dc · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-06T13:12:25.780029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:12:25.780029Z digest=sha256:259bf9ddc3c23eda7c37bbbe2798896b56d580ce5d04d943e8edbef3b374ccd2

Observation 46b94a5b-c9e8-470f-bb59-88695b3faaf7 · outbound

This paper cites Pyright: Static type checker for python.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories Pyright: Static type checker for python

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:12:29.042908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:12:26.036090Z digest=sha256:b688229c216651c4ab002067607c3c23a173ab287b839de70f63c36238452dfa

Observation 624e43b7-e775-4632-adfb-fece35d6f95d · outbound

This paper cites APPL: A Prompt Programming Language for Harmonious Integration of Programs and Large Language Model Prompts.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories APPL: A Prompt Programming Language for Harmonious Integration of Programs and Large Language Model Prompts

Reference 2021

Resolution
verified exact
local_arxiv, observed 2026-08-06T13:12:27.644129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:12:25.496645Z digest=sha256:9d3fe6f7d6ebabc408dc6e69b66ab43111ad5364d21a868a49d85473b50df216

Observation bf30dcd8-a15b-4e87-95b2-92f1631972fc · outbound

This paper cites Evaluating Large Language Models Trained on Code.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories Evaluating Large Language Models Trained on Code

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T13:12:25.366886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:12:25.366886Z digest=sha256:57ce303588d141310a4b21db435929ecad0a6799d8329d3dcf29f312f72c38f0

Observation eb5268bc-b9bb-43a4-894c-7d6bd9650e62 · outbound

This paper cites Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J.

TypyBench: Evaluating LLM Type Inference for Untyped Python Repositories Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:12:29.540548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:12:25.258861Z digest=sha256:cd2faf10475713ef146de40cc8283e1ade910821bf360ab94bc8317bf492c6a9

Pith citing papers

No inbound Pith citation observations are available.