Pith. sign in

Paper Citation Record · LEDGER

AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone?

As of 5 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 1 inbound Pith citation observation for arXiv:2510.05432.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2510.05432 v2

Coverage vector

measured 16 of 16 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-11T11:50:26.030339Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-11T02:48:15.074349Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-05-11T03:05:55.205507Z

Reference resolution

16 of 16 outbound references displayed

  • verified exact8
  • verified fuzzy3
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e0e6b61f-32b9-4429-bce1-7430a7774220 · outbound

This paper cites Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901.

AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone? Language models are few-shot learners.Advances in neural information processing systems, 33:1877–1901

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T09:31:12.007180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T09:30:16.609270Z digest=sha256:f9ec47372e0b32ad5b714c0717f5e00dd9753cdb66858c38b7c7c78bbe8e2d62

Observation 89d4cdaf-50e5-408c-b0b0-e4dee967960d · outbound

This paper cites On the Measure of Intelligence.

AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone? On the Measure of Intelligence

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T09:31:11.237718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T09:30:16.609270Z digest=sha256:ee462d35aa0afc04aca58d898247fffacc71c3e829e6c2cf8219a8033a8bfcba

Observation f8767d78-1f19-4143-b60d-ccc37d77edd3 · outbound

This paper cites CURIE: Evaluating LLMs On Multitask Scientific Long Context Understanding and Reasoning.

AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone? CURIE: Evaluating LLMs On Multitask Scientific Long Context Understanding and Reasoning

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T09:31:11.201411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T09:30:16.609270Z digest=sha256:617d532116f9c4a2c5ce303cea41ddfe1eee67a4b1d8edc155ec557f60ac5557

Observation a0548e15-3c1f-4b27-a8e8-29d3fe291ca8 · outbound

This paper cites MathOdyssey: Benchmarking Mathematical Problem-Solving Skills in Large Language Models Using Odyssey Math Data.

AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone? MathOdyssey: Benchmarking Mathematical Problem-Solving Skills in Large Language Models Using Odyssey Math Data

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-18T09:31:11.248044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T09:30:16.609270Z digest=sha256:3c5200d12ab48ec5e23939d71f5767076b8b7ce328698c05b8a3d492057ddb7d

Observation 3e3f18c2-724b-4cb0-833b-78b453f74951 · outbound

This paper cites G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment.

AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone? G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-18T09:31:11.221679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T09:30:16.609270Z digest=sha256:3d2fe7531831f118bfb75901a7ada8f8d099b32990bcd167c3610ad0ae266cba

Observation aa673644-ee47-4c5a-8833-d4c948f71515 · outbound

This paper cites Entity-based knowledge conflicts in question answering.

AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone? Entity-based knowledge conflicts in question answering

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T09:31:12.000419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T09:30:16.609270Z digest=sha256:5cfc2e19ea3d7f51b5ff7060f103e67c00604648667485499a0cfb8585e74c3a

Observation 6679bc5a-970b-42b4-b066-360504ae336f · outbound

This paper cites BIGbench: A Unified Benchmark for Evaluating Multi-dimensional Social Biases in Text-to-Image Models.

AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone? BIGbench: A Unified Benchmark for Evaluating Multi-dimensional Social Biases in Text-to-Image Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-18T09:31:11.206284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T09:30:16.609270Z digest=sha256:13db0c1082a56b6b93daebd2c7f3a61770409bd8757c7b4575b914ee4ffbe572

Observation 68557ee4-9c99-4f9c-8079-bcc37189df79 · outbound

This paper cites an unresolved cited work.

AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone? Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-05-18T09:31:12.004057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T09:30:16.609270Z digest=sha256:665ad3186d3920b562493f977d363dff6622ea3ddda894f160a1168277218458

Observation 47d06d17-f112-4a03-b87d-3c9e086379ac · outbound

This paper cites In-context Learning and Induction Heads.

AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone? In-context Learning and Induction Heads

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T09:31:11.056757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:d75eba5137c999797ac62a9b2979e3c351e34bac8b6c10d5cfd53f601a23b178

Observation cca66dfa-ebc1-417e-8c22-6c09af738b88 · outbound

This paper cites Grokking: Generalization Beyond Overfitting on Small Algorithmic Datasets.

AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone? Grokking: Generalization Beyond Overfitting on Small Algorithmic Datasets

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-18T09:31:11.227224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T09:30:16.609270Z digest=sha256:39ea2ce700c134aa7678f5b435cac960fcc735e6fe35cb49f613496879b7dff1

Observation be219fc4-03bb-432a-9d62-0976c019161c · outbound

This paper cites Prompt programming for large language models: Beyond the few-shot paradigm.

AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone? Prompt programming for large language models: Beyond the few-shot paradigm

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T09:31:12.010207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T09:30:16.609270Z digest=sha256:8ac0f2e3dbb795d6d906d9023d50133620c467c8076635b1780c5964f2ee1a31

Observation 98bd970a-b86f-4b7d-9f79-6dc6f6bf8d7b · outbound

This paper cites Evaluation Metrics in the Era of GPT-4: Reliably Evaluating Large Language Models on Sequence to Sequence Tasks.

AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone? Evaluation Metrics in the Era of GPT-4: Reliably Evaluating Large Language Models on Sequence to Sequence Tasks

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-18T09:31:11.253621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T09:30:16.609270Z digest=sha256:5d4e1552eae232a2bc2f97afbf086fe27177cca25e8ee2a9794262ae6f5f434f

Observation c38a01d4-0375-4f26-9482-7d42b34410bc · outbound

This paper cites SciBench: Evaluating College-Level Scientific Problem-Solving Abilities of Large Language Models.

AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone? SciBench: Evaluating College-Level Scientific Problem-Solving Abilities of Large Language Models

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-18T09:31:11.242556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T09:30:16.609270Z digest=sha256:e1d08514b06f301cd1ec2aa6590ede2045d53223f79c0799b18ee2f0e2bebbcb

Observation 0faf3c5a-bbc8-48ae-bf7c-831a63c16d12 · outbound

This paper cites Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models.

AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone? Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models

Reference 14

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T09:31:11.232362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T09:30:16.609270Z digest=sha256:4b39d4c318412c10dca7afe264042ffd6324698ef419ed6d31b800e2c8a3c4d9

Observation f7da91b1-f1dd-456a-b1ba-5e84fe50dbd7 · outbound

This paper cites Advancing the Scientific Method with Large Language Models: From Hypothesis to Discovery.

AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone? Advancing the Scientific Method with Large Language Models: From Hypothesis to Discovery

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-18T09:31:11.211183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T09:30:16.609270Z digest=sha256:5135848bbfd0985e8c638b945a0d6599462ed34e5a9b69a276193a269cb91f63

Observation 0c9e4792-1cae-481e-a328-fc9052fa1a82 · outbound

This paper cites Knowledge Augmented Complex Problem Solving with Large Language Models: A Survey.

AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone? Knowledge Augmented Complex Problem Solving with Large Language Models: A Survey

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-18T09:31:11.216601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T09:30:16.609270Z digest=sha256:a3505b6190612f7b0b530f7f64cf6f6e07b761942325f5f635c880fc9d47752b

Pith citing papers

Observation 0537282e-f30d-4deb-a9ee-ef9a65f2e6a2 · inbound

FAME: Forecasting Academic Impact via Continuous-Time Manifold Evolution cites this paper.

FAME: Forecasting Academic Impact via Continuous-Time Manifold Evolution AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone?

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-05-11T03:05:55.209103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T02:48:15.074349Z digest=sha256:64cfb7726ddcc279cd55e4c4b19a2895aa863395add94665c30c7666ac62de57