Pith. sign in

Paper Citation Record · LEDGER

Hypothesis generation and updating in large language models

As of 5 August 2026, this Paper Citation Record lists 100 of 300 outbound references and 1 inbound Pith citation observation for arXiv:2605.05851.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.05851 v1

Coverage vector

measured 100 of 300 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-08T14:43:56.125546Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-30T19:24:34.330570Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 300 outbound references displayed

  • verified exact63
  • verified fuzzy19
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch11

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9030ad06-daeb-4e7e-a634-20715395635e · outbound

This paper cites NVIDIA Nemotron 3: Efficient and Open Intelligence.

Hypothesis generation and updating in large language models NVIDIA Nemotron 3: Efficient and Open Intelligence

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:40:43.189893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:2c36a5e9401fa4bf5637f373eb060e1c0a1aadc8a6f122e05ba18ba24f521ddf

Observation 9b5a28c2-1800-4576-853a-0f2f120064de · outbound

This paper cites Google , author =.

Hypothesis generation and updating in large language models Google , author =

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.865128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:cbb2721d9dd184d0df014805b0a06bc32e35b4a229843aa87b093f77db3d474f

Observation ec88b1fe-2ef9-4b4e-9878-b90fe54d4fde · outbound

This paper cites From ai for science to agentic science: A survey on autonomous scientific discovery.

Hypothesis generation and updating in large language models From ai for science to agentic science: A survey on autonomous scientific discovery

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.274884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:92427329052fd26eedcc76e16bd71818de08e26a4fd106025bf45e1540aaf4a1

Observation 790abe83-e7c2-4b7d-bffd-541e91b301ff · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-05-26T11:37:37.848455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:9f0764d2cc9e0425ce550398f0d6ad6efcc75ee75a1470271b1ee55958f75586

Observation d1fec973-a031-4a46-b0ed-5e094d57f379 · outbound

This paper cites Kosmos: An AI Scientist for Autonomous Discovery.

Hypothesis generation and updating in large language models Kosmos: An AI Scientist for Autonomous Discovery

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T08:40:55.079629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:a440da8e8e649d791044171420467d47438e19a05fe697f3231eb522ce3db67c

Observation a2edbde2-1db6-4964-bed1-03454921ef3d · outbound

This paper cites AI-Researcher: Autonomous Scientific Innovation.

Hypothesis generation and updating in large language models AI-Researcher: Autonomous Scientific Innovation

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-08T21:14:12.390138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:6c1e53740cbe4e1bed835d4a0099d8d46369e784ddbd21113b71e34a9a4d7df6

Observation 3b9fb214-94fc-42ce-8882-7553b7a3ab1f · outbound

This paper cites The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search.

Hypothesis generation and updating in large language models The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-09T01:44:55.508555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:9799128a4ac5ca5f4870c870dda38b01bdaa42b6491a58a427ac377ce3cc84b0

Observation 4702acf5-838b-4c25-8899-c735d731a04e · outbound

This paper cites Towards an AI co-scientist.

Hypothesis generation and updating in large language models Towards an AI co-scientist

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T13:02:46.484476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:12d3af2ad8cd7bb5a88366e51926af7b7f20329f72e8585c0345d8872c444b03

Observation 565ab26d-3941-41b3-ab0c-372112869588 · outbound

This paper cites year =.

Hypothesis generation and updating in large language models year =

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.839265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:1938b6dfea4f34088031d1fd434334bacb3b98820d5075167be24c1b67fbaac9

Observation e838d177-a85e-4a0c-95b3-bb6cf9e5c228 · outbound

This paper cites Cognitive basis of language learning in infants , volume =.

Hypothesis generation and updating in large language models Cognitive basis of language learning in infants , volume =

Reference 10

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.132977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:fec628df78071fe274cd1216189225d286657e709040455135d6a159ba11de36

Observation d7e2bfd7-41ca-4699-a401-61d6c0006fe3 · outbound

This paper cites Journal of Child Language , author =.

Hypothesis generation and updating in large language models Journal of Child Language , author =

Reference 11

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.097757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:db38cd11b846cafff9df93e0a706a02b37d00fb1a829c6f0db69eca7a727988a

Observation 920bce99-57ac-4a11-a7e7-52204fb82e95 · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-05-26T11:37:37.877646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:17e74f541639eb2f21c7f091cc8c3e1f0c6d6d10615514ec8689f32563664517

Observation ec7c6509-430e-4217-b21b-499937208c35 · outbound

This paper cites How LLMs Detect and Correct Their Own Errors: The Role of Internal Confidence Signals.

Hypothesis generation and updating in large language models How LLMs Detect and Correct Their Own Errors: The Role of Internal Confidence Signals

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-08T21:14:12.052674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:f675c071c4d300f383fec9fe7a903be865daf513838ff8fc2ca9f1802a2d91a6

Observation 309979ad-b93f-4dc1-b552-0bf3f3cb9cbb · outbound

This paper cites Reasoning Theater: Disentangling Model Beliefs from Chain-of-Thought.

Hypothesis generation and updating in large language models Reasoning Theater: Disentangling Model Beliefs from Chain-of-Thought

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-29T03:04:37.872643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:ef90d5187c3ce97384a5d5167f2847533523c666b0ea2b5398ebe1a16ba2d078

Observation d101b226-0181-410a-b54c-0d954c171684 · outbound

This paper cites Real-Time Progress Prediction in Reasoning Language Models.

Hypothesis generation and updating in large language models Real-Time Progress Prediction in Reasoning Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-27T02:05:03.856858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:b869b1b686b6399fd8b1f7c5bb8f20427904ff4cec9eed3b332c3d34a90ecb18

Observation e23c75d5-2a11-46e4-85fa-f7bd3e166791 · outbound

This paper cites Calibrating Reasoning in Language Models with Internal Consistency.

Hypothesis generation and updating in large language models Calibrating Reasoning in Language Models with Internal Consistency

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.111952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:914490e409eab983c8a8ca4e44305660707a2adcfa899919c47d37a795cd642d

Observation 3b89ed0e-7218-45c9-9346-85425f3c3574 · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 17

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.014581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:3d4751819946ea2530333fb2cc1e4ca7068e2d29d321d14a6dc7bf556195d4fc

Observation 796152bf-b314-4302-94e4-7aadf23dfc05 · outbound

This paper cites Reasoning with Sampling: Your Base Model is Smarter Than You Think.

Hypothesis generation and updating in large language models Reasoning with Sampling: Your Base Model is Smarter Than You Think

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:16:05.617665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:1f2fd13fd6160483317a8fa41e27197683e7a73e78e8f0f0d8d27ee2941490b2

Observation 7296389a-8f57-4c26-b371-89e5c905f99e · outbound

This paper cites Chen, J., Hu, S., Liu, Z., and Sun, M.

Hypothesis generation and updating in large language models Chen, J., Hu, S., Liu, Z., and Sun, M

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.591282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:558bcc8e234373a0d12706e4c45d7cb5acef8826a0c0b1419f9143af53d1cb88

Observation aedd9192-98de-49cb-b0cc-8a072bbdbf32 · outbound

This paper cites Temporal.

Hypothesis generation and updating in large language models Temporal

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.484443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:ee42c5d9f40b414fdfd1f7d4598e1b63bada3b680580f1f33da2a9f06e3b77eb

Observation f06e2c7e-5456-48d2-83a6-a3975ed6a966 · outbound

This paper cites LLM Reasoning as Trajectories: Step-Specific Representation Geometry and Correctness Signals.

Hypothesis generation and updating in large language models LLM Reasoning as Trajectories: Step-Specific Representation Geometry and Correctness Signals

Reference 21

Resolution
metadata mismatch
local_arxiv, observed 2026-05-08T21:14:12.466420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:17d6d291d3ef45b74c4d2ff5d43ccd692be4df045bd7b5265005a57a774b302b

Observation 738f7e20-5654-48cc-8d9e-dce89599c7aa · outbound

This paper cites Efficient PRM Training Data Synthesis via Formal Verification.

Hypothesis generation and updating in large language models Efficient PRM Training Data Synthesis via Formal Verification

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-08T21:14:12.601324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:61c7774bbcd63f85121d8414f7688c0996f23bc870dedd55e941aee2eca8e8fd

Observation ce6dcc1d-ed4c-4d33-a7ef-ec9fa6b64cdf · outbound

This paper cites Truth as a.

Hypothesis generation and updating in large language models Truth as a

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.858080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:d70f35cc9c3825fd7bfed96276ac891c29b00d84d59986debb4d766f72815058

Observation ae5f2d7e-07d3-4ef5-bae1-48a8449934dd · outbound

This paper cites Closing the confidence-faithfulness gap in large language models.

Hypothesis generation and updating in large language models Closing the confidence-faithfulness gap in large language models

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.002783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:6ed780929af63210616bdbbbc0269dabb7ac045d9cb33aad5c8909aac882f16c

Observation cab58a57-38f9-4728-8732-bf73168dec87 · outbound

This paper cites How do LLMs Compute Verbal Confidence.

Hypothesis generation and updating in large language models How do LLMs Compute Verbal Confidence

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-20T02:04:53.698588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:d78db2ec7a49d00b2a7ca136e0dca830c319623634c452cab7a4d9c0015fa56e

Observation 934401ba-8901-4f67-91ce-b32cc00281ac · outbound

This paper cites ReProbe: Efficient Test-Time Scaling of Multi-Step Reasoning by Probing Internal States of Large Language Models.

Hypothesis generation and updating in large language models ReProbe: Efficient Test-Time Scaling of Multi-Step Reasoning by Probing Internal States of Large Language Models

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-08T21:14:12.439467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:96f92e4c5f6de901b0e65827853181d74b7aed6f2437ae2bb0741d110855fe06

Observation d7a8dd25-1125-4982-8dbc-6ea001b6f201 · outbound

This paper cites Do LLMs Know What They Know? Measuring Metacognitive Efficiency with Signal Detection Theory.

Hypothesis generation and updating in large language models Do LLMs Know What They Know? Measuring Metacognitive Efficiency with Signal Detection Theory

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-15T01:20:58.272624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:1ffcbc133ed8b6f15d55a85ef1075ccd4c28481dc4a04acdd6254c536f83750a

Observation cec467fe-4509-43b4-956d-92da3564b497 · outbound

This paper cites Evidence for.

Hypothesis generation and updating in large language models Evidence for

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.864758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:c4baae2489670014a3e7e696de0aa15d64e3492dbe42939ab685a0b1b64dd126

Observation 11b3bda7-934c-40a8-944d-f32ba7120fa2 · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-05-26T11:37:37.888437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:86f715abfff8f71e8696018a25b058a3428c28b185c7fc537a1b7f416126a3e8

Observation 74dd7402-4515-4cd8-903a-80bffca283c7 · outbound

This paper cites Learning to.

Hypothesis generation and updating in large language models Learning to

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.586341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:8b01f5b83a48afcbc1d139d6486423b8d5656a87669486ea9b82df79f481b87b

Observation 40059f7c-f8a4-46ab-827f-183b3fbbac62 · outbound

This paper cites More Capable, Less Cooperative? When LLMs Fail At Zero-Cost Collaboration.

Hypothesis generation and updating in large language models More Capable, Less Cooperative? When LLMs Fail At Zero-Cost Collaboration

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-08T21:14:12.343277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:a674193a875fa3fa20a4469bb0dd22d5049e2aa3233e41198b8e46f65d1129d1

Observation 943be80b-38f7-4956-8235-374e2932df85 · outbound

This paper cites Lossless data compression by large models , volume =.

Hypothesis generation and updating in large language models Lossless data compression by large models , volume =

Reference 32

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.334953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:52119513bbb646044450491ecdeeb163298d634fb2ab1fdb47187db77ac369a5

Observation 6707a743-e3a0-430a-a3bd-1873db53943b · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-05-26T11:37:37.851930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:abddcbad5e49cc0e339ad3866a2442c76ebb771fd5fb57375cb29d4dd03885da

Observation ebee73ff-dd3a-49be-b827-49aadcdf00df · outbound

This paper cites The No Free Lunch Theorem, Kolmogorov Complexity, and the Role of Inductive Biases in Machine Learning.

Hypothesis generation and updating in large language models The No Free Lunch Theorem, Kolmogorov Complexity, and the Role of Inductive Biases in Machine Learning

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.395122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:9e3769dd3f62b2e1a23c8654693c5b206bc69e3374bc6e99f02e452ddfc8f87b

Observation 167d62cb-ba3c-493e-9cae-c352b6d7ea6e · outbound

This paper cites Language.

Hypothesis generation and updating in large language models Language

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.842558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:5429a41a4635079f9af6bde87507264b01a82355b5254c51220fe90fe1365ebc

Observation c04cc3f6-0291-4638-b927-ec9287d678c1 · outbound

This paper cites On predictability of reinforce- ment learning dynamics for large language models.CoRR, abs/2510.00553.

Hypothesis generation and updating in large language models On predictability of reinforce- ment learning dynamics for large language models.CoRR, abs/2510.00553

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.318454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:9ea89ad254ac26f914d8cb72b4a3f19e9a0e089e52081a87371f26a6e1596f39

Observation 24b3e082-9f2c-4b90-a23d-f7b730fd2790 · outbound

This paper cites Advances in Neural Information Processing Systems , author =.

Hypothesis generation and updating in large language models Advances in Neural Information Processing Systems , author =

Reference 37

Resolution
verified exact
doi, observed 2026-05-08T21:14:11.849183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:04c1fa01e4f9362713dfbde65104b499706921acec3cec2549879f5d5cae8c09

Observation 666ed916-1592-458c-ae35-78a183f763ec · outbound

This paper cites Transactions on Machine Learning Research , author =.

Hypothesis generation and updating in large language models Transactions on Machine Learning Research , author =

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.869270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:72233584ad3eacd75c911201b98861da789d9b41c88a13bd2f052b82417b4fd4

Observation c8c91847-9a8f-442d-8c80-6c4a83d67baa · outbound

This paper cites BIG-Bench Extra Hard.

Hypothesis generation and updating in large language models BIG-Bench Extra Hard

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.648060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:9ba75847d7e0adaa412ff22bd925f3350828b6998b763bea7e027517f360ca55

Observation df13b7df-a3a1-4a93-a922-01e6b55f9af7 · outbound

This paper cites The Lookahead Limitation: Why Multi-Operand Addition is Hard for LLMs.

Hypothesis generation and updating in large language models The Lookahead Limitation: Why Multi-Operand Addition is Hard for LLMs

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.026036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:80721cc31bc954c1702a0b6aa47b13c619c61e214dbd7ef6b2ade42c39558f51

Observation 4cedac66-791a-4cf0-aee5-09055db8150a · outbound

This paper cites CRUXEval: A Benchmark for Code Reasoning, Understanding and Execution.

Hypothesis generation and updating in large language models CRUXEval: A Benchmark for Code Reasoning, Understanding and Execution

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T20:57:16.443569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:031e3e727f649f507c50a09c03d4a2804952514d52b4fe0fea9bdf4779fd6762

Observation 46cc2235-3f58-473f-8c6d-596d72a4c114 · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-05-26T11:37:37.893579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:21829e8a1b09dfb956f90b6daba49932d5d6519f355af8449d66da5767dd4740

Observation 307ff8af-470a-4658-b215-82d2cac54d82 · outbound

This paper cites and Moros-Daval, Yael and Zhang, Seraphina and Zhao, Qinlin and Huang, Yitian and Sun, Luning and Prunty, Jonathan E.

Hypothesis generation and updating in large language models and Moros-Daval, Yael and Zhang, Seraphina and Zhao, Qinlin and Huang, Yitian and Sun, Luning and Prunty, Jonathan E

Reference 43

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.365090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:14de62bce66172eb6919b526f1df2dc26f6805fd0e80476b16c4163e5b6d5cff

Observation 3c8dc81b-6c3d-4e59-9333-7661731bf3b6 · outbound

This paper cites What and.

Hypothesis generation and updating in large language models What and

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.861475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:32a6e08d56374a30d675097627ae3c7eba027485fdec9dcb61bb28f441bc8fea

Observation 10d9403e-fb7d-4245-9450-56cbdf70834f · outbound

This paper cites Revisiting the.

Hypothesis generation and updating in large language models Revisiting the

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:11.905598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:8f0dbec853c397b1001ce69e2b67307a7d579595a0d3c2fec28bbbd56c822b92

Observation 89bbc8d3-5a58-48bd-9d2e-b9b3fd00e777 · outbound

This paper cites and Piantadosi, Steven T.

Hypothesis generation and updating in large language models and Piantadosi, Steven T

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.855949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:2944484a8e32f156ef69fadf4ab1383d1c0ad7ff955e916331f064fd36778587

Observation 62b22744-6742-4ac3-9272-812ec128ea37 · outbound

This paper cites 2023 , keywords =.

Hypothesis generation and updating in large language models 2023 , keywords =

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.862580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:786d3647d20d0f0e7b7489198c878da2b38f3c05d1b171e8161539ab4f73c918

Observation 64ac58bc-3f4a-4b14-87b4-465d3bd8781f · outbound

This paper cites Journal of Open Psychology Data , author =.

Hypothesis generation and updating in large language models Journal of Open Psychology Data , author =

Reference 48

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.360252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:1a33af6ce5b15fb1fde3e0ca5621378d372a5409fce0ae77a8a4b8bb720da201

Observation 842efd26-b153-4b83-a3e7-5e69534230ed · outbound

This paper cites Rules and.

Hypothesis generation and updating in large language models Rules and

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.860176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:f37a4edfe329efc0cc5f1e83ebb9706679b6610f4a7dcedb3d320ec6e6db72ec

Observation 807a7adf-db54-4099-bb4d-ece52fbf9e0f · outbound

This paper cites month = apr, year =.

Hypothesis generation and updating in large language models month = apr, year =

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.208720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:970419f2e86b727fc3fd09d2ecdcbcfaa88b6b5de72147342e23a98b317530b8

Observation b86f6e62-086c-436f-9dd1-89c6ee9f4bb5 · outbound

This paper cites Bayesian teaching enables probabilistic reasoning in large language models , volume =.

Hypothesis generation and updating in large language models Bayesian teaching enables probabilistic reasoning in large language models , volume =

Reference 51

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.033039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:cdeb5475542df8d44bd1635adad97e70eaa35084cc3f9d27201966287ab4ec36

Observation 20b79598-a638-4265-baf1-645ee0c42272 · outbound

This paper cites Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing.

Hypothesis generation and updating in large language models Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-16T06:58:37.155837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:59b2897b661d116b0ec71f6c59c80258846e3544bdc4059a1caff57b2cd89421

Observation 0b494f22-b586-40aa-b01b-f9a62948d817 · outbound

This paper cites On Language Models' Sensitivity to Suspicious Coincidences.

Hypothesis generation and updating in large language models On Language Models' Sensitivity to Suspicious Coincidences

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.424085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:bd5a3fb435c4cdb276374dd8fd62defd22f818f5cc02c0f4beee53281a86d2cd

Observation da5818ff-a8b9-442f-9ed5-04fd2856a259 · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:11.926703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:4d3cf4a372305b9057e2f754ed0a3523ad4e3ddcc41eab25fa4e16f97bcde26b

Observation 606be60e-364a-4153-ad06-b5cf858f1bae · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.349649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:c35f448d70a365472f941a4346ff17b415f1ab2a177435152e1adfec98f9cdfe

Observation a9d14390-04ab-4387-9f8e-a6b2c7d95f72 · outbound

This paper cites Revisiting Uncertainty Estimation and Calibration of Large Language Models.

Hypothesis generation and updating in large language models Revisiting Uncertainty Estimation and Calibration of Large Language Models

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.533436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:968eefde8192b422904ee3578d3daa3da00592d715f401a6e9cf69025192c247

Observation ef6e4ed6-30c8-4bb8-bfb5-d117879768b1 · outbound

This paper cites Kapoor, N.

Hypothesis generation and updating in large language models Kapoor, N

Reference 57

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.614338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:06f2b833174a854056599c2f8864895f729727686f4846899e074b4fd3f9bd63

Observation d7b39099-2378-481a-85d8-d7f7fa4b4d9e · outbound

This paper cites The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity.

Hypothesis generation and updating in large language models The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:10:31.796621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:ec947f599f9347864ee3338c12776912758735e103519a80f8f30d764e6166fc

Observation bd9e3a89-5a5c-4d6a-91de-e68dbff3f2f5 · outbound

This paper cites Contextual Position Encoding: Learning to Count What's Important.

Hypothesis generation and updating in large language models Contextual Position Encoding: Learning to Count What's Important

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.639564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:ad09187fb853441f4b1e579e1d2f5360840bf03d9515c0b25eab740eb6f57249

Observation e413db63-9e03-480d-b264-6c04ac38ec80 · outbound

This paper cites arXiv preprint arXiv:2509.10739 , year=.

Hypothesis generation and updating in large language models arXiv preprint arXiv:2509.10739 , year=

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.330921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:86d96ce83f037c5b74b78f86b3efcf6757ce2551bd084971f2a09af5075fbbda

Observation f00305ee-4de6-4c8f-ba48-55b9b1f716c2 · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-05-26T11:37:37.937338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:8d0bb9e4b87c33b6adc1175542578134eadf325ecc6126f2fc47c579450f90cc

Observation 3de428d9-d1db-4970-95c1-a49be9cfe006 · outbound

This paper cites an unresolved cited work.

Hypothesis generation and updating in large language models Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-05-26T11:37:37.918363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:12cb862573faa7d26533b76a1679bdc7e76313accf025b8e0471f708a6c922c5

Observation 68d7090a-5f83-4c76-a79e-f2fb09b009f5 · outbound

This paper cites Position.

Hypothesis generation and updating in large language models Position

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.915230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:ed5da7b03141eef11af891f12f724bbd0104370e804687183170381e554b0af9

Observation d406970f-48ef-4374-9dee-06a89e1321cb · outbound

This paper cites EleutherAI Blog , author =.

Hypothesis generation and updating in large language models EleutherAI Blog , author =

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.911880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:422502a7e085978cf7f35e51a6b146888fde3e12141ac4226c82658e1c6c29bf

Observation f4a461d7-b74d-43f9-9e19-185f33a674a1 · outbound

This paper cites Embarrassingly Simple Self-Distillation Improves Code Generation.

Hypothesis generation and updating in large language models Embarrassingly Simple Self-Distillation Improves Code Generation

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-06-26T01:15:18.478768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:5b8d4f577ae8765318e4c14a9f3115a5ee54dc180cab6ecc27c258e1859f5f61

Observation 791cac6d-ab96-4894-99a4-67548f373226 · outbound

This paper cites Learning Montezuma's Revenge from a Single Demonstration.

Hypothesis generation and updating in large language models Learning Montezuma's Revenge from a Single Demonstration

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.576022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:0f0784e1dd72956bb8f23691c7950eea5e2527742c6b4a586cbce68e4a423a3d

Observation 63804fe2-61a3-411d-abcd-5750c045f237 · outbound

This paper cites Cooperative Inverse Reinforcement Learning.

Hypothesis generation and updating in large language models Cooperative Inverse Reinforcement Learning

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.404414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:fd1f43a0c4c72255afdf995f4b845c9bbe5b1488887b7d673226fa669186a9a9

Observation cfde71cf-a7da-422d-8b3d-6f14e32563c3 · outbound

This paper cites Test your best methods on our hard.

Hypothesis generation and updating in large language models Test your best methods on our hard

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.853527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:d07973f3e81a32ef21224684d7529bed2872424ddff29cba375b017bf1b94300

Observation 13784bc3-b8cd-4a38-916c-2ebe908fa1a7 · outbound

This paper cites SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization.

Hypothesis generation and updating in large language models SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization

Reference 69

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T00:00:37.467312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:7147aecafe6b55607e89c43cc01e6782acb20c45ca6b51abd542f908f28c9737

Observation b9ed458f-64a0-418d-9f1f-8e7d0ffd7a44 · outbound

This paper cites Cognitive Dark Matter: Measuring What AI Misses.

Hypothesis generation and updating in large language models Cognitive Dark Matter: Measuring What AI Misses

Reference 70

Resolution
metadata mismatch
arxiv_id, observed 2026-08-03T03:15:37.660537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:9a887c10e5417c06c304b68516a36d0468c7739d9cfd5abeacd7507f5f3dab3e

Observation 8f40c6f6-7722-473c-b08e-9b4bcaf10497 · outbound

This paper cites Cognitive models and AI algorithms provide templates for designing language agents.

Hypothesis generation and updating in large language models Cognitive models and AI algorithms provide templates for designing language agents

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.461251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:0b3a276d73c9d3bd1fd23aba7e33a99d923fc021e90d57e18b5323b01a0f7a2f

Observation fcf0aac7-129a-4af7-b8c9-649661eb5185 · outbound

This paper cites year =.

Hypothesis generation and updating in large language models year =

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.905582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:d7c0a9385bfca60883e5e16a43eb08a30774f7997cf596c182ea9bbb2c0071e6

Observation 0edeb54a-7728-4a43-9d32-2e43cd6f93c0 · outbound

This paper cites and Vezhnevets, Alexander Sasha and Diaz, Manfred and Agapiou, John P.

Hypothesis generation and updating in large language models and Vezhnevets, Alexander Sasha and Diaz, Manfred and Agapiou, John P

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.456537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:4f771a9b3bde7637133e477d600bbb1265f5d1a0428bf7ea8c5af3bf811fd6d6

Observation 15508f36-4631-4690-a8ba-d89d2e139d7c · outbound

This paper cites Superposition Yields Robust Neural Scaling.

Hypothesis generation and updating in large language models Superposition Yields Robust Neural Scaling

Reference 74

Resolution
verified exact
local_arxiv, observed 2026-05-08T21:14:11.802350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:9f01f34ef0d29c145d8a24747610deef7f46a1d837ec1d08b7a1fe40608fe311

Observation 5a406431-6aac-4af1-8997-0b6b8df8c5f2 · outbound

This paper cites Science , author =.

Hypothesis generation and updating in large language models Science , author =

Reference 75

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.007531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:a90fdc8f06b9bb6afbe78afa86deeb238eabf03824eed01f8e2f1d41869d706f

Observation b9b3855e-4a3e-49a9-92a0-cf91301bb7e0 · outbound

This paper cites Representations and generalization in artificial and brain neural networks , volume =.

Hypothesis generation and updating in large language models Representations and generalization in artificial and brain neural networks , volume =

Reference 76

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.630733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:25e838aeacb15d164f6c7b5f72e8351393ed52306a358d0fb8a3b13e63eee701

Observation 9271fc38-c9f5-41b5-adff-a55eb6e689c9 · outbound

This paper cites (Joshua Brett) , year =.

Hypothesis generation and updating in large language models (Joshua Brett) , year =

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.908809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:b40d217069147fe9726e1633924a4fc349338d54ed0d01f237d2d5121c86723e

Observation 028f480d-b607-4d22-bfc1-86512fb8de11 · outbound

This paper cites Liar, Liar Pants on Fire.

Hypothesis generation and updating in large language models Liar, Liar Pants on Fire

Reference 78

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.061280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:4ceaf3fc60045e97ee90d6556d15ee41a08e74a6632f582cc1ee7ecdaf09d292

Observation a265d452-25a5-4ff4-94b2-e76cf0c8da30 · outbound

This paper cites Prototypical.

Hypothesis generation and updating in large language models Prototypical

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.926413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:d9130dffd240f38a8e9be640b14671a19ea34d52df043c83d1f47a02b5d60adc

Observation ab757c56-0feb-4d27-a975-13b0aab1da32 · outbound

This paper cites Gemma 3 Technical Report.

Hypothesis generation and updating in large language models Gemma 3 Technical Report

Reference 80

Resolution
metadata mismatch
arxiv_id, observed 2026-05-08T21:14:12.283722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:79e855fe2e3a2d7d38cd167f529313db7f87dfe34ad2a7acc5791ef126479999

Observation 89ff62ea-41cd-441c-baca-d65d5788a0c6 · outbound

This paper cites Qwen3.5.

Hypothesis generation and updating in large language models Qwen3.5

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.930444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:2a8842691ea566a5cf891a9264cf5ae43e1866870cf520f97577361032ba5fb7

Observation 3e31b0bc-65bb-4023-a22c-645cfaabc746 · outbound

This paper cites Memory & Cognition , author =.

Hypothesis generation and updating in large language models Memory & Cognition , author =

Reference 82

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.565162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:9a1af7c84bad24f4f2001b94b2f306aab4554f4240bab827463216361e3c2f7e

Observation d99e07fe-909a-46c7-967b-e4c2d5baea09 · outbound

This paper cites year =.

Hypothesis generation and updating in large language models year =

Reference 83

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.418720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:ba2efe1abee521f490032cc0af0167645b8033e9a7dc90268f466a3ec4bbe9ba

Observation d03c56c5-dc6d-4a67-95bd-d1ad0da2d0e2 · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

Hypothesis generation and updating in large language models Instruction-Following Evaluation for Large Language Models

Reference 84

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:11.784375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:de6f85f681cbe21b5a272b35da56ca53b7d15e07889e79038ffdfb64cd9ff895

Observation d68e8c28-e01c-49e6-a560-079cb7183aaf · outbound

This paper cites month = aug, year =.

Hypothesis generation and updating in large language models month = aug, year =

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.894214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:3d2e861a0dff91028505f01edca2d0c0b370c173d5a544130134609e09e28adf

Observation d2d4ea88-ddfc-4783-b16e-5efd0f920766 · outbound

This paper cites Advances in.

Hypothesis generation and updating in large language models Advances in

Reference 86

Resolution
verified exact
doi, observed 2026-05-08T21:14:11.760282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:4493468beb4c61c93c5141a8b250142ecaa133f1cb09a4be00a45dcbdef89e15

Observation 4c0b3152-88bb-4b2d-aaf0-1a50cacb38a3 · outbound

This paper cites Damien Ernst, Pierre Geurts, and Louis Wehenkel.

Hypothesis generation and updating in large language models Damien Ernst, Pierre Geurts, and Louis Wehenkel

Reference 87

Resolution
metadata mismatch
doi, observed 2026-05-08T21:14:12.550713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:cadb128ec74d2e3ff6de2aa7e9207e495ae4266a9c57d316db03bde3dabd4ca6

Observation 0cf4e194-1fb2-4dd2-ba37-ea58b500fea7 · outbound

This paper cites Proceedings of the National Academy of Sciences , author =.

Hypothesis generation and updating in large language models Proceedings of the National Academy of Sciences , author =

Reference 88

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.469751Z

Source-reported events for the cited work

correction dated 2023-03-10. Source: crossref record 10.1073/pnas.2302618120->10.1073/pnas.1915984117:correction, observed 2026-07-11T03:09:38.508366+00:00. This notice travels one citation hop only.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:18382254af4df8faa074b43d01ad60733b2778762661f27997e02eaa1f8831a8

Observation 1c9bb7b6-aaef-413f-a408-82806c467117 · outbound

This paper cites Murray, Alberto Bernacchia, Nicholas A.

Hypothesis generation and updating in large language models Murray, Alberto Bernacchia, Nicholas A

Reference 89

Resolution
metadata mismatch
doi, observed 2026-05-08T21:14:12.411540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:182e7f0f11213d17e3c2b6f98a946715348a4cf2c290d6c73ab2495a6353b573

Observation bdc6dca4-7d26-43ac-9d01-c83d5d8d5841 · outbound

This paper cites The selective updating of working memory: a predictive coding account , language =.

Hypothesis generation and updating in large language models The selective updating of working memory: a predictive coding account , language =

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T11:37:37.897106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:4a90335b6d3dff6f39515c2a7626678e6d5237251496d8e05d670ccc826f3490

Observation 3415c68f-2f86-40b8-bd0e-c053cf7d5dd6 · outbound

This paper cites Annual Review of Psychology , author =.

Hypothesis generation and updating in large language models Annual Review of Psychology , author =

Reference 91

Resolution
verified exact
doi, observed 2026-05-08T21:14:11.774246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:bea11cc9892320cde8b3a1730ec877f9353df59c502803b2ffd9b29c9a213275

Observation 1c025f57-b5d2-4beb-a5af-89e3231be64d · outbound

This paper cites Current Opinion in Behavioral Sciences , author =.

Hypothesis generation and updating in large language models Current Opinion in Behavioral Sciences , author =

Reference 92

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.223360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:722dc76319e72bda1232f6b1519f9812126db5cee6984322f60d4d716f57d456

Observation bcfe6f02-e0d6-41fa-bf4b-e10fc354eba5 · outbound

This paper cites Nature Communications , author =.

Hypothesis generation and updating in large language models Nature Communications , author =

Reference 93

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.493616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:f0bb625a5ecd4508102ab8b63cb61b71fbac914d468d7453e3db04f14056b239

Observation d6e5f528-aaa3-4aca-9d1e-65333a743ce3 · outbound

This paper cites Nature Communications , author =.

Hypothesis generation and updating in large language models Nature Communications , author =

Reference 94

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.488519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:e2f7125f81e2b4208dde17723057aad0661a1d63088cc9615c71d168aee576d5

Observation 9bfd6ec2-0c9e-4d72-9f29-2d86181cf552 · outbound

This paper cites Trends in Cognitive Sciences , author =.

Hypothesis generation and updating in large language models Trends in Cognitive Sciences , author =

Reference 95

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.188279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:1ccb69e1899af5d45af756c11bdacfcaa89eb8c99aa93130a5d791ceb56bee85

Observation c718251a-5550-4a8c-b5ea-4a5f9a1e85fb · outbound

This paper cites Trends in Cognitive Sciences , author =.

Hypothesis generation and updating in large language models Trends in Cognitive Sciences , author =

Reference 96

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.368517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:16246b8c851f9a0a9758d92d45caac81a4d50b35d789e5c0be27c5aba99e5e61

Observation 9290a1af-cb4d-4235-b91d-249679329683 · outbound

This paper cites The Journal of Neuroscience , author =.

Hypothesis generation and updating in large language models The Journal of Neuroscience , author =

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:14:12.252349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:f50d0fd73e0c549a13256962af0a4441a55cbecb931bb5bbc629ab72b6db6f18

Observation 0622aeeb-d196-40de-9146-28b98fab0968 · outbound

This paper cites month = oct, year =.

Hypothesis generation and updating in large language models month = oct, year =

Reference 98

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.377037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:59cb659549736dddeb58637a83b2fc38b34a11de68126489cc338f485be34368

Observation 49d62c70-6557-494d-a977-d4e717a0a28f · outbound

This paper cites Nature Human Behaviour , author =.

Hypothesis generation and updating in large language models Nature Human Behaviour , author =

Reference 99

Resolution
verified exact
doi, observed 2026-05-08T21:14:12.125013Z

Source-reported events for the cited work

correction dated 2020-10-09. Source: crossref record 10.1038/s41562-020-00993-7->10.1038/s41562-020-00938-0:correction, observed 2026-07-11T03:09:11.611456+00:00. This notice travels one citation hop only.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:1ee0d4f5af836a19184ce5008e040a72553efbbd40bfc1501cda487129f0f4e2

Observation e73daff0-40d0-4e7c-b47e-31bfe1080c67 · outbound

This paper cites Strong inhibitory signaling underlies stable temporal dynamics and working memory in spiking neural networks.

Hypothesis generation and updating in large language models Strong inhibitory signaling underlies stable temporal dynamics and working memory in spiking neural networks

Reference 100

Resolution
metadata mismatch
doi, observed 2026-05-08T21:14:12.415186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T14:43:56.125546Z digest=sha256:7ab8ce16615f759c72b37bfc363eb176171a1d50f1fadec2cad188de4c4f498c

Pith citing papers

Observation ccbae499-94fa-4b49-a548-c657576fade6 · inbound

Thinking Under Uncertainty: Evidence Use and Information-Seeking in Language Models cites this paper.

Thinking Under Uncertainty: Evidence Use and Information-Seeking in Language Models Hypothesis generation and updating in large language models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-30T19:24:34.330570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T19:24:34.330570Z digest=sha256:af258955446b195628e83582eb3a4db2cc3152b0b7c65302147ea265cc68f9d0