Pith. sign in

Paper Citation Record · LEDGER

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering

As of 14 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 1 inbound Pith citation observation for arXiv:2411.12395.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.12395 v1

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T17:39:08.452490Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-22T11:06:38.102348Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T11:11:27.550895Z

Reference resolution

25 of 25 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation dd79d9b1-7380-466f-ab58-c8a12865509f · outbound

This paper cites 56% of college students have used AI on assignments or exams: Bestcolleges,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering 56% of college students have used AI on assignments or exams: Bestcolleges,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.907569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T17:39:08.358990Z digest=sha256:763d29985aa9168071d8a694c2b95d4c471b3f14745a90aa5b57bf60bfc4303d

Observation b7fba134-b158-4159-9e82-7fdd89f28d8f · outbound

This paper cites Manjrekar.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Manjrekar

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.894498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T17:39:08.363513Z digest=sha256:35c9f1fb6dee39c9597aa12648895f8e00273a63d19daf400d7a539df90c8f82

Observation 120c83de-044f-4cd6-9322-4df3b20c7f5a · outbound

This paper cites Siren's Song in the AI Ocean: A Survey on Hallucination in Large Language Models.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Siren's Song in the AI Ocean: A Survey on Hallucination in Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.367529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.367529Z digest=sha256:f12c6990dfcefd1297010a17ae6f8b0e9f71f16e4d8f8a9424cd3e56e2606c48

Observation a434acd6-0841-4cf4-a452-fb353b979dcf · outbound

This paper cites Attention is all you need,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Attention is all you need,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.371709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.371709Z digest=sha256:eb2b77670d423a8ca120d84b17f3a320ea9d7b47bd7dc0a58020b45222f373f7

Observation a06bbfa9-ee43-40a4-8667-340e2f005c4b · outbound

This paper cites Sparks of Artificial General Intelligence: Early experiments with GPT-4.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Sparks of Artificial General Intelligence: Early experiments with GPT-4

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.376028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.376028Z digest=sha256:c44dd9cc53af24653ad076479a312d1f814be559a1f5fd58d8cf9530752ee0c7

Observation 4c81cbad-95bf-491a-9a14-bb9bdbebf415 · outbound

This paper cites Language models are unsupervised multitask learners,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Language models are unsupervised multitask learners,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.381029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.381029Z digest=sha256:05aa7ed65c8dbc95c0566261ac1c734f7813051ec4d323e7fc8f3640bf5d224e

Observation edf07336-dc48-4ab6-9669-655e786389f1 · outbound

This paper cites Language Models are Few-Shot Learners.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Language Models are Few-Shot Learners

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.385458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.385458Z digest=sha256:13fb8049fe9dd27384e706eb65ae6bc6154971f696d7eaf73846d388e1ff869f

Observation 3e003c7c-18cb-4c5c-bd27-11ade923dab2 · outbound

This paper cites GPT-4 Technical Report.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering GPT-4 Technical Report

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.389307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.389307Z digest=sha256:006dd2d341ceb582b470920a8d42beb95adc2267be5b1a416e615609ba3f341e

Observation 03695bbc-4bfc-4842-bc74-c5760b1055bb · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering LLaMA: Open and Efficient Foundation Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.393263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.393263Z digest=sha256:3cf30f3c9bca8604282cda24432e5ee1a3a386d1566ba01114031554383584be

Observation 3e7a26aa-8f41-4065-9a68-49dedbd208c0 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.397186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.397186Z digest=sha256:fcc8cd580fffd5fbdd68ea0ab03076d8e29e50e69e610e85ea2cca7ad873eb30

Observation 42446af3-b266-4cf3-8a5b-9d7181cdb463 · outbound

This paper cites Instruction tuning for large language models: A survey,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Instruction tuning for large language models: A survey,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.401394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.401394Z digest=sha256:11e00b8a17e0f20e9869eb332c4a4eaa503f102b4673e129d76a1a6abb854b71

Observation 115bdd92-df93-46d8-a240-b2764aad0616 · outbound

This paper cites Fighting fire with fire: can chatgpt detect ai-generated text?.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Fighting fire with fire: can chatgpt detect ai-generated text?

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.863853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T17:39:08.405159Z digest=sha256:809a832b8e487122626283930a6bc1c2d3fc4ca6c6046e1b80b6cb071edce6d6

Observation b790d220-7b46-4551-bbe7-eadfd1481f07 · outbound

This paper cites MathBench: Evaluating the Theory and Application Proficiency of LLMs with a Hierarchical Mathematics Benchmark.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering MathBench: Evaluating the Theory and Application Proficiency of LLMs with a Hierarchical Mathematics Benchmark

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.408800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.408800Z digest=sha256:eb8d491a4e6cdb1211c4c7bb268ac4ff59b86292dc1564e1b447770ab2120e14

Observation 3285a012-fc59-4c10-93a6-87c5269e170f · outbound

This paper cites The Butterfly Effect of Altering Prompts: How Small Changes and Jailbreaks Affect Large Language Model Performance.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering The Butterfly Effect of Altering Prompts: How Small Changes and Jailbreaks Affect Large Language Model Performance

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.412464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.412464Z digest=sha256:927a907f8e0a1bf4c796dee32926a5349ac10836f36261441b70bdfeb41ed6c8

Observation 0c58a8ba-7cf8-4d05-b7e5-c781e7ea4439 · outbound

This paper cites Quantifying language models’ sensitivity to spurious features in prompt design or: How i learned to start worrying about prompt formatting,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Quantifying language models’ sensitivity to spurious features in prompt design or: How i learned to start worrying about prompt formatting,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.850880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T17:39:08.416156Z digest=sha256:f433c7284d3e87f04cd5dd1a3b5ed43bbc442503c1f23d12368a9b7ca3372bdf

Observation fcde1c31-717f-4f56-923b-320ea5e2a6f3 · outbound

This paper cites Task Ambiguity in Humans and Language Models.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Task Ambiguity in Humans and Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.419760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.419760Z digest=sha256:fc2196acd3ef927af49c6b0785415b5ed790e6be665701f328d5c96781c8911a

Observation 2d443af0-68d7-495b-acff-79f44d8a38d4 · outbound

This paper cites an unresolved cited work.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:39:08.838025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T17:39:08.423686Z digest=sha256:d67597aa1127a1286ec0cd5c0764509c106731bb99d0b3fa683f37c04fc09988

Observation 60877db6-650f-475a-8a3c-eae3a2054af7 · outbound

This paper cites The winograd schema challenge,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering The winograd schema challenge,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.824823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T17:39:08.427497Z digest=sha256:f0e27eb24eefa5e09de10123f124c97883577e14ff1698a5a5723dcb6235ccc0

Observation 5d86b036-0ab5-48da-9192-963695f02871 · outbound

This paper cites Gpt-4o system card,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Gpt-4o system card,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.810453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T17:39:08.431031Z digest=sha256:f7c9e9e8dec6cfdcc71fcc29857ad2670568cc24cbe0642cafd4ed4f53b63537

Observation f4e1e5be-6b1d-4cca-8047-6bad97f4e02c · outbound

This paper cites Natural questions: a benchmark for question answering research,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Natural questions: a benchmark for question answering research,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.797087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T17:39:08.434415Z digest=sha256:c116ebcdf8ff284107ca8ac6b44fb992c9b1ea7fe104f104e234105d0dbe097c

Observation 350d6fbc-e570-4e7b-b330-34ca0f244b95 · outbound

This paper cites AmbigQA: Answering ambiguous open-domain questions,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering AmbigQA: Answering ambiguous open-domain questions,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.783784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T17:39:08.438058Z digest=sha256:040006847e24416762091c9bede0fbfa4949d0e35d641470fda78195cafaeee0

Observation 7a9ec339-6cfe-4d92-9cba-42f986aae223 · outbound

This paper cites Scope ambiguities in large language models,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Scope ambiguities in large language models,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.770656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T17:39:08.441544Z digest=sha256:99b0669232376bc2228a64271cafa5c6fc266e7845f76f0d68f592712f2c94c6

Observation 39706813-4643-43d5-bcda-7d4ecd053b7f · outbound

This paper cites Mistral 7B.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Mistral 7B

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.445089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.445089Z digest=sha256:d737e38efb0b546e7837116af5edb3e1a95b7a6c2906d38b304105f465b2b7ea

Observation e4113f78-fd4e-47d4-a15b-9c74d6b13237 · outbound

This paper cites An Empirical Study of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering An Empirical Study of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.448804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.448804Z digest=sha256:0ca871768377e8bedb0bdc3996ddef4b0c5b67727eb8e99d8937ec9072c7fcc0

Observation 142c4aa5-8cac-46e6-bf93-9ae4a50f416c · outbound

This paper cites Preserving principal subspaces to reduce catastrophic forgetting in fine-tuning,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Preserving principal subspaces to reduce catastrophic forgetting in fine-tuning,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.757445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T17:39:08.452490Z digest=sha256:e795830c9aa27912cf390c2a566eeac4f5574e7abeccfe5b48f430348d70f4b6

Pith citing papers

Observation 65801f10-9aa9-4555-ae5a-e38a67416de9 · inbound

VisPhyWorld: Probing Physical Reasoning via Code-Driven Video Reconstruction cites this paper.

VisPhyWorld: Probing Physical Reasoning via Code-Driven Video Reconstruction Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-22T11:11:27.554087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-22T11:06:38.102348Z digest=sha256:c47585f30c03b2ddc58b29320d07d344c963838969983724b40f93092705e470