Pith. sign in

Paper Citation Record · LEDGER

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering

As of 23 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 1 inbound Pith citation observation for arXiv:2411.12395.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.12395 v1

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T17:39:08.452490Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-22T11:06:38.102348Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T11:11:27.550895Z

Reference resolution

25 of 25 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation dd79d9b1-7380-466f-ab58-c8a12865509f · outbound

This paper cites 56% of college students have used AI on assignments or exams: Bestcolleges,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering 56% of college students have used AI on assignments or exams: Bestcolleges,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.907569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T17:39:08.358990Z digest=sha256:356a04d948a818361cedffb39043ce16fb075a1e5ffaf712643b398a08429f96

Observation b7fba134-b158-4159-9e82-7fdd89f28d8f · outbound

This paper cites Manjrekar.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Manjrekar

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.894498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T17:39:08.363513Z digest=sha256:15b1cb5053d1af55a56c24c541c464aa7adcb40d443942cea651c8b6af759197

Observation 120c83de-044f-4cd6-9322-4df3b20c7f5a · outbound

This paper cites Siren's Song in the AI Ocean: A Survey on Hallucination in Large Language Models.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Siren's Song in the AI Ocean: A Survey on Hallucination in Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.367529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.367529Z digest=sha256:45b192416b1a140ce61e9cd07d90c1e9df673574f7eb4bfe00b10b881c5e1be8

Observation a434acd6-0841-4cf4-a452-fb353b979dcf · outbound

This paper cites Attention is all you need,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Attention is all you need,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.371709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.371709Z digest=sha256:0078f59d2f46facfac9ff8327ee46460a208cbd6bbfbec412ef635b49e25751e

Observation a06bbfa9-ee43-40a4-8667-340e2f005c4b · outbound

This paper cites Sparks of Artificial General Intelligence: Early experiments with GPT-4.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Sparks of Artificial General Intelligence: Early experiments with GPT-4

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.376028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.376028Z digest=sha256:d095af75ca6d824919086214dd55d29102f9ab7b400ab8d7f7c34422ab189812

Observation 4c81cbad-95bf-491a-9a14-bb9bdbebf415 · outbound

This paper cites Language models are unsupervised multitask learners,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Language models are unsupervised multitask learners,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.381029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.381029Z digest=sha256:8a5fb9176a543b6bfca201716163649eb5ce2b547e32ed14d94894dcc95098ca

Observation edf07336-dc48-4ab6-9669-655e786389f1 · outbound

This paper cites Language Models are Few-Shot Learners.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Language Models are Few-Shot Learners

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.385458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.385458Z digest=sha256:65321121333db00ac81d7693e99256536a30be46de7cf6c25eaa02586e4c5783

Observation 3e003c7c-18cb-4c5c-bd27-11ade923dab2 · outbound

This paper cites GPT-4 Technical Report.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering GPT-4 Technical Report

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.389307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.389307Z digest=sha256:ec1414a48b1ef85c629364735af9098acbb730807a1d34303559459a3a9938fb

Observation 03695bbc-4bfc-4842-bc74-c5760b1055bb · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering LLaMA: Open and Efficient Foundation Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.393263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.393263Z digest=sha256:da4b1c2c6e30001c901074a9c76452fab11286919b1676aed3f2fc1e3fe81923

Observation 3e7a26aa-8f41-4065-9a68-49dedbd208c0 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.397186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.397186Z digest=sha256:005404d5655c6c1043b35394e9e61c5888eda5309120bf0456db7b2e5f295067

Observation 42446af3-b266-4cf3-8a5b-9d7181cdb463 · outbound

This paper cites Instruction tuning for large language models: A survey,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Instruction tuning for large language models: A survey,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.401394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.401394Z digest=sha256:bed182316a94e6ea9fd11935f72b41b2ba768bde12573b43517a08c20aa19dd6

Observation 115bdd92-df93-46d8-a240-b2764aad0616 · outbound

This paper cites Fighting fire with fire: can chatgpt detect ai-generated text?.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Fighting fire with fire: can chatgpt detect ai-generated text?

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.863853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T17:39:08.405159Z digest=sha256:5f697e45bce19ac1b280f363d06e027b23555b793b219ff287fc17b3f00d9374

Observation b790d220-7b46-4551-bbe7-eadfd1481f07 · outbound

This paper cites MathBench: Evaluating the Theory and Application Proficiency of LLMs with a Hierarchical Mathematics Benchmark.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering MathBench: Evaluating the Theory and Application Proficiency of LLMs with a Hierarchical Mathematics Benchmark

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.408800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.408800Z digest=sha256:1a01d01a6af1968341149356224795cb4f5e1fcfb2ad24745c2489bcb68421b2

Observation 3285a012-fc59-4c10-93a6-87c5269e170f · outbound

This paper cites The Butterfly Effect of Altering Prompts: How Small Changes and Jailbreaks Affect Large Language Model Performance.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering The Butterfly Effect of Altering Prompts: How Small Changes and Jailbreaks Affect Large Language Model Performance

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.412464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.412464Z digest=sha256:cc0cabbad49ab05d296929d697246e8295aff054764789792dd9b82287625438

Observation 0c58a8ba-7cf8-4d05-b7e5-c781e7ea4439 · outbound

This paper cites Quantifying language models’ sensitivity to spurious features in prompt design or: How i learned to start worrying about prompt formatting,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Quantifying language models’ sensitivity to spurious features in prompt design or: How i learned to start worrying about prompt formatting,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.850880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T17:39:08.416156Z digest=sha256:d136a5793474aa8a5eac4d6926fda3863a83f19b2cb2f83cd6f5f0b417a4659b

Observation fcde1c31-717f-4f56-923b-320ea5e2a6f3 · outbound

This paper cites Task Ambiguity in Humans and Language Models.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Task Ambiguity in Humans and Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.419760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.419760Z digest=sha256:f4b720993de04418cb5f22e3d780cadd48054ef3fdf3fb0eb94690b18f0a6121

Observation 2d443af0-68d7-495b-acff-79f44d8a38d4 · outbound

This paper cites an unresolved cited work.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:39:08.838025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T17:39:08.423686Z digest=sha256:3cea5d15df223ec7829d07b87144909eaf77649d70fc8e80c89c8d4b39f73fce

Observation 60877db6-650f-475a-8a3c-eae3a2054af7 · outbound

This paper cites The winograd schema challenge,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering The winograd schema challenge,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.824823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T17:39:08.427497Z digest=sha256:47f3ce2c930520412e70a1f1d1090ed1b49f9bbc977dcb1c5252b0b1dc1c4420

Observation 5d86b036-0ab5-48da-9192-963695f02871 · outbound

This paper cites Gpt-4o system card,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Gpt-4o system card,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.810453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T17:39:08.431031Z digest=sha256:86deb3c22bd68887efd6d8cc0f13acd392a161354811d59c4c5d70bba140a3a4

Observation f4e1e5be-6b1d-4cca-8047-6bad97f4e02c · outbound

This paper cites Natural questions: a benchmark for question answering research,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Natural questions: a benchmark for question answering research,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.797087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T17:39:08.434415Z digest=sha256:f6cca2ee1aacd9e79094553eaae9c3563d0d1e5037433ee00594d0dc63995d23

Observation 350d6fbc-e570-4e7b-b330-34ca0f244b95 · outbound

This paper cites AmbigQA: Answering ambiguous open-domain questions,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering AmbigQA: Answering ambiguous open-domain questions,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.783784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T17:39:08.438058Z digest=sha256:7df51ef684bb8c13abdf94dd3be876860bd0609b663a8362c314558d71a8fef5

Observation 7a9ec339-6cfe-4d92-9cba-42f986aae223 · outbound

This paper cites Scope ambiguities in large language models,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Scope ambiguities in large language models,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.770656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T17:39:08.441544Z digest=sha256:fb559d1d9af70024dcb87eb182cf39da51c61a0a18b2213f5cfabdc8596b8310

Observation 39706813-4643-43d5-bcda-7d4ecd053b7f · outbound

This paper cites Mistral 7B.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Mistral 7B

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.445089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.445089Z digest=sha256:322f49ea44f4ab9c46b00cd35e82b769cd9712f5a8a6a2d6e632ddec1f6e07d3

Observation e4113f78-fd4e-47d4-a15b-9c74d6b13237 · outbound

This paper cites An Empirical Study of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering An Empirical Study of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T17:39:08.448804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T17:39:08.448804Z digest=sha256:956fd7ec227e5227701f8af1ce4801e8a529f9d92dd3457a67ec8af5cc1fc85b

Observation 142c4aa5-8cac-46e6-bf93-9ae4a50f416c · outbound

This paper cites Preserving principal subspaces to reduce catastrophic forgetting in fine-tuning,.

Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering Preserving principal subspaces to reduce catastrophic forgetting in fine-tuning,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:39:08.757445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-12T17:39:08.452490Z digest=sha256:a343b08a5c7a79e871cf322fe743c42809433b17ecaa44ff8233b94b72bfacfa

Pith citing papers

Observation 65801f10-9aa9-4555-ae5a-e38a67416de9 · inbound

VisPhyWorld: Probing Physical Reasoning via Code-Driven Video Reconstruction cites this paper.

VisPhyWorld: Probing Physical Reasoning via Code-Driven Video Reconstruction Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-22T11:11:27.554087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-22T11:06:38.102348Z digest=sha256:8708be8112c924604e7870dda01040828bcafdbb51d4142d79285d0863f7f636