Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T12:48:09.853658Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 3 inbound Pith citation observations for arXiv:2507.21476.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T12:48:09.853658Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T17:05:10.690456Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-13T23:08:25.106938Z
27 of 27 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a953727c-fb1c-4391-b553-3239a5d0b3bf · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench Phi-4-reasoning Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12470ab8-9ba2-4b65-8aff-82ea03bf0790 · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench https: //blog.google/technology/google-deepmind/ gemini-model-thinking-updates-march-2025/
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 683f9eb8-8304-4b22-9cc8-ac4ddc7001b2 · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench In Working Notes of CLEF 2025 - Conference and Labs of the Evaluation F orum
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 62bbde9c-ac16-4bc1-92d8-1c737c8671eb · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench BIG-Bench Extra Hard
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80ac73db-40d3-45a8-86ed-31cbec686418 · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench When 'YES' Meets 'BUT': Can Large Models Comprehend Contradictory Humor Through Comparative Reasoning?
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f28ad304-3e3b-4fcb-a08e-165d092eff2d · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench Let's Verify Step by Step
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cc49c60-75f7-49d8-b82f-f6234cad64d4 · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench In Proceed- ings of the 2023 Conference on Empirical Methods in Natural Language Processing, pages 11069–11081
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ad9be2f5-65c2-4393-bae4-9f82ba291152 · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b084aec8-6fee-4087-a879-fb1e266d68de · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench Inverse Scaling: When Bigger Isn't Better
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75acb1a9-c225-43a2-9ff5-3e318ecf04b2 · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench CodeElo: Benchmarking Competition-level Code Generation of LLMs with Human-comparable Elo Ratings
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c712a67-ee2b-4547-b5fb-db7e402e4fd9 · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench GPQA: A Graduate-Level Google-Proof Q&A Benchmark
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06d8d6d1-29d0-44c1-8461-11d8fd80da06 · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7ae82a2-ee3b-4401-854d-6be08571409b · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench Phishing Awareness via Game-Based Learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f61fd422-528b-483d-81a9-585fc134065a · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench Between Underthinking and Overthinking: An Empirical Study of Reasoning Length and correctness in LLMs
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a26467da-b9b2-4995-ac92-11b90ac777cb · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96465f1a-7000-4334-ae1f-ad044ac3a944 · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench In Proceedings of the 2022 Confer- ence on Empirical Methods in Natural Language Processing, pages 2866–2879
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a0c34740-672c-4533-a901-2407262fa64c · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench Signatures of room-temperature superconductivity emerging in two-dimensional domains within the new Bi/Pb-based ceramic cuprate superconductors at ambient pressure
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b61d379f-95df-47da-bca1-d3318f3f7d37 · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench In Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Tech- nologies, pages 4213–4228
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 656c4b5c-aa05-4ebe-8311-f2e126843037 · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench Uniformly rotating vortices for the lake equation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9c75b147-50f5-4294-a4eb-4c281786a722 · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench arXiv preprint arXiv:2502.18080
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b202ac87-475a-49a2-a998-8a4f9e111b8b · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench Humor in AI: Massive Scale Crowd-Sourced Preferences and Benchmarks for Cartoon Captioning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2720ae89-6683-4c2a-a731-9dbeb5c2d960 · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench Bridging the Creativity Understanding Gap: Small-Scale Human Alignment Enables Expert-Level Humor Ranking in LLMs
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66882ed1-a3c6-4913-8e54-06cd911d2f7f · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench AmbigQA: Answering Ambiguous Open-domain Questions
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a21e266-f982-45af-b54a-c35634c4fe36 · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench Solving Quantitative Reasoning Problems with Language Models
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c50d8ef-ea44-4d1b-a529-5caa53466861 · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64403be8-ce55-4562-a051-38c6ef8c9e4b · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3118d44f-b9c8-4fe4-8ee5-0a339bc9a225 · outbound
Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench ARC Prize 2024: Technical Report
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7da47090-9728-4baa-8a07-054361c8b7ce · inbound
Relationship-Centered Care: Relatedness and Responsible Design for Human Connections in Mental-Health Care Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0dc073df-7454-4b5b-beb0-9f15dbc037eb · inbound
HumorRank: A Tournament-Based Leaderboard for Evaluating Humor Generation in Large Language Models Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6af44ac6-42c1-4c0b-87bb-15bc9f946ce6 · inbound
HumorRank: A Tournament-Based Leaderboard for Evaluating Humor Generation in Large Language Models Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.