Pith. sign in

Paper Citation Record · LEDGER

Understanding Benchmark Language Under Weakened Formal Semantics

As of 18 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2509.17455.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.17455 v2

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:51:31.967967Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact3
  • verified fuzzy17
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2de25e7c-f430-4d7c-ac19-78191c13d3bb · outbound

This paper cites GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models.

Understanding Benchmark Language Under Weakened Formal Semantics GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.814803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.814803Z digest=sha256:dd0c5aa0981f9a13aef1775a784c71330f19f82e365061ea6403c846dbfc8fd1

Observation c52771ec-87d9-4339-8ead-85f195999dd7 · outbound

This paper cites (eds.) Advances in Neural Information Processing Systems, vol.

Understanding Benchmark Language Under Weakened Formal Semantics (eds.) Advances in Neural Information Processing Systems, vol

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.598429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.819936Z digest=sha256:24d12dd0dee17901332faef37387a6a6b6b3acb9912d8c970fbc7a2524beb601

Observation 68ba365b-3242-4603-bf1f-99761eaacef5 · outbound

This paper cites https://arxiv.org/abs/2410.

Understanding Benchmark Language Under Weakened Formal Semantics https://arxiv.org/abs/2410

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.588925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.823840Z digest=sha256:88a40b8605473f387ed7de9920539ec87fc1896917f30f9b28220d3f1604a37f

Observation 6a205139-13dd-478c-97da-eb06b3ed7a56 · outbound

This paper cites In: Proceedings of the 36th International Conference on Neural Informa- tion Processing Systems.

Understanding Benchmark Language Under Weakened Formal Semantics In: Proceedings of the 36th International Conference on Neural Informa- tion Processing Systems

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.578889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.827897Z digest=sha256:7f926155cbd7f30d0e2ac2af4a5e28498c20ab4d4379e4cb1ab3b57d78f2da67

Observation c8a2e701-6f54-47d8-97bd-b49e93b97198 · outbound

This paper cites In: Salakhutdinov, R., Kolter, Z., Heller, K., Weller, A., Oliver, N., Scarlett, J., Berkenkamp, F.

Understanding Benchmark Language Under Weakened Formal Semantics In: Salakhutdinov, R., Kolter, Z., Heller, K., Weller, A., Oliver, N., Scarlett, J., Berkenkamp, F

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.566969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.831657Z digest=sha256:f0551da3d73bace16317b16dbe128671be9df46e291c7ea99cda267b297cca5e

Observation 6dd43a1d-73c5-4441-9d23-8ed6acf9aa5c · outbound

This paper cites Program of Thoughts Prompting: Disentangling Computation from Reasoning for Numerical Reasoning Tasks.

Understanding Benchmark Language Under Weakened Formal Semantics Program of Thoughts Prompting: Disentangling Computation from Reasoning for Numerical Reasoning Tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.835390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.835390Z digest=sha256:6b41dc0baab740a0128e4551446af2cec57a3c9c8eeb83168dcf34dd699b669e

Observation ebcceba5-9cf2-4475-a400-ffd2161e8a33 · outbound

This paper cites In: Proceedings of the 40th International Conference on Machine Learning.

Understanding Benchmark Language Under Weakened Formal Semantics In: Proceedings of the 40th International Conference on Machine Learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.554546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.839791Z digest=sha256:88e2912e886c563a673ed4eb205cd1175922d828a18b0a28283f907c66988d7b

Observation f699e4aa-2cd1-4edf-81ef-5f2c55807ddd · outbound

This paper cites In: Larochelle, H., Ranzato, M., Hadsell, R., Balcan, M.F., Lin, H.

Understanding Benchmark Language Under Weakened Formal Semantics In: Larochelle, H., Ranzato, M., Hadsell, R., Balcan, M.F., Lin, H

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.543121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.843649Z digest=sha256:424583481e69eb07a29b139ee39a4f4fec38fa17b28b2c3b86946d1922a53553

Observation 98d8ed54-524a-4b50-95dc-8dc92e32e738 · outbound

This paper cites Science378(6624), 1092–1097 (2022) https://doi.org/10.1126/science.abq1158.

Understanding Benchmark Language Under Weakened Formal Semantics Science378(6624), 1092–1097 (2022) https://doi.org/10.1126/science.abq1158

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.847366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.847366Z digest=sha256:bd4970061f77c314c07e1762699fbe64e7dadc0b0efb225edad2e1629adf616a

Observation 7d3e2b02-2f7b-4e51-888a-7a621175fb05 · outbound

This paper cites SWE-bench: Can Language Models Resolve Real-World GitHub Issues?.

Understanding Benchmark Language Under Weakened Formal Semantics SWE-bench: Can Language Models Resolve Real-World GitHub Issues?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.851013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.851013Z digest=sha256:ae191003b496623f11036a9deabee2bdbb85b41c408f9ffae11500656a6b7963

Observation 631f2b20-37a9-4aaa-8871-bf497a2fec35 · outbound

This paper cites ACM Trans.

Understanding Benchmark Language Under Weakened Formal Semantics ACM Trans

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.855060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.855060Z digest=sha256:e0cd1da986fa8f210b425ada0d66a20e3766347f151c87800b41aa239fc963eb

Observation fbc3c4a2-6e76-4f44-a75a-c51ee6025c3e · outbound

This paper cites In: Thirty-seventh Conference on Neural Information Processing Systems (2023).

Understanding Benchmark Language Under Weakened Formal Semantics In: Thirty-seventh Conference on Neural Information Processing Systems (2023)

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.531058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.858624Z digest=sha256:ddd235d7c5aff439ca3cd3a867e9b7837939539c41f7a2de63cb9a057ae97036

Observation e14ebaa3-c76a-40d5-bd55-9e1614263042 · outbound

This paper cites From Frege to Gödel: A Source Book in Mathematical Logic1931, 1–82 (1879).

Understanding Benchmark Language Under Weakened Formal Semantics From Frege to Gödel: A Source Book in Mathematical Logic1931, 1–82 (1879)

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.519941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.862136Z digest=sha256:cb7c51476ff697d556a56c1c1fee68e3691f82b87fcd1b8b72264b82ffe483bb

Observation 340fa614-08e4-47b5-804e-505955e5ee6e · outbound

This paper cites Linguistics and Philosophy4(2), 159–219 (1981) https://doi.org/10.1007/bf00350139.

Understanding Benchmark Language Under Weakened Formal Semantics Linguistics and Philosophy4(2), 159–219 (1981) https://doi.org/10.1007/bf00350139

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.865693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.865693Z digest=sha256:767a424603206c52e09944c285df98f060afb5acdbdd5f7ad3471b873bb7533d

Observation 4dd9e6a8-41d4-4cd8-af4c-bb911ddfcc6b · outbound

This paper cites Transactions of the Association for Computational Linguistics 10, 1266–1284 (2022) https://doi.org/10.1162/tacl_a_00518.

Understanding Benchmark Language Under Weakened Formal Semantics Transactions of the Association for Computational Linguistics 10, 1266–1284 (2022) https://doi.org/10.1162/tacl_a_00518

Reference 15

Resolution
verified exact
doi, observed 2026-08-15T15:51:32.027947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.869371Z digest=sha256:024a6731645d46783885fa0301335f7fc0cf5499b7c46810884de7fd509e80a0

Observation 91ec05b2-6c00-47b5-9cab-3310e76d1a87 · outbound

This paper cites Edin- burgh Advanced Textbooks in Linguistics, ??? (2016).

Understanding Benchmark Language Under Weakened Formal Semantics Edin- burgh Advanced Textbooks in Linguistics, ??? (2016)

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.509085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.873778Z digest=sha256:33c5aba3d6b7f959b4b826b8abf66ddb2dee4a1e83433c53466e7c9eee751ae8

Observation 49d40da3-c386-4820-a361-a25c3f8aa011 · outbound

This paper cites Computer Law & Security Review46, 105696 (2022) https://doi.org/ 10.1016/j.clsr.2022.105696.

Understanding Benchmark Language Under Weakened Formal Semantics Computer Law & Security Review46, 105696 (2022) https://doi.org/ 10.1016/j.clsr.2022.105696

Reference 17

Resolution
metadata mismatch
raw_fallback, observed 2026-08-15T15:51:32.353164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.877245Z digest=sha256:074b5be70e3660879f8162a59e7c0babc713a39c67be87114f3cb9a519847e7d

Observation 602e1fbf-687e-43c2-b0a1-8c0d44b6f561 · outbound

This paper cites In: Toutanova, K., Rumshisky, A., Zettlemoyer, L., Hakkani-Tur, D., Beltagy, I., Bethard, S., Cotterell, R., Chakraborty, T., Zhou, Y.

Understanding Benchmark Language Under Weakened Formal Semantics In: Toutanova, K., Rumshisky, A., Zettlemoyer, L., Hakkani-Tur, D., Beltagy, I., Bethard, S., Cotterell, R., Chakraborty, T., Zhou, Y

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.880675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.880675Z digest=sha256:709cd63e4c545e7a82157f3290eff4b09755999f22a9254128ae7e2eeec96eab

Observation 3c9c8866-8be6-4e81-a075-0893e691af4a · outbound

This paper cites Think-on-Graph: Deep and Responsible Reasoning of Large Language Model on Knowledge Graph.

Understanding Benchmark Language Under Weakened Formal Semantics Think-on-Graph: Deep and Responsible Reasoning of Large Language Model on Knowledge Graph

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.883664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.883664Z digest=sha256:d687cc67ba839b005d032ab84935a21f42584bfc731df0f98bd8da68cca7443f

Observation 2569a87d-923a-4833-9d3b-bf5008906fcc · outbound

This paper cites In: Moens, M.-F., Huang, X., Specia, L., Yih, S.W.-t.

Understanding Benchmark Language Under Weakened Formal Semantics In: Moens, M.-F., Huang, X., Specia, L., Yih, S.W.-t

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.886681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.886681Z digest=sha256:6a83eee5d29dfc26eb36397bce72fb921a398dac1612fbc733fd734a1c74e16b

Observation b630010b-8321-47a2-a9b3-aef78d127784 · outbound

This paper cites (eds.) Findings of the Association for Computational Linguistics: NAACL 2025, pp.

Understanding Benchmark Language Under Weakened Formal Semantics (eds.) Findings of the Association for Computational Linguistics: NAACL 2025, pp

Reference 21

Resolution
verified exact
doi, observed 2026-08-15T15:51:32.009260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.889348Z digest=sha256:47a8ec99751e932570f0763cba24e448dbcef8f5e0d5386e645f5b270588a0bb

Observation 936836d9-2fbb-45a8-9ecd-fabc17b42395 · outbound

This paper cites An Empirical Study on the Potential of LLMs in Automated Software Refactoring.

Understanding Benchmark Language Under Weakened Formal Semantics An Empirical Study on the Potential of LLMs in Automated Software Refactoring

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.892074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.892074Z digest=sha256:aa567b3a2beac3cba33a1475b48a88b8795dc88447ec22b3841a72e6c3237e17

Observation 5758ca97-4232-48d5-8cea-f69f8e7e73cf · outbound

This paper cites In: Companion Proceedings of the 32nd ACM International Conference on the Foundations of Software Engineering.

Understanding Benchmark Language Under Weakened Formal Semantics In: Companion Proceedings of the 32nd ACM International Conference on the Foundations of Software Engineering

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.895066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.895066Z digest=sha256:ba674659eb757528bb1606649570c1c76c22032bd14dd257b7b4fc0dd3e58309

Observation 51d58336-5cda-474e-8f75-47b0cb2be60f · outbound

This paper cites Automated Software Engg.32(1) (2025) https://doi.org/10.1007/s10515-024-00485-2.

Understanding Benchmark Language Under Weakened Formal Semantics Automated Software Engg.32(1) (2025) https://doi.org/10.1007/s10515-024-00485-2

Reference 24

Resolution
verified exact
doi, observed 2026-08-15T15:51:31.998627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.898481Z digest=sha256:dfa84eb87fe742497ac419c271e9c34b3fcc101dba9baccca90afdd50a13ff30

Observation 5a334b3d-45b7-4490-b603-10f135a8f87a · outbound

This paper cites Interleaving Retrieval with Chain-of-Thought Reasoning for Knowledge-Intensive Multi-Step Questions.

Understanding Benchmark Language Under Weakened Formal Semantics Interleaving Retrieval with Chain-of-Thought Reasoning for Knowledge-Intensive Multi-Step Questions

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.901816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.901816Z digest=sha256:0bab14df6dbd960d73cf2eb3b71c9df035b929572d7d59a5f4c2b75dc251e79c

Observation e319d689-5811-40d0-8158-a427b7f3ed81 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Understanding Benchmark Language Under Weakened Formal Semantics Training Verifiers to Solve Math Word Problems

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.905574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.905574Z digest=sha256:636965578ce63bdf112c51eaa7c6fcc36a5b44e03e08cc01c1fb27c9bb66b158

Observation 5c6105d8-3613-4573-9ae0-a42a773c4324 · outbound

This paper cites an unresolved cited work.

Understanding Benchmark Language Under Weakened Formal Semantics Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:51:32.492230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.909223Z digest=sha256:229d177769b1b74780c12e47b3a927257d31e17dace73b0d31499a67f3ee4ab6

Observation 5765e5eb-a4b1-4f3c-8824-880e97b5b789 · outbound

This paper cites In: The 2023 Conference on Empirical Methods in Natural Language Processing (2023).

Understanding Benchmark Language Under Weakened Formal Semantics In: The 2023 Conference on Empirical Methods in Natural Language Processing (2023)

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.481822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.912511Z digest=sha256:7d3a239d5f690db24556fdc29e007413b6b25fb55c8614aa0b92a9441ccd599e

Observation b0ae1015-e9e0-4bda-86c5-0c35f85e40c3 · outbound

This paper cites Can Large Language Models Infer Causation from Correlation?.

Understanding Benchmark Language Under Weakened Formal Semantics Can Large Language Models Infer Causation from Correlation?

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.915637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.915637Z digest=sha256:2272e1619dde24eb787c2b4e161ff1fc31873c0807ff86535787c35978016e97

Observation b6e6c072-4e2f-4f6d-868f-586cfa97845b · outbound

This paper cites Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them.

Understanding Benchmark Language Under Weakened Formal Semantics Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.919237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.919237Z digest=sha256:aa079b312a7174d01ac791dbb82d066cde5e22c2bbbf6d78f195a9465183cb72

Observation cc2eecc1-32f7-4035-8f4d-87a148b3e677 · outbound

This paper cites BIG-Bench Extra Hard.

Understanding Benchmark Language Under Weakened Formal Semantics BIG-Bench Extra Hard

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.922614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.922614Z digest=sha256:2fd47abd443aa4b96d20ab22d533050d0c70b3b73fbd67912767b5889860e786

Observation 2aaacc5d-86e0-46d1-8f7a-4632953e5c82 · outbound

This paper cites CAIL2018: A Large-Scale Legal Dataset for Judgment Prediction.

Understanding Benchmark Language Under Weakened Formal Semantics CAIL2018: A Large-Scale Legal Dataset for Judgment Prediction

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.926275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.926275Z digest=sha256:b8f88c0c0cbe5d7aa8a6a0899ef29a60763315be0cc3b3bfc7da6ee697c3fbbb

Observation a4b48898-840e-42ae-bc38-13e239884b9f · outbound

This paper cites In: Proceedings of the Annual Conference of the North American Chapter of the Association for Computational Linguistics.

Understanding Benchmark Language Under Weakened Formal Semantics In: Proceedings of the Annual Conference of the North American Chapter of the Association for Computational Linguistics

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.471592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.929814Z digest=sha256:d1401929f51bf2bdecdc980ee224cfaaf49d3d42ce44bb212b186690e09264ac

Observation c3ad653e-6229-4898-8cee-2461dfcfbd8a · outbound

This paper cites ClassActionPrediction: A Challenging Benchmark for Legal Judgment Prediction of Class Action Cases in the US.

Understanding Benchmark Language Under Weakened Formal Semantics ClassActionPrediction: A Challenging Benchmark for Legal Judgment Prediction of Class Action Cases in the US

Reference 34

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T15:51:32.131292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.933161Z digest=sha256:515d14ccc82e046e7bb13b4826c49db2f491a29731ca9b1f348ed41f44fae5c9

Observation 3a649783-5533-4bee-b65f-c15ffa6a029c · outbound

This paper cites an unresolved cited work.

Understanding Benchmark Language Under Weakened Formal Semantics Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.937049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.937049Z digest=sha256:5a9df496441709656c5527a2fdc4521319e2d40a19050eecbcc2c4bd7727d627

Observation 75b5497b-abee-4b24-9e14-f7fa024a2853 · outbound

This paper cites In: Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP), pp.

Understanding Benchmark Language Under Weakened Formal Semantics In: Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP), pp

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.455996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.940675Z digest=sha256:c409b2fc962f349090007b8e07ab066f378a4635c2bc06c3c07ee99a6e2a5290

Observation 0460b5c6-7bc2-47b0-8404-d1fe2a1bb79b · outbound

This paper cites an unresolved cited work.

Understanding Benchmark Language Under Weakened Formal Semantics Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:51:32.446676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.943871Z digest=sha256:f0f36d7f4d4e64687828e6aad5c1ec341ed96f508d0fae9fe20545081f2b61e2

Observation 4ab7248e-d1e8-4325-863b-2d4452db7ac4 · outbound

This paper cites https://github.com/google-research/google-research/tree/master/mbpp.

Understanding Benchmark Language Under Weakened Formal Semantics https://github.com/google-research/google-research/tree/master/mbpp

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.436167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.947174Z digest=sha256:fb983dac5db036a9ea4a8e7d473a7ad8d83ea2502a1cbd69cc81b9df5cd8dd0b

Observation a82f4095-b977-48cc-9126-b7867245044c · outbound

This paper cites https://huggingface.co/datasets/PatrickHaller/the-stack-python-1M.

Understanding Benchmark Language Under Weakened Formal Semantics https://huggingface.co/datasets/PatrickHaller/the-stack-python-1M

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.425791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.950866Z digest=sha256:f4364a0fba3b90f54181b82179e2538893bb1526168939c1f4d0077340ba069a

Observation 7c9af900-b20e-4319-b889-3c04fe19f775 · outbound

This paper cites https://huggingface.co/datasets/notbadai/python_functions_reasoning.

Understanding Benchmark Language Under Weakened Formal Semantics https://huggingface.co/datasets/notbadai/python_functions_reasoning

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.414983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.954341Z digest=sha256:cfb8e18105557fac0dd10cbf463c6ed6df8d9a09d0e69d9881786a4e47512b35

Observation 63cee2cd-60f9-4c6c-8658-d942e70e3a97 · outbound

This paper cites OpenCoder: The Open Cookbook for Top-Tier Code Large Language Models.

Understanding Benchmark Language Under Weakened Formal Semantics OpenCoder: The Open Cookbook for Top-Tier Code Large Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.957576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.957576Z digest=sha256:7518c96396895de09f27fed19926b21f4d2d70dbedd1489296ea162fe19e6995

Observation d076f12f-ff21-45e1-9944-3c955770868c · outbound

This paper cites MIT Press, ??? (2009).

Understanding Benchmark Language Under Weakened Formal Semantics MIT Press, ??? (2009)

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.404402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.960984Z digest=sha256:524385c247548e4db8f67e799b617918fcbdfb554581242e10a03bc4881348e9

Observation 1b39479a-3578-4aa5-9a27-71c97cf341a4 · outbound

This paper cites IEEE Transactions on Software Engineering SE-2(4), 308–320 (1976) https://doi.org/10.1109/TSE.1976.233837.

Understanding Benchmark Language Under Weakened Formal Semantics IEEE Transactions on Software Engineering SE-2(4), 308–320 (1976) https://doi.org/10.1109/TSE.1976.233837

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.964547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.964547Z digest=sha256:0c9abd4b6d962b62922e2c5505e483cf27b5ba129a4e9b2e22ec9e3acd256baf

Observation 93d47042-3557-4b95-96d0-478cacb086b2 · outbound

This paper cites https://docs.python.org/3/ library/ast.html.

Understanding Benchmark Language Under Weakened Formal Semantics https://docs.python.org/3/ library/ast.html

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.393731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T15:51:31.967967Z digest=sha256:fa0d41d6bb04735de5e5261e957e56b4b60e77cd1f88fb7f7c32cf59450951f0

Pith citing papers

No inbound Pith citation observations are available.