Pith. sign in

Paper Citation Record · LEDGER

Understanding Benchmark Language Under Weakened Formal Semantics

As of 18 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2509.17455.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.17455 v2

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:51:31.967967Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact3
  • verified fuzzy17
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2de25e7c-f430-4d7c-ac19-78191c13d3bb · outbound

This paper cites GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models.

Understanding Benchmark Language Under Weakened Formal Semantics GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.814803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.814803Z digest=sha256:dd0c5aa0981f9a13aef1775a784c71330f19f82e365061ea6403c846dbfc8fd1

Observation c52771ec-87d9-4339-8ead-85f195999dd7 · outbound

This paper cites (eds.) Advances in Neural Information Processing Systems, vol.

Understanding Benchmark Language Under Weakened Formal Semantics (eds.) Advances in Neural Information Processing Systems, vol

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.598429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.819936Z digest=sha256:00da920e560d9e0b51aedffb788f507447301dd4fea58d506685f8b375e45376

Observation 68ba365b-3242-4603-bf1f-99761eaacef5 · outbound

This paper cites https://arxiv.org/abs/2410.

Understanding Benchmark Language Under Weakened Formal Semantics https://arxiv.org/abs/2410

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.588925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.823840Z digest=sha256:a22ac4c8d3f24cdd1407a2537e0e04788e2fd92f31e9660edad4adcdc35e67f5

Observation 6a205139-13dd-478c-97da-eb06b3ed7a56 · outbound

This paper cites In: Proceedings of the 36th International Conference on Neural Informa- tion Processing Systems.

Understanding Benchmark Language Under Weakened Formal Semantics In: Proceedings of the 36th International Conference on Neural Informa- tion Processing Systems

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.578889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.827897Z digest=sha256:5f4c2f5a2019371e74387974f42ba650e77b6077d3ec5315def756f05819d9c8

Observation c8a2e701-6f54-47d8-97bd-b49e93b97198 · outbound

This paper cites In: Salakhutdinov, R., Kolter, Z., Heller, K., Weller, A., Oliver, N., Scarlett, J., Berkenkamp, F.

Understanding Benchmark Language Under Weakened Formal Semantics In: Salakhutdinov, R., Kolter, Z., Heller, K., Weller, A., Oliver, N., Scarlett, J., Berkenkamp, F

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.566969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.831657Z digest=sha256:285c8841f487532dca08080fda7cf87c50adeb1c3a84a871d179c56314509954

Observation 6dd43a1d-73c5-4441-9d23-8ed6acf9aa5c · outbound

This paper cites Program of Thoughts Prompting: Disentangling Computation from Reasoning for Numerical Reasoning Tasks.

Understanding Benchmark Language Under Weakened Formal Semantics Program of Thoughts Prompting: Disentangling Computation from Reasoning for Numerical Reasoning Tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.835390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.835390Z digest=sha256:6b41dc0baab740a0128e4551446af2cec57a3c9c8eeb83168dcf34dd699b669e

Observation ebcceba5-9cf2-4475-a400-ffd2161e8a33 · outbound

This paper cites In: Proceedings of the 40th International Conference on Machine Learning.

Understanding Benchmark Language Under Weakened Formal Semantics In: Proceedings of the 40th International Conference on Machine Learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.554546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.839791Z digest=sha256:6a3673d0ed194a6964a5409fb32c18bdae1c4f25bfbfecec4f20fa7038bd4f68

Observation f699e4aa-2cd1-4edf-81ef-5f2c55807ddd · outbound

This paper cites In: Larochelle, H., Ranzato, M., Hadsell, R., Balcan, M.F., Lin, H.

Understanding Benchmark Language Under Weakened Formal Semantics In: Larochelle, H., Ranzato, M., Hadsell, R., Balcan, M.F., Lin, H

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.543121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.843649Z digest=sha256:002e908491702e556fe7d89fb83b43e0c2b70c82283b27b3f89120574270f81a

Observation 98d8ed54-524a-4b50-95dc-8dc92e32e738 · outbound

This paper cites Science378(6624), 1092–1097 (2022) https://doi.org/10.1126/science.abq1158.

Understanding Benchmark Language Under Weakened Formal Semantics Science378(6624), 1092–1097 (2022) https://doi.org/10.1126/science.abq1158

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.847366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.847366Z digest=sha256:bd4970061f77c314c07e1762699fbe64e7dadc0b0efb225edad2e1629adf616a

Observation 7d3e2b02-2f7b-4e51-888a-7a621175fb05 · outbound

This paper cites SWE-bench: Can Language Models Resolve Real-World GitHub Issues?.

Understanding Benchmark Language Under Weakened Formal Semantics SWE-bench: Can Language Models Resolve Real-World GitHub Issues?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.851013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.851013Z digest=sha256:a9fb769e9ea4d208ec167e5e3b91eb250c470e67f1dd5a564027748f93bdbdf3

Observation 631f2b20-37a9-4aaa-8871-bf497a2fec35 · outbound

This paper cites ACM Trans.

Understanding Benchmark Language Under Weakened Formal Semantics ACM Trans

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.855060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.855060Z digest=sha256:e0cd1da986fa8f210b425ada0d66a20e3766347f151c87800b41aa239fc963eb

Observation fbc3c4a2-6e76-4f44-a75a-c51ee6025c3e · outbound

This paper cites In: Thirty-seventh Conference on Neural Information Processing Systems (2023).

Understanding Benchmark Language Under Weakened Formal Semantics In: Thirty-seventh Conference on Neural Information Processing Systems (2023)

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.531058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.858624Z digest=sha256:9ada06db006b93b143e2ad2a64e1f0d71cc333e3e49d66a717d54343f1ead1f4

Observation e14ebaa3-c76a-40d5-bd55-9e1614263042 · outbound

This paper cites From Frege to Gödel: A Source Book in Mathematical Logic1931, 1–82 (1879).

Understanding Benchmark Language Under Weakened Formal Semantics From Frege to Gödel: A Source Book in Mathematical Logic1931, 1–82 (1879)

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.519941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.862136Z digest=sha256:f147dccc8080cb4ff6067d7f7e0fe25dde7439f317dbcc969d56dfa6362967e0

Observation 340fa614-08e4-47b5-804e-505955e5ee6e · outbound

This paper cites Linguistics and Philosophy4(2), 159–219 (1981) https://doi.org/10.1007/bf00350139.

Understanding Benchmark Language Under Weakened Formal Semantics Linguistics and Philosophy4(2), 159–219 (1981) https://doi.org/10.1007/bf00350139

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.865693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.865693Z digest=sha256:767a424603206c52e09944c285df98f060afb5acdbdd5f7ad3471b873bb7533d

Observation 4dd9e6a8-41d4-4cd8-af4c-bb911ddfcc6b · outbound

This paper cites Transactions of the Association for Computational Linguistics 10, 1266–1284 (2022) https://doi.org/10.1162/tacl_a_00518.

Understanding Benchmark Language Under Weakened Formal Semantics Transactions of the Association for Computational Linguistics 10, 1266–1284 (2022) https://doi.org/10.1162/tacl_a_00518

Reference 15

Resolution
verified exact
doi, observed 2026-08-15T15:51:32.027947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.869371Z digest=sha256:0df7a2b00d7c35f7bfa235662c15a33a30c337a9099e257fb38b6e796215d271

Observation 91ec05b2-6c00-47b5-9cab-3310e76d1a87 · outbound

This paper cites Edin- burgh Advanced Textbooks in Linguistics, ??? (2016).

Understanding Benchmark Language Under Weakened Formal Semantics Edin- burgh Advanced Textbooks in Linguistics, ??? (2016)

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.509085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.873778Z digest=sha256:84f00f129056a7067846cca6ca15c6705117ee0c28f953fc3ff9e46e7731ee9b

Observation 49d40da3-c386-4820-a361-a25c3f8aa011 · outbound

This paper cites Computer Law & Security Review46, 105696 (2022) https://doi.org/ 10.1016/j.clsr.2022.105696.

Understanding Benchmark Language Under Weakened Formal Semantics Computer Law & Security Review46, 105696 (2022) https://doi.org/ 10.1016/j.clsr.2022.105696

Reference 17

Resolution
metadata mismatch
raw_fallback, observed 2026-08-15T15:51:32.353164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.877245Z digest=sha256:b42e946e3d99a6e3b396c1a455fc258a7e788c0436487a52cf424471636cbe90

Observation 602e1fbf-687e-43c2-b0a1-8c0d44b6f561 · outbound

This paper cites In: Toutanova, K., Rumshisky, A., Zettlemoyer, L., Hakkani-Tur, D., Beltagy, I., Bethard, S., Cotterell, R., Chakraborty, T., Zhou, Y.

Understanding Benchmark Language Under Weakened Formal Semantics In: Toutanova, K., Rumshisky, A., Zettlemoyer, L., Hakkani-Tur, D., Beltagy, I., Bethard, S., Cotterell, R., Chakraborty, T., Zhou, Y

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.880675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.880675Z digest=sha256:709cd63e4c545e7a82157f3290eff4b09755999f22a9254128ae7e2eeec96eab

Observation 3c9c8866-8be6-4e81-a075-0893e691af4a · outbound

This paper cites Think-on-Graph: Deep and Responsible Reasoning of Large Language Model on Knowledge Graph.

Understanding Benchmark Language Under Weakened Formal Semantics Think-on-Graph: Deep and Responsible Reasoning of Large Language Model on Knowledge Graph

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.883664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.883664Z digest=sha256:d687cc67ba839b005d032ab84935a21f42584bfc731df0f98bd8da68cca7443f

Observation 2569a87d-923a-4833-9d3b-bf5008906fcc · outbound

This paper cites In: Moens, M.-F., Huang, X., Specia, L., Yih, S.W.-t.

Understanding Benchmark Language Under Weakened Formal Semantics In: Moens, M.-F., Huang, X., Specia, L., Yih, S.W.-t

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.886681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.886681Z digest=sha256:6a83eee5d29dfc26eb36397bce72fb921a398dac1612fbc733fd734a1c74e16b

Observation b630010b-8321-47a2-a9b3-aef78d127784 · outbound

This paper cites (eds.) Findings of the Association for Computational Linguistics: NAACL 2025, pp.

Understanding Benchmark Language Under Weakened Formal Semantics (eds.) Findings of the Association for Computational Linguistics: NAACL 2025, pp

Reference 21

Resolution
verified exact
doi, observed 2026-08-15T15:51:32.009260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.889348Z digest=sha256:1bbbb5a327d1dbe577305c384689fe805fdc9f4cf837ceea4d5fc3902fb5528a

Observation 936836d9-2fbb-45a8-9ecd-fabc17b42395 · outbound

This paper cites An Empirical Study on the Potential of LLMs in Automated Software Refactoring.

Understanding Benchmark Language Under Weakened Formal Semantics An Empirical Study on the Potential of LLMs in Automated Software Refactoring

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.892074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.892074Z digest=sha256:aa567b3a2beac3cba33a1475b48a88b8795dc88447ec22b3841a72e6c3237e17

Observation 5758ca97-4232-48d5-8cea-f69f8e7e73cf · outbound

This paper cites In: Companion Proceedings of the 32nd ACM International Conference on the Foundations of Software Engineering.

Understanding Benchmark Language Under Weakened Formal Semantics In: Companion Proceedings of the 32nd ACM International Conference on the Foundations of Software Engineering

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.895066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.895066Z digest=sha256:ba674659eb757528bb1606649570c1c76c22032bd14dd257b7b4fc0dd3e58309

Observation 51d58336-5cda-474e-8f75-47b0cb2be60f · outbound

This paper cites Automated Software Engg.32(1) (2025) https://doi.org/10.1007/s10515-024-00485-2.

Understanding Benchmark Language Under Weakened Formal Semantics Automated Software Engg.32(1) (2025) https://doi.org/10.1007/s10515-024-00485-2

Reference 24

Resolution
verified exact
doi, observed 2026-08-15T15:51:31.998627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.898481Z digest=sha256:fcc3b88b1fff52abf8824ced4779f1496c66828606b6966813c2a20e9a267c99

Observation 5a334b3d-45b7-4490-b603-10f135a8f87a · outbound

This paper cites Interleaving Retrieval with Chain-of-Thought Reasoning for Knowledge-Intensive Multi-Step Questions.

Understanding Benchmark Language Under Weakened Formal Semantics Interleaving Retrieval with Chain-of-Thought Reasoning for Knowledge-Intensive Multi-Step Questions

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.901816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.901816Z digest=sha256:0bab14df6dbd960d73cf2eb3b71c9df035b929572d7d59a5f4c2b75dc251e79c

Observation e319d689-5811-40d0-8158-a427b7f3ed81 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Understanding Benchmark Language Under Weakened Formal Semantics Training Verifiers to Solve Math Word Problems

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.905574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.905574Z digest=sha256:636965578ce63bdf112c51eaa7c6fcc36a5b44e03e08cc01c1fb27c9bb66b158

Observation 5c6105d8-3613-4573-9ae0-a42a773c4324 · outbound

This paper cites an unresolved cited work.

Understanding Benchmark Language Under Weakened Formal Semantics Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:51:32.492230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.909223Z digest=sha256:80b8bcdb15f2fa8bff0e8228ecacdd76203dd3ac2af0da0a8a1e78c93547000d

Observation 5765e5eb-a4b1-4f3c-8824-880e97b5b789 · outbound

This paper cites In: The 2023 Conference on Empirical Methods in Natural Language Processing (2023).

Understanding Benchmark Language Under Weakened Formal Semantics In: The 2023 Conference on Empirical Methods in Natural Language Processing (2023)

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.481822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.912511Z digest=sha256:960c0176844e0147e8bb334472efc9e218904e46ea637d3d1c76dcc8ec347d7a

Observation b0ae1015-e9e0-4bda-86c5-0c35f85e40c3 · outbound

This paper cites Can Large Language Models Infer Causation from Correlation?.

Understanding Benchmark Language Under Weakened Formal Semantics Can Large Language Models Infer Causation from Correlation?

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.915637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.915637Z digest=sha256:2272e1619dde24eb787c2b4e161ff1fc31873c0807ff86535787c35978016e97

Observation b6e6c072-4e2f-4f6d-868f-586cfa97845b · outbound

This paper cites Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them.

Understanding Benchmark Language Under Weakened Formal Semantics Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.919237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.919237Z digest=sha256:aa079b312a7174d01ac791dbb82d066cde5e22c2bbbf6d78f195a9465183cb72

Observation cc2eecc1-32f7-4035-8f4d-87a148b3e677 · outbound

This paper cites BIG-Bench Extra Hard.

Understanding Benchmark Language Under Weakened Formal Semantics BIG-Bench Extra Hard

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.922614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.922614Z digest=sha256:dc5749bea46b4a61e093b84c9f3e224d535c7cd6f805bf231f52daec937ffb47

Observation 2aaacc5d-86e0-46d1-8f7a-4632953e5c82 · outbound

This paper cites CAIL2018: A Large-Scale Legal Dataset for Judgment Prediction.

Understanding Benchmark Language Under Weakened Formal Semantics CAIL2018: A Large-Scale Legal Dataset for Judgment Prediction

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.926275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.926275Z digest=sha256:b8f88c0c0cbe5d7aa8a6a0899ef29a60763315be0cc3b3bfc7da6ee697c3fbbb

Observation a4b48898-840e-42ae-bc38-13e239884b9f · outbound

This paper cites In: Proceedings of the Annual Conference of the North American Chapter of the Association for Computational Linguistics.

Understanding Benchmark Language Under Weakened Formal Semantics In: Proceedings of the Annual Conference of the North American Chapter of the Association for Computational Linguistics

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.471592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.929814Z digest=sha256:b5e2a032d63cdff8fcfdbed7d9b8ec5867196791eaf5cc2a8789ad685904a394

Observation c3ad653e-6229-4898-8cee-2461dfcfbd8a · outbound

This paper cites ClassActionPrediction: A Challenging Benchmark for Legal Judgment Prediction of Class Action Cases in the US.

Understanding Benchmark Language Under Weakened Formal Semantics ClassActionPrediction: A Challenging Benchmark for Legal Judgment Prediction of Class Action Cases in the US

Reference 34

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T15:51:32.131292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.933161Z digest=sha256:b1d63a2d834d275fb2da39c56c130984c66abe926177bc164a23eb5ad856aa51

Observation 3a649783-5533-4bee-b65f-c15ffa6a029c · outbound

This paper cites an unresolved cited work.

Understanding Benchmark Language Under Weakened Formal Semantics Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.937049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.937049Z digest=sha256:5a9df496441709656c5527a2fdc4521319e2d40a19050eecbcc2c4bd7727d627

Observation 75b5497b-abee-4b24-9e14-f7fa024a2853 · outbound

This paper cites In: Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP), pp.

Understanding Benchmark Language Under Weakened Formal Semantics In: Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP), pp

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.455996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.940675Z digest=sha256:a4525498f34ae486be17218141d5eceb4f3e8e2db328354e60c5c461d1978a35

Observation 0460b5c6-7bc2-47b0-8404-d1fe2a1bb79b · outbound

This paper cites an unresolved cited work.

Understanding Benchmark Language Under Weakened Formal Semantics Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:51:32.446676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.943871Z digest=sha256:12c9d9e865854a897b23b2904b76ea4b6c0677f041273929efb3dcb3cfe9abe6

Observation 4ab7248e-d1e8-4325-863b-2d4452db7ac4 · outbound

This paper cites https://github.com/google-research/google-research/tree/master/mbpp.

Understanding Benchmark Language Under Weakened Formal Semantics https://github.com/google-research/google-research/tree/master/mbpp

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.436167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.947174Z digest=sha256:f89da1d104e4a745fc694b53c96d78dbdb6fc81ad77fd91a1682b0441c7da409

Observation a82f4095-b977-48cc-9126-b7867245044c · outbound

This paper cites https://huggingface.co/datasets/PatrickHaller/the-stack-python-1M.

Understanding Benchmark Language Under Weakened Formal Semantics https://huggingface.co/datasets/PatrickHaller/the-stack-python-1M

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.425791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.950866Z digest=sha256:669348dfa30d341595db22b55b1345ab1b2adecf48ca4d35dc85223ab4e8c85f

Observation 7c9af900-b20e-4319-b889-3c04fe19f775 · outbound

This paper cites https://huggingface.co/datasets/notbadai/python_functions_reasoning.

Understanding Benchmark Language Under Weakened Formal Semantics https://huggingface.co/datasets/notbadai/python_functions_reasoning

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.414983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.954341Z digest=sha256:d6da043b1c8c799e07db8f8fa07716e3e5ad16f0c1ccd38dd072fd34c7159694

Observation 63cee2cd-60f9-4c6c-8658-d942e70e3a97 · outbound

This paper cites OpenCoder: The Open Cookbook for Top-Tier Code Large Language Models.

Understanding Benchmark Language Under Weakened Formal Semantics OpenCoder: The Open Cookbook for Top-Tier Code Large Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.957576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.957576Z digest=sha256:7518c96396895de09f27fed19926b21f4d2d70dbedd1489296ea162fe19e6995

Observation d076f12f-ff21-45e1-9944-3c955770868c · outbound

This paper cites MIT Press, ??? (2009).

Understanding Benchmark Language Under Weakened Formal Semantics MIT Press, ??? (2009)

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.404402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.960984Z digest=sha256:444b318cf184ae5879a20174034264e716fe6e25d1c3791aa9f38bb7182e8984

Observation 1b39479a-3578-4aa5-9a27-71c97cf341a4 · outbound

This paper cites IEEE Transactions on Software Engineering SE-2(4), 308–320 (1976) https://doi.org/10.1109/TSE.1976.233837.

Understanding Benchmark Language Under Weakened Formal Semantics IEEE Transactions on Software Engineering SE-2(4), 308–320 (1976) https://doi.org/10.1109/TSE.1976.233837

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:31.964547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:31.964547Z digest=sha256:0c9abd4b6d962b62922e2c5505e483cf27b5ba129a4e9b2e22ec9e3acd256baf

Observation 93d47042-3557-4b95-96d0-478cacb086b2 · outbound

This paper cites https://docs.python.org/3/ library/ast.html.

Understanding Benchmark Language Under Weakened Formal Semantics https://docs.python.org/3/ library/ast.html

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:32.393731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:31.967967Z digest=sha256:2e2ee02752771dcb2bb9c7539e499e78e8267036cc0b40e03a315c48a6afe780

Pith citing papers

No inbound Pith citation observations are available.