Pith. sign in

Paper Citation Record · LEDGER

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables

As of 9 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 0 inbound Pith citation observations for arXiv:2608.04077.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04077 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T00:39:34.850226Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

50 of 50 outbound references displayed

  • verified exact3
  • verified fuzzy11
  • unresolved36
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0e0489e4-4def-4311-89a1-b690d341cb8c · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Constitutional AI: Harmlessness from AI Feedback

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.692929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.692929Z digest=sha256:5985d9c48ef52a8591a8c65db2274b432dca2f3c8489cf8caf8fc66017e24254

Observation ff28b7f4-c0a7-4882-ad26-d0f5c327341b · outbound

This paper cites A.; and Terry, M.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables A.; and Terry, M

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.785146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.697760Z digest=sha256:504710f7b94085d0c588ce90cff7875ed00da41cd2fe82f33e0b9b9d51c0398a

Observation e5fbf2a1-d75e-486b-8828-2b3768a59c10 · outbound

This paper cites DISC-FinLLM: A Chinese Financial Large Language Model based on Multiple Experts Fine-tuning.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables DISC-FinLLM: A Chinese Financial Large Language Model based on Multiple Experts Fine-tuning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.701168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.701168Z digest=sha256:c904c7246d67f953d94b95a2f0b2d7447ca34767cfc10ba4c792aef5629824bc

Observation 292d2183-e124-4e8b-9926-67cdb67e7ef9 · outbound

This paper cites N.; Li, T.; Li, D.; Zhu, B.; Zhang, H.; Jordan, M.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables N.; Li, T.; Li, D.; Zhu, B.; Zhang, H.; Jordan, M

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.775289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.703990Z digest=sha256:3abe40ba52849efb31ae1edd4e1ab56158674b7da881da37a610d1e412a7cc05

Observation 715481de-f1aa-4784-9716-0375167468ba · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.767171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.706593Z digest=sha256:3a6bd2850f92d60484efc582e3fa219c7c57ed21ba932b5a434daacd096f368b

Observation 205a2af6-4d1f-450b-a694-701186f1ceb7 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 6

Resolution
verified exact
raw_fallback, observed 2026-08-08T00:39:35.591321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.710181Z digest=sha256:ce96fb6b18dcb6a8237c50699c1e75f7f4ba9936f5866a3aa9dbcc3e1572e12d

Observation 9ef16854-2f69-480b-b599-fe7f05b64e5b · outbound

This paper cites E.; R \'e , C.; Chilton, A.; Narayana, A.; Chohlas-Wood, A.; Peters, A.; Waldon, B.; Rockmore, D.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables E.; R \'e , C.; Chilton, A.; Narayana, A.; Chohlas-Wood, A.; Peters, A.; Waldon, B.; Rockmore, D

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.759686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.712947Z digest=sha256:38d2bad410e420738baf66d3b1f1d7c7353bed7275e70639c1f6d9d338f66396

Observation 3dcb4e1f-1932-4953-8c32-5dc31fcccc48 · outbound

This paper cites FinEval: A Chinese Financial Domain Knowledge Evaluation Benchmark for Large Language Models.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables FinEval: A Chinese Financial Domain Knowledge Evaluation Benchmark for Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.715710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.715710Z digest=sha256:697b47e66c20b83f73b1f91f90d0caf97b5a3cabdfd214b6aff25b8e664e187f

Observation 9ad630f9-9cd5-44a8-a45f-06550b4afa92 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.719154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.719154Z digest=sha256:9d6a858d0e199c84b060867807289984f5023df605576acb17fec94823a23510

Observation 82722c69-3480-4d5e-89a3-fa737e5c1de5 · outbound

This paper cites FinanceBench: A New Benchmark for Financial Question Answering.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables FinanceBench: A New Benchmark for Financial Question Answering

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.722411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.722411Z digest=sha256:2f20ebb709c303d1a8910864333fd551531067aaab1b63a2a6137f2248922a09

Observation 43397b8a-7e53-494a-83ad-ebd2ccb9a496 · outbound

This paper cites E.; Yang, J.; Wettig, A.; Yao, S.; Pei, K.; Press, O.; and Narasimhan, K.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables E.; Yang, J.; Wettig, A.; Yao, S.; Pei, K.; Press, O.; and Narasimhan, K

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.748658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.725837Z digest=sha256:86caa4f8be6e092b9ff36e4247855f6fc039b3d29f055d4056309dc58098c4c4

Observation c3727f3e-79cd-4e4f-af44-ff5d5e9f5e04 · outbound

This paper cites What Disease does this Patient Have? A Large-scale Open Domain Question Answering Dataset from Medical Exams.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables What Disease does this Patient Have? A Large-scale Open Domain Question Answering Dataset from Medical Exams

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.729453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.729453Z digest=sha256:760b0edc909ea320349415c6575acd6139ec4bdca5e27554d8da24a27bda749e

Observation 37f84774-50fb-457a-b2b2-7dfaa74d9b33 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.740711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.733674Z digest=sha256:08aea604203831a34485ff78cea4edaf01caa2685953d516a4523da39337bc0e

Observation fec3250d-6c14-44e8-aec2-606e17cdbf85 · outbound

This paper cites RewardBench: Evaluating Reward Models for Language Modeling.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables RewardBench: Evaluating Reward Models for Language Modeling

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.737024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.737024Z digest=sha256:efcd81760fe3d3688b97c14ef9ec877edb71eb86a06be6075d80355d3ce268ca

Observation 2dfb6c3b-3496-4058-a2d0-7fc9d6b18843 · outbound

This paper cites CFBenchmark: Chinese Financial Assistant Benchmark for Large Language Model.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables CFBenchmark: Chinese Financial Assistant Benchmark for Large Language Model

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.741055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.741055Z digest=sha256:65dfb2b0fa6e9693030151fcc96eb59895aada2e9c3f5e844933c4ad860cdd49

Observation b91c7869-32d7-4f27-9cd2-585456c72a57 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.745767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.745767Z digest=sha256:825fe7cdc2ad571d490e2e8860e0788fd2977a665afca90335ad384db9158bc1

Observation 680c13f9-7a68-4ed1-8936-1974e04d9c9c · outbound

This paper cites ARES: Automated Rubric Synthesis for Scalable LLM Reinforcement Learning.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables ARES: Automated Rubric Synthesis for Scalable LLM Reinforcement Learning

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-08T00:39:35.432825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.748768Z digest=sha256:809bcdbd3852b1857d7393ab4fbe7b30e8ae9586f5787120cebe6b7466a991ad

Observation 41fee5ac-b157-482e-aa9a-8150980958ec · outbound

This paper cites JobBench: Aligning Agent Work With Human Will.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables JobBench: Aligning Agent Work With Human Will

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.752504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.752504Z digest=sha256:936ad700d5f618690e4b9bc4482d6b966fec8df0942cb9a0890e014d42f2f4f5

Observation 8eed73fb-c9a4-42e9-b4ce-fbf3a853ffdd · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.733076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.756993Z digest=sha256:df4d0363f022838b0baa20886c222e8d1c3cab690ce8057903207f9e49ff8353

Observation 31e9824c-c8ef-4403-bb60-113d4d9226ec · outbound

This paper cites Y.; Deng, Y.; Chandu, K.; Brahman, F.; Bhagavatula, C.; and Choi, Y.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Y.; Deng, Y.; Chandu, K.; Brahman, F.; Bhagavatula, C.; and Choi, Y

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.724302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.760575Z digest=sha256:6aed24f222b4c396ace952c560be622d75e15e48cad0d1aea64fd88cc94a2776

Observation cf680bf0-6096-43d9-9e00-a943cae13ef3 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.765278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.765278Z digest=sha256:b5654072cdea7208fbdc6d4241d2d91aa2c0ae95d596db4b8832aad8556d6663

Observation b2ea64d1-7ae5-4b21-a539-d14d4e553fec · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.716019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.768687Z digest=sha256:2b16cf55d46ddf24e9cade29e2c234dc284d5d557267acaa9a250689296c4338

Observation 81cf1719-acc9-4d98-acbd-df68591dabf3 · outbound

This paper cites AgentBench: Evaluating LLMs as Agents.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables AgentBench: Evaluating LLMs as Agents

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.772089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.772089Z digest=sha256:55c2048ad0e992afbde87924e1f0b12ce8df9ee2b38cb2743a9bf24fb87317af

Observation bdd8ea85-1aa0-410a-9251-d3d142ddec80 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.707601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.775217Z digest=sha256:d1ae102729d531e1f80c29181e384e1a147bfc174eb56a166652ba495c1edde2

Observation 9b3690fa-dd86-4197-a962-8c27a9431167 · outbound

This paper cites FinResearchBench II: A Deep Research Benchmark with Consensus-Derived Gold Rubrics for Distinguishing Financial Report Quality.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables FinResearchBench II: A Deep Research Benchmark with Consensus-Derived Gold Rubrics for Distinguishing Financial Report Quality

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.778096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.778096Z digest=sha256:d8169f58b9ef2e26eae0faf26c465d86d3887418a4f0fa31820dc8abcc68759b

Observation c1246f93-9f84-41db-b992-c951877e8222 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.699142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.782085Z digest=sha256:df21e2999afca1d81712edca4600eec93717a94d245eda9f616a0d73e9e012e1

Observation 3038f6ce-47d4-42b2-8b90-f7921749749c · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.691815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.784771Z digest=sha256:f12fec0e9c02dc1d1f5f8c321289d2d16c11a9db229815846ed3a7df3a8071cf

Observation 3d34ca5f-c1e0-454d-9ae9-f11d1ebe5fba · outbound

This paper cites GPT-4 Technical Report.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables GPT-4 Technical Report

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.787813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.787813Z digest=sha256:c728779fc998c0207f1555806a32155bf4e178b7d3a3809774fc5ad7170ebfba

Observation 555ecf1d-a730-4e56-91ba-934108541986 · outbound

This paper cites L.; Mishkin, P.; Zhang, C.; Agarwal, S.; Slama, K.; Ray, A.; et al.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables L.; Mishkin, P.; Zhang, C.; Agarwal, S.; Slama, K.; Ray, A.; et al

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.683610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.790805Z digest=sha256:bcd5f68e6a76bca63d40dd7f7cce6ea24d9c8f1ba5b8213eb05f8289f00edc21

Observation 77965a63-7f2f-4934-883d-dbd1bbc416a3 · outbound

This paper cites GDPval: Evaluating AI Model Performance on Real-World Economically Valuable Tasks.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables GDPval: Evaluating AI Model Performance on Real-World Economically Valuable Tasks

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.793186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.793186Z digest=sha256:8dd4e539e042d96896c22784c24ad25bcefdc4bd248593a89c24d9f3c7d8baa9

Observation 2ba8720e-de77-441d-bc25-0bb5612bbbaa · outbound

This paper cites D.; Ermon, S.; and Finn, C.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables D.; Ermon, S.; and Finn, C

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.796664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.796664Z digest=sha256:a5ff51410daff36d8a6ac41b95bd6b7bfa245d8a6c17ff584c9477e1e5b61b7a

Observation e1975965-a9ce-4382-af5d-568ceb3cc374 · outbound

This paper cites L.; Stickland, A.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables L.; Stickland, A

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.671807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.799748Z digest=sha256:b1f5928824d490572b01cafc6b8fad20d603742066b417757f91808e76ca1d43

Observation ef025533-8f6e-40f5-8f6d-d8f3d3a31e42 · outbound

This paper cites S.; Chawla, K.; Eidnani, D.; Shah, A.; Du, W.; Chava, S.; Raman, N.; Smiley, C.; Chen, J.; and Yang, D.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables S.; Chawla, K.; Eidnani, D.; Shah, A.; Du, W.; Chava, S.; Raman, N.; Smiley, C.; Chen, J.; and Yang, D

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.662462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.802371Z digest=sha256:31db2d5980fcfddbdf29aa980105f27ad321ca99b3c5765984b502c7236306ae

Observation 18828a47-4c2e-4f53-b9fe-e21e2de1e0bf · outbound

This paper cites F.; Qiu, X.; Whitehouse, C.; Alazraki, L.; Goel, S.; Barbieri, F.; Willi, T.; Mathur, A.; and Leontiadis, I.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables F.; Qiu, X.; Whitehouse, C.; Alazraki, L.; Goel, S.; Barbieri, F.; Willi, T.; Mathur, A.; and Leontiadis, I

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.805634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.805634Z digest=sha256:daa5b5f552810844058126148e431694e071cec13ce236a00ecb880591bef3c7

Observation 2d9b5fde-e5b8-4b90-a9de-e53d3588eb76 · outbound

This paper cites Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.809274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.809274Z digest=sha256:793f5c35484aade4b3def974062935951f2ae270b392a148279907cc7e3e374f

Observation eaf13a77-dd98-40bb-a2d9-b349356bf4b4 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.812234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.812234Z digest=sha256:62b58728249a2b3505ad063823906767920b717de1177f801daa896abbb943a9

Observation c240abeb-97cf-457e-9f62-0faf239c4e3a · outbound

This paper cites R.; Zhang, S.; Sun, Y.; and Wang, W.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables R.; Zhang, S.; Sun, Y.; and Wang, W

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.652830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.814612Z digest=sha256:6f29210704e8e052d41774d50e46012c571dbdb4cffdec8c44ec8cb9a6107705

Observation 994903cb-0052-4517-bc69-44ce23c6314d · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.644264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.818031Z digest=sha256:645b3b89823d1e1183101c06c22dc7249a765d132f7fdba42b16d09dcc614fea

Observation 0fb0c3e0-5527-4f5f-9a73-0914d091ade0 · outbound

This paper cites BloombergGPT: A Large Language Model for Finance.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables BloombergGPT: A Large Language Model for Finance

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.821009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.821009Z digest=sha256:eb1b36d2b41b11a36f74c1b76101bcebaa71d0244ccd64890a4ee11a8e629564

Observation 0694b070-9be6-4301-bf50-1b91f6d33db8 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.823780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.823780Z digest=sha256:91b6c8e21a58d61ceaf1fed655f72d34eeb03d0f2bdbfbbb253cdfedcd360e4d

Observation 29850e5e-9dfc-47d3-81c9-da242a46b030 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.636713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.826266Z digest=sha256:9146d8a9590c5f2727887ffbb70d21855396a1a00d87acedfc4a1694fd42737f

Observation 520d8522-0943-466a-8dfe-a5182b9142b1 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.629569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.828757Z digest=sha256:2b652f04b4724a9144bca9dcfa6d6a6e59f8c2019f7005de7b9d7ed67e7945bb

Observation 96c4e350-66c7-4c5e-87f5-fc76ddad77ec · outbound

This paper cites J.; Cheng, Z.; Shin, D.; Lei, F.; et al.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables J.; Cheng, Z.; Shin, D.; Lei, F.; et al

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.621353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.831495Z digest=sha256:52b9dd81eaa39a7e85e5ef8cf4aa7a330612bd26476871c1a0e543d72f15c651

Observation 2bbe066f-1684-466d-9b35-57fe113b1f3b · outbound

This paper cites TheAgentCompany: Benchmarking LLM Agents on Consequential Real World Tasks.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables TheAgentCompany: Benchmarking LLM Agents on Consequential Real World Tasks

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.834018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.834018Z digest=sha256:792830cb74290a68bb5aeb05a28f935e0c638f3525ae69b26a875f3742dbb794

Observation 5614d325-e0e8-413b-bb38-f520a481b47b · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.836759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.836759Z digest=sha256:8a2b1050b14e09c07ba1ce19268aa8b91f05a83a7b49518f0999124b9323770c

Observation 55fabdb5-d135-4096-96fd-b7aa4720107c · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.839268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.839268Z digest=sha256:357cbfbd0bf62fdbfece46a09d356ab062cf881d5374236e69e80e0ec5fdcc3e

Observation 2729b7a3-2b0c-43b9-9a03-93ffc16555c0 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.841766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.841766Z digest=sha256:5838bb92f6b56f711f47c04ac5c9af5ef36eca6799afa266bdcb51738b5d050e

Observation 5d682386-6df6-4e0c-9e19-d3180f0d3798 · outbound

This paper cites P.; Zhang, H.; Gonzalez, J.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables P.; Zhang, H.; Gonzalez, J

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.613211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.844682Z digest=sha256:6130aae66a29c99adcc9715d1c690604eaca1fed85fd12a4db46c849309ff14c

Observation 2351c981-e091-4280-9b88-18fdaca87113 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 49

Resolution
verified exact
raw_fallback, observed 2026-08-08T00:39:34.942147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.847515Z digest=sha256:ed1873a328db0939d9c98c67e6e85214366edd3cad3ae063aa4397688762d260

Observation 323f78cf-7f11-45ea-8cd1-d91a038ad57a · outbound

This paper cites F.; Zhu, H.; Zhou, X.; Lo, R.; Sridhar, A.; Cheng, X.; Ou, T.; Bisk, Y.; Fried, D.; Alon, U.; and Neubig, G.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables F.; Zhu, H.; Zhou, X.; Lo, R.; Sridhar, A.; Cheng, X.; Ou, T.; Bisk, Y.; Fried, D.; Alon, U.; and Neubig, G

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.850226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.850226Z digest=sha256:7586e602bdbda5f4497bae41e7c83ce4de8bc670af2fd561e05e0f27298a3195

Pith citing papers

No inbound Pith citation observations are available.