Pith. sign in

Paper Citation Record · LEDGER

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation

As of 8 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 4 inbound Pith citation observations for arXiv:2505.24714.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24714 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:35:34.053651Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T00:48:03.603561Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T00:32:54.660795Z

Reference resolution

45 of 45 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved45
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 46a9a3f4-bc1b-479c-9f5b-d9650a48eb46 · outbound

This paper cites online" 'onlinestring :=.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.250153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.250153Z digest=sha256:92f5bd7440ce9e5b201c638571c9bdd0ec0673b604b1e5ce5e4294b6897ecef9

Observation 30626c9e-a431-4b6a-ade8-8bff8cd64751 · outbound

This paper cites write newline.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.300258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.300258Z digest=sha256:f7416ae7de2ad7789df5a2a1e68e7082f61c6d518bff52b767a33b7ebae80337

Observation 3c9b4a6e-8e04-4a02-b2e1-8392f64ba389 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.385403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.385403Z digest=sha256:cf79ae3613ff6cfe7a536bdb25516c8f30dba8d3d1dafc813c74e81f8124de6a

Observation ae996e5e-503d-4314-81d5-d6a4ec8ae4d5 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.451544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.451544Z digest=sha256:4832206f8d9ce9915adf097775f5046b3ae6d47eabe0dd939d39443fccd2e388

Observation b4758da1-1d6f-401b-94bf-f2e350990edb · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:36.636367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:30.577706Z digest=sha256:967e9e543ad0801fe20a09cbcb761d499d8a359737a822420f16c56b42b306e8

Observation a096e340-dcb0-43a5-81c2-c4c9c9467ecd · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:36.414991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:30.669831Z digest=sha256:d5db727e3166bf457f147f22217299754f404b96b85f586463dda709ec72b1fa

Observation b0e2a5d7-7c2f-4d10-875d-aea697ff2226 · outbound

This paper cites FinTextQA: A Dataset for Long-form Financial Question Answering.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation FinTextQA: A Dataset for Long-form Financial Question Answering

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.786174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.786174Z digest=sha256:83fa60ccb231c23b586a01e2a3227920f5f687656bbf6027d75db988ccb8a233

Observation 1e1b1bab-2e11-48bb-b6bc-8b78c245ef86 · outbound

This paper cites Are We on the Right Way for Evaluating Large Vision-Language Models?.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Are We on the Right Way for Evaluating Large Vision-Language Models?

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.868617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.868617Z digest=sha256:d58c0015e2bbf8504e1b381a4ef610c3733d8db99a00c613f13f9e0a2aa8e214

Observation 8049285d-3b31-4852-aa3d-7f1bc9756a48 · outbound

This paper cites Microsoft COCO Captions: Data Collection and Evaluation Server.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Microsoft COCO Captions: Data Collection and Evaluation Server

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.975702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.975702Z digest=sha256:f8980eac0934937cdd81646216725fad79429c6e15480ff39f27fd8c615f5aeb

Observation e828c486-00f1-42da-82cd-bf068d589372 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:36.204405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:31.063802Z digest=sha256:3f70b13a5965c77993d241e637558e74d0872d7b732394b53f525310d4eb37fc

Observation c0a46116-05ce-42c4-bc55-125d8ee42754 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.132886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.132886Z digest=sha256:9ff9db55a95ad30105240ff7eac1d2de997cd4e39b2d4cae4a2d57e18b6a8732

Observation f450c3a2-f6b5-401a-9db7-6da32c58d97d · outbound

This paper cites VITA: Towards Open-Source Interactive Omni Multimodal LLM.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation VITA: Towards Open-Source Interactive Omni Multimodal LLM

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.216234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.216234Z digest=sha256:80f127fa4b48cf43ef0347f564b7f7511cbf2db9b36b3c2dc76d8709aba9d50d

Observation 4b08dc92-9a58-40c5-949b-d6060879c27e · outbound

This paper cites VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.300260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.300260Z digest=sha256:4829693c2b76953b7f4907c57a067b44f7876db989a88fcb3bdb732a766635d1

Observation f31ac3a0-02af-4b1c-aa28-ff6583550664 · outbound

This paper cites MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.362263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.362263Z digest=sha256:7b2534b1884ffc2da86d52f6d91ce1e4f2282431c64bd286cabad26bbe4ba6af

Observation c2de8741-0ef5-443b-aa22-0d47aaaafa92 · outbound

This paper cites MME-Finance: A Multimodal Finance Benchmark for Expert-level Understanding and Reasoning.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MME-Finance: A Multimodal Finance Benchmark for Expert-level Understanding and Reasoning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.447032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.447032Z digest=sha256:6c020bef74dda20d9fadaea809dea4e06fe085d553421ae593482c68ff4b72f4

Observation 018c1fa4-a5f3-4832-9495-398357670cb4 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.544870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.544870Z digest=sha256:2f0f4f8956075f58a8f72d73a3210ca94e43e5336ab7314644fc71cffcc0c299

Observation 2ae10eca-e090-40ea-a05c-866acca43fcc · outbound

This paper cites MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.621908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.621908Z digest=sha256:3d781bdfc34957c0ec2ce216d5d07f1f1230563b385d9d2585c71de19a6c04eb

Observation 43819a59-7a02-4009-b212-5a7a357ef73b · outbound

This paper cites MMEvalPro: Calibrating Multimodal Benchmarks Towards Trustworthy and Efficient Evaluation.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MMEvalPro: Calibrating Multimodal Benchmarks Towards Trustworthy and Efficient Evaluation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.673244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.673244Z digest=sha256:fed64037d71319ee8c6764c98a9b5fb0ed95b4ca85ec90d32c07ee4fcba5c7c7

Observation f7924ab8-8164-475b-9820-5429feaa179a · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.753764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.753764Z digest=sha256:0f44919b71a5da64ac132746eff945f9797f78b67e3640c7c2512f8125ea4ccf

Observation 7b19a732-f152-4a6b-91da-508b6892bccd · outbound

This paper cites FinanceBench: A New Benchmark for Financial Question Answering.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation FinanceBench: A New Benchmark for Financial Question Answering

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.894547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.894547Z digest=sha256:71c208da2e261c48fc9a4c37b83d487948b18d1df951c4122a56a7cabf6a72a4

Observation d43e62f6-bfd3-4309-a92f-346af62da9ac · outbound

This paper cites CFBenchmark: Chinese Financial Assistant Benchmark for Large Language Model.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation CFBenchmark: Chinese Financial Assistant Benchmark for Large Language Model

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.015781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.015781Z digest=sha256:38997b0a77da200eacda371fe322be66e24a5222835b0ae77739bc7d62744397

Observation a6a1a634-b303-4b6e-b6c7-e8b03cf126b3 · outbound

This paper cites SEED-Bench-2-Plus: Benchmarking Multimodal Large Language Models with Text-Rich Visual Comprehension.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation SEED-Bench-2-Plus: Benchmarking Multimodal Large Language Models with Text-Rich Visual Comprehension

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.130801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.130801Z digest=sha256:c96d6d4da4c7ce3776946533976c70941e77ecba66f32fb5e72f17341e081582

Observation 7ce7526c-13bd-4e7c-9562-7336fb5a67de · outbound

This paper cites SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.258980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.258980Z digest=sha256:77451e1eeae8aea3883ed62b6f8b26d4f6a74e4d8fec89723435d3d8a15edc42

Observation ce077ca2-8eea-4988-a970-464b19c96896 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:35.831999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:32.331654Z digest=sha256:163c86f297fc8b704c3cdec241e056f42202796191ed18f272f0a59cb2f91e1d

Observation 6230c13c-89c2-4300-ba64-c992c82b2b83 · outbound

This paper cites VisualBERT: A Simple and Performant Baseline for Vision and Language.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation VisualBERT: A Simple and Performant Baseline for Vision and Language

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.435193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.435193Z digest=sha256:2a7bcb6056d5aa0e9175a829774fdfad31a4f12a36b77dc6ac4aa6948a6291ad

Observation 02595493-e29f-42e0-b415-cde8ecc35ec0 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.531062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.531062Z digest=sha256:1ff8e2a6f448de92ace20f26724c7e65e781539faff69b72c628f5eb973f704f

Observation fc712e4a-7263-4d74-8a77-eb2b39696964 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.627555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.627555Z digest=sha256:3956e01506c12e0198bec9647178cd343c4c08722dc9dc64e0f0a9e5963d9438

Observation 351ae998-d17c-4a67-a52e-c13e5fa00d65 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:35.416557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:32.733322Z digest=sha256:17a8657a9d879d9d7b7d3e7838b4d2ca4f5398817b4da8f38e810ac613a6aaa2

Observation b77ffe11-d12e-4b0c-bf52-1ce72333bc35 · outbound

This paper cites MMC: Advancing Multimodal Chart Understanding with Large-scale Instruction Tuning.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MMC: Advancing Multimodal Chart Understanding with Large-scale Instruction Tuning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.804493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.804493Z digest=sha256:dc941f8875fe7e35cf2168acfa3c9dd17347d37527bc074163d214c6d5acfe2f

Observation 765647e4-25e9-4de1-829d-3e45ce599fdf · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:35.082683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:32.878031Z digest=sha256:f9ef45723c345db10a1c31a6ccd1294d4ad8b0bdb833704f0e59139fe939b559

Observation 1eaefa05-4451-47a2-a956-17b1f11dbf87 · outbound

This paper cites Large Language Model Agent: A Survey on Methodology, Applications and Challenges.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Large Language Model Agent: A Survey on Methodology, Applications and Challenges

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.951314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.951314Z digest=sha256:8bb69f55f1bfadd7665943fbf56b481388666ab288ea911c64a0bd959728ab88

Observation c0c17bf7-1135-4f0d-80d9-d471607fb763 · outbound

This paper cites ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.032851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.032851Z digest=sha256:de435f06c514fb631f1bfe8ac26e0eaa44ac64576522535d61a4ecdf25101e63

Observation a649a890-a2ae-41f9-949f-fc4452089ea5 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.100252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.100252Z digest=sha256:3d28224c5b21891124706f6cfe0f18c9fe347ac1895e623131c03dad818cdddb

Observation cdd1da3a-cfd2-4f28-87f5-2a14b13734e3 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:34.794795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:33.188206Z digest=sha256:5f905c1ac32f1c7485e503b767f201e1992a759983c74770242b19cb009c2389

Observation 8b4dde1d-9671-4586-9790-51849ae05df1 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:34.601957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:33.295608Z digest=sha256:53b1188f793f02bfbedcd81d35e2df7fc7649b1d46495074abcf318a2f30c529

Observation 73083d0d-814f-4b9a-8f00-a97cb649ec96 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Gemini: A Family of Highly Capable Multimodal Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.363156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.363156Z digest=sha256:bb71773c45cdf49fa12766738b94166aea3439fc5594bba72a8d6316847e3597

Observation d24b0338-a1f2-4828-aa74-2e83e0c92aaa · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.424499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.424499Z digest=sha256:5addba89cf37b3535000363425056bebd0742e6fe23a09b5b3216194180e46c6

Observation 9338ce06-acb8-4c3a-bb8c-256df082e970 · outbound

This paper cites CogVLM: Visual Expert for Pretrained Language Models.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation CogVLM: Visual Expert for Pretrained Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.482668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.482668Z digest=sha256:c13e7fc6c40f9e6b58bcb63b061ba38222e6fbc9a0f7a117448eeffc2a69941a

Observation bb422c51-2059-4a9a-a4db-d8f2096415f7 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.549590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.549590Z digest=sha256:61166988c04d7d9f2e43e6dc753c14e998d4b5bae67b87e0d4ed47cdcf58a33f

Observation 80fd2abe-bd96-4da6-85ee-8b42611c8015 · outbound

This paper cites FinBen: A Holistic Financial Benchmark for Large Language Models.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation FinBen: A Holistic Financial Benchmark for Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.646547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.646547Z digest=sha256:d2cfec8fec5dcc1f7cbc8ca1c07bcb3fe598b6cb7f94050e9d3d0b16a5546bb2

Observation aef175a3-a5cb-4f3d-92c4-6abd39b60f83 · outbound

This paper cites Qwen2.5 Technical Report.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Qwen2.5 Technical Report

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.745859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.745859Z digest=sha256:ffa04fc8138f60f31208eb374b1c33630f7165fe8b14ab0decf74a81e621fcb9

Observation 801eec10-b278-4453-aa22-7d28771b1e1c · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:34.535561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T12:35:33.819813Z digest=sha256:ac68e3e8d102724349b1ebcb4bf2c6b923cb1d00e4cf699ef3b7de9ebb014aa5

Observation c696ed45-e664-4492-a530-0ce48ece7a2e · outbound

This paper cites MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.911928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.911928Z digest=sha256:286f941846f6dfbc4e90d704e1bcc7685edf75995113793a8003c8791ac0abea

Observation ae8b3625-fc31-4879-9886-38472b69f91b · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.972789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.972789Z digest=sha256:bf1f20fedb46c9d82f3e529c6dc4ddfa4fc4ac24a3fa786082dcd7bff4515f5b

Observation 46802361-d266-44d6-b044-e80cf70fa281 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:34.053651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:34.053651Z digest=sha256:46125ab8a2b6d48427b504227b9f06b634ae498268a232929f93a1f62742ebab

Pith citing papers

Observation f432e913-cf41-4b9d-acbe-fd43187802aa · inbound

FinCriticalED: A Visual Benchmark for Financial Fact-Level OCR cites this paper.

FinCriticalED: A Visual Benchmark for Financial Fact-Level OCR FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:20:11.563240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T20:19:52.701262Z digest=sha256:49fccc2b355e5ef02ed96d2ab0356c5f1515cf06791439c31aad19ed60958dd7

Observation 67bbb98f-4303-4c9a-9bf4-b8765afe70a7 · inbound

Strat-LLM: Stratified Strategy Alignment for LLM-based Stock Trading with Real-time Multi-Source Signals cites this paper.

Strat-LLM: Stratified Strategy Alignment for LLM-based Stock Trading with Real-time Multi-Source Signals FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:56:08.512517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T10:42:08.721879Z digest=sha256:8fdde53344e82aa73057d9a3796d8243c7970b984c2e111ed4bcf9f46f6c1d8e

Observation 9629dd8f-8fb5-47aa-bb10-a92f24c2ff51 · inbound

FinDocMRE: A Benchmark for Document-Level Financial Multimodal Reasoning Evaluation cites this paper.

FinDocMRE: A Benchmark for Document-Level Financial Multimodal Reasoning Evaluation FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-20T00:32:54.664118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T00:28:48.871454Z digest=sha256:3fa1a90d8cce9849b10d21c712476d20fa9ef9d5279ed06482b4fc0fe25532b2

Observation 909eae90-6689-45d6-94bf-624b8d8cd79b · inbound

Are the Financial Reasoning from LLMs Credible? A Real World Test over Long-Horizon Statements cites this paper.

Are the Financial Reasoning from LLMs Credible? A Real World Test over Long-Horizon Statements FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T00:48:03.603561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:48:03.603561Z digest=sha256:cbcc04ce0e4db187d6ec603a71a8bb8205218e75ed32d5848d9375658d9fbc26