Pith. sign in

Paper Citation Record · LEDGER

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation

As of 17 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 5 inbound Pith citation observations for arXiv:2505.24714.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24714 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:35:34.053651Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T18:53:01.866207Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T00:32:54.660795Z

Reference resolution

45 of 45 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved45
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 46a9a3f4-bc1b-479c-9f5b-d9650a48eb46 · outbound

This paper cites online" 'onlinestring :=.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.250153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.250153Z digest=sha256:92f5bd7440ce9e5b201c638571c9bdd0ec0673b604b1e5ce5e4294b6897ecef9

Observation 30626c9e-a431-4b6a-ade8-8bff8cd64751 · outbound

This paper cites write newline.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.300258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.300258Z digest=sha256:f7416ae7de2ad7789df5a2a1e68e7082f61c6d518bff52b767a33b7ebae80337

Observation 3c9b4a6e-8e04-4a02-b2e1-8392f64ba389 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.385403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.385403Z digest=sha256:e9492e7b0d43ef8b51aeeb2d36d38b07c868cc61d5fe225f798c9a3638b49df8

Observation ae996e5e-503d-4314-81d5-d6a4ec8ae4d5 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.451544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.451544Z digest=sha256:4832206f8d9ce9915adf097775f5046b3ae6d47eabe0dd939d39443fccd2e388

Observation b4758da1-1d6f-401b-94bf-f2e350990edb · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:36.636367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T12:35:30.577706Z digest=sha256:5f1c9e6fedba4cff83b77456007c8045531e29c7f1daaf4c6e663e192ebfc71f

Observation a096e340-dcb0-43a5-81c2-c4c9c9467ecd · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:36.414991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T12:35:30.669831Z digest=sha256:347af82a4ecde49f7bd6aa7156fe7ad653e9d9270e2079410642cc3c9c5611a8

Observation b0e2a5d7-7c2f-4d10-875d-aea697ff2226 · outbound

This paper cites FinTextQA: A Dataset for Long-form Financial Question Answering.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation FinTextQA: A Dataset for Long-form Financial Question Answering

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.786174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.786174Z digest=sha256:8152b7e52b1192d764cde902a8d505408ece3d69091998d92e6f87c3104e0a2c

Observation 1e1b1bab-2e11-48bb-b6bc-8b78c245ef86 · outbound

This paper cites Are We on the Right Way for Evaluating Large Vision-Language Models?.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Are We on the Right Way for Evaluating Large Vision-Language Models?

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.868617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.868617Z digest=sha256:d58c0015e2bbf8504e1b381a4ef610c3733d8db99a00c613f13f9e0a2aa8e214

Observation 8049285d-3b31-4852-aa3d-7f1bc9756a48 · outbound

This paper cites Microsoft COCO Captions: Data Collection and Evaluation Server.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Microsoft COCO Captions: Data Collection and Evaluation Server

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.975702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.975702Z digest=sha256:bde2b8b595e4d508041d64894f56e2a3c8a97fc919b8184835f3a851e6952ac5

Observation e828c486-00f1-42da-82cd-bf068d589372 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:36.204405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T12:35:31.063802Z digest=sha256:da7b2a973d6cb5db1e6c1b03cdd0749582ad5722581abe4ce87a42204a362797

Observation c0a46116-05ce-42c4-bc55-125d8ee42754 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.132886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.132886Z digest=sha256:9ff9db55a95ad30105240ff7eac1d2de997cd4e39b2d4cae4a2d57e18b6a8732

Observation f450c3a2-f6b5-401a-9db7-6da32c58d97d · outbound

This paper cites VITA: Towards Open-Source Interactive Omni Multimodal LLM.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation VITA: Towards Open-Source Interactive Omni Multimodal LLM

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.216234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.216234Z digest=sha256:a5b1e0b5a0c46f9cbd40c6f0305cf22238e844640f2e67e03293ed62949b3050

Observation 4b08dc92-9a58-40c5-949b-d6060879c27e · outbound

This paper cites VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.300260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.300260Z digest=sha256:8b1309bbeb12a14c0f6effec321c5f38d5ecb772cddb0a77d9fd0236462810f7

Observation f31ac3a0-02af-4b1c-aa28-ff6583550664 · outbound

This paper cites MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.362263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.362263Z digest=sha256:f033248db23be10c2489efa58b1e218fa30263d78399cbbc3505e18b35117339

Observation c2de8741-0ef5-443b-aa22-0d47aaaafa92 · outbound

This paper cites MME-Finance: A Multimodal Finance Benchmark for Expert-level Understanding and Reasoning.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MME-Finance: A Multimodal Finance Benchmark for Expert-level Understanding and Reasoning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.447032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.447032Z digest=sha256:a24870209508dfac11e7c9473b806ef6c2d85ad3573339d4dc8b6de71dfe7b79

Observation 018c1fa4-a5f3-4832-9495-398357670cb4 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.544870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.544870Z digest=sha256:2f0f4f8956075f58a8f72d73a3210ca94e43e5336ab7314644fc71cffcc0c299

Observation 2ae10eca-e090-40ea-a05c-866acca43fcc · outbound

This paper cites MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.621908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.621908Z digest=sha256:abacf285a24854ee0ddf70ceb0d4dc856b3a8e3481658692981680fd0347776a

Observation 43819a59-7a02-4009-b212-5a7a357ef73b · outbound

This paper cites MMEvalPro: Calibrating Multimodal Benchmarks Towards Trustworthy and Efficient Evaluation.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MMEvalPro: Calibrating Multimodal Benchmarks Towards Trustworthy and Efficient Evaluation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.673244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.673244Z digest=sha256:1f0bf5778c920fea3ba54565770becffc8c165f761450b9871e32f306b71ddac

Observation f7924ab8-8164-475b-9820-5429feaa179a · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.753764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.753764Z digest=sha256:0f44919b71a5da64ac132746eff945f9797f78b67e3640c7c2512f8125ea4ccf

Observation 7b19a732-f152-4a6b-91da-508b6892bccd · outbound

This paper cites FinanceBench: A New Benchmark for Financial Question Answering.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation FinanceBench: A New Benchmark for Financial Question Answering

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.894547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.894547Z digest=sha256:1273861aac47ccea4db85574d4b51e946dea312d4f2c57a94170c25235be542f

Observation d43e62f6-bfd3-4309-a92f-346af62da9ac · outbound

This paper cites CFBenchmark: Chinese Financial Assistant Benchmark for Large Language Model.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation CFBenchmark: Chinese Financial Assistant Benchmark for Large Language Model

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.015781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.015781Z digest=sha256:db93ebfee0f02583dbea4605f74d53bf5819eb98985e2c96d46f6a0ca4ac5ce4

Observation a6a1a634-b303-4b6e-b6c7-e8b03cf126b3 · outbound

This paper cites SEED-Bench-2-Plus: Benchmarking Multimodal Large Language Models with Text-Rich Visual Comprehension.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation SEED-Bench-2-Plus: Benchmarking Multimodal Large Language Models with Text-Rich Visual Comprehension

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.130801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.130801Z digest=sha256:ee32556498dda59fddda0f0957bf4029d7a871c8eb7fa745787d507f9b50f11b

Observation 7ce7526c-13bd-4e7c-9562-7336fb5a67de · outbound

This paper cites SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.258980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.258980Z digest=sha256:77451e1eeae8aea3883ed62b6f8b26d4f6a74e4d8fec89723435d3d8a15edc42

Observation ce077ca2-8eea-4988-a970-464b19c96896 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:35.831999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T12:35:32.331654Z digest=sha256:e47acf42e02f7aa7a7084dc463c30e15eba3e40b96333f97eb9943d8e0c9814e

Observation 6230c13c-89c2-4300-ba64-c992c82b2b83 · outbound

This paper cites VisualBERT: A Simple and Performant Baseline for Vision and Language.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation VisualBERT: A Simple and Performant Baseline for Vision and Language

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.435193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.435193Z digest=sha256:2a7bcb6056d5aa0e9175a829774fdfad31a4f12a36b77dc6ac4aa6948a6291ad

Observation 02595493-e29f-42e0-b415-cde8ecc35ec0 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.531062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.531062Z digest=sha256:1ff8e2a6f448de92ace20f26724c7e65e781539faff69b72c628f5eb973f704f

Observation fc712e4a-7263-4d74-8a77-eb2b39696964 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.627555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.627555Z digest=sha256:3956e01506c12e0198bec9647178cd343c4c08722dc9dc64e0f0a9e5963d9438

Observation 351ae998-d17c-4a67-a52e-c13e5fa00d65 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:35.416557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T12:35:32.733322Z digest=sha256:1c0a4e6734fcb9ac068975f06e7935899205a74ebb144564d3d8f16f14251a8b

Observation b77ffe11-d12e-4b0c-bf52-1ce72333bc35 · outbound

This paper cites MMC: Advancing Multimodal Chart Understanding with Large-scale Instruction Tuning.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MMC: Advancing Multimodal Chart Understanding with Large-scale Instruction Tuning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.804493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.804493Z digest=sha256:6943700a4d0159495522a64fadb6c0012bafcdca8afcaf97c39b16f480a7182d

Observation 765647e4-25e9-4de1-829d-3e45ce599fdf · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:35.082683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T12:35:32.878031Z digest=sha256:bfcc3a2f9a3dc2569e63a5119ed380e1dd09b74b6a5eeb8630166cff927cd7ad

Observation 1eaefa05-4451-47a2-a956-17b1f11dbf87 · outbound

This paper cites Large Language Model Agent: A Survey on Methodology, Applications and Challenges.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Large Language Model Agent: A Survey on Methodology, Applications and Challenges

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.951314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.951314Z digest=sha256:8bb69f55f1bfadd7665943fbf56b481388666ab288ea911c64a0bd959728ab88

Observation c0c17bf7-1135-4f0d-80d9-d471607fb763 · outbound

This paper cites ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.032851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.032851Z digest=sha256:e2dcd2a9c8f83b8e05a022fe0f0703ae2a1926ab31ca35512aa0ab8967bd84e3

Observation a649a890-a2ae-41f9-949f-fc4452089ea5 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.100252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.100252Z digest=sha256:3d28224c5b21891124706f6cfe0f18c9fe347ac1895e623131c03dad818cdddb

Observation cdd1da3a-cfd2-4f28-87f5-2a14b13734e3 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:34.794795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T12:35:33.188206Z digest=sha256:520b6362572c440f39d44ced59f30166e80f81a0cc0fdb0b76aec6870eeb788e

Observation 8b4dde1d-9671-4586-9790-51849ae05df1 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:34.601957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T12:35:33.295608Z digest=sha256:9de9991f0cc5d2e6d4cba9ad541e18af35e60cacabc9079fb5a15ae752cf8919

Observation 73083d0d-814f-4b9a-8f00-a97cb649ec96 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Gemini: A Family of Highly Capable Multimodal Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.363156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.363156Z digest=sha256:bb71773c45cdf49fa12766738b94166aea3439fc5594bba72a8d6316847e3597

Observation d24b0338-a1f2-4828-aa74-2e83e0c92aaa · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.424499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.424499Z digest=sha256:5addba89cf37b3535000363425056bebd0742e6fe23a09b5b3216194180e46c6

Observation 9338ce06-acb8-4c3a-bb8c-256df082e970 · outbound

This paper cites CogVLM: Visual Expert for Pretrained Language Models.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation CogVLM: Visual Expert for Pretrained Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.482668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.482668Z digest=sha256:c13e7fc6c40f9e6b58bcb63b061ba38222e6fbc9a0f7a117448eeffc2a69941a

Observation bb422c51-2059-4a9a-a4db-d8f2096415f7 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.549590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.549590Z digest=sha256:61166988c04d7d9f2e43e6dc753c14e998d4b5bae67b87e0d4ed47cdcf58a33f

Observation 80fd2abe-bd96-4da6-85ee-8b42611c8015 · outbound

This paper cites FinBen: A Holistic Financial Benchmark for Large Language Models.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation FinBen: A Holistic Financial Benchmark for Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.646547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.646547Z digest=sha256:adb7b0ab4e245fdcf19005adba291044faf909ba704f03441ba375ef1db3ff7f

Observation aef175a3-a5cb-4f3d-92c4-6abd39b60f83 · outbound

This paper cites Qwen2.5 Technical Report.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Qwen2.5 Technical Report

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.745859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.745859Z digest=sha256:f7a21cc944a6df9fd3c4a5d319968a8e92ce36191e33b4384d6638205900adb2

Observation 801eec10-b278-4453-aa22-7d28771b1e1c · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:34.535561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T12:35:33.819813Z digest=sha256:8a0726dbe093dc6228fa826b8b7f11468f15acd983dd7e68f6f9094770677fa2

Observation c696ed45-e664-4492-a530-0ce48ece7a2e · outbound

This paper cites MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.911928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.911928Z digest=sha256:f1245ff02f790bc4b8760defcfe023c7b35ed59b269569836f511299c13d4838

Observation ae8b3625-fc31-4879-9886-38472b69f91b · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.972789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.972789Z digest=sha256:bf1f20fedb46c9d82f3e529c6dc4ddfa4fc4ac24a3fa786082dcd7bff4515f5b

Observation 46802361-d266-44d6-b044-e80cf70fa281 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:34.053651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:34.053651Z digest=sha256:46125ab8a2b6d48427b504227b9f06b634ae498268a232929f93a1f62742ebab

Pith citing papers

Observation f432e913-cf41-4b9d-acbe-fd43187802aa · inbound

FinCriticalED: A Visual Benchmark for Financial Fact-Level OCR cites this paper.

FinCriticalED: A Visual Benchmark for Financial Fact-Level OCR FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:20:11.563240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-17T20:19:52.701262Z digest=sha256:a70c5cd4a4d5d7003327a53a37ba8b8d301290ab6d96aaf11002fe11d35dd225

Observation 67bbb98f-4303-4c9a-9bf4-b8765afe70a7 · inbound

Strat-LLM: Stratified Strategy Alignment for LLM-based Stock Trading with Real-time Multi-Source Signals cites this paper.

Strat-LLM: Stratified Strategy Alignment for LLM-based Stock Trading with Real-time Multi-Source Signals FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:56:08.512517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-08T10:42:08.721879Z digest=sha256:b84058069750e5faaaf6ddfc30623ad6d04444a64f0fc958e2ef2482d5579a1e

Observation 9629dd8f-8fb5-47aa-bb10-a92f24c2ff51 · inbound

FinDocMRE: A Benchmark for Document-Level Financial Multimodal Reasoning Evaluation cites this paper.

FinDocMRE: A Benchmark for Document-Level Financial Multimodal Reasoning Evaluation FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-20T00:32:54.664118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-20T00:28:48.871454Z digest=sha256:bc31e71bbe16d884407d10db7ab0c0632463863cb3e923cddb158073dcaee64b

Observation 909eae90-6689-45d6-94bf-624b8d8cd79b · inbound

Are the Financial Reasoning from LLMs Credible? A Real World Test over Long-Horizon Statements cites this paper.

Are the Financial Reasoning from LLMs Credible? A Real World Test over Long-Horizon Statements FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T00:48:03.603561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:48:03.603561Z digest=sha256:9068de29a72ef037ed6cc33998d84ac3a5c88ef53fc52528e4d33cf174a855da

Observation 95dc9f97-1f97-4eb9-948b-acd3850c3991 · inbound

FinReportBench: Measuring and Improving Institution-Grade Financial Report Generation cites this paper.

FinReportBench: Measuring and Improving Institution-Grade Financial Report Generation FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T18:53:01.866207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T18:53:01.866207Z digest=sha256:5e2c2096870c22adf9374932ddd75a9eeb54d2d7689742667c71e2155d708ccb