Pith. sign in

Paper Citation Record · LEDGER

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation

As of 22 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 5 inbound Pith citation observations for arXiv:2505.24714.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24714 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:35:34.053651Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T18:53:01.866207Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T00:32:54.660795Z

Reference resolution

45 of 45 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved45
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 46a9a3f4-bc1b-479c-9f5b-d9650a48eb46 · outbound

This paper cites online" 'onlinestring :=.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.250153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.250153Z digest=sha256:65a6684a0d53e498118c05d8beab99f7a48d2b61ab322d61197852ebd2b2688b

Observation 30626c9e-a431-4b6a-ade8-8bff8cd64751 · outbound

This paper cites write newline.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.300258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.300258Z digest=sha256:ac2d6807b537d2cf7f0e5d041dc0d3f488c789939403b676b301cd054aa0be5b

Observation 3c9b4a6e-8e04-4a02-b2e1-8392f64ba389 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.385403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.385403Z digest=sha256:e981a9fea0f7e37847427b6477632046585c2d3cbb32f73157417db62f26b472

Observation ae996e5e-503d-4314-81d5-d6a4ec8ae4d5 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.451544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.451544Z digest=sha256:9b0085900aee17935c18ba1110524660d74d2d5ee547de5c6a0c4d523f0cdb7f

Observation b4758da1-1d6f-401b-94bf-f2e350990edb · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:36.636367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-07T12:35:30.577706Z digest=sha256:f594bd923336671dfeb9a9ca21ffd001c130557043ac8b093077ccce57526fb6

Observation a096e340-dcb0-43a5-81c2-c4c9c9467ecd · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:36.414991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-07T12:35:30.669831Z digest=sha256:05b552504d05658f802e933923ba4c837bd73a89317b1326b0bad8fa1342f3a7

Observation b0e2a5d7-7c2f-4d10-875d-aea697ff2226 · outbound

This paper cites FinTextQA: A Dataset for Long-form Financial Question Answering.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation FinTextQA: A Dataset for Long-form Financial Question Answering

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.786174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.786174Z digest=sha256:3a923ae9aad376fa8ad7ec8b0dce6224587a696b7a2f457a11fd3eacd12ce959

Observation 1e1b1bab-2e11-48bb-b6bc-8b78c245ef86 · outbound

This paper cites Are We on the Right Way for Evaluating Large Vision-Language Models?.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Are We on the Right Way for Evaluating Large Vision-Language Models?

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.868617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.868617Z digest=sha256:a578286bb8f0a87732f7bde06343ba28523bd03bf25c6d7963ae23099c1bad96

Observation 8049285d-3b31-4852-aa3d-7f1bc9756a48 · outbound

This paper cites Microsoft COCO Captions: Data Collection and Evaluation Server.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Microsoft COCO Captions: Data Collection and Evaluation Server

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.975702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:30.975702Z digest=sha256:a5e36da3606c3f84ab1307b8bc10981bf22becde9bfc990cbe4de329186d7401

Observation e828c486-00f1-42da-82cd-bf068d589372 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:36.204405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-07T12:35:31.063802Z digest=sha256:7d89617aea72c319af538462a44d980ca773ca530dc92d425d498d0c04264335

Observation c0a46116-05ce-42c4-bc55-125d8ee42754 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.132886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.132886Z digest=sha256:55b603b65081d71a51214909f1f9771c072ed66b5cf98ea806e743932322ce10

Observation f450c3a2-f6b5-401a-9db7-6da32c58d97d · outbound

This paper cites VITA: Towards Open-Source Interactive Omni Multimodal LLM.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation VITA: Towards Open-Source Interactive Omni Multimodal LLM

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.216234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.216234Z digest=sha256:872d76fa30a750f5f2ca2b9d2e3a2a67405d3649bde76e171bed7c9dfdd47656

Observation 4b08dc92-9a58-40c5-949b-d6060879c27e · outbound

This paper cites VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.300260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.300260Z digest=sha256:df463842315a064ceb61ef2bff7f9009ac61b586a518cbda92fd0536a19ea86f

Observation f31ac3a0-02af-4b1c-aa28-ff6583550664 · outbound

This paper cites MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.362263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.362263Z digest=sha256:f542ed98a9e5760a54447e56bf53dfb1cc87a2a7b91512869d6cfddbef74052d

Observation c2de8741-0ef5-443b-aa22-0d47aaaafa92 · outbound

This paper cites MME-Finance: A Multimodal Finance Benchmark for Expert-level Understanding and Reasoning.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MME-Finance: A Multimodal Finance Benchmark for Expert-level Understanding and Reasoning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.447032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.447032Z digest=sha256:c51245e59e929c240c16d52816e795d775be6b0e604c5b789008765457d97a27

Observation 018c1fa4-a5f3-4832-9495-398357670cb4 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.544870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.544870Z digest=sha256:9cb5e2140c2d040b27494fd5ddf442598004641aac0035e1ac81da2050aa3050

Observation 2ae10eca-e090-40ea-a05c-866acca43fcc · outbound

This paper cites MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.621908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.621908Z digest=sha256:8a903788b7126d5aaf5aae4489ca8979315b8773d9e5845808de1cc3a6ed6dac

Observation 43819a59-7a02-4009-b212-5a7a357ef73b · outbound

This paper cites MMEvalPro: Calibrating Multimodal Benchmarks Towards Trustworthy and Efficient Evaluation.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MMEvalPro: Calibrating Multimodal Benchmarks Towards Trustworthy and Efficient Evaluation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.673244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.673244Z digest=sha256:e7fa2a0828db4dc0d82bebc7166c6388a422b0d933e39fbc8d06fbf43b8c0d65

Observation f7924ab8-8164-475b-9820-5429feaa179a · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.753764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.753764Z digest=sha256:76bf222a49399542dbc0c5ed052d62dbbf23550cd302b22dcbb75e1e7f9cb792

Observation 7b19a732-f152-4a6b-91da-508b6892bccd · outbound

This paper cites FinanceBench: A New Benchmark for Financial Question Answering.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation FinanceBench: A New Benchmark for Financial Question Answering

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:31.894547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:31.894547Z digest=sha256:9e15abf57ec606c4db74a1b9b243159df53141870e0666cabb1760c83fa01e5e

Observation d43e62f6-bfd3-4309-a92f-346af62da9ac · outbound

This paper cites CFBenchmark: Chinese Financial Assistant Benchmark for Large Language Model.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation CFBenchmark: Chinese Financial Assistant Benchmark for Large Language Model

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.015781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.015781Z digest=sha256:5c345cf2908c72420710fd2ab03cd015c8682875f44ad470ff14528c59f03304

Observation a6a1a634-b303-4b6e-b6c7-e8b03cf126b3 · outbound

This paper cites SEED-Bench-2-Plus: Benchmarking Multimodal Large Language Models with Text-Rich Visual Comprehension.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation SEED-Bench-2-Plus: Benchmarking Multimodal Large Language Models with Text-Rich Visual Comprehension

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.130801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.130801Z digest=sha256:5af5e74e50cb8e43a6e275861235ac440a92521bc3e804c4a22db2520668e994

Observation 7ce7526c-13bd-4e7c-9562-7336fb5a67de · outbound

This paper cites SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.258980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.258980Z digest=sha256:dfe363a3570b25ee5697bfef5f3c8f04402a12ef11ee44b6857b60068d220761

Observation ce077ca2-8eea-4988-a970-464b19c96896 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:35.831999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-07T12:35:32.331654Z digest=sha256:a74fcee2cff49f3be13b706ad2f3657dd9707f0cdd32e611404d6eb45933cdc7

Observation 6230c13c-89c2-4300-ba64-c992c82b2b83 · outbound

This paper cites VisualBERT: A Simple and Performant Baseline for Vision and Language.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation VisualBERT: A Simple and Performant Baseline for Vision and Language

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.435193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.435193Z digest=sha256:411e379aef08ea3e71c76cce626deba8346ff5497ee5c29a35388473def7727a

Observation 02595493-e29f-42e0-b415-cde8ecc35ec0 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.531062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.531062Z digest=sha256:51535ed33afa62390d39fbfcb6415bdecef8329b6855064663393d5911eec56b

Observation fc712e4a-7263-4d74-8a77-eb2b39696964 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.627555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.627555Z digest=sha256:3a4bb2382de68e67cf09705b13b161952b2388cb68b3fd912647c35f7f25c8c4

Observation 351ae998-d17c-4a67-a52e-c13e5fa00d65 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:35.416557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-07T12:35:32.733322Z digest=sha256:22029fb8ae894fb1ec0ad5e42e2d8f7e3894bae2e942f8783b6cd45d53b47003

Observation b77ffe11-d12e-4b0c-bf52-1ce72333bc35 · outbound

This paper cites MMC: Advancing Multimodal Chart Understanding with Large-scale Instruction Tuning.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MMC: Advancing Multimodal Chart Understanding with Large-scale Instruction Tuning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.804493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.804493Z digest=sha256:bd75e47041e518a3d2f82a3121c023d091870091bc82d7c56f07905bc033b1d2

Observation 765647e4-25e9-4de1-829d-3e45ce599fdf · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:35.082683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-07T12:35:32.878031Z digest=sha256:a27107bbb29fd57f4c6d7d0e9bad70dbeceb749a76a336175654c206464442de

Observation 1eaefa05-4451-47a2-a956-17b1f11dbf87 · outbound

This paper cites Large Language Model Agent: A Survey on Methodology, Applications and Challenges.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Large Language Model Agent: A Survey on Methodology, Applications and Challenges

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:32.951314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:32.951314Z digest=sha256:545c0d575b0657c24653ca87c8f580d5a3453a06ce70072d1dbc7434d7da76aa

Observation c0c17bf7-1135-4f0d-80d9-d471607fb763 · outbound

This paper cites ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.032851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.032851Z digest=sha256:6ee361a46cd6a185cd26337ed8bab48436701cedb14718009fa5caf6ba9f85e7

Observation a649a890-a2ae-41f9-949f-fc4452089ea5 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.100252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.100252Z digest=sha256:b5633b73ed65201032725e6e3acb69d896af8b514f79c1afa6b01ab3f65a5b5a

Observation cdd1da3a-cfd2-4f28-87f5-2a14b13734e3 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:34.794795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-07T12:35:33.188206Z digest=sha256:a99753ee7c305b441a48cb6e6cfaf34cf61c04fe5556cf3f1e281ed92fed3b6f

Observation 8b4dde1d-9671-4586-9790-51849ae05df1 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:34.601957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-07T12:35:33.295608Z digest=sha256:6272bb192a2dee9d6596762c66198923dbf0e6e962fd9a01521cc2d340f90e99

Observation 73083d0d-814f-4b9a-8f00-a97cb649ec96 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Gemini: A Family of Highly Capable Multimodal Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.363156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.363156Z digest=sha256:50b01595f885c56297a0ebc2a24da6370e6056c29e551507c89b8a74df5ab30e

Observation d24b0338-a1f2-4828-aa74-2e83e0c92aaa · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.424499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.424499Z digest=sha256:d98b52d5b2442b950b213544daa460b6f03185f9211a5814d911b5340dd8a7c6

Observation 9338ce06-acb8-4c3a-bb8c-256df082e970 · outbound

This paper cites CogVLM: Visual Expert for Pretrained Language Models.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation CogVLM: Visual Expert for Pretrained Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.482668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.482668Z digest=sha256:9197758af54540a316a82677bf0c717c8210607653c8d113b5997c44e3f98782

Observation bb422c51-2059-4a9a-a4db-d8f2096415f7 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.549590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.549590Z digest=sha256:08bf91007cf036edbcd09be755f4cfd5a9ad3d145731005c74b8b40ddfb03311

Observation 80fd2abe-bd96-4da6-85ee-8b42611c8015 · outbound

This paper cites FinBen: A Holistic Financial Benchmark for Large Language Models.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation FinBen: A Holistic Financial Benchmark for Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.646547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.646547Z digest=sha256:eb39543991673f60ed074ae983e2224206a078d071be0cac5220a4d210335d69

Observation aef175a3-a5cb-4f3d-92c4-6abd39b60f83 · outbound

This paper cites Qwen2.5 Technical Report.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Qwen2.5 Technical Report

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.745859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.745859Z digest=sha256:6aba9dd3be5e5c9f9cae7fcc48d10727dc74ca71fd31ab377b30148b4b4da3a4

Observation 801eec10-b278-4453-aa22-7d28771b1e1c · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:35:34.535561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-07T12:35:33.819813Z digest=sha256:85e2651cba3fc2db76df8f38f165b901621f558abc0bd15d29fd068ff0bf04e4

Observation c696ed45-e664-4492-a530-0ce48ece7a2e · outbound

This paper cites MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.911928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.911928Z digest=sha256:4cc611c67e3873665ada19c192f30b6f4285f9a8160f6196940f8989971592ca

Observation ae8b3625-fc31-4879-9886-38472b69f91b · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:33.972789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:33.972789Z digest=sha256:31e5946cedd2f7f7d17b64b8f5c15b8473aee333a891003a749c836c50c02442

Observation 46802361-d266-44d6-b044-e80cf70fa281 · outbound

This paper cites an unresolved cited work.

FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:34.053651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:34.053651Z digest=sha256:3bffd39f7696efe710f14f463aecdd49467849e0ec0d54addafb86a6890a0d4c

Pith citing papers

Observation f432e913-cf41-4b9d-acbe-fd43187802aa · inbound

FinCriticalED: A Visual Benchmark for Financial Fact-Level OCR cites this paper.

FinCriticalED: A Visual Benchmark for Financial Fact-Level OCR FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:20:11.563240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T20:19:52.701262Z digest=sha256:628121cf5c172b25f492724eea676a768b4b069895a6a047c819db7c2292574f

Observation 67bbb98f-4303-4c9a-9bf4-b8765afe70a7 · inbound

Strat-LLM: Stratified Strategy Alignment for LLM-based Stock Trading with Real-time Multi-Source Signals cites this paper.

Strat-LLM: Stratified Strategy Alignment for LLM-based Stock Trading with Real-time Multi-Source Signals FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:56:08.512517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-08T10:42:08.721879Z digest=sha256:60c32d3a841727abcbb950675466f8771075906e989607f55aa1ac3146bd72f5

Observation 9629dd8f-8fb5-47aa-bb10-a92f24c2ff51 · inbound

FinDocMRE: A Benchmark for Document-Level Financial Multimodal Reasoning Evaluation cites this paper.

FinDocMRE: A Benchmark for Document-Level Financial Multimodal Reasoning Evaluation FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-20T00:32:54.664118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-20T00:28:48.871454Z digest=sha256:297c88eabc9cf44be7c940c0c37f20d569884e76b64fa88c7c44e343c2a0a425

Observation 909eae90-6689-45d6-94bf-624b8d8cd79b · inbound

Are the Financial Reasoning from LLMs Credible? A Real World Test over Long-Horizon Statements cites this paper.

Are the Financial Reasoning from LLMs Credible? A Real World Test over Long-Horizon Statements FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T00:48:03.603561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:48:03.603561Z digest=sha256:69f7098864ca37a5526a7476e1d07367e64ae950d7489957dc8247068d427398

Observation 95dc9f97-1f97-4eb9-948b-acd3850c3991 · inbound

FinReportBench: Measuring and Improving Institution-Grade Financial Report Generation cites this paper.

FinReportBench: Measuring and Improving Institution-Grade Financial Report Generation FinMME: Benchmark Dataset for Financial Multi-Modal Reasoning Evaluation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T18:53:01.866207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T18:53:01.866207Z digest=sha256:f1685a93a8c2a7e7e2e4297015a6a0d6189833682643d82599fe306d8b916caa