Pith. sign in

Paper Citation Record · LEDGER

MINERVA: Evaluating Complex Video Reasoning

As of 18 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 8 inbound Pith citation observations for arXiv:2505.00681.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.00681 v1

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:40:19.235637Z

measured 73 of 73 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:41:39.724979Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T19:07:17.504041Z

Reference resolution

65 of 65 outbound references displayed

  • verified exact0
  • verified fuzzy15
  • unresolved47
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c2a7edd7-abbb-4144-808b-8d55946811ca · outbound

This paper cites GPT-4 Technical Report.

MINERVA: Evaluating Complex Video Reasoning GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:18.995236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:18.995236Z digest=sha256:053debc7b23ef7cd4f9e96aff8088ac3faf8242210aace50ebea6a60e2f85fb3

Observation d6e60dba-a1b7-4c1e-a808-51b285e868d9 · outbound

This paper cites https://openai.com/index/gpt-4-1/,.

MINERVA: Evaluating Complex Video Reasoning https://openai.com/index/gpt-4-1/,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:40:20.040680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:18.999633Z digest=sha256:7526ff806af64d4958f36a95a7e290d9ef212c1f63545cc2db117d1f8f367f51

Observation f0428ccc-c09a-4032-a55e-7ff3e854e549 · outbound

This paper cites https://openai.com/index/learning- to-reason-with-llms , 2025.

MINERVA: Evaluating Complex Video Reasoning https://openai.com/index/learning- to-reason-with-llms , 2025

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:40:20.018651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.007499Z digest=sha256:4130d27c84d5be171a44272b80a515a920f5bfabf8a38c6e16bbaf5a906766d2

Observation f56051d5-86aa-41b2-a7d5-97265eaa8f3b · outbound

This paper cites Claude 3.5 sonnet v2.

MINERVA: Evaluating Complex Video Reasoning Claude 3.5 sonnet v2

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:40:20.007091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.011112Z digest=sha256:20dfdf2bbd18a2885d45d1c2aa214eb0f3997f4fa527c2fd34b9f28ecf7102e5

Observation a80824e8-ad65-4d11-bbf4-d5fec0ebc769 · outbound

This paper cites In- finiBench: A comprehensive benchmark for large multimodal models in very long video understanding.arXiv preprint arXiv:2406.19875, 2024.

MINERVA: Evaluating Complex Video Reasoning In- finiBench: A comprehensive benchmark for large multimodal models in very long video understanding.arXiv preprint arXiv:2406.19875, 2024

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.015073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.015073Z digest=sha256:93d3695c1fe805840e9afa6f94096ad060140ecdf9a9e317e960007555e61dbd

Observation a9b9895b-35d4-4c58-b0b6-8548d926f46f · outbound

This paper cites Qwen2.5-VL Technical Report.

MINERVA: Evaluating Complex Video Reasoning Qwen2.5-VL Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.019103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.019103Z digest=sha256:8369a452851a6ca6613ba31b4e7fd65d8c9c66805296369be3a34b5eacf043d9

Observation a2544e8b-e3ca-4e4f-904a-808be547e498 · outbound

This paper cites TemporalBench: Benchmarking Fine-grained Temporal Understanding for Multimodal Video Models.

MINERVA: Evaluating Complex Video Reasoning TemporalBench: Benchmarking Fine-grained Temporal Understanding for Multimodal Video Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.023361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.023361Z digest=sha256:2f20420057d9d7375ccaed6aec6a1bbaedbdd42c2fedd417acc412faae46b5ab

Observation 8ff75016-971d-46fa-80c0-bbbc4f9c8912 · outbound

This paper cites HourVideo: 1-Hour Video-Language Understanding.

MINERVA: Evaluating Complex Video Reasoning HourVideo: 1-Hour Video-Language Understanding

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.028130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.028130Z digest=sha256:de3ab870ef228bb764aa80d4dfd0de18d25a720d9d0237e3a41fabf6e7cbf7ae

Observation eb175d54-5dac-45a1-9f43-0c970240ec07 · outbound

This paper cites Mllm-as-a-judge: Assessing multimodal llm- as-a-judge with vision-language benchmark.

MINERVA: Evaluating Complex Video Reasoning Mllm-as-a-judge: Assessing multimodal llm- as-a-judge with vision-language benchmark

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:40:19.995524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.032589Z digest=sha256:c4daa75708399ebe60a66a7d7afc1d8a9b97bd5d9f4845e057148c30f8785635

Observation 03c3afe2-d8b3-4a0d-a9c5-04e5758ce265 · outbound

This paper cites Cg-bench: Clue-grounded question answering benchmark for long video understanding.

MINERVA: Evaluating Complex Video Reasoning Cg-bench: Clue-grounded question answering benchmark for long video understanding

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:40:19.984836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.036738Z digest=sha256:6bd881483afeed78e0f9fecb808c3175914f6dfce2ac18673638556f6ce2753a

Observation 8a4a9d7e-edc7-43bb-92fd-b818933fd4d9 · outbound

This paper cites Lost in Time: A New Temporal Benchmark for VideoLLMs.

MINERVA: Evaluating Complex Video Reasoning Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.041344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.041344Z digest=sha256:706ba6c4939d6782f9b661dd9cdf533e6ea508d5e2c948b47f694cb85039211f

Observation 059487b9-5d39-4bf1-a5b3-cf632dbff861 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

MINERVA: Evaluating Complex Video Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.045631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.045631Z digest=sha256:d2932b2e537ebb43ab41a35f0e2f4698c49172940f81894b76f73467ca52c340

Observation 2876843f-257c-44d1-b456-691a0b01e14c · outbound

This paper cites On the Limitations of Reference-Free Evaluations of Generated Text.

MINERVA: Evaluating Complex Video Reasoning On the Limitations of Reference-Free Evaluations of Generated Text

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.050067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.050067Z digest=sha256:620d6de05ed6d15186bb14c31fb0ed6329f707818f8852d44c70b388508ebf52

Observation 6e4e7687-4e6f-46c0-abad-da28cc3ae30d · outbound

This paper cites Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition.

MINERVA: Evaluating Complex Video Reasoning Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.054393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.054393Z digest=sha256:eb6b775d095a42a09f42c7351e9691807c2250342dbcb691f0458f1c19be2867

Observation 20ddc919-ba6f-4ae3-b22f-9921ba5164ed · outbound

This paper cites Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis.

MINERVA: Evaluating Complex Video Reasoning Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.058011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.058011Z digest=sha256:e98f12a9528f0b2cb2e56dee1ef69375b7845cec76b9b376667a5bcf0f2d6fe7

Observation 60ac4818-1537-4650-b108-8cbb6df33a99 · outbound

This paper cites ROSCOE: A Suite of Metrics for Scoring Step-by-Step Reasoning.

MINERVA: Evaluating Complex Video Reasoning ROSCOE: A Suite of Metrics for Scoring Step-by-Step Reasoning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.061742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.061742Z digest=sha256:77dcdb7b7d74ef3f498257cf58d3dbdcbfaf6ea5b6ef4bc9471c7562cb62baca

Observation 8654661c-ebd7-4a0f-a1e0-fbfb98c166cc · outbound

This paper cites something something.

MINERVA: Evaluating Complex Video Reasoning something something

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.065568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.065568Z digest=sha256:6fa024f278b0842d54ce83f2fa80eab533086e4f628933607df274398f8e4ab2

Observation daafd92f-7778-4657-841c-37cdd492c9a2 · outbound

This paper cites VideoEspresso: A Large-Scale Chain-of-Thought Dataset for Fine-Grained Video Reasoning via Core Frame Selection.

MINERVA: Evaluating Complex Video Reasoning VideoEspresso: A Large-Scale Chain-of-Thought Dataset for Fine-Grained Video Reasoning via Core Frame Selection

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.069310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.069310Z digest=sha256:5d998fa25d781cd29867c9fa8e061e64887d1abcb726f4be946a99e70310f33e

Observation a2ac3aeb-bb24-43f5-b8eb-1e400fb35b00 · outbound

This paper cites LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models.

MINERVA: Evaluating Complex Video Reasoning LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.072960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.072960Z digest=sha256:20dda676b91c5dd8473a33b62a7203d4fdb342ca3471bc031b35ce40665dad60

Observation cc53432c-3ca6-49c9-a288-4ba8f8c160da · outbound

This paper cites Sugarcrepe: Fixing hackable benchmarks for vision-language compositionality.Advances in neural information processing systems, 36:31096–31116,.

MINERVA: Evaluating Complex Video Reasoning Sugarcrepe: Fixing hackable benchmarks for vision-language compositionality.Advances in neural information processing systems, 36:31096–31116,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.076871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.076871Z digest=sha256:bec2a15512e1991f3272bc6926cbf9377e3932cd74d1caa305e86e17df2c2d46

Observation b7257963-1245-4fd0-ab17-52cecb2efa6e · outbound

This paper cites Lita: Language instructed temporal-localization assistant.

MINERVA: Evaluating Complex Video Reasoning Lita: Language instructed temporal-localization assistant

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.081090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.081090Z digest=sha256:627a7ad0147856c8d8fb63822b10e76a665c349d2baf9f530d70de03650f8d84

Observation 61f1cfd6-bf88-4d79-a432-5872c68f542d · outbound

This paper cites Large Language Models Cannot Self-Correct Reasoning Yet.

MINERVA: Evaluating Complex Video Reasoning Large Language Models Cannot Self-Correct Reasoning Yet

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.084640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.084640Z digest=sha256:5d411a0ca959c4cdc64daf7b9859acbf0c62e55b5d84912b813ba5874951ef8f

Observation 809c7e9f-14d9-49f5-b5dd-7dbf87fd2c5b · outbound

This paper cites OpenAI o1 System Card.

MINERVA: Evaluating Complex Video Reasoning OpenAI o1 System Card

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.088386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.088386Z digest=sha256:58427d5b7ffc73bbda3c49f5e6878b51c42826a132279a0530554ff93eb93279

Observation 5469f1c0-51ad-4d2d-9759-60ef72c52d75 · outbound

This paper cites Scaling Scaling Laws with Board Games.

MINERVA: Evaluating Complex Video Reasoning Scaling Scaling Laws with Board Games

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.091931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.091931Z digest=sha256:32e874bbe02a5669ff2e2803b054d8547feac262aef9b9ce7727271ab64dd11a

Observation 11c189a4-27c5-4ef1-93e6-4c18dd521163 · outbound

This paper cites When can llms actually correct their own mistakes? a critical survey of self-correction of llms.Transactions of the Association for Computational Linguistics, 12:1417–1440,.

MINERVA: Evaluating Complex Video Reasoning When can llms actually correct their own mistakes? a critical survey of self-correction of llms.Transactions of the Association for Computational Linguistics, 12:1417–1440,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:40:19.955113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.095646Z digest=sha256:01ca9697dfceb350c06526ce2f69af69a3454bce6028e4b970d4cb935bb3c7b2

Observation 6df0ffb0-8bcc-4de6-a961-fa4de3247b00 · outbound

This paper cites The Kinetics Human Action Video Dataset.

MINERVA: Evaluating Complex Video Reasoning The Kinetics Human Action Video Dataset

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.099607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.099607Z digest=sha256:03c012c2898242a201a81acbb82cd2347ac4f0144ee95a4c7f71576bb2a8110e

Observation f221d178-42eb-4797-aa28-353866954554 · outbound

This paper cites Large language models are zero-shot reasoners.Advances in neural information process- ing systems, 35:22199–22213, 2022.

MINERVA: Evaluating Complex Video Reasoning Large language models are zero-shot reasoners.Advances in neural information process- ing systems, 35:22199–22213, 2022

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:40:19.944297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.103825Z digest=sha256:50a4eca14315d16df989d6214d4b757895600720f0eea98a3d449f21d0ae85f7

Observation c859e00f-5cca-4994-acb3-5c26240841b1 · outbound

This paper cites Adversarial filters of dataset biases.

MINERVA: Evaluating Complex Video Reasoning Adversarial filters of dataset biases

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.107247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.107247Z digest=sha256:d14dc78fc59b2fd24bf5dc46126431b4ec990f7a638225f9c86e499c82b393b4

Observation 6e2c68df-91d9-4275-b987-114009d35643 · outbound

This paper cites VideoVista: A Versatile Benchmark for Video Understanding and Reasoning.

MINERVA: Evaluating Complex Video Reasoning VideoVista: A Versatile Benchmark for Video Understanding and Reasoning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.110718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.110718Z digest=sha256:75ddd799f5c11cf9089fb3c4d697b7b2776803c9baa761e307dade3b37f695ca

Observation ea0c572c-831c-4f3b-be4d-16e0902c1e79 · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

MINERVA: Evaluating Complex Video Reasoning Rouge: A package for automatic evaluation of summaries

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:40:19.926451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.114453Z digest=sha256:a6d8faaee64a57ee12109343123f79470127f3c88d11a814e728482d6deca7e2

Observation fd282045-113e-4d7e-96bd-9f7053b719c6 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023.

MINERVA: Evaluating Complex Video Reasoning Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.117897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.117897Z digest=sha256:5637f85d7fbbce611eb8e34a851027eca297c437a5c51a1b5ea2d286c575d7c1

Observation da1ebe5e-3684-46bb-a256-17da29bf03c6 · outbound

This paper cites E.T. Bench: Towards Open-Ended Event-Level Video-Language Understanding.

MINERVA: Evaluating Complex Video Reasoning E.T. Bench: Towards Open-Ended Event-Level Video-Language Understanding

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.121484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.121484Z digest=sha256:dde6423879cb8aa3a87f84031f846661725ed51429f387d8765549fef5e6acaf

Observation ae9dbe84-bc9a-4cb4-bfd0-89cba52e3698 · outbound

This paper cites Beyond Accuracy: Evaluating the Reasoning Behavior of Large Language Models -- A Survey.

MINERVA: Evaluating Complex Video Reasoning Beyond Accuracy: Evaluating the Reasoning Behavior of Large Language Models -- A Survey

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.125322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.125322Z digest=sha256:f11726a260bf99a063e93fa3f31db3c3ba223adb353dcdd0bad681c78ce244e8

Observation 540726ed-300e-482f-989c-2a5a522d5a9c · outbound

This paper cites Neptune: The Long Orbit to Benchmarking Long Video Understanding.

MINERVA: Evaluating Complex Video Reasoning Neptune: The Long Orbit to Benchmarking Long Video Understanding

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.128996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.128996Z digest=sha256:7a5124832fe213db1817195e300e7a424c2ad6014c3f6da168179298f70f0910

Observation ef94c780-8914-4dd0-975a-9485cf8c32f3 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

MINERVA: Evaluating Complex Video Reasoning Bleu: a method for automatic evaluation of machine translation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.132938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.132938Z digest=sha256:e621d5e7dee0bf5d55f0300602366507b76ab723240b69e6cb30021c5a22813f

Observation 20d2029b-c48e-4687-8554-93afe8b4d416 · outbound

This paper cites Per- ception test: A diagnostic benchmark for multimodal video models.Advances in Neural Information Processing Systems, 36, 2024.

MINERVA: Evaluating Complex Video Reasoning Per- ception test: A diagnostic benchmark for multimodal video models.Advances in Neural Information Processing Systems, 36, 2024

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:40:19.901489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.136506Z digest=sha256:d2448a6ab0b0a9e0b14e2340ccd8c168045c459adcd3eb2252abf9500e644fe3

Observation 3024974b-e2c8-4826-ac94-5d257b215237 · outbound

This paper cites ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness.

MINERVA: Evaluating Complex Video Reasoning ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.139964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.139964Z digest=sha256:ff05248113fe6c3ce8b874e548d5c5278bfc86f5dbb0c2e97ae3ed7ff310776d

Observation 18cff72b-07a3-4749-98d3-c88eb1f12408 · outbound

This paper cites CinePile: A Long Video Question Answering Dataset and Benchmark.

MINERVA: Evaluating Complex Video Reasoning CinePile: A Long Video Question Answering Dataset and Benchmark

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.143794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.143794Z digest=sha256:1c1422fd34fece43f19bef17888a1f034e78d6eed0ff5327d9ad582ce12aa0dc

Observation 7c404739-601a-48af-ac09-b0544bca5cd3 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

MINERVA: Evaluating Complex Video Reasoning Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.147600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.147600Z digest=sha256:410280c495e1b90e363f598d3afb946573008cb9d8762d1834c2b3acdaaa40d8

Observation d4d8f8d6-ddeb-4b86-b9db-ab07140005b1 · outbound

This paper cites Scienceqa: A novel resource for question answering on scholarly articles.International Journal on Digital Libraries, 23(3):289–301, 2022.

MINERVA: Evaluating Complex Video Reasoning Scienceqa: A novel resource for question answering on scholarly articles.International Journal on Digital Libraries, 23(3):289–301, 2022

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.151118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.151118Z digest=sha256:ae28d655cb8df8d13ffb21648d62d466613ba78d7f645cf2d15cbb37884d5a9e

Observation 92aed85b-b1b3-4d53-9e84-ce319b244102 · outbound

This paper cites Visual cot: Advancing multi-modal language models with a comprehen- sive dataset and benchmark for chain-of-thought reasoning.

MINERVA: Evaluating Complex Video Reasoning Visual cot: Advancing multi-modal language models with a comprehen- sive dataset and benchmark for chain-of-thought reasoning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.154785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.154785Z digest=sha256:5f6fb0cf4f276d247c0141a7bd75d4df46607f4d3f0b9a68b9d208c575660529

Observation 648a3e6e-ab88-41d8-8795-14a4d724dd8e · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

MINERVA: Evaluating Complex Video Reasoning Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.158225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.158225Z digest=sha256:587a81d2cb81528fc6471ec64edd7d8411533934d7c87efb0eb06119d1cb9716

Observation 86eea503-1204-4b06-b007-61fb207f444f · outbound

This paper cites Gemini 2.5: Our most intelligent ai model.

MINERVA: Evaluating Complex Video Reasoning Gemini 2.5: Our most intelligent ai model

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:40:19.877272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.161989Z digest=sha256:db981c7ad6020b7f8c26860cb6eb80416eaf92629fbb239714cbc325e4fe60bb

Observation c38a9dce-18cc-4943-9de6-4a5d398c5ce1 · outbound

This paper cites LLMs cannot find reasoning errors, but can correct them given the error location.

MINERVA: Evaluating Complex Video Reasoning LLMs cannot find reasoning errors, but can correct them given the error location

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.165361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.165361Z digest=sha256:54e24f596d1eda23e2cd3d73bbfaffa2c92621d2f228305a6f9e6a3f2349172a

Observation c39e7ba0-82e8-48bd-a07b-2201bd12e6ea · outbound

This paper cites LVBench: An Extreme Long Video Understanding Benchmark.

MINERVA: Evaluating Complex Video Reasoning LVBench: An Extreme Long Video Understanding Benchmark

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.168789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.168789Z digest=sha256:606f77c651b6366ac16274fe791da4d4baee43deedf0726de7c4dff9fdb9f258

Observation 5990f448-635a-409d-91a9-73dcaffe18ac · outbound

This paper cites InternVideo: General Video Foundation Models via Generative and Discriminative Learning.

MINERVA: Evaluating Complex Video Reasoning InternVideo: General Video Foundation Models via Generative and Discriminative Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.172186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.172186Z digest=sha256:73370cddb35be6bad6caaf7edd65722db1fb512977980860dc597b167ae3cb87

Observation eaa66bf0-5deb-4057-a273-14e059fdfefb · outbound

This paper cites VideoCoT: A Video Chain-of-Thought Dataset with Active Annotation Tool.

MINERVA: Evaluating Complex Video Reasoning VideoCoT: A Video Chain-of-Thought Dataset with Active Annotation Tool

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.175729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.175729Z digest=sha256:2397a5435458af4bd84f4e1b5ecd3ec6101c9222ba8d8662ce221b6bf337d413

Observation 0c7dcbe5-b8c6-41c8-9121-8d4ac7f5f1e9 · outbound

This paper cites Chain-of- thought prompting elicits reasoning in large language models.

MINERVA: Evaluating Complex Video Reasoning Chain-of- thought prompting elicits reasoning in large language models

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:40:19.867089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.179327Z digest=sha256:174d1386fc3d85273bc0d3f60428b0ced478b72364983681325defc31cb092cf

Observation 15517e25-f80e-46fb-81d1-cbe8f470d3e8 · outbound

This paper cites LLaVA-Critic: Learning to Evaluate Multimodal Models.

MINERVA: Evaluating Complex Video Reasoning LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.182501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.182501Z digest=sha256:236749df9638312292149bcc118783e1a545beb4c9af95454ade72f422ce734e

Observation 7d7d8ba8-6e45-44f7-bd01-8550833123f0 · outbound

This paper cites Activitynet-qa: A dataset for understanding complex web videos via question answering.

MINERVA: Evaluating Complex Video Reasoning Activitynet-qa: A dataset for understanding complex web videos via question answering

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:40:19.856734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.186191Z digest=sha256:33bcb2b54449ecf772cb921be3f434836fc24e78b8f30a2d8a52ede5509107c2

Observation 9bf137ab-b37a-4409-a098-2556b6ead520 · outbound

This paper cites VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding.

MINERVA: Evaluating Complex Video Reasoning VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.189541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.189541Z digest=sha256:13041fcbec76594814cba7d4d84671ae6c54dc4a7f9399ed967588960e722922

Observation aac7b765-c82f-4f39-bb9d-9f141e575d82 · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena.

MINERVA: Evaluating Complex Video Reasoning Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.193114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.193114Z digest=sha256:a7a96f78cb5ded4eb1662ddc139b841abd7e8dd67aa3782621be08f214240ec7

Observation ab0f3609-0d5f-4000-b411-8f76c711b25c · outbound

This paper cites All raters are native English speakers with graduate degrees.

MINERVA: Evaluating Complex Video Reasoning All raters are native English speakers with graduate degrees

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:40:19.836233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.200004Z digest=sha256:35cb1ef44f2ff5ab88bf4916c7529f0d2964351b2899b86dc474f682dc41403c

Observation 7efac0d4-3887-4fca-a888-8784f5bf8d46 · outbound

This paper cites an unresolved cited work.

MINERVA: Evaluating Complex Video Reasoning Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-16T04:40:19.826174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.203344Z digest=sha256:763d88085d5b01dbb4b8c41f9aeabf7ca8f037008f1abe7674adf0b9b0c630f3

Observation b96e2f0d-a34f-4cbd-a2e7-532797e4ade7 · outbound

This paper cites an unresolved cited work.

MINERVA: Evaluating Complex Video Reasoning Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-16T04:40:19.815781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.206756Z digest=sha256:2c2b90773883a2798654ef85397fc04c3540184e82d92fc94f4d0806b3c18e70

Observation e1d021a9-7331-42a1-a441-acf6d205ca25 · outbound

This paper cites an unresolved cited work.

MINERVA: Evaluating Complex Video Reasoning Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-16T04:40:19.804898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.210199Z digest=sha256:c4b0da93a719f4f956a9a4622d57b7ba56302f827528f8c47d6050c533ee6d5c

Observation 76ff8f4c-9be0-4df2-81a4-b4d7227b3300 · outbound

This paper cites Final Answer: (X).

MINERVA: Evaluating Complex Video Reasoning Final Answer: (X)

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:40:19.794736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.214114Z digest=sha256:f12311b9bd182ebf3a2c120eb952947356034707a8ffe682badb3b6b9d41f49b

Observation e99261d7-ebc1-4810-9ea8-3d76059d5e09 · outbound

This paper cites Perceptual correctness.

MINERVA: Evaluating Complex Video Reasoning Perceptual correctness

Reference 60

Resolution
malformed identifier
raw_fallback, observed 2026-08-16T04:40:19.783930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.217722Z digest=sha256:3a027aea9762724a04b6521d62506aef88c2467fc999c86d18dcc578b6308c64

Observation df54c089-3d2e-41b9-babd-c882f9d97928 · outbound

This paper cites an unresolved cited work.

MINERVA: Evaluating Complex Video Reasoning Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-16T04:40:19.772237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.221381Z digest=sha256:8227037e2e69a9dd04094b0b9a0a62ceb2987d4b6506ac43f2c1f03d1686e54d

Observation 7a6283fc-e17d-463d-abfc-c773e116f2e3 · outbound

This paper cites an unresolved cited work.

MINERVA: Evaluating Complex Video Reasoning Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-08-16T04:40:19.761028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.225286Z digest=sha256:665c0ea3cb68125ce7021c01b528c65de15c2b15687225bdb60fee59532be76f

Observation ad56fed8-f637-4e4f-90a0-31bcceacc9e8 · outbound

This paper cites Whoever winsout of you two enters the rumble last.

MINERVA: Evaluating Complex Video Reasoning Whoever winsout of you two enters the rumble last

Reference 63

Resolution
malformed identifier
raw_fallback, observed 2026-08-16T04:40:19.748761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.228533Z digest=sha256:41999b272ceafc0116e5578420531e4bed595d08cefe2fccd55b3177de2815ce

Observation 61959b2f-2864-4433-afcf-d809e19778ad · outbound

This paper cites 5", "8", and.

MINERVA: Evaluating Complex Video Reasoning 5", "8", and

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:40:19.737900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.232041Z digest=sha256:e67bd475cce5e75af6da9419b31452a7128ab27c5e51ef5bbf63a548a47b233b

Observation 23d92538-48b9-48da-b2b9-f782c2ef649f · outbound

This paper cites Afterdiscarding the shots that scored, the video shows ten shotstaken that were saved: 00:16, 00:17, 00:23, 00:25, 00:27,01:01, 01:37, 01:40, 01:45, 02:05.

MINERVA: Evaluating Complex Video Reasoning Afterdiscarding the shots that scored, the video shows ten shotstaken that were saved: 00:16, 00:17, 00:23, 00:25, 00:27,01:01, 01:37, 01:40, 01:45, 02:05

Reference 65

Resolution
malformed identifier
raw_fallback, observed 2026-08-16T04:40:19.726381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.235637Z digest=sha256:4db27debc6eddc81024b9878f9fd858cef34a15fd94976e7b43151fcec2c8585

Observation fe612967-a65f-4b18-9a26-5f2e766811a0 · outbound

This paper cites an unresolved cited work.

MINERVA: Evaluating Complex Video Reasoning Unresolved cited work

Reference 2023

Resolution
unresolved
raw_fallback, observed 2026-08-16T04:40:19.846669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.196749Z digest=sha256:18542cd3ccf9c75d023d433c5c910e08a343af90aea7505e4517d8971ee53e8a

Observation d64c3964-61f8-4b66-901f-955ec99ebaa5 · outbound

This paper cites an unresolved cited work.

MINERVA: Evaluating Complex Video Reasoning Unresolved cited work

Reference 2025

Resolution
unresolved
raw_fallback, observed 2026-08-16T04:40:20.029499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-16T04:40:19.003774Z digest=sha256:70e8da08e2bce5947bfa77f409d9e52a3cc23aad1afe0107bf0005053445d9e8

Pith citing papers

Observation d340bd51-3c6c-4172-b062-c3df7d13eb43 · inbound

ARGUS: Hallucination and Omission Evaluation in Video-LLMs cites this paper.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs MINERVA: Evaluating Complex Video Reasoning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.724979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.724979Z digest=sha256:3f43602394c2895cb838266c6517a5273f6dee7f5f2eb05a9c52ed636ddf5f22

Observation fa2c85b1-5b33-456a-b356-b8be16d1da35 · inbound

Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities cites this paper.

Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities MINERVA: Evaluating Complex Video Reasoning

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-19T05:52:07.692564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-19T05:48:02.828938Z digest=sha256:1e2b7a931bec6c671f20602293924bb0f986d34a4ede15b9a65c438135d4d67e

Observation 2e0c78fd-7bd0-440b-9d12-66746b729a4c · inbound

Seed1.8 Model Card: Towards Generalized Real-World Agency cites this paper.

Seed1.8 Model Card: Towards Generalized Real-World Agency MINERVA: Evaluating Complex Video Reasoning

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:45:14.460537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T07:44:02.827006Z digest=sha256:00eff555cc9b6a1421551a08592455ff1083647a07c257f29f3b8a93bf8e02ce

Observation c23bfe44-a4f2-43da-aea9-732d14197a0d · inbound

Efficient Inference for Large Vision-Language Models: Bottlenecks, Techniques, and Prospects cites this paper.

Efficient Inference for Large Vision-Language Models: Bottlenecks, Techniques, and Prospects MINERVA: Evaluating Complex Video Reasoning

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:45:50.972350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T18:54:04.104227Z digest=sha256:16ed5d446cb89f3bf113c19ed9b53831592480bcc2f64bc87335a06c9a112e8a

Observation d5346c66-48ae-419f-850c-b4dbb559166b · inbound

EasyVideoR1: Easier RL for Video Understanding cites this paper.

EasyVideoR1: Easier RL for Video Understanding MINERVA: Evaluating Complex Video Reasoning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-10T07:47:12.647236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T07:41:27.231098Z digest=sha256:e6830c247ab7ca6037b437b8a69ff082f411d41e82d1598bbb3dfb542c93fadd

Observation cc3db42b-c11c-4ba3-b5cb-899a8aecf110 · inbound

Act2See: Emergent Active Visual Perception for Video Reasoning cites this paper.

Act2See: Emergent Active Visual Perception for Video Reasoning MINERVA: Evaluating Complex Video Reasoning

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-09T05:45:22.878586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-08T19:34:53.683729Z digest=sha256:5ca42bb49f143962069da33baa26224ab7c80643b6c16f7e9e69b385176a898e

Observation 121f92b2-96a2-4daa-84c1-8074191c7aef · inbound

Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding cites this paper.

Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding MINERVA: Evaluating Complex Video Reasoning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-19T16:03:08.086591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-19T16:02:53.887605Z digest=sha256:64b208985ef6a642c8be1fabd16c6673ed3672145befa90c3471ebb49c71289e

Observation d37d8fed-89ad-4aa7-a4a4-c69f78f61b9f · inbound

Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity cites this paper.

Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity MINERVA: Evaluating Complex Video Reasoning

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:07:17.505528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-02T18:57:46.841456Z digest=sha256:f11c273d8bdef1cf6c79be551cf1e849011b998860284b9ca5367402d9b3a774