Pith. sign in

Paper Citation Record · LEDGER

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism

As of 19 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 1 inbound Pith citation observation for arXiv:2506.08691.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08691 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:09:24.020605Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T19:36:57.231932Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T21:37:25.289403Z

Reference resolution

45 of 45 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved45
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 23ace000-230b-4d70-a2f6-a22c80d74a04 · outbound

This paper cites M$^3$CoT: A Novel Benchmark for Multi-Domain Multi-step Multi-modal Chain-of-Thought.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism M$^3$CoT: A Novel Benchmark for Multi-Domain Multi-step Multi-modal Chain-of-Thought

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:18.196208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:18.196208Z digest=sha256:ce2f29b9a236b603451a0808619eb0c5671b404d59c086df907b4f73c2e73f20

Observation 7c80bb66-02a8-4420-a3c7-08d9ebd843db · outbound

This paper cites Vision-Language Models Can Self-Improve Reasoning via Reflection.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Vision-Language Models Can Self-Improve Reasoning via Reflection

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:18.339751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:18.339751Z digest=sha256:088914de48a277919665e49fea9d1b7d3cb163f4b458aec4191f39d12b73c52c

Observation a88b7e8b-8b0b-4109-b718-e29bf56f8044 · outbound

This paper cites an unresolved cited work.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:09:26.869391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:09:18.463000Z digest=sha256:d9e0bb23dca77aa032878a359a418bbb084f5792f46322394ddb5d8eb83fb94f

Observation 4234ff1d-a7b8-47fb-b9ec-4092ced0f56f · outbound

This paper cites Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:18.616592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:18.616592Z digest=sha256:fac34389807310850d4eb91345f98d6b8fdec8a0de0e74e97234efc9fe8208aa

Observation 1ff5ce33-eb00-477a-9bfe-dedaac308c02 · outbound

This paper cites an unresolved cited work.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:09:26.688387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:09:18.772129Z digest=sha256:b676fb904a93590f03b3498ad0dd7dded71c2f8a01663f757daaf1d0ca62901b

Observation 545bdbd5-21ae-42ef-983b-af63ca588c00 · outbound

This paper cites MAmmoTH-VL: Eliciting Multimodal Reasoning with Instruction Tuning at Scale.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism MAmmoTH-VL: Eliciting Multimodal Reasoning with Instruction Tuning at Scale

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:18.908261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:18.908261Z digest=sha256:e3aa3aecd4962b44e3f72893e03a7b87c2885b7319838f42ceb90528557bd49a

Observation 6b82c278-4165-4991-bafd-c1790c45ff77 · outbound

This paper cites Reasoning with Language Model is Planning with World Model.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Reasoning with Language Model is Planning with World Model

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:18.989445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:18.989445Z digest=sha256:d3efc0a65de96fa36cb5ee5a29486fd24083ea98e281e70af8f429e89d82ff59

Observation d9b9faf8-5635-468d-ad9f-915166704426 · outbound

This paper cites an unresolved cited work.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:09:26.473220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:09:19.196937Z digest=sha256:7a637d6eaec9da0f8c05eec7e0f5925ca0e1d6561e7d8cf3928bd056bff5f53e

Observation 22e7a611-396f-4de5-9f91-6664fcbf3176 · outbound

This paper cites an unresolved cited work.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:19.387591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:19.387591Z digest=sha256:4337ebb8ca9d4242d26e1638719cc950631a050741c7b0bffb1ff5305412530a

Observation a4cc418c-f9fe-4f52-9f47-9608d2518715 · outbound

This paper cites Enhancing LLM Reasoning with Reward-guided Tree Search.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:19.582648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:19.582648Z digest=sha256:bb883630156d6b45501f40ab77cbda731b93c8b716fe0c106bf91738ea6c97f7

Observation cf9a2bda-4e86-44c3-87b2-93b8c94e6932 · outbound

This paper cites an unresolved cited work.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:19.753354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:19.753354Z digest=sha256:d2061d54c1544dc9f0f310f6d2a1acb71316cbfdc7f8d28fdd4c8bf94275cc66

Observation e48b2f53-69e0-4b15-b07c-c0d92a0647c5 · outbound

This paper cites an unresolved cited work.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:19.869927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:19.869927Z digest=sha256:4f91bc3d42fd5415a77a29dc2036b69f3e6d1c327c6c139054ee2ee5f9f136d1

Observation 0838fab1-c0a1-4f3b-bca0-ad1570aa6a54 · outbound

This paper cites an unresolved cited work.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:09:26.276145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:09:20.016775Z digest=sha256:be1d1c5b6eebda5af6aae7f2230869864535eeab3b872a9e18fcf416bf93c1c0

Observation 703522a3-f8db-4ca3-bffe-b10263c8ba16 · outbound

This paper cites Let's Verify Step by Step.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Let's Verify Step by Step

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:20.155515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:20.155515Z digest=sha256:b8d6aa9db39cbe198ed4c0fd24aa7febeb5b565f76372254f77a6e4dbba678c3

Observation 53f62501-e248-4ba2-b045-4974a4ac66bb · outbound

This paper cites ChartThinker: A Contextual Chain-of-Thought Approach to Optimized Chart Summarization.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism ChartThinker: A Contextual Chain-of-Thought Approach to Optimized Chart Summarization

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:20.285362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:20.285362Z digest=sha256:64e1495dcee003436bd7240b9fc1aa7ecad3018a93ae56fdab07cd81a6c944bf

Observation 3053ebc7-247d-47bd-a3ba-0dca31b5db42 · outbound

This paper cites Large Language Model Guided Tree-of-Thought.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Large Language Model Guided Tree-of-Thought

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:20.396124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:20.396124Z digest=sha256:e42ba5df86de69d08869354b060e07867eaee2e3147cef85481456cd9555c2b5

Observation 2343b7c5-30cf-4720-be51-62f1247f0507 · outbound

This paper cites MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:20.523908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:20.523908Z digest=sha256:0cf0e85d6d371dbb75bae6b7677305115cad3bb0b280138f132c0f6e3ea156ac

Observation cf928d51-7b97-4148-b96b-7546f86bebab · outbound

This paper cites an unresolved cited work.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:20.636332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:20.636332Z digest=sha256:35046bb475359fa66ab499675ad49fc0eafed745b4c5e55fb084e77169b34065

Observation 02824006-03e0-44bd-8de7-a77bf2d4f914 · outbound

This paper cites an unresolved cited work.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:09:26.094013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:09:20.747025Z digest=sha256:1bfa45ff962ed08d7e7bd2f9c037ce33567fb24751d3f9c4c952ffe486cb02ce

Observation 64f3840c-b3e4-4392-9489-7bf1e8cf09d0 · outbound

This paper cites an unresolved cited work.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:09:25.899175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:09:20.887191Z digest=sha256:ff4fe01cbe165c3447fe0d57c0f55d820daf0781a3c8bd6b9adc2c2cbcb82d26

Observation 243c15b1-9f15-4dcf-aac3-c2fadbf53751 · outbound

This paper cites an unresolved cited work.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:09:25.683150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:09:21.057769Z digest=sha256:d6f1e594f1b01d490b5752e7ec6bc240fd5450d6ab2c8987c471d28869eb04f4

Observation 1a6999fa-faf2-47bb-a125-bdb61b8312a0 · outbound

This paper cites Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:21.174185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:21.174185Z digest=sha256:0d347dfb29ee6a1e4c0bb8b5a3cd79ef5f1c3f2e07a541699d3f30bd8e3c9d6c

Observation 51686c0b-61a3-46c9-975f-c39cbefbb205 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:21.380191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:21.380191Z digest=sha256:d09e9bd08247cd280bffad286fae7c64e8158dec49bc1ceab3d2332216a80717

Observation d22d9979-53ec-47a4-af77-8cb356a00a67 · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:21.545744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:21.545744Z digest=sha256:af57a50964d8027dcccc2809849647fc49c974e87f0e105e8c47300d730f63e2

Observation 59fbe118-6147-4a95-9996-0156d9c7c597 · outbound

This paper cites Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:21.647142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:21.647142Z digest=sha256:e032540aadce20e1576fd4e5bd9085a66ecd56dfd11e151adce2669e330256e1

Observation bc132b93-c994-4b01-8c8a-2b858c3fdb7f · outbound

This paper cites CharXiv: Charting Gaps in Realistic Chart Understanding in Multimodal LLMs.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism CharXiv: Charting Gaps in Realistic Chart Understanding in Multimodal LLMs

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:21.813886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:21.813886Z digest=sha256:808b9ec62669423503dd14061d9129993dbb10cc22238ff553210dd737102667

Observation ecb2c3a1-578c-4746-853a-24f3db9e00e9 · outbound

This paper cites an unresolved cited work.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:22.026151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:22.026151Z digest=sha256:d282454ab0e618c54bfba4299536f709a2a37f1c13c6935734ac58807ab375e1

Observation a92ce3da-0fea-4661-b610-a2ca319d9984 · outbound

This paper cites ChartInsights: Evaluating Multimodal Large Language Models for Low-Level Chart Question Answering.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism ChartInsights: Evaluating Multimodal Large Language Models for Low-Level Chart Question Answering

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:22.134598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:22.134598Z digest=sha256:f930d940d61cd540ede22f872fd19f5f293b196f18472fee9470a7cee6ab7165

Observation 7e1acb8e-75d2-4094-aa64-372b6ae565b4 · outbound

This paper cites Number it: Temporal Grounding Videos like Flipping Manga.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Number it: Temporal Grounding Videos like Flipping Manga

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:22.277066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:22.277066Z digest=sha256:12d3338e925bca7a889a2daa745b57db8049097e07e1f088d28edaa8f12b931f

Observation caded182-4d53-4ea1-bd7c-7174b021f941 · outbound

This paper cites an unresolved cited work.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:09:25.508723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:09:22.476292Z digest=sha256:4495afbafea8ef8be07e76b158763fc87863c198b8ff47ddda1b64f08fccc8be

Observation 0ee7b07b-56ff-4d14-bb91-6822defa8d9e · outbound

This paper cites ChartBench: A Benchmark for Complex Visual Reasoning in Charts.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism ChartBench: A Benchmark for Complex Visual Reasoning in Charts

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:22.590714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:22.590714Z digest=sha256:4df60e6845805a9e3c682053185399182e19e880e00d590fa66b42f06cb0cd42

Observation 9fe46db7-7db3-4442-8aa6-ac1809a3e171 · outbound

This paper cites Qwen2 Technical Report.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Qwen2 Technical Report

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:22.792599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:22.792599Z digest=sha256:b2b5fc8aa254a79cdaf69f6b95de40e0f4232b309867a9abf8a876f6b267b16c

Observation 1aabe9c5-9116-472c-81e7-422112b144eb · outbound

This paper cites an unresolved cited work.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:09:25.334215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:09:22.870253Z digest=sha256:fa1e88d9e0e746b7173a22f8a2ae1a8270b5bf35d3b86e7dd26ad448b105d0eb

Observation 2a01d426-545e-4ba5-b37a-309808586dc8 · outbound

This paper cites an unresolved cited work.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:22.962916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:22.962916Z digest=sha256:0ee032635b9545cfe5ce7af369e728843ac4944992f8593f876438b7e066f76f

Observation adadffbe-0eba-41be-8c18-fa335bfa63c7 · outbound

This paper cites an unresolved cited work.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:23.042401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:23.042401Z digest=sha256:aad6042aa6b200b225b4847b723b82290c940907d51be879f302c8b9d8dd9f9d

Observation b712d1f3-ae63-4f03-beab-9d97bf352f50 · outbound

This paper cites an unresolved cited work.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:09:25.128428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:09:23.127849Z digest=sha256:bc18e4584c80be2004811115e174f8bb331a62722ef523326bd785905e45435a

Observation f4a3a9e7-5c69-4fc8-bffa-56f9115de103 · outbound

This paper cites ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:23.240182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:23.240182Z digest=sha256:58c41802c15f1a923d3725f1b27516f79ef634b347cc6c4069021ce335c4e2ce

Observation f87f5801-3e66-45b7-a293-8186d484b3df · outbound

This paper cites LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:23.349125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:23.349125Z digest=sha256:a2b5952ed58ed95c13ca47d47f37a874e5e2ebf1a3d91f6a110225ed575d2934

Observation 4c15f74a-18aa-48b8-a829-3e63b4303cae · outbound

This paper cites an unresolved cited work.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:09:24.940863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:09:23.429504Z digest=sha256:a5fcd7ecff1d2d7d1975042c11228b1ec9f3a2faa7f5058c53f019dbb6ba5b6b

Observation ba1fb63a-f1e6-4641-aa47-f9c3ad038dcd · outbound

This paper cites Automatic Chain of Thought Prompting in Large Language Models.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Automatic Chain of Thought Prompting in Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:23.513166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:23.513166Z digest=sha256:c3d691caffb7a7cbe9b00c28a9d50b5e79aae313c201fa8957705341d3353cd1

Observation 133a0312-fd14-48e7-a0bb-71019c467bc8 · outbound

This paper cites Multimodal Chain-of-Thought Reasoning in Language Models.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Multimodal Chain-of-Thought Reasoning in Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:23.620651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:23.620651Z digest=sha256:bae91782f419886ac8d9284b8444ef261eb552a3c4e85a7c8f7a394c6a799d5b

Observation 00d1e608-90c7-4f33-8808-8973a85ceb53 · outbound

This paper cites Benchmarking Multi-Image Understanding in Vision and Language Models: Perception, Knowledge, Reasoning, and Multi-Hop Reasoning.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Benchmarking Multi-Image Understanding in Vision and Language Models: Perception, Knowledge, Reasoning, and Multi-Hop Reasoning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:23.728013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:23.728013Z digest=sha256:90a099ce7744f8499341e7a07d61526db029a3f9fa5a944e7e6cc83d0937bd70

Observation 4c852993-e185-4769-a192-ca990e3e4728 · outbound

This paper cites an unresolved cited work.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:09:24.769823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:09:23.812478Z digest=sha256:ec2036ee1a848908214db3560423cd7c2a2cf9e3f10bd3d04a96e63b5c55db5b

Observation b55daa42-dcb9-4598-b6f7-af149e0545d6 · outbound

This paper cites online" 'onlinestring :=.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism online" 'onlinestring :=

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:23.911059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:23.911059Z digest=sha256:44c8aaa3be7b8eb5f82a2ae157148fce5533a926476631cbf9b6d2ec4b5b68bd

Observation d416c5df-3027-4e69-a81a-896a28d1c2ca · outbound

This paper cites write newline.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism write newline

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:24.020605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:24.020605Z digest=sha256:5fc35a742c23aa75b6d5be8eeab8d1e8995c09004f9be939b16d387c8e8faa1e

Pith citing papers

Observation cc0b75d9-5201-44ca-bc33-ab92b9cfcfb7 · inbound

Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning cites this paper.

Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism

Reference 109

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:37:25.290672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T19:36:57.231932Z digest=sha256:8b8bfa18559e2cf83cb42c9984f98dc90ca7016fbcd0cb746439cda8ba6165a2