Pith. sign in

Paper Citation Record · LEDGER

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought

As of 21 August 2026, this Paper Citation Record lists 98 of 98 outbound references and 0 inbound Pith citation observations for arXiv:2507.02984.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.02984 v2

Coverage vector

measured 98 of 98 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:19:42.977063Z

measured 98 of 98 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

98 of 98 outbound references displayed

  • verified exact0
  • verified fuzzy22
  • unresolved75
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f416be21-c02a-4b9b-af2c-fcf74533376d · outbound

This paper cites MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:33.852245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:33.852245Z digest=sha256:e70d3e22522b88249b8ab9997236e5c7e65525149b752019e7ebda0cbfcb0edc

Observation 6b577f96-c747-4113-88ce-7e575130aeed · outbound

This paper cites Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:33.912791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:33.912791Z digest=sha256:1b2a061ca04f38324b3393f840f9711b8cd41519e67c910e8ce3026202cd6df1

Observation 13ec6268-f581-4eb7-a91e-dd40699eb75b · outbound

This paper cites Kosmos-2: Grounding Multimodal Large Language Models to the World.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Kosmos-2: Grounding Multimodal Large Language Models to the World

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.013688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.013688Z digest=sha256:955e2ae9a039e39866716a57a90d1a1c66be5a448354df8b59beed2c19e7dc29

Observation fe18d58a-759a-4aeb-a4ab-e676cca4ece0 · outbound

This paper cites CogVLM2: Visual Language Models for Image and Video Understanding.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought CogVLM2: Visual Language Models for Image and Video Understanding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.077710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.077710Z digest=sha256:88c9a23ce7c463bc08ca0b7965151320631a265be46219d3af3624c33bf59223

Observation 9f277b7f-a5e5-4af2-999b-fbd7e98ae66f · outbound

This paper cites In: International Conference on Machine Learning, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: International Conference on Machine Learning, pp

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.177966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.177966Z digest=sha256:04aa12797f17fb7243f338a684cb520cdbc5af5f7d60da05ba941b4592c94869

Observation 1f93425b-bfc9-47e1-a7d1-4d58c7ce2b8d · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.243662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.243662Z digest=sha256:3c247ab7119cbcb64615fab3deef5c338710a103031ddcd9d029e3f4e6327b75

Observation 9a18fa4d-9cd2-47c2-b3ad-e2f59471e96d · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.353123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.353123Z digest=sha256:f8b077a986a3851b63958d04c2aba70b1781dcd67fd9d19a8bc1f330a9a527e3

Observation 927fd270-58f6-46c2-ade2-ded5b8f8ed54 · outbound

This paper cites SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.422630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.422630Z digest=sha256:261432fac38c3e415dd732c7383b1c6519c86064cc054485521fc090f71bfb7c

Observation 3d524dfb-0e36-4805-ba3e-b1bbf50ed920 · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.505690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.505690Z digest=sha256:d80a63ecaba6213d04478465fe1d138bc72ffe5fbec36712531e4af649357c93

Observation 6c9756ba-a469-4149-b741-3c25008706bd · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.629464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.629464Z digest=sha256:094c777752a796cd8bf6f0500159fc4b898629882c7f1ff0d0f09acc65e02e00

Observation e3ec2494-dc05-4060-ae76-a6cb5f1f5a56 · outbound

This paper cites Yi: Open Foundation Models by 01.AI.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Yi: Open Foundation Models by 01.AI

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.703051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.703051Z digest=sha256:c861319646ceca47902cb0e0b82485225cd52731a22c58a802c034c3e92def39

Observation cc9022a1-7334-4aff-8f1d-9af90499af99 · outbound

This paper cites https://llava-vl.github.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought https://llava-vl.github

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.772600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.772600Z digest=sha256:6408ad9c5428428db8d708e413ca67b3b21ad0b176a543a6a4ddb35b628b19a2

Observation 055f57a2-4f6f-4f21-b0e6-8ab138540039 · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.859505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.859505Z digest=sha256:c8dad5f88db5bc30ad30aa8990c4e7a10dbce849f3b7a118d4d91483508c79f3

Observation 2422e848-f067-473f-bb2f-4a07fd237b05 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.937712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.937712Z digest=sha256:b224f13be7badccdbc8a641865dc557465ef9a5decc6ceeab1493bcfb22fc203

Observation 4708ab48-5e6b-485b-9155-412f188f8d23 · outbound

This paper cites AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.022823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.022823Z digest=sha256:18b61cc1bcfcf64973cf6329203102ae5c8d5d4253381fe2ede076961e61f4d9

Observation 7900ed98-6e53-4bac-8741-2e33725432fe · outbound

This paper cites NeurIPS (2023).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought NeurIPS (2023)

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.110223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.110223Z digest=sha256:29f220dc938c1968f664923dcb0fb597185f406fe7e39ec4f8e7df2295de40b6

Observation b2ad41e5-2682-428e-b7ab-be0596cb1577 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.203253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.203253Z digest=sha256:edc5611d5b001bf978c14bf66778ad892903f32ec91aeabc2070d64d61142af1

Observation e4225a55-a0a2-487c-9e2e-1be11a7eea43 · outbound

This paper cites In: Proceedings of the IEEE International Conference on Computer Vision, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the IEEE International Conference on Computer Vision, pp

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.260301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.260301Z digest=sha256:7df51494734d6427aaad332c04e8cc55897dfa3d43669d366a59444ced60e1a0

Observation 26b4a820-0ff1-4e90-b6c6-3dfd42f199ec · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.355158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.355158Z digest=sha256:e60f324a623f6e211ec78650d026f1a4073c5254e405dc012d3051bd2b5d9ab6

Observation 04819b31-00e6-4373-a1cb-084c8649815d · outbound

This paper cites MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.419588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.419588Z digest=sha256:a810c53f75be0e4202a83b9f2e17864123369dc4d68718ba44cacb3efa67ed66

Observation 89997c7a-e103-4a9d-ad83-bba8518244d1 · outbound

This paper cites In: International Conference on Learning Representations (ICLR) (2024).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: International Conference on Learning Representations (ICLR) (2024)

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.492421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.492421Z digest=sha256:9f228b2dabd77bba95e3713af8bf87aba05cbafe7286cf34dc849b2115fb3cd2

Observation 284e0a8b-399e-456d-b674-0fcf5fb39f21 · outbound

This paper cites ACL (2024).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought ACL (2024)

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.556816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.556816Z digest=sha256:e27d546735d5c2a23f79e47ff72443bdb701dcad22053a0c0eca7bdfca77e2a7

Observation f07419c3-9f53-422c-82e6-d0314e8f14f2 · outbound

This paper cites NeurIPS (2022).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought NeurIPS (2022)

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.650695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.650695Z digest=sha256:d9c4985acbcf4851da2ed5261fcb710c704b298501f7f7ab39466d427a2dce65

Observation d64192ec-5a48-479b-98c3-4c1ae57b6b90 · outbound

This paper cites Multimodal Chain-of-Thought Reasoning in Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Multimodal Chain-of-Thought Reasoning in Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.695869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.695869Z digest=sha256:005cca267ceaa986360374a45bc5154980511805528c680c42f1c36570ff0cd5

Observation 5dd3c81e-b9b3-49bf-8400-c58b711e5f56 · outbound

This paper cites NeurIPS (2023).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought NeurIPS (2023)

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.779999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.779999Z digest=sha256:57d0511cbdb97214d1fc7e2378d9cfb5f2423eac8ce25a410b927a3220e045f4

Observation 089aa579-19a6-41ed-875b-9525f7e724b3 · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence, vol.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the AAAI Conference on Artificial Intelligence, vol

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.846251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.846251Z digest=sha256:3c65fcb15733808d4a75c92c09d7131480a78a746a571f1aa4b6ae9497c0389a

Observation 9fe9cbd2-4458-4d38-98ff-8964cf14f0fc · outbound

This paper cites ACM MM (2024) 16.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought ACM MM (2024) 16

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.912123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.912123Z digest=sha256:cfffe4e68e2a9f4d149142d48eb0acf57d347939d556ecfa4e24a234e53d43f2

Observation 207edfa8-d3a2-4a54-a274-d4ba0697749a · outbound

This paper cites In: CVPR (2024).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: CVPR (2024)

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.992476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.992476Z digest=sha256:3b548f9378e8dcf9fb9cf09c70ab59b69558ff62d28c920155a9b5e56693b908

Observation 518bea90-bf82-478e-9b14-1bafb206c194 · outbound

This paper cites The Role of Chain-of-Thought in Complex Vision-Language Reasoning Task.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought The Role of Chain-of-Thought in Complex Vision-Language Reasoning Task

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.092686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.092686Z digest=sha256:2b719d300280188862c7773ff2ca8bd29547fea36facdba9b7149898c20f6079

Observation 5b546750-8f74-4020-b4b0-484e32544102 · outbound

This paper cites Enhancing Large Vision Language Models with Self-Training on Image Comprehension.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Enhancing Large Vision Language Models with Self-Training on Image Comprehension

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.169281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.169281Z digest=sha256:39b58ecbf8aa4109c67fb4e00821ef0894e1f8a0ef39e107f799486150ad6918

Observation dcc6de35-d0ab-40e1-9116-b0127168f179 · outbound

This paper cites ISR-DPO: Aligning Large Multimodal Models for Videos by Iterative Self-Retrospective DPO.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought ISR-DPO: Aligning Large Multimodal Models for Videos by Iterative Self-Retrospective DPO

Reference 31

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T21:19:43.391954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:36.242869Z digest=sha256:c95b72143778451a6cfbe3c7fd85570d19e8b9a7459d159ce4d1d2d0148f1718

Observation 898c1883-d398-4829-a80a-277d07070c79 · outbound

This paper cites Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.307866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.307866Z digest=sha256:5de9e09b2f0df514d9f1122cc5222dcdfc3adeba61b16ca96a0f8eb49d53c87e

Observation f00da12f-f109-47a1-9460-1bdec9af2d54 · outbound

This paper cites Beyond Human Data: Aligning Multimodal Large Language Models by Iterative Self-Evolution.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Beyond Human Data: Aligning Multimodal Large Language Models by Iterative Self-Evolution

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.373666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.373666Z digest=sha256:ae8e06e753b59bb408bd8b1d4f314c38e0c43d5a725c0498edaa1a277f06e827

Observation c9fb9dea-dadb-44bc-9070-4029abaf468e · outbound

This paper cites In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.451178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.451178Z digest=sha256:8a2dff2aa87f3fe134d2eae9cdd3b42fcf7694b8062809d4e5f8af11a41614b5

Observation 56362f2f-5a59-4b3c-83fc-daa77b6a2a8b · outbound

This paper cites In: Computer Vision–ECCV 2016: 14th European Conference, Amsterdam, The Netherlands, October 11–14, 2016, Proceedings, Part IV 14, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Computer Vision–ECCV 2016: 14th European Conference, Amsterdam, The Netherlands, October 11–14, 2016, Proceedings, Part IV 14, pp

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.528716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.528716Z digest=sha256:3e3f707d32e2dcf84ed8f263f58515572195e67304bb01211d97c13c1576a387

Observation 18401eb3-51b1-4d45-bc32-17a14552b1ca · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.608901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.608901Z digest=sha256:998e54658c6b3ed955afa4bd69af68b26f47a8e379781dcb03ed14274cd4a75e

Observation e0aec378-e836-4353-a392-163c639c9984 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Measuring Mathematical Problem Solving With the MATH Dataset

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.653815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.653815Z digest=sha256:18141d4e0428aa411ee28e761814ac02985f5cddffef2151913a52a987cff145

Observation b1d80a42-65fc-418e-a042-88622768a4dd · outbound

This paper cites In: European Conference on Computer Vision, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: European Conference on Computer Vision, pp

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.720980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.720980Z digest=sha256:7f6d496bd18a488b723139ae1c40dee733aeabe01de2e04e8f179cc04521cc51

Observation 90347a6c-eb5a-4ac5-9b91-3134c0d0cba6 · outbound

This paper cites CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.765868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.765868Z digest=sha256:88650d91d5eb10b6f71ed4d288246dc613d422ca398b7a5bd5245ddecda2e469

Observation ce0619b4-d8bf-4350-9aa7-441eacbc1a92 · outbound

This paper cites Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.818352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.818352Z digest=sha256:84c5194c105c8e83596c716227654df436f71f6a16e6ec14717601c4ec6cdcae

Observation 488629da-b6f5-4919-b848-b543690d8c37 · outbound

This paper cites Journal of machine learning research 21(140), 1–67 (2020).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Journal of machine learning research 21(140), 1–67 (2020)

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.279871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:36.869481Z digest=sha256:ab94ec6cdceb4625be1de0f03534276f1b9df06093d56ae54551c4ffadfb5cdc

Observation a9ac94fc-c90b-4f57-b292-963270d15c1e · outbound

This paper cites : Training language models to follow instructions with human feedback.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought : Training language models to follow instructions with human feedback

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.268677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:36.945519Z digest=sha256:d5802ef370353fbdb9188328e58ac144892535bc15d97b72ed3e7ab21c841418

Observation 8ad19dd6-f942-4ace-90d9-dab0c91f6bc5 · outbound

This paper cites LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.992645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.992645Z digest=sha256:003148807949ba6fa46e4efbff8cfcb5e48f5c5a50345cbcb469c3534d8d64b6

Observation 009094eb-d556-4465-a87d-d4e5bc3d9846 · outbound

This paper cites In: Proceedings of the 37th International Conference on Neural Information Processing Systems (2024).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the 37th International Conference on Neural Information Processing Systems (2024)

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.257314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:37.046109Z digest=sha256:e0b970e5b1223ca46b46ce6cc861009cc6a553e08451fe4996a29819b71869ce

Observation a56f1546-e8ca-4dee-91e8-64d5b58d6023 · outbound

This paper cites https://openai.com/research/ gpt-4v-system-card.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought https://openai.com/research/ gpt-4v-system-card

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.245815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:37.124103Z digest=sha256:628c46d75a9755b4ecb1b0e1dc9c6f53a72da603abe37cb3f574268541fb86bc

Observation 1f54acb1-611e-49c7-ba95-7ebfe0a7f823 · outbound

This paper cites Direct Language Model Alignment from Online AI Feedback.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Direct Language Model Alignment from Online AI Feedback

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.182865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.182865Z digest=sha256:e3c9a5a13381e9e4578c84ebc7e856fd3f8ec1640e6c1f9c4a116be9be798bb0

Observation 7c78759c-5123-4548-b72a-5f8e56dc45c4 · outbound

This paper cites Self-Rewarding Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Self-Rewarding Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.231697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.231697Z digest=sha256:aeb92d5fd15cabef5e8b532da358b1af9bda04d64d11812320ed64b312ca3beb

Observation 15b86385-82d0-4368-b077-1c142b2587e1 · outbound

This paper cites Human Alignment of Large Language Models through Online Preference Optimisation.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Human Alignment of Large Language Models through Online Preference Optimisation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.284453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.284453Z digest=sha256:3681159341312b8525c398328c88acc1e67f67f6e1f5f6a3482e17f39e13cbdf

Observation 67e203a4-c003-42bd-a4d7-b46b6906879b · outbound

This paper cites RLHF Workflow: From Reward Modeling to Online RLHF.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought RLHF Workflow: From Reward Modeling to Online RLHF

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.331130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.331130Z digest=sha256:bac0467cd0cd4b2c2013102dca02638cf6adef5dd0a54a11c711e51957a2ab84

Observation 884b4a97-d57d-43a0-98d2-87d6e9ac026f · outbound

This paper cites Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.398160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.398160Z digest=sha256:a14feffb4e4c3769e6e7f76f52fa11e1217305717d887c54930647f3ab7b4a99

Observation 07f27bc1-2ecd-417a-b148-5556b19f72db · outbound

This paper cites Advances in Neural Information Processing Systems 35, 15476–15488 (2022).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Advances in Neural Information Processing Systems 35, 15476–15488 (2022)

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.234825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:37.456187Z digest=sha256:c6c77096c3dfe6a226c3ab5eb3de0ab4cf96584e48b0c9d2c6bf1bda419633d1

Observation d2bfe77f-6232-485e-94f8-8b69dbbaa289 · outbound

This paper cites Iterative Reasoning Preference Optimization.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Iterative Reasoning Preference Optimization

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.514013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.514013Z digest=sha256:e44bb2ffbac4006f4f10104eafa92b7cc076df612d393f8b051974ba8b3b1502

Observation 1160b51d-4ce0-42a1-a8d0-d48d17eccc78 · outbound

This paper cites arXiv preprint arXiv:2405.17220 (2024).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought arXiv preprint arXiv:2405.17220 (2024)

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.569087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.569087Z digest=sha256:83e92bfdeeab1b254b3a2482682a92a81fcc7129c4c19ebfe374db98392bd937

Observation 064b4759-aea7-4fb9-ab43-37bc3456dccd · outbound

This paper cites Calibrated Self-Rewarding Vision Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Calibrated Self-Rewarding Vision Language Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.631397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.631397Z digest=sha256:a7741437e8b5d0b75359969bcd5e68c159bf1a39c7b07ecc47c30df6b9ee45d5

Observation 38fbfaff-8ca6-4452-b784-652d4f837eca · outbound

This paper cites ACM MM (2024).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought ACM MM (2024)

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.222385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:37.684578Z digest=sha256:e84652924508b7d321b2e4ce26672ac2808da31a10dca1c19a148c8ec014043b

Observation 7c2861a9-00e8-44db-8dbe-f02ff5886416 · outbound

This paper cites Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.737245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.737245Z digest=sha256:8672b82a1c52af02d634551829eefc3b226bd9d9ea727cb134dd1abc58123a2d

Observation df92a167-3d00-4554-abbe-a401f8be36b3 · outbound

This paper cites In: Findings of the Association for Computational Linguistics: ACL 2022, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Findings of the Association for Computational Linguistics: ACL 2022, pp

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.212095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:37.787387Z digest=sha256:6dc6a01acd49a4b72c38640675acfb01041fc31ba29c46f72153e88ae0927864

Observation a6837e4b-3184-4146-8a75-ec98fa7cac91 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.200325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:37.847899Z digest=sha256:1ef875cf5ee50ce3b97ac991f9c2bca363541b7a182ec4c985eefbb00eb32584

Observation 76c29561-5953-47e2-9bd0-8e5a3f1e83af · outbound

This paper cites NeurIPS (2022).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought NeurIPS (2022)

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.187971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:37.904281Z digest=sha256:12bd873efef68562fab94831bcc99c363921065522e4000e64d553c19115291a

Observation 506ad1ae-9ee1-4c99-bd5d-de37afc5029a · outbound

This paper cites Advances in neural information processing systems 33, 6840–6851 (2020).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Advances in neural information processing systems 33, 6840–6851 (2020)

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.978169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.978169Z digest=sha256:4f234f550f2858f2bc8a5bafb653bab3ee645eacc31337268c3fbe2710d5fcd0

Observation 902e720f-9e47-4ccb-9474-628f406e1f2c · outbound

This paper cites In: Findings of the Association for Computational Linguistics: EMNLP 2024, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Findings of the Association for Computational Linguistics: EMNLP 2024, pp

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.169529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:38.078828Z digest=sha256:761a5342d81fcf2611c967b03a3aa8634769d8c3ea8f79902cd1722fae9b888c

Observation b62d63ba-ff73-417d-8a51-463c0431f0e5 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern 19 Recognition, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern 19 Recognition, pp

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.159105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:38.187584Z digest=sha256:e49a657541cc8755d1113f7e61b8d979f8f3cb943275a15dd7973b632bc175d9

Observation b747dfa7-a7fa-4294-97d0-637d11820d2a · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:38.241871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:38.241871Z digest=sha256:51210492746d2c6046d77e2fcbd39c50a0bd05135944a94967b23b29b02ce834

Observation 363b3e55-d3a6-48f2-a8fb-aa67c8c7ff9f · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Gemini: A Family of Highly Capable Multimodal Models

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:38.244912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:38.244912Z digest=sha256:09ea830a65144e07ed0547be7875b3f18f55e9719b5fb28f255d03273050c81d

Observation e61e0aa2-a79c-49ca-a8a6-0e3d3924ef27 · outbound

This paper cites Mixed Preference Optimization: Reinforcement Learning with Data Selection and Better Reference Model.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Mixed Preference Optimization: Reinforcement Learning with Data Selection and Better Reference Model

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:38.248448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:38.248448Z digest=sha256:f11715efe92e676a155df0eed7325324b7a5bc778281ed1aaf1c9a0c88f04478

Observation 27597d7a-d8a5-40f7-a23d-2c0e71ada57e · outbound

This paper cites GPT-4 Technical Report.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought GPT-4 Technical Report

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:38.285085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:38.285085Z digest=sha256:02999557736eccc2dd1272a0a74130b75311e218725945aa98a0f39d71fbdee1

Observation d137d441-3785-40b9-8f92-a61768b12a98 · outbound

This paper cites The Llama 3 Herd of Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought The Llama 3 Herd of Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:38.412010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:38.412010Z digest=sha256:40fdffdbedd565cb39d22c11683c05ad49c7fdc88194430619a1359c00875119

Observation b7f589be-bd18-4eec-babf-f731bf6cb5c3 · outbound

This paper cites Sub-answers:.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Sub-answers:

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.146448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:38.524221Z digest=sha256:92a50a4776a160014d897d645528e0b8e796ab7cfb01aa2f1c4d1e670a118e91

Observation 30d283c8-525d-4455-9304-d74c5c41f5a0 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.134735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:38.640246Z digest=sha256:1df789bd9a28262eadf1035165b67f81d2d645d9bf9e8f036f02d49da0d3f9a6

Observation e81b333b-9878-4b08-9774-621548c8a829 · outbound

This paper cites Uncertain.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Uncertain

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.123412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:38.712756Z digest=sha256:e0ba1aa7956111a5d0da7d560bcbd489c4ad5e8e74d537a7785a73d139d23205

Observation ff08c9b2-35bd-46ab-aee7-14311f515173 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.111855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:38.821657Z digest=sha256:aaf398906e4abd1b113d19e7a718c0da61912d326db311bb6b3ceb83d1c9a769

Observation 28c47846-b7a7-4869-bd4c-9f3c93f91451 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.101612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:38.986073Z digest=sha256:6d4592915aa55c68913e4278a8b7064e68b9790a9db08ce9bd87b692d8bab7e5

Observation 0934a1a5-231d-4d5a-84a4-3258eaca149e · outbound

This paper cites The formula for the circumference C of a circle is given by: C = 2πr, where r is the radius of the circle.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought The formula for the circumference C of a circle is given by: C = 2πr, where r is the radius of the circle

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.090837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:39.149496Z digest=sha256:232e9f1dcd8c12aa123cfc25b9dcda443542ada8634219fc24cb14b39f84eec6

Observation 21841584-672d-4710-bfc0-f5a8b1dbf116 · outbound

This paper cites - Each side of the square is equal to the height of the rectangle, which is given as 32 units.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought - Each side of the square is equal to the height of the rectangle, which is given as 32 units

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.078402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:39.353941Z digest=sha256:01b908181cc05d4ccaa7f3c5058774cb502a0bdea18b22a065ab49800ea1fcbb

Observation fe9dc88b-a511-49e8-b5a3-7bb538349296 · outbound

This paper cites - The diagonal of a square with side length s is given by s\sqrt{2}.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought - The diagonal of a square with side length s is given by s\sqrt{2}

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.066148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:39.555704Z digest=sha256:e169f109b0245c6e23af64d1907b88f0a774124f59a3448808921f134db84fb4

Observation 75072758-ec69-4e69-ae37-13309c87ea6b · outbound

This paper cites objects": [ {.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought objects": [ {

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.055324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:39.697546Z digest=sha256:09ee12d447e099f27915870137b3134c4e61757a68e996823742854b39535b0e

Observation ce7acc8a-7b65-4e0e-a640-cca607d7fb03 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 77

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.045568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:39.904367Z digest=sha256:a9312e11d80ba8857d6431b7b17dda821b2864cb4ee109113463ea7dece4cc04

Observation 7b657692-e593-4303-be7e-ab141c48ca68 · outbound

This paper cites objects": [ {.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought objects": [ {

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.036316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:40.066958Z digest=sha256:8584e238c6051a116094be301c63824bf0fa9cc29fa14a35b5ab38919f6a455b

Observation ee271c14-4e87-4c91-bd70-311482409aea · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 79

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.026104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:40.197225Z digest=sha256:f303ccf9bd8dcfdfed10407489e0286a6d14a1ac0de7f54459b4753fa101aadf

Observation f88a98fc-b701-4a4d-bfd9-f731d866947b · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 80

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.016416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:40.353407Z digest=sha256:ff7b58b9a9b8c0d3b004dba57927c67a5ddff3b50b8e76ce84bc7b2806d7a59d

Observation 1883a4f1-7c59-4490-bf2f-b2cf1a354f1c · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.007140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:40.505961Z digest=sha256:d58e7b921327a9538309452c5317ac19feb5ca73f7381e848973dee1a7d604c3

Observation 2ad79d24-12a9-441e-91b7-affaaf63f698 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 82

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.997450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:40.674246Z digest=sha256:393c1b8627e5759535d4ddec4e6575b99418273c633690cfd70037eea55de53a

Observation d0b55959-a5ab-490a-98d2-0e3beca719df · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 83

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.988969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:40.913540Z digest=sha256:ada8c9a0aac21cca3b795b183d117fe869fe2aad7c2fde43c842b54bebb109b4

Observation c222e408-bde4-4718-89c1-13c4082e2752 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 84

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.981093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:41.053299Z digest=sha256:094a220f6059d73d17663707ff4d7a7acb88bcc945d7a9f37226a65f8dcc3052

Observation e7512077-b686-4bd8-a9b3-5cea7199c551 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 85

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.972431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:41.231965Z digest=sha256:7889921db7b0344e979cf19d07343bc8e9c8f14088170c4867070143aeea1a9a

Observation 56bc3bf0-545e-492a-9f06-abd4b9e3ffb5 · outbound

This paper cites objects": [{.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought objects": [{

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:44.964871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:41.411657Z digest=sha256:de508cdbf03989ab867314b09a4d87bcdd062304863c138d9cf4d5f3d1b3b56f

Observation f684a584-5991-4a67-a1aa-b443994bcc49 · outbound

This paper cites **E** - Early blastocyst 3.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought **E** - Early blastocyst 3

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:44.957213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:41.562213Z digest=sha256:39772c4a838270e190d28a07db8cf7761008357e4dc5163fcd614fa5120f7512

Observation 11dbcc7f-07d4-4bfd-9fd0-89dcdfd5d89f · outbound

This paper cites (D) No.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought (D) No

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:44.948916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:41.753763Z digest=sha256:f593f3faf59526e5989e1754b958f22249c06a656c3cf69b54fc00f8da0e8fa1

Observation 7448bd87-3e1f-46a6-9bd6-fc8b96006fee · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 89

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.939965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:41.888306Z digest=sha256:1b3005b0f6d8193d58cbd5fcf60b4e793889ff945172524ba2871bd65d7b2681

Observation fc7a62d8-836e-4b9c-85a4-3c8aa8f6a1cd · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 90

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.930051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:42.048251Z digest=sha256:f59d97656c702d77b786a4ca77a768a6458ddb7b114954628876a2a872065cfd

Observation e6b28067-6d4a-4129-a304-bffe378bbf2b · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 91

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.920005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:42.250601Z digest=sha256:a4da8ddc4d088a070ead4b37824dda115ef0360a49ff75f196efadbe448e3cc9

Observation 52c84aee-7ee7-4961-a5b7-29a97ddc5936 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 92

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.850249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:42.419828Z digest=sha256:1eea25253b686bd2ebeb820fa82d85cee08eabfeed80fb6b06355cd292d03b63

Observation ab215d8c-386d-424d-8cac-6d20dfabc0da · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 93

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.638857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:42.629321Z digest=sha256:3f37592d542febea2ae0d57e5f870152269aa745221bf4268191c8df598a42df

Observation bdeacc47-2d6a-4866-ae5d-8a665d175dc6 · outbound

This paper cites objects": [ {.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought objects": [ {

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:44.429933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:42.743404Z digest=sha256:48315be35bc96c85470b5887384248ce64e013f9ead0fd73738bd530d9ae5efc

Observation 35053a44-ae4e-4854-9b81-98b7c152a411 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 95

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.238427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:42.783517Z digest=sha256:cecc761f8cd6297198752a07b8a0bca904e54d0e333a79e4ffd118ea3aec0d7c

Observation 5b2c26af-e05c-445d-ad92-00dc4fc1c7ec · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 96

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.075489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:42.836296Z digest=sha256:f4140cdac96bd6deae4413a2435dc952aeef6f6516bebad25e6b6c0356417c0e

Observation 0700ebef-9854-4246-8fec-40276333546d · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 97

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:43.777934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:42.927117Z digest=sha256:5c86ed9cdd17d235b99e0e803c636d48f8db5f280c7a095b94ad21ded47bd9b8

Observation 5da17c2f-9656-45aa-af6a-8b064383b604 · outbound

This paper cites The correct answer is: (C) smaller than Qwen-VL-7B + SMART : The derivative of the function y = log_2(x) is \frac{1}{x ln(2)}.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought The correct answer is: (C) smaller than Qwen-VL-7B + SMART : The derivative of the function y = log_2(x) is \frac{1}{x ln(2)}

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:43.630021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T21:19:42.977063Z digest=sha256:a3062008c86be30c7e499103e62b8bb4417637548039b98c1d973a52734e7f37

Pith citing papers

No inbound Pith citation observations are available.