Pith. sign in

Paper Citation Record · LEDGER

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning

As of 8 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2508.18687.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.18687 v1

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T16:21:43.229838Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

32 of 32 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0d56dbef-d2c6-4400-93da-305577108b73 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.069676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.069676Z digest=sha256:ebf84daff613849e41303fc3f49974ceabc49abbd69d915f8e7a5076297e880b

Observation 874730ed-2367-469f-bf19-a068bfbae59d · outbound

This paper cites Bioengineering (2023).

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Bioengineering (2023)

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T16:21:43.750709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T16:21:43.076104Z digest=sha256:4ed5abce415275e1c5fea77aac36beee51db74aa7292f1abb16c282fd5e806e4

Observation 81d137b3-ba22-44e8-8a5c-0dda87616818 · outbound

This paper cites Stable LM 2 1.6B Technical Report.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Stable LM 2 1.6B Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.081686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.081686Z digest=sha256:d8bd41cc229243a9e6a9a3ab1625569c7f5155c6bb7d1faf4526b0eb233e3361

Observation 856c2817-485b-4fd5-93b5-fd73e42c9f16 · outbound

This paper cites HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.088270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.088270Z digest=sha256:2c9ba84d18f21cdbc722952d99755f80a7c7499776e5cec6c339c7ff536ff27c

Observation 654ea1b8-186c-4fda-b34c-6ae2f4e6c556 · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.094255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.094255Z digest=sha256:4af165e4dfc5576bbd816866314ae00c235bd1dbd9c55ab994ae8240dbc659c3

Observation 35404bd6-df0c-4732-b267-bbc2486977b6 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (June 2019).

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (June 2019)

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T16:21:43.736485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T16:21:43.099994Z digest=sha256:cd299974ed5803ddbd046c4122dec40c58234f0ca8f6513240f9d93a9839795f

Observation 4fcd4a0c-82f5-4899-97cb-2edef63093a3 · outbound

This paper cites an unresolved cited work.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-05T16:21:43.721162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T16:21:43.105329Z digest=sha256:036094563a1ed4c9c3fcc18ebc00dd07f2b58dc836b834e10642bb0013050874

Observation a2af4b97-a9d2-4d81-bc67-fef2b11bbc37 · outbound

This paper cites PathVQA: 30000+ Questions for Medical Visual Question Answering.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning PathVQA: 30000+ Questions for Medical Visual Question Answering

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.109874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.109874Z digest=sha256:ff10fc3185a43e70d0968e7d5302daa1ca4d503903794bc020ce5e99b992c229

Observation af194243-1dd1-4209-8d07-1ff654613c4f · outbound

This paper cites GPT-4o System Card.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning GPT-4o System Card

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.114782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.114782Z digest=sha256:26e744ecf46ccb08aa0620dc1c2b22353e193703163165fbc171b9150bc59e19

Observation ded7381b-d726-490c-96a9-da1e5a36551c · outbound

This paper cites CAPO: Reinforcing Consistent Reasoning in Medical Decision-Making.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning CAPO: Reinforcing Consistent Reasoning in Medical Decision-Making

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T16:21:43.493258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T16:21:43.119934Z digest=sha256:0f878c5678d95c29350b0fff700d891329da5567e5c395c55d74c7cccea102db

Observation 99b54aae-8589-477b-8e8c-086b0b51fb87 · outbound

This paper cites OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning OmniV-Med: Scaling Medical Vision-Language Model for Universal Visual Understanding

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.125834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.125834Z digest=sha256:ffa779c001a59019df296eb022370c271c0d759c4394ad231ad28a6d361a5d80

Observation 842011c5-775f-46b1-9dac-e03e853ca6aa · outbound

This paper cites Modality-Fair Preference Optimization for Trustworthy MLLM Alignment.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Modality-Fair Preference Optimization for Trustworthy MLLM Alignment

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.131742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.131742Z digest=sha256:b9869b3dad27e1d6188a1c0a346f400925e4714d7cfdb1f1b48666225745443b

Observation 824ab91f-d14b-4f32-a7e4-eed74b56c12e · outbound

This paper cites HSCR: Hierarchical Self-Contrastive Rewarding for Aligning Medical Vision Language Models.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning HSCR: Hierarchical Self-Contrastive Rewarding for Aligning Medical Vision Language Models

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T16:21:43.441693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T16:21:43.136812Z digest=sha256:c0d09f1c8d33f2088e5a237a0a352e00d220d5357a1b81490f732081cda3378a

Observation 94835c0f-3e84-4c54-8a0f-e2081b6cdb4f · outbound

This paper cites In: Find- ingsoftheAssociationforComputationalLinguistics:EMNLP2024.pp.3843–3860 (2024).

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning In: Find- ingsoftheAssociationforComputationalLinguistics:EMNLP2024.pp.3843–3860 (2024)

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T16:21:43.704488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T16:21:43.141682Z digest=sha256:cd0dc6e728eac0de4ca5fd91a0d7a6631169ba917b82d5ddb3adb03fa5f7b512

Observation f77f716f-6263-4232-bdff-30f7e95c68e4 · outbound

This paper cites Advances in neural information processing systems33, 18661–18673 (2020).

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Advances in neural information processing systems33, 18661–18673 (2020)

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.146007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.146007Z digest=sha256:0c28e63160e686a544607ce71ab5d180e9cbc7dd1e9a2c5563f944e41e9604d5

Observation d715c7cf-4939-4695-a727-68261835492d · outbound

This paper cites Scientific data 5(1), 1–10 (2018).

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Scientific data 5(1), 1–10 (2018)

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.150464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.150464Z digest=sha256:d71f58e6e6ea333ca995d3dcabfce6cfec17d7388518adb7979a4ca2b4888bd9

Observation 0572ec1c-ded5-419b-8a4b-4299eeab4699 · outbound

This paper cites Advances in Neural Information Processing Systems36 (2024).

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Advances in Neural Information Processing Systems36 (2024)

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T16:21:43.669282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T16:21:43.154662Z digest=sha256:cf5f65d47160d1c79bf00edaac057b50a2f468ccbb429208a145bd18e6b29bf1

Observation 5fbe4744-d153-480e-a805-bed249f89ba5 · outbound

This paper cites Self-supervised vision-language pretraining for Medical visual question answering.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Self-supervised vision-language pretraining for Medical visual question answering

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.159467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.159467Z digest=sha256:7f6d13b50ed411eaba453030bac577ee13c413e41c70bfe4cd3aed725b5336af

Observation 70745962-1e3d-4423-ac72-54c6a5d21985 · outbound

This paper cites IEEE (2021).

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning IEEE (2021)

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T16:21:43.654054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T16:21:43.164536Z digest=sha256:c18d02ff267d6afa4dc3a52a54aad7f017bd0f1bdee064da85639241de983afa

Observation f214512f-ec9d-428b-ab54-bfc31d4fa80e · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR).

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.169297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.169297Z digest=sha256:b07582adf0e6c7b7a580c352bf8f7a1a7a0cad6707f31e112bb76cd924e3cdbc

Observation 9408c02f-6a75-459a-bb0a-cf1fc51d0c4a · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T16:21:43.629787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T16:21:43.175304Z digest=sha256:560bdd1710bde2fde9e3dbcd89703d07818c94b0fbb1cd8b432e3422227f49e5

Observation ca3a2c91-7e22-4437-a36e-e40dd505295a · outbound

This paper cites MedCoT: Medical Chain of Thought via Hierarchical Expert.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning MedCoT: Medical Chain of Thought via Hierarchical Expert

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.179799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.179799Z digest=sha256:98644b4ca254d85d5af549857bfd8988cba84199f36732af6647878f054e695e

Observation 74af216d-72f2-47fe-9127-c179c3ad6d18 · outbound

This paper cites Q2ATransformer: Improving Medical VQA via an Answer Querying Decoder.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Q2ATransformer: Improving Medical VQA via an Answer Querying Decoder

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.184749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.184749Z digest=sha256:c05ca8e05cf5ffea99fc595c9d7bda2b7ee82a83cf58239e788a71f784f748be

Observation 34b7e499-6ab9-435f-949f-d673064c32be · outbound

This paper cites In: Machine Learning for Health (ML4H).

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning In: Machine Learning for Health (ML4H)

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.189927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.189927Z digest=sha256:3e19d958794e03829bcaf72b42c6b122e3746874c5f42bc3a54000e328495a60

Observation b21ae348-288f-4af6-a30b-26bb5361d0fb · outbound

This paper cites Sunny and Dark Outside?! Improving Answer Consistency in VQA through Entailed Question Generation.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Sunny and Dark Outside?! Improving Answer Consistency in VQA through Entailed Question Generation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.194719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.194719Z digest=sha256:e308d624d8e89382ea8975aecbbf718c1b709bbaf8ea0ddf24edbcd7dd7c198e

Observation 694e967a-fb44-4d47-90f9-98f13fa09ce7 · outbound

This paper cites Towards Expert-Level Medical Question Answering with Large Language Models.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Towards Expert-Level Medical Question Answering with Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.199298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.199298Z digest=sha256:343fbb4b98e233b95b1ad91fb3cdbad4e7027ae3e6170e43d2fcf7cc0ec80d30

Observation c46c0616-a5cb-43c2-888f-13cef66c9cd9 · outbound

This paper cites Open-Ended Medical Visual Question Answering Through Prefix Tuning of Language Models.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning Open-Ended Medical Visual Question Answering Through Prefix Tuning of Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.204129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.204129Z digest=sha256:cfd6d24216aa64aaf4440e05f1d4d4e3f0a50130435512c4741aadd81d95af54

Observation 1378a18a-d0e6-461a-bcc4-512d825f694d · outbound

This paper cites STLLaVA-Med: Self-Training Large Language and Vision Assistant for Medical Question-Answering.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning STLLaVA-Med: Self-Training Large Language and Vision Assistant for Medical Question-Answering

Reference 28

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T16:21:43.325896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T16:21:43.209676Z digest=sha256:1ec89858e85eea15efb9e021738e39e2842200c61acccafca43c7a94247ad32f

Observation e873239e-c954-40da-ad86-1aa1f2e1c26a · outbound

This paper cites In: European Conference on Computer Vision.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning In: European Conference on Computer Vision

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T16:21:43.604315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T16:21:43.214578Z digest=sha256:0a6f2d4ae83806debcdcabd62188f532ee0c9277e192896367d573c73fdc68f7

Observation f02ac44f-b821-4240-910b-61c7fd4ba075 · outbound

This paper cites BiomedGPT: A Generalist Vision-Language Foundation Model for Diverse Biomedical Tasks.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning BiomedGPT: A Generalist Vision-Language Foundation Model for Diverse Biomedical Tasks

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.219267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.219267Z digest=sha256:7202a1722b2ecbf896891575479545787f236d05dc3900f4a14b49c3cf4ef04c

Observation 0681a122-b310-45ab-8758-383c3c0fb0bf · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.224387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.224387Z digest=sha256:618082675bb2d03ee484d954f37d13f6ed3a1622d1d8b149db472649f654ca5b

Observation ea78d15b-2a13-4fa5-b3dd-391b32e2c7ee · outbound

This paper cites PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering.

Knowing or Guessing? Robust Medical Visual Question Answering via Joint Consistency and Contrastive Learning PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:43.229838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:21:43.229838Z digest=sha256:16144fbbb9ae09b9bd71420f66e1236278f03fe3c657772f45eb06254c4f6b24

Pith citing papers

No inbound Pith citation observations are available.