Pith. sign in

Paper Citation Record · LEDGER

Silkie: Preference Distillation for Large Visual Language Models

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 44 inbound Pith citation observations for arXiv:2312.10665.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.10665 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 44 of 44 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:35:53.595711Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

3
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 43b959f8-b5da-4dbb-89c9-e9adc90adbeb · inbound

PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering cites this paper.

PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering Silkie: Preference Distillation for Large Visual Language Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-15T23:08:21.802896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-15T23:08:21.673251Z digest=sha256:27fae679a8c2ee38fa141d2ad2f91f3630a4dde5794e416f71342a5dcffefc78

Observation 33f6ea4b-dd79-4aba-b544-0a251b110e92 · inbound

A Survey on Multimodal Large Language Models cites this paper.

A Survey on Multimodal Large Language Models Silkie: Preference Distillation for Large Visual Language Models

Reference 117

Resolution
verified exact
arxiv_id, observed 2026-05-16T02:56:41.820282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-16T02:56:41.658658Z digest=sha256:dc65d751e277d89186f7928fbad5447726cf3fb361f33402cf906f98c2fd35db

Observation a9f82327-f41d-408d-a9eb-0d0f0314a681 · inbound

Aligning Modalities in Vision Large Language Models via Preference Fine-tuning cites this paper.

Aligning Modalities in Vision Large Language Models via Preference Fine-tuning Silkie: Preference Distillation for Large Visual Language Models

Reference 161

Resolution
verified exact
arxiv_id, observed 2026-05-17T10:58:53.398463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-17T10:58:53.215887Z digest=sha256:164042457bdcfa4540b477951a6df099df91c796c30840bd1dbc950c63e580c4

Observation 2c650b40-4f4a-4d41-bb7a-cc3b06e5423c · inbound

ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models cites this paper.

ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Silkie: Preference Distillation for Large Visual Language Models

Reference 106

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:20:21.522610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-23T22:20:21.427717Z digest=sha256:3b364756d4c7f71eb965bc31a042d262acaa7a141fac3c2ce1da202f6ef491ca

Observation c39c5c11-79f8-4f80-8e00-22f304f144db · inbound

A Survey on Knowledge Distillation of Large Language Models cites this paper.

A Survey on Knowledge Distillation of Large Language Models Silkie: Preference Distillation for Large Visual Language Models

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T23:31:11.672692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-17T23:31:11.213552Z digest=sha256:904e72babf23308d834df434dbb1e9c38380b1f2c4dd2e8e67060f4e43980252

Observation 233cdd86-cc0c-4292-a444-a54f32b7ff39 · inbound

OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments cites this paper.

OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments Silkie: Preference Distillation for Large Visual Language Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:19:32.561482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-13T01:19:32.406859Z digest=sha256:6ac6361bad05a186ea44dbf6616b9f00b860ca0e5dfbf440bc0329baa8842c6d

Observation acdea6cc-2661-40b4-829c-5593e707556c · inbound

Hallucination of Multimodal Large Language Models: A Survey cites this paper.

Hallucination of Multimodal Large Language Models: A Survey Silkie: Preference Distillation for Large Visual Language Models

Reference 105

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:33:33.939949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-11T12:33:32.631346Z digest=sha256:7e0842f9ab6ca79337f7cba7ab0a536b93e4564616b30baa8af2d825d6b29acb

Observation 13a8d59a-a86e-4d9e-90e9-2a40123a5b4b · inbound

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output cites this paper.

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output Silkie: Preference Distillation for Large Visual Language Models

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-17T10:46:28.649205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-17T10:46:28.447347Z digest=sha256:3f5c1f5f139f6e0ae222323f74924231441d697ac25d6e7bb28f8c0077da4b37

Observation 60a72d31-2619-46f2-9ead-8271ae6e6c22 · inbound

Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization cites this paper.

Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization Silkie: Preference Distillation for Large Visual Language Models

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:16:17.256298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-16T09:16:17.150383Z digest=sha256:ba102a9ac503918a7aaac7e8c93bea324c5a49484da5bd84c5551aa40e29678e

Observation 4a215b56-61d5-4b59-9c3e-91dad1ac794f · inbound

Decompose and Leverage Preferences from Expert Models for Improving Trustworthiness of MLLMs cites this paper.

Decompose and Leverage Preferences from Expert Models for Improving Trustworthiness of MLLMs Silkie: Preference Distillation for Large Visual Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T16:19:45.731027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:19:45.731027Z digest=sha256:fc3e749e56799b5dcc3fc030a1c145d798b640551abe064b583fd7fdce4c3c77

Observation cd70d805-43f4-47f7-b63b-c1baf50175af · inbound

MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs cites this paper.

MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs Silkie: Preference Distillation for Large Visual Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T14:31:36.751542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:31:36.751542Z digest=sha256:488bb33fd5ddf6df9e4c5ded50a41bfe4d62fee702e5101de24eb656c42556dc

Observation 9332e79f-d7ce-46ff-82ee-da5bfda82e81 · inbound

Video-Text Dataset Construction from Multi-AI Feedback: Promoting Weak-to-Strong Preference Learning for Video Large Language Models cites this paper.

Video-Text Dataset Construction from Multi-AI Feedback: Promoting Weak-to-Strong Preference Learning for Video Large Language Models Silkie: Preference Distillation for Large Visual Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T13:31:10.635516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:31:10.635516Z digest=sha256:e9a502928613ec48f1e09d625e0f75a8a61856fd388c570dbaf086ecf192f9b3

Observation 379e3394-eb4e-4230-89de-79b16e85e071 · inbound

Efficient Self-Improvement in Multimodal Large Language Models: A Model-Level Judge-Free Approach cites this paper.

Efficient Self-Improvement in Multimodal Large Language Models: A Model-Level Judge-Free Approach Silkie: Preference Distillation for Large Visual Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T12:42:49.210147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:42:49.210147Z digest=sha256:a8e95db9c0144277592320053dc9a01b8f77f8a2ea4eb335b6e400ce8cf310c4

Observation a01a592e-44bc-406f-a3c8-a5ab89e09ed0 · inbound

Critic-V: VLM Critics Help Catch VLM Errors in Multimodal Reasoning cites this paper.

Critic-V: VLM Critics Help Catch VLM Errors in Multimodal Reasoning Silkie: Preference Distillation for Large Visual Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T11:27:33.204022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:27:33.204022Z digest=sha256:f485ea71b70641a835b7aac8dddb05df3feee0939f0d40feade78a3da152842a

Observation 6c9fc335-538d-4d28-a41f-de1ee3788a70 · inbound

Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment cites this paper.

Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment Silkie: Preference Distillation for Large Visual Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T11:03:01.124098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:03:01.124098Z digest=sha256:106ac5705b572412ad44f237659bf0825ef143f64ed762c185eab6f17f0eea24

Observation 752bf7ef-9dcf-48ea-a232-c45f7efea877 · inbound

EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation cites this paper.

EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Silkie: Preference Distillation for Large Visual Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T21:15:45.306173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:15:45.306173Z digest=sha256:253b239e8170343f413ecc9b64bb5700a241c814faa3bc118e8bf32184b860da

Observation a7882aee-c046-4125-ad1b-239d7c18578f · inbound

Pruning All-Rounder: Rethinking and Improving Inference Efficiency for Large Vision Language Models cites this paper.

Pruning All-Rounder: Rethinking and Improving Inference Efficiency for Large Vision Language Models Silkie: Preference Distillation for Large Visual Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:46.513264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:46.513264Z digest=sha256:6d8448db234ab647d79942ad8fbd445103e4c9e47c7586ef417b4415b2160f7b

Observation 747aa7de-c4fb-409c-abd2-71b910309c24 · inbound

Beyond Human Data: Aligning Multimodal Large Language Models by Iterative Self-Evolution cites this paper.

Beyond Human Data: Aligning Multimodal Large Language Models by Iterative Self-Evolution Silkie: Preference Distillation for Large Visual Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T11:18:35.611659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:18:35.611659Z digest=sha256:5fba61f099570be462a9a69fdbb1605c91e394d9fee4b730256a039138f70025

Observation 37c30236-9aaf-494a-96cc-95933a978e63 · inbound

Cross-Modal Attention Calibration for LVLM Hallucination Mitigation cites this paper.

Cross-Modal Attention Calibration for LVLM Hallucination Mitigation Silkie: Preference Distillation for Large Visual Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:38.096551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:38.096551Z digest=sha256:32e28d4e845b6acdf50705b94a5b23d336cb952440d23cc3b72d6739ae3e38bd

Observation eab059b0-6c42-4853-bc9d-45afcf1cf908 · inbound

AVTrustBench: Assessing and Enhancing Reliability and Robustness in Audio-Visual LLMs cites this paper.

AVTrustBench: Assessing and Enhancing Reliability and Robustness in Audio-Visual LLMs Silkie: Preference Distillation for Large Visual Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T22:19:53.387458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:19:53.387458Z digest=sha256:c51d0d3e7ffc5ab3a29cfb4365f66b883b71a4920b5fd4802cfeaaaaeafc17a3

Observation 044b9f2b-85a2-4e0a-a707-b6eddc44a1cd · inbound

Socratic Questioning: Learn to Self-guide Multimodal Reasoning in the Wild cites this paper.

Socratic Questioning: Learn to Self-guide Multimodal Reasoning in the Wild Silkie: Preference Distillation for Large Visual Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T22:06:30.452437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:06:30.452437Z digest=sha256:dfb6e4beab5c50c3469c2afff5b63d3ee559724cf896154380d41fd120a6a9bd

Observation 8e534273-46d3-4802-b811-757a84634a33 · inbound

Temporal Preference Optimization for Long-Form Video Understanding cites this paper.

Temporal Preference Optimization for Long-Form Video Understanding Silkie: Preference Distillation for Large Visual Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T15:35:30.176415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:35:30.176415Z digest=sha256:e62e44ca0546bc9db4e15add168cbaa5257939fe161a2ace2c3c2c038e18d029

Observation f0d45872-f93d-4c1d-9d9a-f4c369859f78 · inbound

CHiP: Cross-modal Hierarchical Direct Preference Optimization for Multimodal LLMs cites this paper.

CHiP: Cross-modal Hierarchical Direct Preference Optimization for Multimodal LLMs Silkie: Preference Distillation for Large Visual Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T11:53:55.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:53:55.428207Z digest=sha256:4e48360826875c5831fc7d37100e3d791877cf49fc0ccad59eee4fb25d919dfa

Observation 1a02d2a3-5dbb-4a20-85be-16e009ade606 · inbound

MM-RLHF: The Next Step Forward in Multimodal LLM Alignment cites this paper.

MM-RLHF: The Next Step Forward in Multimodal LLM Alignment Silkie: Preference Distillation for Large Visual Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T18:23:50.053172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T18:23:50.053172Z digest=sha256:9b43c11916d24f19f065171c93bb40f1ae9a911943da673b22fd7241653672b4

Observation 3a983374-25b0-4703-b8b5-674f6d89abe8 · inbound

ASPO: Adaptive Sentence-Level Preference Optimization for Fine-Grained Multimodal Reasoning cites this paper.

ASPO: Adaptive Sentence-Level Preference Optimization for Fine-Grained Multimodal Reasoning Silkie: Preference Distillation for Large Visual Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:22:59.485200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:22:59.485200Z digest=sha256:6cbc1b203502423ebd5e2f6f7260c4b654eab7be96629bbf6b8e5ab442f84fb3

Observation 33fe8ba4-4dc3-4704-aafd-3cccd388b715 · inbound

LPOI: Listwise Preference Optimization for Vision Language Models cites this paper.

LPOI: Listwise Preference Optimization for Vision Language Models Silkie: Preference Distillation for Large Visual Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T13:45:55.178315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:45:55.178315Z digest=sha256:3db381bdfcfbb83c27f37e2f0caa149f5992115d971f1484fd2a2b9695e2055f

Observation 0eb33f72-d27a-4a76-930a-3606de24e0fc · inbound

GThinker: Towards General Multimodal Reasoning via Cue-Guided Rethinking cites this paper.

GThinker: Towards General Multimodal Reasoning via Cue-Guided Rethinking Silkie: Preference Distillation for Large Visual Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:55:12.792022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:55:12.792022Z digest=sha256:ed5e80ec4e4d9dd0e042002d8113e9f4034dc69f1493c8a641b56b29f04f5877

Observation 2851502b-aa0c-47f8-bb4b-5bbecad3798b · inbound

DPO Learning with LLMs-Judge Signal for Computer Use Agents cites this paper.

DPO Learning with LLMs-Judge Signal for Computer Use Agents Silkie: Preference Distillation for Large Visual Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:11:47.407418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:11:47.407418Z digest=sha256:138f975824a32dc0a79b85ee02c0302fa3c0bd42fabc1544abaf759e0c32aac9

Observation 96db0a19-af0d-4def-b013-5187b34ca691 · inbound

LeanPO: Lean Preference Optimization for Likelihood Alignment in Video-LLMs cites this paper.

LeanPO: Lean Preference Optimization for Likelihood Alignment in Video-LLMs Silkie: Preference Distillation for Large Visual Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:49.027723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:49.027723Z digest=sha256:f77ea9bfbee8cf13adca9c9ddabc38dcd5e8c31caa6abb1b416cfaa8767a7458

Observation c32024ec-ed48-4892-8f1a-7e24be6b056e · inbound

PostAlign: Multimodal Grounding as a Corrective Lens for MLLMs cites this paper.

PostAlign: Multimodal Grounding as a Corrective Lens for MLLMs Silkie: Preference Distillation for Large Visual Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:19.350818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:19.350818Z digest=sha256:8ab55a6c6c25a435920f7ea0650f817980cb44587369508c08fe4a089815e0d6

Observation 197ee628-b76b-4860-9504-86827db0aa01 · inbound

Bridging the Gap in Vision Language Models in Identifying Unsafe Concepts Across Modalities cites this paper.

Bridging the Gap in Vision Language Models in Identifying Unsafe Concepts Across Modalities Silkie: Preference Distillation for Large Visual Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T17:21:35.238949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:21:35.238949Z digest=sha256:ea79c5122d0f8c9589941487f4aa0d207e4c82e3c74b094670ad8524a2839cfa

Observation a5c9280e-2c3a-4799-be94-002e5bd50402 · inbound

Mitigating Object Hallucinations via Sentence-Level Early Intervention cites this paper.

Mitigating Object Hallucinations via Sentence-Level Early Intervention Silkie: Preference Distillation for Large Visual Language Models

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-25T08:35:32.553471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-25T08:31:24.173135Z digest=sha256:955db749a6c5f0100c6c65623ccb9eb2e0bc14807d82def1b623ba64ef7d580d

Observation da879e89-1fa4-4893-b6cd-b2135a2a2a45 · inbound

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey cites this paper.

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey Silkie: Preference Distillation for Large Visual Language Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-05T20:28:48.508649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:28:48.508649Z digest=sha256:73b93b4c2c5c835939844dcb3d5abd1d62eb0d8db314eae039635011a5ab5901

Observation 82399be2-2ebe-412b-b0e7-3220ae1b890c · inbound

Improving Large Vision and Language Models by Learning from a Panel of Peers cites this paper.

Improving Large Vision and Language Models by Learning from a Panel of Peers Silkie: Preference Distillation for Large Visual Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T12:27:27.380128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:27:27.380128Z digest=sha256:5f1e5a7bfc6b9096b776965d0895b6acc0f02ad3e981ed3d186a5069acc42f46

Observation bcfcfc53-0e48-4a26-963a-ef15269b830b · inbound

Magic-MM-Embedding: Towards Visual-Token-Efficient Universal Multimodal Embedding with MLLMs cites this paper.

Magic-MM-Embedding: Towards Visual-Token-Efficient Universal Multimodal Embedding with MLLMs Silkie: Preference Distillation for Large Visual Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-03T04:20:54.615173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:20:54.615173Z digest=sha256:c64725bbe1d9bbab5fb17264d988469e08713e1ef48a5c5f533f8c83c2ae4de9

Observation 8fc7bf25-77a8-4a10-837c-680aa97badec · inbound

Topo-R1: Detecting Topological Anomalies via Vision-Language Models cites this paper.

Topo-R1: Detecting Topological Anomalies via Vision-Language Models Silkie: Preference Distillation for Large Visual Language Models

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T11:45:32.848017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-15T11:41:27.021776Z digest=sha256:5d2095292b8fae9bc1724cacfa74b0100189c4efc206a3742f54e435a048d304

Observation e06f9429-3404-4e78-868f-c91be94d3765 · inbound

You Only Judge Once: Multi-response Reward Modeling in a Single Forward Pass cites this paper.

You Only Judge Once: Multi-response Reward Modeling in a Single Forward Pass Silkie: Preference Distillation for Large Visual Language Models

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:20:59.335015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T16:06:57.084550Z digest=sha256:419d1aa286751270111dc4333d202d016161b0fa965cc09dd61cb66c95c5e970

Observation 9a57f30b-4b80-4ea1-ab84-b05383e431fc · inbound

Visual Preference Optimization with Rubric Rewards cites this paper.

Visual Preference Optimization with Rubric Rewards Silkie: Preference Distillation for Large Visual Language Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:56:00.855516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T15:45:52.980881Z digest=sha256:c0fc0ad66826951e31a87d6f585c36722b92e38f2f50a39f50d94dbd4fcfb967

Observation 529af750-587c-4766-b290-e9a46de366ea · inbound

SignDPO: Multi-level Direct Preference Optimisation for Skeleton-based Gloss-free Sign Language Translation cites this paper.

SignDPO: Multi-level Direct Preference Optimisation for Skeleton-based Gloss-free Sign Language Translation Silkie: Preference Distillation for Large Visual Language Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:01:06.492898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T04:17:32.036750Z digest=sha256:57030b84c01c1ddd225f4b757fb8d589db5d7d0c7b6c1f54b2a1cb11883f1eb4

Observation bcfe1c6d-7f72-4077-a0f3-a9edb50883ab · inbound

Online Self-Calibration Against Hallucination in Vision-Language Models cites this paper.

Online Self-Calibration Against Hallucination in Vision-Language Models Silkie: Preference Distillation for Large Visual Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:16:10.669193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-09T20:20:27.931679Z digest=sha256:2e01fb6045b88fd62e89518f34cdeaa6a5369e7541332a3c0355b78d5ace8a89

Observation d2a7ace1-8370-4b84-a5ae-b6f5fe6768e0 · inbound

Deep Pre-Alignment for VLMs cites this paper.

Deep Pre-Alignment for VLMs Silkie: Preference Distillation for Large Visual Language Models

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-19T16:27:39.101866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-19T16:26:41.094936Z digest=sha256:15824c0a7d50dd3f83f2868d8aa3b16d4b8327ec5a6632e808f0b90dbcfcba03

Observation 792dae05-a8e5-47e6-8fdd-9ba7f06406cf · inbound

Toward Native Multimodal Modeling: A Roadmap cites this paper.

Toward Native Multimodal Modeling: A Roadmap Silkie: Preference Distillation for Large Visual Language Models

Reference 171

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:04:01.994211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T22:58:38.610609Z digest=sha256:230073373e58d0621d5e4c96a40f0ab283e893b6098109125fa6b26bcb6e6920

Observation 34a67d23-2985-4dfa-943f-853f97b59a81 · inbound

MAPL: Multi-Objective Preference Learning for Robot Locomotion cites this paper.

MAPL: Multi-Objective Preference Learning for Robot Locomotion Silkie: Preference Distillation for Large Visual Language Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:40:06.232152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-25T21:11:00.753949Z digest=sha256:a6d5735cdb9dd378b9207b759ea2a762e8bd87755cf5f6da65c89ddf47eee262

Observation e0ddff48-fbb3-49de-ae9f-37bda997c476 · inbound

VADER: Adaptive Debiasing for Hallucination Mitigation in Video Large Language Models cites this paper.

VADER: Adaptive Debiasing for Hallucination Mitigation in Video Large Language Models Silkie: Preference Distillation for Large Visual Language Models

Reference 262

Resolution
unresolved
no resolver link, observed 2026-08-14T04:35:53.595711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:35:53.595711Z digest=sha256:acf1144f20f62d4b479d295f3a9a0308b7d15e73641c7cdecb91fca0df1da287