Pith. sign in

Paper Citation Record · LEDGER

Silkie: Preference Distillation for Large Visual Language Models

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 43 inbound Pith citation observations for arXiv:2312.10665.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.10665 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 43 of 43 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T16:19:45.731027Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

3
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 43b959f8-b5da-4dbb-89c9-e9adc90adbeb · inbound

PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering cites this paper.

PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering Silkie: Preference Distillation for Large Visual Language Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-15T23:08:21.802896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-15T23:08:21.673251Z digest=sha256:7147d6de1a249c43f23f9cc36cdfb0c01c4673207142739f2a2fb43a89740ea5

Observation 33f6ea4b-dd79-4aba-b544-0a251b110e92 · inbound

A Survey on Multimodal Large Language Models cites this paper.

A Survey on Multimodal Large Language Models Silkie: Preference Distillation for Large Visual Language Models

Reference 117

Resolution
verified exact
arxiv_id, observed 2026-05-16T02:56:41.820282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-16T02:56:41.658658Z digest=sha256:2b80be9523a4a90d8db87ab81d106053887be48af64e018e30a8db9d83fdc4ae

Observation a9f82327-f41d-408d-a9eb-0d0f0314a681 · inbound

Aligning Modalities in Vision Large Language Models via Preference Fine-tuning cites this paper.

Aligning Modalities in Vision Large Language Models via Preference Fine-tuning Silkie: Preference Distillation for Large Visual Language Models

Reference 161

Resolution
verified exact
arxiv_id, observed 2026-05-17T10:58:53.398463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-17T10:58:53.215887Z digest=sha256:c01c34633bd4406017804c96d14685391add5deec16f69d1abaebc07e5c1d3d0

Observation 2c650b40-4f4a-4d41-bb7a-cc3b06e5423c · inbound

ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models cites this paper.

ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models Silkie: Preference Distillation for Large Visual Language Models

Reference 106

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:20:21.522610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-23T22:20:21.427717Z digest=sha256:3227ec39ddd16741f4f5b5c837fd7778d25df0a9382046059952fb3ba94d2e2d

Observation c39c5c11-79f8-4f80-8e00-22f304f144db · inbound

A Survey on Knowledge Distillation of Large Language Models cites this paper.

A Survey on Knowledge Distillation of Large Language Models Silkie: Preference Distillation for Large Visual Language Models

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T23:31:11.672692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-17T23:31:11.213552Z digest=sha256:ba09a663683ef587137900a11d33796a8b4b1a7caf7a0cd52224d7ad466d06b8

Observation 233cdd86-cc0c-4292-a444-a54f32b7ff39 · inbound

OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments cites this paper.

OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments Silkie: Preference Distillation for Large Visual Language Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:19:32.561482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-13T01:19:32.406859Z digest=sha256:038d3bd5fadf18ed7f053fedc84456c94d8cf3abf4b8de707041425e75ff9a5b

Observation acdea6cc-2661-40b4-829c-5593e707556c · inbound

Hallucination of Multimodal Large Language Models: A Survey cites this paper.

Hallucination of Multimodal Large Language Models: A Survey Silkie: Preference Distillation for Large Visual Language Models

Reference 105

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:33:33.939949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-11T12:33:32.631346Z digest=sha256:9b1f578b30a563f03bcaea9c14e398ffa3b4febdef84238185d858f4d3c75bcf

Observation 13a8d59a-a86e-4d9e-90e9-2a40123a5b4b · inbound

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output cites this paper.

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output Silkie: Preference Distillation for Large Visual Language Models

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-17T10:46:28.649205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-17T10:46:28.447347Z digest=sha256:3d37054069856d438fd4a7255fa1a1832afb3f607306462cd5641d7944690f78

Observation 60a72d31-2619-46f2-9ead-8271ae6e6c22 · inbound

Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization cites this paper.

Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization Silkie: Preference Distillation for Large Visual Language Models

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:16:17.256298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-16T09:16:17.150383Z digest=sha256:59117875a0271e680fe5f2cf13c4f63aed3cfa7df7781d0c0a0d86ae1c288c3a

Observation 4a215b56-61d5-4b59-9c3e-91dad1ac794f · inbound

Decompose and Leverage Preferences from Expert Models for Improving Trustworthiness of MLLMs cites this paper.

Decompose and Leverage Preferences from Expert Models for Improving Trustworthiness of MLLMs Silkie: Preference Distillation for Large Visual Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T16:19:45.731027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:19:45.731027Z digest=sha256:028ef8c05cb8c5d37985a7381eeb204ee176c5c4ccda1f6a229e1f7751d08079

Observation cd70d805-43f4-47f7-b63b-c1baf50175af · inbound

MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs cites this paper.

MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs Silkie: Preference Distillation for Large Visual Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T14:31:36.751542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:31:36.751542Z digest=sha256:437213ec169c843eeed628b5da68e92ab38c138a29e321f2e90621ecc3ca61fe

Observation 9332e79f-d7ce-46ff-82ee-da5bfda82e81 · inbound

Video-Text Dataset Construction from Multi-AI Feedback: Promoting Weak-to-Strong Preference Learning for Video Large Language Models cites this paper.

Video-Text Dataset Construction from Multi-AI Feedback: Promoting Weak-to-Strong Preference Learning for Video Large Language Models Silkie: Preference Distillation for Large Visual Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T13:31:10.635516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:31:10.635516Z digest=sha256:019d28fce7e24b36f305e132a39edbd6ca68a287d0a0e0d90140abda3ddd09dd

Observation 379e3394-eb4e-4230-89de-79b16e85e071 · inbound

Efficient Self-Improvement in Multimodal Large Language Models: A Model-Level Judge-Free Approach cites this paper.

Efficient Self-Improvement in Multimodal Large Language Models: A Model-Level Judge-Free Approach Silkie: Preference Distillation for Large Visual Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T12:42:49.210147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:42:49.210147Z digest=sha256:baeac3de07c84f26f3e8ed9130eeb385cc827c7a003df5aff9243268cbde9825

Observation a01a592e-44bc-406f-a3c8-a5ab89e09ed0 · inbound

Critic-V: VLM Critics Help Catch VLM Errors in Multimodal Reasoning cites this paper.

Critic-V: VLM Critics Help Catch VLM Errors in Multimodal Reasoning Silkie: Preference Distillation for Large Visual Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T11:27:33.204022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:27:33.204022Z digest=sha256:a7493d6295c0a1ff078de54bf033e89566e008b26e3aa29acd5c9808dea00198

Observation 6c9fc335-538d-4d28-a41f-de1ee3788a70 · inbound

Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment cites this paper.

Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment Silkie: Preference Distillation for Large Visual Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T11:03:01.124098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:03:01.124098Z digest=sha256:96e24212ba137d34fc0f4445c002729f59facb9a5dceee7177cc8ee2c75c6908

Observation 752bf7ef-9dcf-48ea-a232-c45f7efea877 · inbound

EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation cites this paper.

EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Silkie: Preference Distillation for Large Visual Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T21:15:45.306173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:15:45.306173Z digest=sha256:9b0360d7c61b92ff916a2241aac6f7bba4202ef1f7259b887c57185b46696311

Observation a7882aee-c046-4125-ad1b-239d7c18578f · inbound

Pruning All-Rounder: Rethinking and Improving Inference Efficiency for Large Vision Language Models cites this paper.

Pruning All-Rounder: Rethinking and Improving Inference Efficiency for Large Vision Language Models Silkie: Preference Distillation for Large Visual Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T19:42:46.513264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:42:46.513264Z digest=sha256:ec464d01d4342a6929fb73da6ff35a213bd47b1563915f5630e7272fa5865d27

Observation 747aa7de-c4fb-409c-abd2-71b910309c24 · inbound

Beyond Human Data: Aligning Multimodal Large Language Models by Iterative Self-Evolution cites this paper.

Beyond Human Data: Aligning Multimodal Large Language Models by Iterative Self-Evolution Silkie: Preference Distillation for Large Visual Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T11:18:35.611659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:18:35.611659Z digest=sha256:8292aea2c10dc1e3390e41310e6a5b7a4334f69706e906e07407f68898fccbc0

Observation 37c30236-9aaf-494a-96cc-95933a978e63 · inbound

Cross-Modal Attention Calibration for LVLM Hallucination Mitigation cites this paper.

Cross-Modal Attention Calibration for LVLM Hallucination Mitigation Silkie: Preference Distillation for Large Visual Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T22:20:38.096551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:20:38.096551Z digest=sha256:71e336cc3d752f113d27f0fc31874c300b9eca3b611d80cacd0bb3818ca5ff9c

Observation eab059b0-6c42-4853-bc9d-45afcf1cf908 · inbound

AVTrustBench: Assessing and Enhancing Reliability and Robustness in Audio-Visual LLMs cites this paper.

AVTrustBench: Assessing and Enhancing Reliability and Robustness in Audio-Visual LLMs Silkie: Preference Distillation for Large Visual Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T22:19:53.387458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:19:53.387458Z digest=sha256:4004d1148b1a794795fd8356dd13bb0457b7255bacac14b5d65a5e31f1f7f1d5

Observation 044b9f2b-85a2-4e0a-a707-b6eddc44a1cd · inbound

Socratic Questioning: Learn to Self-guide Multimodal Reasoning in the Wild cites this paper.

Socratic Questioning: Learn to Self-guide Multimodal Reasoning in the Wild Silkie: Preference Distillation for Large Visual Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T22:06:30.452437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:06:30.452437Z digest=sha256:a94a2eb0d4b0d7346c2bf10b1104b998ea48de760b1b9687e7b4e7e333b63391

Observation 8e534273-46d3-4802-b811-757a84634a33 · inbound

Temporal Preference Optimization for Long-Form Video Understanding cites this paper.

Temporal Preference Optimization for Long-Form Video Understanding Silkie: Preference Distillation for Large Visual Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T15:35:30.176415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:35:30.176415Z digest=sha256:94e830b169bdc1f4f1aa99c33a0e995de02f4fe31655ee42bbf224b45883d6cf

Observation f0d45872-f93d-4c1d-9d9a-f4c369859f78 · inbound

CHiP: Cross-modal Hierarchical Direct Preference Optimization for Multimodal LLMs cites this paper.

CHiP: Cross-modal Hierarchical Direct Preference Optimization for Multimodal LLMs Silkie: Preference Distillation for Large Visual Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T11:53:55.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:53:55.428207Z digest=sha256:84ad67fc2673815bf623e42d4e038f0ade840e0bfbb8728edc8e43aa8cf024f2

Observation 1a02d2a3-5dbb-4a20-85be-16e009ade606 · inbound

MM-RLHF: The Next Step Forward in Multimodal LLM Alignment cites this paper.

MM-RLHF: The Next Step Forward in Multimodal LLM Alignment Silkie: Preference Distillation for Large Visual Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T18:23:50.053172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T18:23:50.053172Z digest=sha256:a5555430096475db4ce9e2513c32a488eeb8a4b04cbcf1dc07900090fc1a4686

Observation 3a983374-25b0-4703-b8b5-674f6d89abe8 · inbound

ASPO: Adaptive Sentence-Level Preference Optimization for Fine-Grained Multimodal Reasoning cites this paper.

ASPO: Adaptive Sentence-Level Preference Optimization for Fine-Grained Multimodal Reasoning Silkie: Preference Distillation for Large Visual Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:22:59.485200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:22:59.485200Z digest=sha256:071782f6ad97984e3358d8569e70fa5adf93dbee61373eb7ad088d154d90d67f

Observation 33fe8ba4-4dc3-4704-aafd-3cccd388b715 · inbound

LPOI: Listwise Preference Optimization for Vision Language Models cites this paper.

LPOI: Listwise Preference Optimization for Vision Language Models Silkie: Preference Distillation for Large Visual Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T13:45:55.178315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:45:55.178315Z digest=sha256:e9dfd7d74833d52305d2c525cea16fd4225b9d8bd74a423cb92378f5acf0bbc8

Observation 0eb33f72-d27a-4a76-930a-3606de24e0fc · inbound

GThinker: Towards General Multimodal Reasoning via Cue-Guided Rethinking cites this paper.

GThinker: Towards General Multimodal Reasoning via Cue-Guided Rethinking Silkie: Preference Distillation for Large Visual Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:55:12.792022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:55:12.792022Z digest=sha256:ca7ea72aae27646a695ca8d292a388532940e626668526985dc1c5c6af04170a

Observation 2851502b-aa0c-47f8-bb4b-5bbecad3798b · inbound

DPO Learning with LLMs-Judge Signal for Computer Use Agents cites this paper.

DPO Learning with LLMs-Judge Signal for Computer Use Agents Silkie: Preference Distillation for Large Visual Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:11:47.407418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:11:47.407418Z digest=sha256:6175ca2b7236199d63fff31b8332e7f885a9d16f2ffbe72115e318acee6c0f9e

Observation 96db0a19-af0d-4def-b013-5187b34ca691 · inbound

LeanPO: Lean Preference Optimization for Likelihood Alignment in Video-LLMs cites this paper.

LeanPO: Lean Preference Optimization for Likelihood Alignment in Video-LLMs Silkie: Preference Distillation for Large Visual Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:49.027723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:49.027723Z digest=sha256:9b2b20acd2d17328bbc9c11f4a5641d0d5dfb4260b9bec72ff7d69a8d5f62b71

Observation c32024ec-ed48-4892-8f1a-7e24be6b056e · inbound

PostAlign: Multimodal Grounding as a Corrective Lens for MLLMs cites this paper.

PostAlign: Multimodal Grounding as a Corrective Lens for MLLMs Silkie: Preference Distillation for Large Visual Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:19.350818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:19.350818Z digest=sha256:7b2549b196446e8e5de2a7cea9f4bfa0637b94717609b8eb07e2a4a81a53ae80

Observation 197ee628-b76b-4860-9504-86827db0aa01 · inbound

Bridging the Gap in Vision Language Models in Identifying Unsafe Concepts Across Modalities cites this paper.

Bridging the Gap in Vision Language Models in Identifying Unsafe Concepts Across Modalities Silkie: Preference Distillation for Large Visual Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T17:21:35.238949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:21:35.238949Z digest=sha256:c353a6877451ff6ac071d824330a8c72168e47bdc5e12a8ea202793893e5c7c0

Observation a5c9280e-2c3a-4799-be94-002e5bd50402 · inbound

Mitigating Object Hallucinations via Sentence-Level Early Intervention cites this paper.

Mitigating Object Hallucinations via Sentence-Level Early Intervention Silkie: Preference Distillation for Large Visual Language Models

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-25T08:35:32.553471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-25T08:31:24.173135Z digest=sha256:00cc370b7cd74467a04485043a8ed43dfded48ed8530171592fdf0ec190cb3a3

Observation da879e89-1fa4-4893-b6cd-b2135a2a2a45 · inbound

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey cites this paper.

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey Silkie: Preference Distillation for Large Visual Language Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-05T20:28:48.508649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:28:48.508649Z digest=sha256:5528535f961b6e9fa00ed878a69bd48ae3b6d892edfbb6fb1d58af742e06a19a

Observation 82399be2-2ebe-412b-b0e7-3220ae1b890c · inbound

Improving Large Vision and Language Models by Learning from a Panel of Peers cites this paper.

Improving Large Vision and Language Models by Learning from a Panel of Peers Silkie: Preference Distillation for Large Visual Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T12:27:27.380128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:27:27.380128Z digest=sha256:ef3bbca583899a8bf9c9f141fccf1c3db6f8bb6ddda4287a456e152958d34383

Observation bcfcfc53-0e48-4a26-963a-ef15269b830b · inbound

Magic-MM-Embedding: Towards Visual-Token-Efficient Universal Multimodal Embedding with MLLMs cites this paper.

Magic-MM-Embedding: Towards Visual-Token-Efficient Universal Multimodal Embedding with MLLMs Silkie: Preference Distillation for Large Visual Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-03T04:20:54.615173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:20:54.615173Z digest=sha256:421d7eb09f5ceaefa893e170e7271ae5be6deb456afa70791c24b711e4b83a65

Observation 8fc7bf25-77a8-4a10-837c-680aa97badec · inbound

Topo-R1: Detecting Topological Anomalies via Vision-Language Models cites this paper.

Topo-R1: Detecting Topological Anomalies via Vision-Language Models Silkie: Preference Distillation for Large Visual Language Models

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T11:45:32.848017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-15T11:41:27.021776Z digest=sha256:b58d56a254b7c8655da5b517dfe26f549bc4ee8024c0e281bfc1ca2f19096676

Observation e06f9429-3404-4e78-868f-c91be94d3765 · inbound

You Only Judge Once: Multi-response Reward Modeling in a Single Forward Pass cites this paper.

You Only Judge Once: Multi-response Reward Modeling in a Single Forward Pass Silkie: Preference Distillation for Large Visual Language Models

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:20:59.335015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T16:06:57.084550Z digest=sha256:e84786b40d26b787b69d34be7745066169d99f4cd8700d6459103f603f8b19bb

Observation 9a57f30b-4b80-4ea1-ab84-b05383e431fc · inbound

Visual Preference Optimization with Rubric Rewards cites this paper.

Visual Preference Optimization with Rubric Rewards Silkie: Preference Distillation for Large Visual Language Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:56:00.855516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T15:45:52.980881Z digest=sha256:2951f7e61ff6c5e3fdfd362edc8e82344d7a96691f566aac5c18e78e8702d3d8

Observation 529af750-587c-4766-b290-e9a46de366ea · inbound

SignDPO: Multi-level Direct Preference Optimisation for Skeleton-based Gloss-free Sign Language Translation cites this paper.

SignDPO: Multi-level Direct Preference Optimisation for Skeleton-based Gloss-free Sign Language Translation Silkie: Preference Distillation for Large Visual Language Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:01:06.492898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T04:17:32.036750Z digest=sha256:469d86a9673c09fab8ca7fdc5864c62bfc7e4b83b3b8be07b85755c98aa4c717

Observation bcfe1c6d-7f72-4077-a0f3-a9edb50883ab · inbound

Online Self-Calibration Against Hallucination in Vision-Language Models cites this paper.

Online Self-Calibration Against Hallucination in Vision-Language Models Silkie: Preference Distillation for Large Visual Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:16:10.669193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-09T20:20:27.931679Z digest=sha256:246d7029458ff28cfce691a98c32442709bb15a710f4b250f3e50549047bcf0b

Observation d2a7ace1-8370-4b84-a5ae-b6f5fe6768e0 · inbound

Deep Pre-Alignment for VLMs cites this paper.

Deep Pre-Alignment for VLMs Silkie: Preference Distillation for Large Visual Language Models

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-19T16:27:39.101866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-19T16:26:41.094936Z digest=sha256:cf3e3497615c903bfabc53696fe16546b06ebe8a40f85d366445f47efe27b686

Observation 792dae05-a8e5-47e6-8fdd-9ba7f06406cf · inbound

Toward Native Multimodal Modeling: A Roadmap cites this paper.

Toward Native Multimodal Modeling: A Roadmap Silkie: Preference Distillation for Large Visual Language Models

Reference 171

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:04:01.994211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-29T22:58:38.610609Z digest=sha256:9af5b2fe30188500265642b0b1aaa0237788b1563d20b771ce6d2d747218af1a

Observation 34a67d23-2985-4dfa-943f-853f97b59a81 · inbound

MAPL: Multi-Objective Preference Learning for Robot Locomotion cites this paper.

MAPL: Multi-Objective Preference Learning for Robot Locomotion Silkie: Preference Distillation for Large Visual Language Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:40:06.232152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T21:11:00.753949Z digest=sha256:7a7828a18852cd12a46c5e399ff10245b718f3fd598f85f06f50b2898263230b