Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T19:52:50.394161Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 0 inbound Pith citation observations for arXiv:2508.11616.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T19:52:50.394161Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
59 of 59 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d1a4b941-605e-47a9-ab80-e1429cdd34d7 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1ba0cee-9f31-4e8b-a0fc-4456da44a26f · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Understanding Alignment in Multimodal LLMs: A Comprehensive Study
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9195a9d8-747a-4bec-887e-5f8ea3d54aba · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Hallucination of Multimodal Large Language Models: A Survey
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c45988db-9f13-4a82-8f9f-bfc81ab5dda1 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding PaliGemma: A versatile 3B VLM for transfer
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37c161bf-3b75-408d-9fc4-df691ab59ba6 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding An Introduction to Vision-Language Modeling
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e014888-65c9-4a83-a39c-9005d33696d5 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Rank analysis of incomplete block designs: I
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ef37e90e-3dbf-4593-9d25-0db8535f987b · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Large Language Monkeys: Scaling Inference Compute with Repeated Sampling
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 334f6ea5-faa9-4989-9a8a-1ab7459ff0ae · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Language Models are Few-Shot Learners
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46ac2cda-50aa-4279-9c0f-b634d3061524 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding The (r) evolution of multi- modal large language models: A survey
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 439aac8e-d30a-4625-ab01-fee4a7b80ec8 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding End-to- end object detection with transformers
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0d560a61-1267-4238-8b4c-a8e4fc62b192 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Plug and play language models: A simple approach to controlled text generation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d48f62ed-f031-4991-b33f-ff70f1455e3b · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0243cfc-04f9-4ed4-85e3-df8eba7f9d7a · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Reward-augmented decod- ing: Efficient controlled text generation with a unidirectional reward model
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d9812ea6-4b95-47ed-be5f-1e9a817c7c31 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding The Llama 3 Herd of Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ff4a560-5341-427f-a45f-3b3a180ba014 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Multi-modal hal- lucination control by visual information grounding
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 76f764cb-197a-4e02-8bac-ad017610ec3a · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Value Augmented Sampling for Language Model Alignment and Personalization
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd8a1b6f-a981-41f8-bc3d-56d9f6a80c33 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Sugarcrepe: Fixing hackable benchmarks for vision-language compositionality
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3e8ce828-3c3f-4b84-93ee-523f77de5faf · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Lora: Low- rank adaptation of large language models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 826fd636-5466-42e6-addf-efbb1a1c091e · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Args: Alignment as reward-guided search
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b3ec0fdf-f59d-440d-9450-2477156f23d6 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding RewardBench: Evaluating Reward Models for Language Modeling
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 372d28b3-8d50-408a-8274-a70465e5fb1d · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Mitigating object hal- lucinations in large vision-language models through visual contrastive decoding
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 19bba8e2-9a60-4c24-98e9-d6aaad167b33 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Sequential monte carlo steering of large lan- guage models using probabilistic programs
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ef43729f-048b-4a7a-89d6-9bb851ee4dc9 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Cascade reward sampling for efficient decoding-time align- ment
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 58053729-0009-4ac4-bf4f-2ff1c2de74c5 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Vlfeedback: A large-scale ai feedback dataset for large vision-language models alignment
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 67a17eee-0953-42c8-9389-3abfde6e5385 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Evaluating object hallucination in large vision-language models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6597e2da-aee1-4f69-a328-ff6636a2ded5 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Mitigating hallucination in large multi-modal models via robust instruction tuning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 29aaa85c-6b44-4b2f-87b7-622186dbcf1f · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Improved baselines with visual instruction tuning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 04e659d9-92b8-4c54-b0fa-96b3b42aafcd · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Don’t throw away your value model! generating more preferable text with value-guided monte-carlo tree search decoding
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f205440f-a614-4cb6-a324-07957d941101 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Investigating and mitigating object hallucinations in pretrained vision-language (clip) models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 02a7cfff-3c27-4b07-ad5b-22851ff919eb · outbound
Controlling Multimodal LLMs via Reward-guided Decoding SmolVLM: Redefining small and efficient multimodal models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61c74f47-441a-413f-af26-9950f4aa96e3 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Scaling open-vocabulary object detection
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 942d4a95-69c1-4301-a756-35226a71d2fa · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Controlled decod- ing from language models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 22c9f9d3-3275-46a8-8896-5c8e209054f7 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Training language models to follow instructions with human feedback
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 53fa1001-81ea-4ce1-a5ef-a22c76750746 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Learning transferable visual models from natural language supervi- sion
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7dcd835-ec49-40c6-82ba-5f5a96b0f0f8 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Direct preference optimization: Your language model is secretly a reward model
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0f17a343-f98b-446c-878e-faeae2c5b3f3 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding A critical look at to- kenwise reward-guided text generation
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0dfb7e8f-4626-4524-b70f-0fed2ea3b02b · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Sentence-bert: Sentence embeddings using siamese bert-networks
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8fb41234-61de-4a1d-be41-294783ee2b63 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Object hallucination in image cap- tioning
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 38a8d27a-7220-42b2-932b-d2eb8fd01b5d · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Mitigating Object Hallucination in MLLMs via Data-augmented Phrase-level Alignment
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8e5d578-55d5-4073-9049-6297fb93ae26 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad339d84-e493-4831-89cc-7096e72b5e41 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding PaliGemma 2: A Family of Versatile VLMs for Transfer
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bccd353f-0b69-4cef-9847-1bc0e1faafd3 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Aligning Large Multimodal Models with Factually Augmented RLHF
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e80dd9b1-5646-4091-99f2-c0e139688bce · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Gemini: A Family of Highly Capable Multimodal Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7c51513-c48c-4e4b-b0b8-a817735f8afd · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Contrastive Region Guidance: Improving Grounding in Vision-Language Models without Training
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a82b1c07-6d07-4831-ad0d-e46326757cc6 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding mDPO: Conditional Preference Optimization for Multimodal Large Language Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5bee5c2-4487-4ccc-a420-ecd5e506f822 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1387abe-36ad-4a24-8474-a8aa06df1608 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Fudge: Controlled text genera- tion with future discriminators
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1dadae72-8f9d-4a43-98a3-363fc9f4b752 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Woodpecker: Hallucination Correction for Multimodal Large Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84cbdba1-2f16-4ba6-b72f-98472b1082cb · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Rlhf-v: Towards trustworthy mllms via behavior alignment from fine-grained correctional hu- man feedback
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ae856375-f6d8-44ce-8fe2-1619c0597096 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Rlaif-v: Aligning mllms through open-source ai feedback for super gpt-4v trustworthiness
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b6617e6-dee4-468c-9fce-02b3789331cf · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Less is More: Mitigating Multimodal Hallucination from an EOS Decision Perspective
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a10d5500-02b9-4cc8-b43e-a2fbc8b451a7 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Self-Correcting Decoding with Generative Feedback for Mitigating Hallucinations in Large Vision-Language Models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fbbeb2a-fc82-43fe-b537-59b5b329929d · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Multimodal chain-of-thought reasoning in language models
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9a6e7ddc-b304-4ad9-ba1e-0ff10c7e9035 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Mitigating Object Hallucination in Large Vision-Language Models via Image-Grounded Guidance
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 641aad0b-c7e6-4738-bd4b-3560e69da43c · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa86e7c7-b053-4781-b60f-33adc81f7be3 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Aligning modalities in vision large lan- guage models via preference fine-tuning
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 5369d77b-72fb-4a3e-b16e-483daf65cdb6 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Analyzing and mitigating object hallucination in large vision-language models
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d23d70d0-a0c0-4148-8f1a-b3ad230b8e11 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Calibrated Self-Rewarding Vision Language Models
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c058615-4b94-457e-9c0b-7ae63f453588 · outbound
Controlling Multimodal LLMs via Reward-guided Decoding Describe this image in detail
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
No inbound Pith citation observations are available.