Pith. sign in

Paper Citation Record · LEDGER

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning

As of 18 August 2026, this Paper Citation Record lists 72 of 72 outbound references and 0 inbound Pith citation observations for arXiv:2608.10513.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.10513 v1

Coverage vector

measured 72 of 72 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T14:24:49.761179Z

measured 72 of 72 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

72 of 72 outbound references displayed

  • verified exact2
  • verified fuzzy14
  • unresolved56
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6025fc5e-cb13-4b9e-a1f7-bcc2fbfedd4b · outbound

This paper cites Communication, Simulation, and Intelligent Agents: Implications of Personal Intelligent Machines for Medical Education.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Communication, Simulation, and Intelligent Agents: Implications of Personal Intelligent Machines for Medical Education

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.286087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.286087Z digest=sha256:133668eafc0a11a0ec793a2df9abac89723b7ccc08135084301518c200c71510

Observation aafe544d-f17a-4557-a3b7-65bcaa860b07 · outbound

This paper cites Classification Problem Solving.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Classification Problem Solving

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.291857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.291857Z digest=sha256:86f4d0160b59370cb223bfabcee722df3adea4bf60ac84e3b57c1964ac52a909

Observation 03f39d12-37c9-475e-8996-782a28b02c46 · outbound

This paper cites , title =.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning , title =

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.297126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.297126Z digest=sha256:d478a9bb6591d7948e0bddaa22deccb0e5afa0a2f612125ef291383d0d0cde29

Observation 9b684ccd-2da1-4a2e-bd08-f0b86098ad39 · outbound

This paper cites New Ways to Make Microcircuits Smaller---Duplicate Entry.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning New Ways to Make Microcircuits Smaller---Duplicate Entry

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.302190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.302190Z digest=sha256:f3848d69a50ebb56c5f09e03b9f5f2b47088ae244baf7a773303b5ef160ee4b0

Observation 6e7c5bc3-0353-4ce4-a8bb-05eedb4e6db7 · outbound

This paper cites Clancey and Glenn Rennels , abstract =.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Clancey and Glenn Rennels , abstract =

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.307325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.307325Z digest=sha256:6452c7a34b79a5fe80edb5b9eb217c64e2700f5ce02a3df006d874323d41cad3

Observation b370780d-21ed-4d8e-922a-391605fab0d1 · outbound

This paper cites and Rennels, Glenn R.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning and Rennels, Glenn R

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.312726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.312726Z digest=sha256:3917572132319fc9fcf5229da0b896ff031a06bdaae2d32d25149f246682673e

Observation 7c2213ad-8c12-4d7b-a616-e70bc73431ad · outbound

This paper cites Poligon: A System for Parallel Problem Solving.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Poligon: A System for Parallel Problem Solving

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.318979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.318979Z digest=sha256:e86a8d632a49a90694ffe144198310f515a05d772f496f1a145cc18e15ba4145

Observation 1af58472-120b-4390-a2f6-3e45ba91d558 · outbound

This paper cites Transfer of Rule-Based Expertise through a Tutorial Dialogue.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Transfer of Rule-Based Expertise through a Tutorial Dialogue

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.324337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.324337Z digest=sha256:70fb66954a38ff819463596772a4c315d9e1f47cfecc0fb6c685561f973d1d3f

Observation 30e37e8f-9630-4387-a5dc-7562ba936e85 · outbound

This paper cites The Engineering of Qualitative Models.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning The Engineering of Qualitative Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.330017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.330017Z digest=sha256:133039f636cbbca2fd0f7917d072759b59d6fe2a78eed203c8ceb273e9f51c37

Observation 8dd50b5f-f6b2-47c0-b120-2f58e7e202f2 · outbound

This paper cites 2023 , eprint=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning 2023 , eprint=

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.334945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.334945Z digest=sha256:2f4d4e2056e87ce749b0f4315803be7d53731922a1710edc7f80e680362bd825

Observation 8412b579-2ad8-4563-9fc2-71154bb189a0 · outbound

This paper cites Pluto: The 'Other' Red Planet.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Pluto: The 'Other' Red Planet

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.340125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.340125Z digest=sha256:9790aa676dcd7b1c7865c04f2837ee66facabd25697c7bd4edeaa646097330a3

Observation 228938b6-472c-4002-bb00-eaa25e47f0d7 · outbound

This paper cites European Conference on Computer Vision , pages=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning European Conference on Computer Vision , pages=

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:51.161453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.366719Z digest=sha256:4f933a2e87ba01fd715e5863fc740e023e4ccfb459936a43cde7ef8115444c29

Observation 08e7d26c-4bf3-49e6-8c3b-0f8a7c03056b · outbound

This paper cites Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages=

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:51.144189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.387989Z digest=sha256:acbe3c150f98b45b45b5559749450cae64b3d022716ed2a9571afa98a885b5f7

Observation 83099ffa-5463-4801-8505-c206be410151 · outbound

This paper cites Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages=

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:51.127311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.392957Z digest=sha256:92493f45b831439949f13d4ea8253a74656e2bffe4434783d847a94fba7dee90

Observation 3dd945eb-064e-4277-833e-943d8cc322aa · outbound

This paper cites International Conference on Learning Representations , year=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning International Conference on Learning Representations , year=

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:51.110729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.419372Z digest=sha256:f4bc86281d5bdc7d1f422f1efe544ad0619e5a7e2bea94a956239d377a9ea371

Observation c0c9e85f-a084-4bbb-96e2-1570013db0a5 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:51.093297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.429522Z digest=sha256:09eae1868b4138cc42a7af978f08d79779cf92eb2501ece4fc73bf918593df1c

Observation 4e7208e8-6538-4df6-b419-8ca1cf767783 · outbound

This paper cites 2024 , organization=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning 2024 , organization=

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:51.076144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.434452Z digest=sha256:fc6d5d1747847d35d8540aa3b8aaacfe1953191b4a3fb0cf2127d54fa3d5591a

Observation 3b98b798-fc97-4541-bc10-1eec8dbbd7f8 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Advances in Neural Information Processing Systems , volume=

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:51.058351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.440055Z digest=sha256:6b07e6a9f6ce65e4a3f70451a992d50c673e28e60338ad896f3f33f61d4596de

Observation af1397e5-5deb-4e3a-8ee1-3a02f280d7c2 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:51.038568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.445335Z digest=sha256:21816138097ab1451ca15d2097edd18e74bf616ba51475f764deeab017da01bc

Observation bd1c348e-53c4-4dfd-ac82-5f65f525fdf2 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:51.019929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.455750Z digest=sha256:e5106d1bb9ee9b07b8f5d90f331e2d96b46fa9e277c8d22bd9014c5bdbde38f4

Observation 7c801524-ce3b-49e5-af20-dae48fee7937 · outbound

This paper cites Immune: Improving Safety Against Jailbreaks in Multi-modal.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Immune: Improving Safety Against Jailbreaks in Multi-modal

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:51.003543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.460698Z digest=sha256:8746f088135c0ea5a88a004386ab935ddfb310cd5a8009834629086cc9cd1ffb

Observation 217032c7-0fc6-4961-983f-9dfe298829f1 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.985880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.477037Z digest=sha256:789a46983b87c51d04144f0604208def04ac4d5ebb927e3d0f21a5d867a4d8c8

Observation ec4efce1-7f79-4e07-9aa5-0179b1da6d55 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.968545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.481819Z digest=sha256:f8350fa3ebb463c7d5e13c03fda891f3fd6b9d0a41e7031e2bd3b35f99161b65

Observation a63deb8f-bea4-4fa6-b41b-7cc2d9cdeff8 · outbound

This paper cites Computer Vision -- ECCV 2024 , year=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Computer Vision -- ECCV 2024 , year=

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:50.952194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.486867Z digest=sha256:5f04d2433888c0d0f2535a4db073bd1655164de6c28acca29fc7b2d3a798ea66

Observation 953d98a7-bbe0-4ebd-b7b5-21f38ce009f0 · outbound

This paper cites Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:50.935088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.491449Z digest=sha256:9077fb4c7785c0b9bd7bc9d60ef9bca64e1fbec8b308c8e383dc3a5b17220823

Observation 2c142d4a-a0a0-47df-913c-31da69c8f934 · outbound

This paper cites 2025 , eprint=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning 2025 , eprint=

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:50.918592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.496220Z digest=sha256:504df4a6dc75ab24ad2dafed2fc0fa08ed1f815eb2c2d0e2fdfca869fcebe8af

Observation 034ce379-cce3-4d28-9fb0-a351d1e9e384 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.901331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.506246Z digest=sha256:5f0d99bc26d7d0410fd0ab994e1bc1d19f2fcad33978caec5d26beb1b2b4625b

Observation 4d562cb7-31a3-43b9-a9dc-e10ec02ef7d5 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Advances in Neural Information Processing Systems , volume=

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.516080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.516080Z digest=sha256:74d336f5104ff95474ade671cb5ef1e1e8835e4240d913d43ffa90acb03beb3e

Observation 3d8d9528-f3c7-4256-862a-3c5a8ccba9fa · outbound

This paper cites 2025 , eprint=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning 2025 , eprint=

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.546118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.546118Z digest=sha256:ffda952c31eeeb4becbf4acc1f09159ab52276a0f8810f299fb911ae61c10dc9

Observation 42b41b8a-ddef-45ae-a798-53896366a1fd · outbound

This paper cites 2025 , howpublished=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning 2025 , howpublished=

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:50.863889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.551124Z digest=sha256:d1c39e73ae411ba7167cfb0b0dad3c95b2960d70596d683ffeb9d71dc8edc530

Observation b99c6442-d704-4675-9462-6145ec9c58e7 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.848497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.561077Z digest=sha256:7a01c85d59ea3556f6747e7bf10df3e08bf107eb3891a24e704b027e76aa4077

Observation 7af71618-73f6-4217-a0d6-926fb2b4d874 · outbound

This paper cites S.; Dong, Y.; Roy-Chowdhury, A.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning S.; Dong, Y.; Roy-Chowdhury, A

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.566112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.566112Z digest=sha256:8e23dd8080c1d81d699223b4de80250af00eb71375a7a9c10f952a7a3fb0a0eb

Observation fcac46c4-dee3-4c64-9f6c-298a3dc99766 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.833135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.570974Z digest=sha256:2688c5087254f09bfb17cda69ad55ff8668d22f003392cc6b638254b2a992049

Observation ead0cd05-33b7-4a54-9afc-9b2b4d1ec861 · outbound

This paper cites Are We on the Right Way for Evaluating Large Vision-Language Models?.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Are We on the Right Way for Evaluating Large Vision-Language Models?

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.575768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.575768Z digest=sha256:1430a3beb26790818d4129d8e5a78ad9233215c118c1e91a7610c64e314510c8

Observation 59813561-4a5a-4fbc-88bb-df1b70ebd447 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.580381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.580381Z digest=sha256:a88941103a06f1adf449e58708b13e15b45a3762fa5aa2ca71c4cb66f8bfd8c8

Observation 22e5aa0d-b540-4990-9f7d-b19fcaec8339 · outbound

This paper cites BLINK: Multimodal Large Language Models Can See but Not Perceive.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning BLINK: Multimodal Large Language Models Can See but Not Perceive

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.585173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.585173Z digest=sha256:dd06886042c15378274018cb8d24c1329a06f3e239c087f9a53252a8ad9b2681

Observation 5948ede7-3092-4529-985d-e8cfa1c54525 · outbound

This paper cites Gemini Robotics: Bringing AI into the Physical World.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Gemini Robotics: Bringing AI into the Physical World

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.589605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.589605Z digest=sha256:621cb1caa0e5285dea40976431d011d9a95f7ce357416ac12f3a70580ed18ece

Observation 31627a61-ebd4-4a9b-ab3a-ae7869c2a350 · outbound

This paper cites S.; Chakraborty, S.; Singh, V.; Guan, T.; Wang, M.; Velasquez, A.; Beirami, A.; Huang, F.; Manocha, D.; and Bedi, A.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning S.; Chakraborty, S.; Singh, V.; Guan, T.; Wang, M.; Velasquez, A.; Beirami, A.; Huang, F.; Manocha, D.; and Bedi, A

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:50.817152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.594546Z digest=sha256:a8608e29c513653fa458e775bd17f95159c1732f5d6ddf0ec339ba74106b7f4a

Observation 28abac41-ddee-48b4-b94f-af4761b95fc6 · outbound

This paper cites FigStep: Jailbreaking Large Vision-Language Models via Typographic Visual Prompts.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning FigStep: Jailbreaking Large Vision-Language Models via Typographic Visual Prompts

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.599621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.599621Z digest=sha256:425497214471a2e51539b2874a6bb5cc34dd2e0234ceb7f8161722f4439bd26b

Observation 25918594-f841-43f1-8862-2212b3e99ad6 · outbound

This paper cites T.; and Zhang, Y.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning T.; and Zhang, Y

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:50.800740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.604537Z digest=sha256:4fd59e7c8382dfff7175921d5fdf468c2470ed04104fb65e81451f86c3f05d79

Observation 4dd94b19-98cd-4cea-987c-4013b0862739 · outbound

This paper cites Tit-for-Tat: Safeguarding Large Vision-Language Models Against Jailbreak Attacks via Adversarial Defense.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Tit-for-Tat: Safeguarding Large Vision-Language Models Against Jailbreak Attacks via Adversarial Defense

Reference 64

Resolution
verified exact
local_arxiv, observed 2026-08-15T14:24:50.323213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.609578Z digest=sha256:ba3441e6b1c599b54fe52a162fe4d5be04a2892a5d7d120872d7c464e56d5149

Observation 5ac197ff-34c5-4570-82fe-84b6764a67ae · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.784835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.614405Z digest=sha256:9805bb7bfed2496af171cd1c62d810fc044dc72bc53f9f3dad9181081fc823cf

Observation 2dc6e744-3892-41f5-815d-f8b9a5e4ff39 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.769285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.619496Z digest=sha256:1f8dd824f281424d4abf95ec2a776ff6d25ba9214ed32ac900e9433dcef22b33

Observation b9bf0e8c-3aaa-4677-8afe-b2c4d7fd2db9 · outbound

This paper cites How Does Vision-Language Adaptation Impact the Safety of Vision Language Models?.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning How Does Vision-Language Adaptation Impact the Safety of Vision Language Models?

Reference 67

Resolution
verified exact
local_arxiv, observed 2026-08-15T14:24:50.599930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.624109Z digest=sha256:4c6b625a2efdde4748e113edefb884819e700f12d0cfba87cefe0d0d31108a1e

Observation b330890a-6d8b-4c62-a945-a091af957280 · outbound

This paper cites Images are Achilles' Heel of Alignment: Exploiting Visual Vulnerabilities for Jailbreaking Multimodal Large Language Models.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Images are Achilles' Heel of Alignment: Exploiting Visual Vulnerabilities for Jailbreaking Multimodal Large Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.628739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.628739Z digest=sha256:6fbafb2210c496d4e5f9388991bf20d6366bd522304011d02acdec5c2f028959

Observation 8c6746d9-388b-4d77-992a-e3c437860635 · outbound

This paper cites GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.633399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.633399Z digest=sha256:1aafd910aa5e9f5d1d7c1367bbf2fbe05e700f6ee30625164f4d2659105ae3e9

Observation a1c14e02-0921-4471-83ab-86e644c7fb38 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.753489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.638422Z digest=sha256:32c06f558f6b3cfec017e85db1633707ed70c4b8897c3a068fbc877720f067c4

Observation 269c60e0-8b68-425e-9851-c51d0534576f · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.736673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.643490Z digest=sha256:f890ac0d1b4cc1f909296d740b048ea208486abc59b202b934833322b257b898

Observation 0566dc98-f71a-4d25-95ae-b9bf57f6730c · outbound

This paper cites UniGuard: Towards Universal Safety Guardrails for Jailbreak Attacks on Multimodal Large Language Models.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning UniGuard: Towards Universal Safety Guardrails for Jailbreak Attacks on Multimodal Large Language Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.648328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.648328Z digest=sha256:956cdb2d8ea808fdcd466cbc28033d67d0001a8be984983448c4d81acefdaca8

Observation 490d3cb1-abf5-456a-8af5-94c4a05e0634 · outbound

This paper cites Learning To See But Forgetting To Follow: Visual Instruction Tuning Makes LLMs More Prone To Jailbreak Attacks.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Learning To See But Forgetting To Follow: Visual Instruction Tuning Makes LLMs More Prone To Jailbreak Attacks

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.653520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.653520Z digest=sha256:9de59a7f87ddbd1a685e3568a5ef974607a110acdd24a7d12e0afe702d3c0063

Observation f1bdd4e5-71c7-40b5-a5a0-d9479f72aa14 · outbound

This paper cites MLLM-Protector: Ensuring MLLM's Safety without Hurting Performance.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning MLLM-Protector: Ensuring MLLM's Safety without Hurting Performance

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.658164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.658164Z digest=sha256:2417235eb066c971358a449d37dfe0ec50d246f997c99bc33b50c45057dd78b5

Observation b468061a-8828-426a-a25a-39e7fb2a1094 · outbound

This paper cites Image Textualization: An Automatic Framework for Creating Accurate and Detailed Image Descriptions.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Image Textualization: An Automatic Framework for Creating Accurate and Detailed Image Descriptions

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.663318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.663318Z digest=sha256:15cbe778367ff7b698128304af25da2ac93b5ffd72ea6d5bf41db3272bd844dd

Observation ff2c9902-8178-4ac7-b812-54158f74d295 · outbound

This paper cites Visual Adversarial Examples Jailbreak Aligned Large Language Models.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.667991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.667991Z digest=sha256:1b1c64b5d8487be19ab489a38be0732df7c83b9dfd16580aa2e0bd0e8567eadf

Observation 80bcc7fb-bb9c-4d73-a9a3-fce47d6851a4 · outbound

This paper cites D.; and Finn, C.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning D.; and Finn, C

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:50.719088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.672709Z digest=sha256:e18e46815ae8fdebe96226aa49b0b7859115004fb48d59a8e50afc38face0703

Observation a084824b-141e-4c09-b436-0b33272b7c61 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.678351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.678351Z digest=sha256:7d77a1814594f13ace49af80a875b441284536b8e6927497e2aea5b03191e020

Observation ee26d16d-fc61-46ae-9fbc-4cce92410aa5 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 79

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.702794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.683390Z digest=sha256:4fe037745cb913bc9dbd827aeb4f8d3330f7974f15732af5735e373479c4b7d5

Observation f619761b-f2a6-471e-88ec-11824f613c4f · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.688116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.688116Z digest=sha256:72cd0420f1d8e9f7fb05d54de05fa4f17631269fb115db1e65de2994593b3f03

Observation fc5301ea-d58a-41b1-b04c-c916f780f957 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.685891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.692606Z digest=sha256:d6a3b4ef2786546edf04762b852aa936e15e36f748c381503666fcf0076fe109

Observation 35a00d1f-c63b-41cb-a266-798754d1fa60 · outbound

This paper cites Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.697447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.697447Z digest=sha256:324493d2990d236b6f6d0712a68797e08c9c0891c2f055645a171819d1953a73

Observation 319c5e91-143d-4896-964a-00b3ea1e1ccb · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 83

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.669292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.702249Z digest=sha256:06f22afe02e2f6baf7dd3430f65e00da1b49e5f71c96200c054550834dedfa0f

Observation 60a371c4-1909-450e-ae9c-02d9c15c21ed · outbound

This paper cites InferAligner: Inference-Time Alignment for Harmlessness through Cross-Model Guidance.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning InferAligner: Inference-Time Alignment for Harmlessness through Cross-Model Guidance

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.707221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.707221Z digest=sha256:d6eff25cb73d4efdb9c30473b1cb77946a22d85fcbe2bbdce47cc7be224ffdd0

Observation b252a0b5-1046-4f9f-8096-c4836462f329 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 85

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.652790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.712315Z digest=sha256:7413953bc85b97f73313ca21a4d645ea2fd0023d46f32660926ac927ac9cc229

Observation 47cbe7ae-0820-4309-b7a4-a69d2c725db3 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.717212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.717212Z digest=sha256:923d5a5b677d6c909625227345192131bc6148ded2904e9776422c0a2799c4cc

Observation 2c287b8c-e440-414c-8279-c9fa5299dea6 · outbound

This paper cites A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.722202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.722202Z digest=sha256:111b61b164f6fe903104f5457c61286d22aa4602975c68454c0a3520b88a29c8

Observation d9b1a76d-8748-4a6d-ad38-cde6c6a62f53 · outbound

This paper cites MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.727045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.727045Z digest=sha256:085244683a57e39ef31ffd1d269e22fb7cb346d960d08dae0ffdc3c886f62f65

Observation 0124c92d-8536-4a13-87e3-6d0eaf31fbb5 · outbound

This paper cites SPA-VL: A Comprehensive Safety Preference Alignment Dataset for Vision Language Model.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning SPA-VL: A Comprehensive Safety Preference Alignment Dataset for Vision Language Model

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.731853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.731853Z digest=sha256:aa7b7e0a1d6b91b8d13168fd42e572795c3c7cb920b9477b4a284c19c05a1322

Observation bff65539-c313-49db-a9ae-bb5b9155bc60 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.736633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.736633Z digest=sha256:e3de9b70a280d75a546b4d2a520e7acfff7a61622554e587ffd23a4461469072

Observation c10d224c-4731-4766-8304-f2d4adcf8da4 · outbound

This paper cites MM-RLHF: The Next Step Forward in Multimodal LLM Alignment.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning MM-RLHF: The Next Step Forward in Multimodal LLM Alignment

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.741420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.741420Z digest=sha256:e6503376d4ed9ac57b7d887fa208a943fde48ce5c47506501b4ede1a9ad04c0a

Observation 96f4e123-ea8d-4458-b644-8b37f42e79f0 · outbound

This paper cites Multimodal Situational Safety.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Multimodal Situational Safety

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.746172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.746172Z digest=sha256:c6cead562dc6e96637db905d2fd57ae93380ef1fbed21c0c93fef588dc62e712

Observation c33304f9-7846-4db1-818e-036cdd66476b · outbound

This paper cites Safety Fine-Tuning at (Almost) No Cost: A Baseline for Vision Large Language Models.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Safety Fine-Tuning at (Almost) No Cost: A Baseline for Vision Large Language Models

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.751283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.751283Z digest=sha256:d760cf4f4e386c8b4fba09a43541ed129d39fd757e1c36bd8aa3827fa2176deb

Observation 0d7bae29-0166-43e4-84d7-a3053f47ca88 · outbound

This paper cites Understanding and Rectifying Safety Perception Distortion in VLMs.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Understanding and Rectifying Safety Perception Distortion in VLMs

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.755989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.755989Z digest=sha256:2f316c53d869d8407ad9986d92643b99182adb7ddf757e284f02a14f7f4f1a9a

Observation fcbbe27f-914f-4b57-9fd5-cf34162e1e2e · outbound

This paper cites Image-to-Text Logic Jailbreak: Your Imagination can Help You Do Anything.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Image-to-Text Logic Jailbreak: Your Imagination can Help You Do Anything

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.761179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.761179Z digest=sha256:88057d77c4b1aa6fac780922a70d81046f8b3843d68960edd36fb3685221c75a

Pith citing papers

No inbound Pith citation observations are available.