Pith. sign in

Paper Citation Record · LEDGER

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision

As of 14 August 2026, this Paper Citation Record lists 100 of 118 outbound references and 1 inbound Pith citation observation for arXiv:2501.04568.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.04568 v2

Coverage vector

measured 100 of 118 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:34:35.594979Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T19:41:58.673348Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T22:35:51.417589Z

Reference resolution

100 of 118 outbound references displayed

  • verified exact1
  • verified fuzzy8
  • unresolved91
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 878090a4-d779-44d2-b6f8-1d0a6d23754b · outbound

This paper cites GPT-4 Technical Report.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.931476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.931476Z digest=sha256:033dccfa75f739831e301d31505a31dd1d11e5f94e140d70f6d0fefa5f37c57d

Observation 32d9879f-8e91-4db7-90ee-899051b96cc1 · outbound

This paper cites Agrawal, K.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Agrawal, K

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.937048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.937048Z digest=sha256:15cd981e965e069b28f35505ebfdcf882f9f3efdf4a915fc8a9283c0f0e5192e

Observation 3694a681-b2d0-4114-8ced-70895428e77c · outbound

This paper cites Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.941481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.941481Z digest=sha256:e60240e528fe216dda6565e68214efbd8ff75521f823464385a82f057c60157e

Observation 4ffd809b-61fd-4843-96ab-870d088a292e · outbound

This paper cites What learning algorithm is in-context learning? Investigations with linear models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision What learning algorithm is in-context learning? Investigations with linear models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.947407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.947407Z digest=sha256:d40b4ecc863fac562475b475350c004a90380710c775b2ab5c114682b6429d39

Observation 7474ee78-909d-4701-9a8f-384dc7bab561 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.953338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.953338Z digest=sha256:204301bcaed70a9facc90b9ae7cad4e79afdf664609b7bc5d381afa8e1444bd2

Observation 71d04822-d2d4-4bc0-9c9a-8ad64dc95e1a · outbound

This paper cites Anthony, Z.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Anthony, Z

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.958760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.958760Z digest=sha256:14c3f7cd0f576c3ef3889d27df0037a1834b674fe8153e237c1fe7087c988e30

Observation cdab2c24-dd18-40d9-89c0-c0f18369e1fe · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Constitutional AI: Harmlessness from AI Feedback

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.964208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.964208Z digest=sha256:dfd8b3bbd5f33c9f1c09606992a5d38fac3922b4e22e29266dddfc092d53f13b

Observation e6885b9d-e499-418e-8676-15396940c355 · outbound

This paper cites Hallucination of Multimodal Large Language Models: A Survey.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Hallucination of Multimodal Large Language Models: A Survey

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.973845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.973845Z digest=sha256:5d7820a036cc44ec176738800f4c3797a3109e5cc59f1eb4bb2413a78e7466ab

Observation eb961fdf-bffd-453e-9911-eaa1254e266c · outbound

This paper cites Banerjee and A.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Banerjee and A

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.978603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.978603Z digest=sha256:c151fdcf435bbaef83f6e68af5b0fa2fe840fa590e61a1ba775c92f5d7b7e4e7

Observation 34624016-5e5f-4eb1-980f-e0ba1c339fe0 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.982785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.982785Z digest=sha256:4233561dff582a3f039c3603e8bcaf198c44b07f7fd6fc721a0bea59f523d1a7

Observation a2d0aaeb-d4cd-4ca1-8fbe-5d464f3bd4b3 · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision PaliGemma: A versatile 3B VLM for transfer

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.987003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.987003Z digest=sha256:198f39afda65a8f9751b7fcd99010bc09d847a96433f627514ee155d25c66a32

Observation 9d087100-76dd-4a0e-950f-c2b6f169a9fd · outbound

This paper cites An Introduction to Vision-Language Modeling.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision An Introduction to Vision-Language Modeling

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.991443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.991443Z digest=sha256:70dd6802f0a13ea5a56d2af752d7f1025c7389f7341ccc010b4eb2765cc74719

Observation c495a7b7-1008-4907-8641-d6b95aabe1bc · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.995902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.995902Z digest=sha256:e2f9d0fab4a8aa8bffb5c73feb4fab8bdcf3a92abe0ec22970a186e177bb3827

Observation 5bf2d01e-dbba-43ac-9d0e-333859fa2f8a · outbound

This paper cites Carion, F.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Carion, F

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.000515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.000515Z digest=sha256:b125c42d364b80f5d73a31f1b10f353a2d30dd7d88ce1827560a3ec1dd0b797a

Observation b5c61565-a26f-4a92-9daf-179399d0330c · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.004387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.004387Z digest=sha256:eb43affb24a53dde65f683fc76cf867ffbe1394501318e22c390a16364214407

Observation 89342bce-3e91-4aac-bf88-d82c344e4c1a · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.008817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.008817Z digest=sha256:400447abb7b8461d7733f5dffcafb2183d191d063fd1a54bf27e1613d6485ca0

Observation 850fcc1d-5643-4686-9f1d-0fddf665bf05 · outbound

This paper cites Chiang, Z.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Chiang, Z

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.013082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.013082Z digest=sha256:094c10d6ac144635bc5d06d3fc73eac165cf22871ed336839870b4e686739573

Observation 2046ea3d-8e12-45da-97ec-a0f27ca6eb92 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.017220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.017220Z digest=sha256:8d090e38fe83694dcf2a3ca818d00e4c485015be4e1900b49ae09f28a0b0f343

Observation fc4d0547-f5f6-462a-a702-3fa95d0eca58 · outbound

This paper cites Collerton, J.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Collerton, J

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.021685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.021685Z digest=sha256:36d2e1db873223f3a197c4b81290d6ca9bf3ca9ad81c05b269292d30efa9a9f9

Observation 56f71cab-a7d4-463f-8650-3f89730ac78a · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.025848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.025848Z digest=sha256:6b5b3b6790f41b85c3da8b5a654ae2c256a63c2238df094db95821ff740766d9

Observation 9adc9177-a2b2-424f-8f65-f755a66aa877 · outbound

This paper cites Evaluating Large-Vocabulary Object Detectors: The Devil is in the Details.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Evaluating Large-Vocabulary Object Detectors: The Devil is in the Details

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.030102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.030102Z digest=sha256:8513d3198106cc2c406fae337ccd46416d637e9f4689aaa6543efc7a9298a81b

Observation a2f69265-b911-483d-a4de-259bd3085b10 · outbound

This paper cites RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.034916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.034916Z digest=sha256:da578bb70d25d6adbdb2c10baa93dc8949b232698d58829576e753252d3e28cc

Observation db55ea79-7b36-47a2-b88d-100efad54b0a · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.039873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.039873Z digest=sha256:84f500572a88e7d42171e9b616f3e14f3d0e2c566b4e2ce1df4d0eddbabf5d47

Observation a2c326b4-a744-4bb0-bdc7-95995fa1fafa · outbound

This paper cites Favero, L.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Favero, L

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.044221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.044221Z digest=sha256:01a8a5721ee13a3921c4a8818ef0b86e5b349c4de49ad37fc73c6ca8a707e8ab

Observation 5052f0c0-c76e-46bb-bea3-dac339e76a4f · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.048986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.048986Z digest=sha256:9df6f0324780fb5fc3ca8dd6754c87baac5835dfb19a5f4d5d50880642e2367d

Observation 5cae555b-9c03-4fa3-ae0f-af6141fd5145 · outbound

This paper cites Girshick, J.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Girshick, J

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.053903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.053903Z digest=sha256:e649753962691e1acd17befcf17d18651f7f306e98baaf57442b4ffa23465b1d

Observation b32a54f5-af38-41f5-a890-2daa442ea00c · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.058226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.058226Z digest=sha256:843622f7bed82faf44aa56a7522622850ba0e3bb962349862a999b7116a65389

Observation 698ca582-b5e9-483e-950e-fd79b10bf7c0 · outbound

This paper cites Aligning Language Models with Preferences through f-divergence Minimization.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Aligning Language Models with Preferences through f-divergence Minimization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.062540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.062540Z digest=sha256:59cc6a0f9dd5edc886823dceb607954fce2428e0359f01321b5ae8f920fd6656

Observation f577cb62-6a42-4cc3-a61f-f0f2f6702c37 · outbound

This paper cites Reinforced Self-Training (ReST) for Language Modeling.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Reinforced Self-Training (ReST) for Language Modeling

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.067108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.067108Z digest=sha256:e4381e963293dc3ff5938a71d1834b2152aa357038edfa4c5609f761d1b97fe4

Observation 68106966-7b72-4cba-a9b4-f035876b65f1 · outbound

This paper cites Gupta, P.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Gupta, P

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.071728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.071728Z digest=sha256:18813fe43143d5277557de8d51eaaf3f08b49b9cd0034d26b369ad781234f75a

Observation 2eca5a40-89a6-41a3-854f-d06155598b7b · outbound

This paper cites Hattie and H.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Hattie and H

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.075865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.075865Z digest=sha256:7aeed28d7699e2073956c5cf63aad63c4ca7ababcf05af5dea0c96745fc64e36

Observation 1e843c24-4400-42a9-8c2d-29f2ad650c82 · outbound

This paper cites Foundation Model for Advancing Healthcare: Challenges, Opportunities, and Future Directions.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Foundation Model for Advancing Healthcare: Challenges, Opportunities, and Future Directions

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.080759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.080759Z digest=sha256:6b5d1db2430354dbafbbf453669c300c6f464aac2eeb0cb0b070df84367f5739

Observation 3b03dd56-acf9-45ee-a62e-ba5f1eb5c64b · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.085552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.085552Z digest=sha256:e398c971546b026b1473b1d0ca5e4c2197ac8f15275f679035b5302255be896a

Observation b6fd9dc4-654f-4009-9412-cb8348171412 · outbound

This paper cites Meta-Learning in Neural Networks: A Survey.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Meta-Learning in Neural Networks: A Survey

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.090793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.090793Z digest=sha256:a611cf388921e4f698919e65ddae3a84061a78b97bde0da80d3180b853b59a27

Observation 8cb71eeb-2f44-4129-8371-3e4fec8a6962 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision LoRA: Low-Rank Adaptation of Large Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.095790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.095790Z digest=sha256:932d68c5bdb13d945371f17a4f5d9d8b55eb884417b757054288d473d7f72615

Observation 86a9a661-7f35-4a24-97f5-3305cf2a3f51 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.100708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.100708Z digest=sha256:0ab4e49bd65d378986eb6669a247a27cc38ffd18343bc631e342c62aff99e953

Observation 23b50291-d88b-466a-8124-7a0bce5b3ec8 · outbound

This paper cites Mistral 7B.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Mistral 7B

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.104820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.104820Z digest=sha256:070dacd9d691247384c875971f9d8f2f5a5b7c50016738bd2a10da4c75143e8f

Observation 037b3905-0924-4610-a01c-dcac07746564 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.109247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.109247Z digest=sha256:99a9ea54b9dfc0967ddd471e6263224b9b9b06db4f8b0fad3f5a99c30b0cb77f

Observation 610cf59a-1b5b-4248-8419-0feacd8cdb6a · outbound

This paper cites Self-MoE: Towards Compositional Large Language Models with Self-Specialized Experts.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Self-MoE: Towards Compositional Large Language Models with Self-Specialized Experts

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.113712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.113712Z digest=sha256:3cd7e4af9fc3dc2890e4037692fe53c947711892f48a4f94f6c9860ee5dd4038

Observation 2f54bac0-1e87-4bd7-88f0-3e37e75f886f · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.118025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.118025Z digest=sha256:86e6d11086ddba9f34f6a404b4f3fab951177e11c8539519cd62b0bad94907bc

Observation 9eb64d4d-e5bd-4d12-8e1d-ca38b30f3a2e · outbound

This paper cites Kazemzadeh, V.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Kazemzadeh, V

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.122652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.122652Z digest=sha256:c81d4ee096312c2b09218fea7d251c1545ffb25ed43d5f132c371c9c307305b0

Observation 3b02740c-0921-4a76-9d19-b0fcbcad3f8d · outbound

This paper cites Kiefer and L.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Kiefer and L

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.126990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.126990Z digest=sha256:4b1cc2d288508d867b84eb9b07ca80253f965069562edb0e07b29570cf24d452

Observation d327302b-3b75-4d4d-8caa-a6c84789e86c · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.131178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.131178Z digest=sha256:a7171d42adbc197ef1647462535bac6792571d1693a0119e6823d1ed584c6e25

Observation 1b8325a1-a876-4e4a-96ef-2575662909ff · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.135371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.135371Z digest=sha256:312228266c6971eab19cd22ce8c4c1550bd7bdb1381962b46de396b675978e58

Observation ee277cb2-d287-40fa-b6ce-2d5dba39fffc · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision LLaVA-OneVision: Easy Visual Task Transfer

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.139694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.139694Z digest=sha256:17b7b3271b5a38aae75bbc6cb906a6e8c05ce10d3f82293a7290a2e58ed3baa9

Observation cb2fafe6-72c7-4b7b-8f58-a838f84a00eb · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.144811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.144811Z digest=sha256:b6b48f9f8feecbcbeb955f01059e692630747ff5b18995edd493559c63eacadd

Observation bc133f54-8fdc-4192-8e62-de5ed743dd83 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.149220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.149220Z digest=sha256:dfdc845e26c88664ae872f99273e3a07805db936ff17e8720f924b311e11ffab

Observation 5631aa0a-b0c0-4041-b00e-ce2faefa1066 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.153357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.153357Z digest=sha256:3319bc84692dbcf73cb2196fcc0c9c6526e93142f7eab83c8696e22bdf176e6d

Observation 78737c36-eb42-4e24-9a99-722fe21f42f0 · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Evaluating Object Hallucination in Large Vision-Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.159313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.159313Z digest=sha256:bbe380d4fce1f0713ba8d4ed0e305b5d7a493f7f1061ce04bcdcbd8e52321ff9

Observation 5ba40cf8-7280-4d1b-a358-e40855b3a362 · outbound

This paper cites Meta-SGD: Learning to Learn Quickly for Few-Shot Learning.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Meta-SGD: Learning to Learn Quickly for Few-Shot Learning

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.164496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.164496Z digest=sha256:2d446e635fd48a73fcc2ec0970b24c12b20debb5d597f83c5221860e36bacc8a

Observation 867056cb-d36c-432d-a46a-5e5117686c13 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.169519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.169519Z digest=sha256:371d32902a24c86bc222f08b3e183f971a76577bc23127f91bb5062f9a60e025

Observation 39b8c9e1-c9b8-4868-a147-d4d1fd770d66 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.174740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.174740Z digest=sha256:737d997b5ae73f3e8db68acb1cc3677864052efd2839ad6e3dbfa2cc48b2e877

Observation 17b11415-edfa-4406-9077-25d3f6030d6a · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.178963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.178963Z digest=sha256:8ffa2e40f07f5a8cdbe80a40366d956a5b7191d9846be7c2ce59f9c77fbb0f26

Observation 0c16ae80-772e-4bc9-ac9e-e8efef9a8723 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.183587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.183587Z digest=sha256:d0d0a8da8d1db22272e1d0dc99bccf9336ac6f593d058636cd41eab873ebb3be

Observation b45e967d-0689-43c1-951d-b377fa5b91ba · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.189707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.189707Z digest=sha256:031860a901761582cabfd16ad986d7633e0cc32f3f3b475ec65b60c89b33def7

Observation ea07be76-999d-477b-8c8a-fc21358cd12d · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.195100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.195100Z digest=sha256:83fc691c0b44c756a0aead7193d39f9289a3e709e7ff18a2a14208f1f5ae2ddf

Observation 61207065-261c-4e15-85d2-d48cc2b14958 · outbound

This paper cites Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.372992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.372992Z digest=sha256:1d771bd7e37af22a5ab28be3ebde3dcb71807aeedb7512c74f432e1a9d38bd42

Observation 845a080a-1088-4a7c-b38e-b51a555704dc · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.378684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.378684Z digest=sha256:2c4de6279e3408b7e7234375a76455bb3113357f405f78c8480e5348430dc391

Observation 91a31114-c52b-481b-9360-3f8f2a83cffa · outbound

This paper cites Madaan, N.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Madaan, N

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.383422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.383422Z digest=sha256:49a1d0061e51f9875062b16fa308c103874f37c9e0c8bd91c01d7010bb4e8c92

Observation d3da44bd-ce66-4737-8f13-cb77737f3385 · outbound

This paper cites Oquab, L.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Oquab, L

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.388417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.388417Z digest=sha256:438f42ea6abd16454687e725af40ec764fa330711993eca2ffed6f27c2da7be9

Observation 83b172f2-c784-43df-af09-ab8179ac6f8e · outbound

This paper cites Training language models to follow instructions with human feedback.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Training language models to follow instructions with human feedback

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.393161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.393161Z digest=sha256:61077e92e74afeb62edf305944ed3039f84d9f1d6ad543c0252b0f07daac7f41

Observation f129c070-7340-42e9-9204-5b344f73b326 · outbound

This paper cites Papineni, S.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Papineni, S

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.398610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.398610Z digest=sha256:f4c8bc4c5858d6a8ac8e9a45575c06f06330853719b903ad61c17928b33dd978

Observation 46d47fc1-be05-47af-b5c2-de0f89acb80e · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:37.027583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T21:34:35.404501Z digest=sha256:e671df21051e2c2e38e61ebf8875cbae79d28e2d97e9f9230f320d476795f50f

Observation 61e5d573-b8a4-4fba-b07a-c8bb7e0464d1 · outbound

This paper cites Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.410160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.410160Z digest=sha256:90988e3195341ed429cdc9c8bd760249aef91f5309c6f1b0b8e217d3da6ff910

Observation 57d8e92c-08d4-493a-aade-2acd2b7ec86e · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:37.008257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T21:34:35.417009Z digest=sha256:390baf6e858b34cf2331f71cc02875a65a94551e15c5b0885c77f84425141f50

Observation 0f4293f2-33f2-494f-9a2c-110b6374ac53 · outbound

This paper cites Peters and S.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Peters and S

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.421983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.421983Z digest=sha256:55b43e99c4d41ddf46a89acf5813514e3c71d443468d8a0d6662a02274e67707

Observation f11008db-3ecd-42dc-b569-bac88f915e96 · outbound

This paper cites From Concept to Manufacturing: Evaluating Vision-Language Models for Engineering Design.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision From Concept to Manufacturing: Evaluating Vision-Language Models for Engineering Design

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.427447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.427447Z digest=sha256:4d72d5f31f6f644c51f01b11e8e52f298fbc87337f9dc47edb8029fdf705a734

Observation 9fe61691-2c1a-4d29-99a5-9929d98c2684 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.435022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.435022Z digest=sha256:f08aea21b6182562b6643590019588d71bc1e2dc56338f4f4454385bfe029e08

Observation a41a9d37-9ca0-461b-9249-475131ffde94 · outbound

This paper cites Radford, J.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Radford, J

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.440302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.440302Z digest=sha256:7582390fa384cbaf01344ce2d36d63404ac634b9dc4ef9901e52de39e7d9a474

Observation be003b50-e0a8-4d17-8689-fcba51ad3b71 · outbound

This paper cites Rafailov, A.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Rafailov, A

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.956900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T21:34:35.445372Z digest=sha256:7b257016c559c072fabe49f22f49a240efe04c431aac040b4e9b209ea0e02994

Observation 1ab5e6f0-e331-4e55-beb0-0b0d50c531e6 · outbound

This paper cites Vision language models are blind: Failing to translate detailed visual features into words.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Vision language models are blind: Failing to translate detailed visual features into words

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.450573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.450573Z digest=sha256:1a88ce52d38810be318c10b3eab13f15e40b302158ebda9e45b4e9184fc4f7bb

Observation 126f1b1d-4ab7-4e86-8581-e6c2e14c3c57 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:36.940756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T21:34:35.455262Z digest=sha256:eaa683c7aab58dab26e9df5282445b9468ccbc98a456a3baaca6637575cc3adb

Observation e5621796-6727-4ed6-b694-2dfa5474df98 · outbound

This paper cites Tapered Off-Policy REINFORCE: Stable and efficient reinforcement learning for LLMs.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Tapered Off-Policy REINFORCE: Stable and efficient reinforcement learning for LLMs

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.460399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.460399Z digest=sha256:b02b1f472d7a252cd37445ad1143de10a33ff4d8f6050c6a153a4b4161981401

Observation 10f49899-a4d2-4c1b-a778-eec86572a4d2 · outbound

This paper cites Saikh, T.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Saikh, T

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.925038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T21:34:35.465293Z digest=sha256:88f043a909d830676ef1d5a04eef3c2cbafec61cada195fa0744b2af0bd82deb

Observation c953b626-e876-4aa4-940e-8fb7a6d1ba7a · outbound

This paper cites debiaSAE: Benchmarking and Mitigating Vision-Language Model Bias.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision debiaSAE: Benchmarking and Mitigating Vision-Language Model Bias

Reference 75

Resolution
verified exact
local_arxiv, observed 2026-08-10T21:34:36.163169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T21:34:35.469776Z digest=sha256:f9e3111438f8b8d7a67f485f2ac60d037eb65a1d570b5cc05b1386548a09e4d7

Observation 28e199e8-3641-487a-a636-1492a94fbc00 · outbound

This paper cites Schaul and J.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Schaul and J

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.907427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T21:34:35.475072Z digest=sha256:a60927a96ca2603cc97696de61e3202b30c424ede6bf75cf755b6772b4bd8953

Observation 2ae740d4-afdc-4fcc-b674-b3b7deaec2e7 · outbound

This paper cites Schmidhuber.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Schmidhuber

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.889969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T21:34:35.479819Z digest=sha256:307a13c634260e863eb8091ca6b9e56411c1af743539d8c3314c87a567b7fd0d

Observation e215bbec-0b46-4a8b-8181-8509df3e1371 · outbound

This paper cites Schmidhuber.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Schmidhuber

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.875814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T21:34:35.485458Z digest=sha256:0efdf392a3a5eb6fb7e4d1b0c9fd80fedd9f44103008d76eb17c96dc7e8998bb

Observation fcabd321-b44f-4286-88ca-3da27a5b362f · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.490872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.490872Z digest=sha256:aaca49a5825f9cbfaa2a669c69255e0ed2b0e3d58b0d7fe87b36123ab3a8982f

Observation 46d187af-f7c2-4f15-ab81-a909bd26285e · outbound

This paper cites Shinn, F.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Shinn, F

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.847748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T21:34:35.497061Z digest=sha256:f959743b8061c4dc119b8557b5dbe6f5080c5256838654e9ff9046880219f3e1

Observation 2e1b7217-daaf-4198-8a5a-83d67092de2e · outbound

This paper cites Silver, A.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Silver, A

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.826926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T21:34:35.501869Z digest=sha256:ce60556aece896122952a78fce87a0f2bda8ee672548ea96edc1f9007a7054a9

Observation 30721518-8d62-404a-a00c-90492671ca67 · outbound

This paper cites Silver, J.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Silver, J

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.506314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.506314Z digest=sha256:115856de5bff9f0143cef5c950af1c502778aa3e162edef9906d08be726ed2ee

Observation 622fd121-1d72-4c7a-8940-df7fe757edb6 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 83

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:36.794520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T21:34:35.514137Z digest=sha256:dd422df84687268ab65471d6f4e35a457ebd90572180de39db55e8cc0d2f5bd9

Observation de6afcbe-64cb-4519-85b2-2a193a16f7bf · outbound

This paper cites LAB: Large-Scale Alignment for ChatBots.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision LAB: Large-Scale Alignment for ChatBots

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.518320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.518320Z digest=sha256:ab60a82df620e22ecbd14fe76e04ff77cd9c39a0f0093721556be4783db576b5

Observation 5aa47013-f822-4ba3-bd82-29e57d74f91e · outbound

This paper cites Aligning Large Multimodal Models with Factually Augmented RLHF.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Aligning Large Multimodal Models with Factually Augmented RLHF

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.523168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.523168Z digest=sha256:381d179db68f06c7710d42a1ead04c28911f7602e57fd13b03652659a13a942c

Observation b88f8223-f487-4038-8072-35647fa28bd9 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 86

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:36.778848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T21:34:35.528275Z digest=sha256:47c510821cf17a5511a55ff240c56f8b97bbb5b8df8a43610348e576ade82308

Observation 38292c0a-2459-4ddd-80c4-20e34317a08d · outbound

This paper cites Tenenbaum and E.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Tenenbaum and E

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.764526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T21:34:35.532809Z digest=sha256:567fd9dc72907d18beb3bdd56486892093190da0acc6c2d758ef3eda892a2bb1

Observation 49f63fe2-1ca2-4fc9-8902-aafa98facb7d · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.538223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.538223Z digest=sha256:83f4e91d8f182891f37dfa32c5f651a6bda110f7779cefbae81ddab4c0f5950c

Observation 07da72ab-9485-4eff-b109-686cc87aff44 · outbound

This paper cites Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.543022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.543022Z digest=sha256:5e5e9f953892391fca2438cba733542883122d278f5bfc5ec188d94be9c721d7

Observation 84fd5bb2-6076-4f32-ba69-019c6427ff41 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision LLaMA: Open and Efficient Foundation Language Models

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.547860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.547860Z digest=sha256:3c80c6fcb0bbb84b14157cfa527e4bfd1b54488091c7ef490b8e98b16059c1a8

Observation 9c8b10f1-ac13-4b87-921f-85d627c45a9b · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 91

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:36.742039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T21:34:35.552315Z digest=sha256:fc861f70b98d219f7f15b818f82c859ecb186f2d1769f70f80414ea57ed2a34f

Observation df5ac0f1-faca-4b31-aae2-791b4ddaeca2 · outbound

This paper cites Vedantam, C.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Vedantam, C

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.556464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.556464Z digest=sha256:7128467c191f63e16730881d9e5b29c3231f05395f7a2ac22680becca8e7ca89

Observation ca17a4fe-520b-41ec-9472-18fda0b16f5b · outbound

This paper cites Contrastive Region Guidance: Improving Grounding in Vision-Language Models without Training.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Contrastive Region Guidance: Improving Grounding in Vision-Language Models without Training

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.560452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.560452Z digest=sha256:bf563ef7d448a0bbcf2660dd24ddd625d1d21ae606fbf16ea1004c55dd20639a

Observation 407d1acc-13bc-4ac8-89f7-d4b7fceafdf1 · outbound

This paper cites Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.564572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.564572Z digest=sha256:b67727c9af24232186965594c48299014723c75183aedc4ced1c8d058594e236

Observation 8066381b-7bc6-4b24-8c16-dff82742db3d · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.569125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.569125Z digest=sha256:a70387a4a9c2c948819a33e36b8c073a99e6a98774f0fdeedca7b5c7741d34c1

Observation fb3bd056-b885-4e09-8458-b7ca1a314824 · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.573795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.573795Z digest=sha256:d8f737c36c237edc32985a7909ac7435ff1734cb123ea2b99ba7d5f3181f694f

Observation eb07f118-bb7e-425d-9a30-e11a15fd5a70 · outbound

This paper cites An Explanation of In-context Learning as Implicit Bayesian Inference.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision An Explanation of In-context Learning as Implicit Bayesian Inference

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.579540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.579540Z digest=sha256:9b7ee4d1693a71f7c788903d7d4d587bc8efb0418c923a6cc101662244b9823d

Observation 8d81d9d4-a9d7-4acd-8e69-4539c64f4c12 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 98

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:36.719148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T21:34:35.584821Z digest=sha256:d9109371c6cb92a6a6c2ecc9c39bf020de90755de323ab4b7fd8d56967c7c8c4

Observation f3ccbcf4-4a0d-4a35-bdf9-792dde0fd88b · outbound

This paper cites Qwen2 Technical Report.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Qwen2 Technical Report

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.589733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.589733Z digest=sha256:b8f25da6e886fa4718e960f6e2cf446270e0f4ea76abffe18e8bc3c7b0bd54a9

Observation 77549c59-4f78-4c74-b2bc-8523a32e7d24 · outbound

This paper cites ReAct: Synergizing Reasoning and Acting in Language Models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision ReAct: Synergizing Reasoning and Acting in Language Models

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.594979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.594979Z digest=sha256:8e31253d96fa103ff924661899126a731912610f427380f69f246c09b3c1af2b

Pith citing papers

Observation e39f026f-64ca-46ac-9e31-8b1b235e05e1 · inbound

Grounding Hierarchical Vision-Language-Action Models Through Explicit Language-Action Alignment cites this paper.

Grounding Hierarchical Vision-Language-Action Models Through Explicit Language-Action Alignment Feedback-Driven Vision-Language Alignment with Minimal Human Supervision

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:35:51.423869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T19:41:58.673348Z digest=sha256:6ccfadb6b9f7d4443494253a72a6c9dfcefd7cd2e2ba68a0f1556d1a1d89c20c