Pith. sign in

Paper Citation Record · LEDGER

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision

As of 14 August 2026, this Paper Citation Record lists 100 of 118 outbound references and 1 inbound Pith citation observation for arXiv:2501.04568.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.04568 v2

Coverage vector

measured 100 of 118 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:34:35.594979Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T19:41:58.673348Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T22:35:51.417589Z

Reference resolution

100 of 118 outbound references displayed

  • verified exact1
  • verified fuzzy8
  • unresolved91
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 878090a4-d779-44d2-b6f8-1d0a6d23754b · outbound

This paper cites GPT-4 Technical Report.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.931476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.931476Z digest=sha256:2d4e943794e00dc1c54270e95a5073a59d0cb0c9cc2ce197dd3457d0178a3f72

Observation 32d9879f-8e91-4db7-90ee-899051b96cc1 · outbound

This paper cites Agrawal, K.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Agrawal, K

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.937048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.937048Z digest=sha256:dde366a770dfdb66b4e42adb810d8fcb03b3e21bddfe429d0096eacc40ad9e70

Observation 3694a681-b2d0-4114-8ced-70895428e77c · outbound

This paper cites Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.941481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.941481Z digest=sha256:ea7f632be1fa027f3665b1cdf09513dfb1a2dc5fdbc0aaccf89a3e099a73cf52

Observation 4ffd809b-61fd-4843-96ab-870d088a292e · outbound

This paper cites What learning algorithm is in-context learning? Investigations with linear models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision What learning algorithm is in-context learning? Investigations with linear models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.947407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.947407Z digest=sha256:3c3385c7812fb2f4c6f218fe637611c90479e4389ed80c8d158fd469d6c4b3d3

Observation 7474ee78-909d-4701-9a8f-384dc7bab561 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.953338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.953338Z digest=sha256:0c37e27ef0d2140544f86d3c5f0ea9d931ab813a621b18ed3c534bec5e88af78

Observation 71d04822-d2d4-4bc0-9c9a-8ad64dc95e1a · outbound

This paper cites Anthony, Z.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Anthony, Z

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.958760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.958760Z digest=sha256:70e9b621dc16d62023abc5a1da406d391496be58154d1ea75197b9824e31cec9

Observation cdab2c24-dd18-40d9-89c0-c0f18369e1fe · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Constitutional AI: Harmlessness from AI Feedback

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.964208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.964208Z digest=sha256:7d6ea4fdd607027f41b2caa21cfc95c5b7ee166aa3d217f9ea6872af458033e9

Observation e6885b9d-e499-418e-8676-15396940c355 · outbound

This paper cites Hallucination of Multimodal Large Language Models: A Survey.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Hallucination of Multimodal Large Language Models: A Survey

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.973845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.973845Z digest=sha256:98442b1d4628d3788df1fcd402e605d657fcc112085551398dc46adb62623a0f

Observation eb961fdf-bffd-453e-9911-eaa1254e266c · outbound

This paper cites Banerjee and A.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Banerjee and A

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.978603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.978603Z digest=sha256:795f57b73b0b5d124ee3297afc651b35fdf3302e7528beb06cdff5cef370c1fd

Observation 34624016-5e5f-4eb1-980f-e0ba1c339fe0 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.982785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.982785Z digest=sha256:fc211db2ca5b8d07ae503284e87d13a98f079f8c0c34505dfbaeb9742b06f601

Observation a2d0aaeb-d4cd-4ca1-8fbe-5d464f3bd4b3 · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision PaliGemma: A versatile 3B VLM for transfer

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.987003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.987003Z digest=sha256:4c64bb54b4f72a8a825791d312ef5dc5a5c1f122d2b4900afde3a99467677e72

Observation 9d087100-76dd-4a0e-950f-c2b6f169a9fd · outbound

This paper cites An Introduction to Vision-Language Modeling.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision An Introduction to Vision-Language Modeling

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.991443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.991443Z digest=sha256:ba3dd70391a489c5db3e4c64f993ea2ab84cf9dc19c37183a8938df5d7c6060a

Observation c495a7b7-1008-4907-8641-d6b95aabe1bc · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.995902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.995902Z digest=sha256:5abe5cd8a6d6efe256fa8642f118d3d9c7f424c0468b7156dc7f655e6f95aa11

Observation 5bf2d01e-dbba-43ac-9d0e-333859fa2f8a · outbound

This paper cites Carion, F.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Carion, F

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.000515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.000515Z digest=sha256:2168d59c4e21ad6ff41b6e9711b9f5fdff8153c303b71b3007228a96c2e8f9ff

Observation b5c61565-a26f-4a92-9daf-179399d0330c · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.004387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.004387Z digest=sha256:dd3f3d2e52585d51f168d82e5e05f464e23f983926ace8f1b286821a198e1cfd

Observation 89342bce-3e91-4aac-bf88-d82c344e4c1a · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.008817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.008817Z digest=sha256:a5cf37d7056ef880b97d302e496a0e7469ce0b37a125eddcc988fba314e00dd0

Observation 850fcc1d-5643-4686-9f1d-0fddf665bf05 · outbound

This paper cites Chiang, Z.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Chiang, Z

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.013082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.013082Z digest=sha256:c2e0db7803f78ce46e92a2baf699ab5619405c91551283a12239c7add144c801

Observation 2046ea3d-8e12-45da-97ec-a0f27ca6eb92 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.017220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.017220Z digest=sha256:30ab1528729f4b12a9a9c67f925a0838fa1c79ae55e7d8b5ac31bfee7b94007a

Observation fc4d0547-f5f6-462a-a702-3fa95d0eca58 · outbound

This paper cites Collerton, J.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Collerton, J

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.021685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.021685Z digest=sha256:b61eb91132d77230850930c75c81875e02a89d98145c83371542efc8a521cbaa

Observation 56f71cab-a7d4-463f-8650-3f89730ac78a · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.025848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.025848Z digest=sha256:5f7522eab8ac1298413de42d7dc1cfb51cc1c94c7cb84f88d026e91874d586b5

Observation 9adc9177-a2b2-424f-8f65-f755a66aa877 · outbound

This paper cites Evaluating Large-Vocabulary Object Detectors: The Devil is in the Details.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Evaluating Large-Vocabulary Object Detectors: The Devil is in the Details

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.030102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.030102Z digest=sha256:667057a0f04847635dd19a399e253335fca8d67aee9ec59a80a0eb455744925b

Observation a2f69265-b911-483d-a4de-259bd3085b10 · outbound

This paper cites RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.034916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.034916Z digest=sha256:343a82b5325d8b9c341070766d45d9323dd663e7612a6ecb618de05fdc0bf11f

Observation db55ea79-7b36-47a2-b88d-100efad54b0a · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.039873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.039873Z digest=sha256:b6720647894fc513aa272568c1761e94b715782f85baf6079b7af68d39952ed1

Observation a2c326b4-a744-4bb0-bdc7-95995fa1fafa · outbound

This paper cites Favero, L.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Favero, L

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.044221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.044221Z digest=sha256:443cc61c6b8c12a0cbb3f09d8d55878ec1489f04e679ef14a0aa0000c741bbc1

Observation 5052f0c0-c76e-46bb-bea3-dac339e76a4f · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.048986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.048986Z digest=sha256:50119b3a53b0a84a1be5dc7fd6644ed76d8758405e70fcf80f6b2b829213e5b3

Observation 5cae555b-9c03-4fa3-ae0f-af6141fd5145 · outbound

This paper cites Girshick, J.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Girshick, J

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.053903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.053903Z digest=sha256:a7be0481e8c22a1b3f9158ff444b572aaabc0e94d7f65f326531fad17db17cc2

Observation b32a54f5-af38-41f5-a890-2daa442ea00c · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.058226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.058226Z digest=sha256:66992503cba1d76dd2fd835ca93f6b660d1196d2aae5b13aa3ea1d7edc74737c

Observation 698ca582-b5e9-483e-950e-fd79b10bf7c0 · outbound

This paper cites Aligning Language Models with Preferences through f-divergence Minimization.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Aligning Language Models with Preferences through f-divergence Minimization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.062540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.062540Z digest=sha256:ae2c2117d0e5c2d9695e026225f47d613a59cabe21b515ad3b5578ea30da4acc

Observation f577cb62-6a42-4cc3-a61f-f0f2f6702c37 · outbound

This paper cites Reinforced Self-Training (ReST) for Language Modeling.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Reinforced Self-Training (ReST) for Language Modeling

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.067108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.067108Z digest=sha256:5bb39535d8e4913cf2a06597703704ae3988a7cffcdd865c07e3fbda2e08c7a2

Observation 68106966-7b72-4cba-a9b4-f035876b65f1 · outbound

This paper cites Gupta, P.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Gupta, P

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.071728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.071728Z digest=sha256:905de0669d5b9fa516d90e8ff7735a2eb2f199c1ce37cd65fa38fd5f704b5c21

Observation 2eca5a40-89a6-41a3-854f-d06155598b7b · outbound

This paper cites Hattie and H.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Hattie and H

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.075865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.075865Z digest=sha256:60a65253aace3c82bb858c4ce97807858d5ca321ce0cc2e8ec9f2f6460f7982c

Observation 1e843c24-4400-42a9-8c2d-29f2ad650c82 · outbound

This paper cites Foundation Model for Advancing Healthcare: Challenges, Opportunities, and Future Directions.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Foundation Model for Advancing Healthcare: Challenges, Opportunities, and Future Directions

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.080759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.080759Z digest=sha256:7dfceab70cf395643fa4f99988a51598c2d00a283707733660dff072bc4c4485

Observation 3b03dd56-acf9-45ee-a62e-ba5f1eb5c64b · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.085552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.085552Z digest=sha256:122c1cafabad60f80dfe32af5b87245a4fbf9eb3ef1e9abe294b4c4149826958

Observation b6fd9dc4-654f-4009-9412-cb8348171412 · outbound

This paper cites Meta-Learning in Neural Networks: A Survey.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Meta-Learning in Neural Networks: A Survey

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.090793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.090793Z digest=sha256:fb0c627da629d40211052cb390e376df36775a24f7a19d34328c60c3ea1bf0ce

Observation 8cb71eeb-2f44-4129-8371-3e4fec8a6962 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision LoRA: Low-Rank Adaptation of Large Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.095790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.095790Z digest=sha256:1fa9f0c7bb9f32d284637f479cc85fe6cff28771defb51dbaddd1317721ec912

Observation 86a9a661-7f35-4a24-97f5-3305cf2a3f51 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.100708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.100708Z digest=sha256:4bd1d62c8f9fff78195f12aaf886362ee91d4fd4982a125f8ed3ee8531a49787

Observation 23b50291-d88b-466a-8124-7a0bce5b3ec8 · outbound

This paper cites Mistral 7B.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Mistral 7B

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.104820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.104820Z digest=sha256:e97417ca5d75598f78b857eef9bc02ef8c51fc61518762263571a736576d79d6

Observation 037b3905-0924-4610-a01c-dcac07746564 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.109247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.109247Z digest=sha256:5b66356fdfb2296d398eba4b668efab8efce6a352ff2db0aa158dd3b7942c2e0

Observation 610cf59a-1b5b-4248-8419-0feacd8cdb6a · outbound

This paper cites Self-MoE: Towards Compositional Large Language Models with Self-Specialized Experts.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Self-MoE: Towards Compositional Large Language Models with Self-Specialized Experts

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.113712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.113712Z digest=sha256:0e3afe9573be13e2271c80f644fff688b8b8b0f7c3e9837fd18c592cae6d8ef5

Observation 2f54bac0-1e87-4bd7-88f0-3e37e75f886f · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.118025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.118025Z digest=sha256:0000bbf0b9382e0979b0ef59fcd4472d0a8290562e39b9edac103732a2e9be31

Observation 9eb64d4d-e5bd-4d12-8e1d-ca38b30f3a2e · outbound

This paper cites Kazemzadeh, V.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Kazemzadeh, V

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.122652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.122652Z digest=sha256:11f6c8842fea65b60b44c1258c0a54d8b55964495cf282604e2b05a7a655f31e

Observation 3b02740c-0921-4a76-9d19-b0fcbcad3f8d · outbound

This paper cites Kiefer and L.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Kiefer and L

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.126990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.126990Z digest=sha256:930e06fa78273df4ced3f9c2f9d869bd5d71c4d2c8edbf9f2ea291cd5587442d

Observation d327302b-3b75-4d4d-8caa-a6c84789e86c · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.131178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.131178Z digest=sha256:28bf16805d2eafa1c9e478478e50eaacb55da4de6399aace5b088886fcd6bda7

Observation 1b8325a1-a876-4e4a-96ef-2575662909ff · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.135371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.135371Z digest=sha256:992180ed904953c83a6fc57117ea79e27ef15382a2a12013365b68b6a2378f63

Observation ee277cb2-d287-40fa-b6ce-2d5dba39fffc · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision LLaVA-OneVision: Easy Visual Task Transfer

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.139694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.139694Z digest=sha256:74adebe6130fce35b0977a0165fa8eb89cf7f2d6e2fe20c0acda830611e1d8a3

Observation cb2fafe6-72c7-4b7b-8f58-a838f84a00eb · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.144811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.144811Z digest=sha256:787d9c116715b53aa601adbee1b5c10b4fa52c3c9c61e0dca48e4626c2b0772b

Observation bc133f54-8fdc-4192-8e62-de5ed743dd83 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.149220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.149220Z digest=sha256:633cd18bff35bd8214e7a99b4c8c1f9438957a16ef1cef8128d6fd1c4445f902

Observation 5631aa0a-b0c0-4041-b00e-ce2faefa1066 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.153357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.153357Z digest=sha256:d979bf916b1091f7629d92ed25919b489800e82a5a49e9d58515de800a237d33

Observation 78737c36-eb42-4e24-9a99-722fe21f42f0 · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Evaluating Object Hallucination in Large Vision-Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.159313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.159313Z digest=sha256:e7f4383388eae451276c15a4ad61867c5513167bdbc8a0b5078842098beffd99

Observation 5ba40cf8-7280-4d1b-a358-e40855b3a362 · outbound

This paper cites Meta-SGD: Learning to Learn Quickly for Few-Shot Learning.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Meta-SGD: Learning to Learn Quickly for Few-Shot Learning

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.164496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.164496Z digest=sha256:0236c18876d06be797b187c08a2de848b1d9deab91fa131074cdbc1965291f5c

Observation 867056cb-d36c-432d-a46a-5e5117686c13 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.169519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.169519Z digest=sha256:ba2dcb85b6eadd7c87d59d9ab76576d5a62feb3d23c9961532317785094cb934

Observation 39b8c9e1-c9b8-4868-a147-d4d1fd770d66 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.174740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.174740Z digest=sha256:2fcf29a1bc4bd2564071a60bda13afcea5bcd54b7fd2a87966b0b31b6a64477c

Observation 17b11415-edfa-4406-9077-25d3f6030d6a · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.178963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.178963Z digest=sha256:c9c1ba3b46409b662a360c9ecc4d8b66c4681503fcd44e2ae0ae77bee1b273cf

Observation 0c16ae80-772e-4bc9-ac9e-e8efef9a8723 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.183587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.183587Z digest=sha256:dd604cc831e898bc328e70189cb34f45dd0366f208ad3c75e7729de9c501538a

Observation b45e967d-0689-43c1-951d-b377fa5b91ba · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.189707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.189707Z digest=sha256:7138799368508a04ad26eb7caea665845a63b566a1ba5a81f9aa56bfdc47eba5

Observation ea07be76-999d-477b-8c8a-fc21358cd12d · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.195100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.195100Z digest=sha256:e95212380a1fa2f6c91c1fde1b9fcc957f0d6acc28dc0ee2dc6976ddd6b310c2

Observation 61207065-261c-4e15-85d2-d48cc2b14958 · outbound

This paper cites Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.372992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.372992Z digest=sha256:e0920bc08c908e997f3af8b56a0e0ab3a3d8bc38c01164a2a7fbac06ea0ea0ec

Observation 845a080a-1088-4a7c-b38e-b51a555704dc · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.378684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.378684Z digest=sha256:70221590842246d1a06f28083d7e70065e50083b5075c13a6862ed62439f213d

Observation 91a31114-c52b-481b-9360-3f8f2a83cffa · outbound

This paper cites Madaan, N.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Madaan, N

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.383422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.383422Z digest=sha256:cf6583cbfeeb4c83a2e09ca79d7b67f7abce71068f2af945c3a60d569fe9bbd7

Observation d3da44bd-ce66-4737-8f13-cb77737f3385 · outbound

This paper cites Oquab, L.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Oquab, L

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.388417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.388417Z digest=sha256:6a702404a93a9984e0e609849b1a366e8de442de3f75eeb3bc972a7812bb5ac1

Observation 83b172f2-c784-43df-af09-ab8179ac6f8e · outbound

This paper cites Training language models to follow instructions with human feedback.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Training language models to follow instructions with human feedback

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.393161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.393161Z digest=sha256:0e12ad5e8be18c620a8c04280f5c263dee5c43dbaa09ea63a8aabb8b0a395caf

Observation f129c070-7340-42e9-9204-5b344f73b326 · outbound

This paper cites Papineni, S.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Papineni, S

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.398610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.398610Z digest=sha256:ed3e08f279ee69b865b98293ff562c3e5a41500818296de913e47060bef93520

Observation 46d47fc1-be05-47af-b5c2-de0f89acb80e · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:37.027583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:34:35.404501Z digest=sha256:03c06bf68ea3099276b39ac478a74a4c4d6f4058c79343afc8769395984a6f7a

Observation 61e5d573-b8a4-4fba-b07a-c8bb7e0464d1 · outbound

This paper cites Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.410160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.410160Z digest=sha256:27dc4dab0074fae36b2b1a0384dd6c112ad2e6f1679f1fcb559622f2a5329da0

Observation 57d8e92c-08d4-493a-aade-2acd2b7ec86e · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:37.008257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:34:35.417009Z digest=sha256:d0f064edb9c50291e9d642452363c5a93832e70b256fc9d53c6e59afdffea08b

Observation 0f4293f2-33f2-494f-9a2c-110b6374ac53 · outbound

This paper cites Peters and S.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Peters and S

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.421983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.421983Z digest=sha256:47853721b63f3d0b46a857348534ef6663d6f92e293ec9b2abcb72708ff46938

Observation f11008db-3ecd-42dc-b569-bac88f915e96 · outbound

This paper cites From Concept to Manufacturing: Evaluating Vision-Language Models for Engineering Design.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision From Concept to Manufacturing: Evaluating Vision-Language Models for Engineering Design

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.427447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.427447Z digest=sha256:5e9571e78bdf38ae55fadb5cba3ba921a7950ebfdb53f16fa12444b072e358e5

Observation 9fe61691-2c1a-4d29-99a5-9929d98c2684 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.435022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.435022Z digest=sha256:923f1fcceea2f0edb640be2dfa1d19edb422c481339a4fb326130c1f53b1c852

Observation a41a9d37-9ca0-461b-9249-475131ffde94 · outbound

This paper cites Radford, J.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Radford, J

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.440302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.440302Z digest=sha256:29be335a8cbf19a5cfb622eae595cbe40ed12a6d00d84aa9b1c7fd634949b86c

Observation be003b50-e0a8-4d17-8689-fcba51ad3b71 · outbound

This paper cites Rafailov, A.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Rafailov, A

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.956900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:34:35.445372Z digest=sha256:dfb22455d7a434c693652185e3be28c31a6db9463fe1306f0713b10bb07510b4

Observation 1ab5e6f0-e331-4e55-beb0-0b0d50c531e6 · outbound

This paper cites Vision language models are blind: Failing to translate detailed visual features into words.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Vision language models are blind: Failing to translate detailed visual features into words

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.450573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.450573Z digest=sha256:959e10ae8e7865fa0721a0d9cbab90200635440deadb379fb67739e928cde55c

Observation 126f1b1d-4ab7-4e86-8581-e6c2e14c3c57 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:36.940756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:34:35.455262Z digest=sha256:c350637923108796247064ede572d14217f895f1f307d7d7dbf68919593b379d

Observation e5621796-6727-4ed6-b694-2dfa5474df98 · outbound

This paper cites Tapered Off-Policy REINFORCE: Stable and efficient reinforcement learning for LLMs.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Tapered Off-Policy REINFORCE: Stable and efficient reinforcement learning for LLMs

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.460399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.460399Z digest=sha256:7693453f03100edc007f85f266661f56eeb4248578e601f5076241f34d775181

Observation 10f49899-a4d2-4c1b-a778-eec86572a4d2 · outbound

This paper cites Saikh, T.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Saikh, T

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.925038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:34:35.465293Z digest=sha256:51401dd050de13222e7e361a95412175f98a62bae3b00670155d059b5003548b

Observation c953b626-e876-4aa4-940e-8fb7a6d1ba7a · outbound

This paper cites debiaSAE: Benchmarking and Mitigating Vision-Language Model Bias.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision debiaSAE: Benchmarking and Mitigating Vision-Language Model Bias

Reference 75

Resolution
verified exact
local_arxiv, observed 2026-08-10T21:34:36.163169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:34:35.469776Z digest=sha256:519f277c5ff1517c370ab66e85c2f02978ba46c46e42c1136d2dc30b6de3f098

Observation 28e199e8-3641-487a-a636-1492a94fbc00 · outbound

This paper cites Schaul and J.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Schaul and J

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.907427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:34:35.475072Z digest=sha256:d2929f8a32c709eca270e52ec89090d74db9a219469cf4701d03183f08e7c9bc

Observation 2ae740d4-afdc-4fcc-b674-b3b7deaec2e7 · outbound

This paper cites Schmidhuber.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Schmidhuber

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.889969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:34:35.479819Z digest=sha256:2a74974460f3bb500eb39edb390b25b341acc2ab31ce0bf08015af4db4d04fa2

Observation e215bbec-0b46-4a8b-8181-8509df3e1371 · outbound

This paper cites Schmidhuber.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Schmidhuber

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.875814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:34:35.485458Z digest=sha256:f515940686770dfbd43c38eae4e0c3a1d81952ebbf4da24128a6fc93da6b3856

Observation fcabd321-b44f-4286-88ca-3da27a5b362f · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.490872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.490872Z digest=sha256:5f521ee0e97b86d20943bc10f69fd25d0f2bf4a345c2a00f85620c785699aa39

Observation 46d187af-f7c2-4f15-ab81-a909bd26285e · outbound

This paper cites Shinn, F.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Shinn, F

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.847748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:34:35.497061Z digest=sha256:a9e39eccc0f0d0c3350bc217463d5e4a9e71a566dac776cf1db2af2cd033892c

Observation 2e1b7217-daaf-4198-8a5a-83d67092de2e · outbound

This paper cites Silver, A.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Silver, A

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.826926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:34:35.501869Z digest=sha256:01f04b414b104a10eb1c75f677c50479d7bc575b77f86116deaef92448567ce6

Observation 30721518-8d62-404a-a00c-90492671ca67 · outbound

This paper cites Silver, J.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Silver, J

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.506314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.506314Z digest=sha256:aabb33cf4846e8df91ec697638c17ca8032e77295d663cf1cabab4e3c5ab14a0

Observation 622fd121-1d72-4c7a-8940-df7fe757edb6 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 83

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:36.794520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:34:35.514137Z digest=sha256:5edaf37d2ced4a80ae2cf20e4d73eb784942ec11a422aba4ddd8d173d6b2ba8d

Observation de6afcbe-64cb-4519-85b2-2a193a16f7bf · outbound

This paper cites LAB: Large-Scale Alignment for ChatBots.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision LAB: Large-Scale Alignment for ChatBots

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.518320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.518320Z digest=sha256:30ed66c06dc47c3d5a6430e7e8cadf8adc5da064fd4a19c6aa046dc8d4dc2d7e

Observation 5aa47013-f822-4ba3-bd82-29e57d74f91e · outbound

This paper cites Aligning Large Multimodal Models with Factually Augmented RLHF.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Aligning Large Multimodal Models with Factually Augmented RLHF

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.523168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.523168Z digest=sha256:06d89210c4d592ea1179ab7b99e53dcbcc13952002d0b1d0c01de64bc237388c

Observation b88f8223-f487-4038-8072-35647fa28bd9 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 86

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:36.778848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:34:35.528275Z digest=sha256:076cec6975ffbd7bfae6204611111196178975c3c94a2dc318e6a8e37805783d

Observation 38292c0a-2459-4ddd-80c4-20e34317a08d · outbound

This paper cites Tenenbaum and E.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Tenenbaum and E

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.764526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:34:35.532809Z digest=sha256:97bbfe73093d0fc9fbcde544f7b92679c50b78cd529d5f559eb544e9615e5833

Observation 49f63fe2-1ca2-4fc9-8902-aafa98facb7d · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.538223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.538223Z digest=sha256:f80d1ae0152622782241708ad20a5f31598c1f1cfd50aa2eb6f48692e2e6a41d

Observation 07da72ab-9485-4eff-b109-686cc87aff44 · outbound

This paper cites Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.543022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.543022Z digest=sha256:28b5bb1478e088dd062f1eb12173961327270e5326806b2b910d1c43be605937

Observation 84fd5bb2-6076-4f32-ba69-019c6427ff41 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision LLaMA: Open and Efficient Foundation Language Models

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.547860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.547860Z digest=sha256:9c690d90ce2127b3389b0314cd2635cd3aa013c913547e218b05da341c0b85e7

Observation 9c8b10f1-ac13-4b87-921f-85d627c45a9b · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 91

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:36.742039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:34:35.552315Z digest=sha256:4906f88b9e5c9faaa3b24d45d1ffac33cfea7922a9e73b32f4acda72fff1ab5c

Observation df5ac0f1-faca-4b31-aae2-791b4ddaeca2 · outbound

This paper cites Vedantam, C.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Vedantam, C

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.556464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.556464Z digest=sha256:a2e5789d7e5319d164ef99b52a6b7af9deff20bd265d8493a039a987bf5833b1

Observation ca17a4fe-520b-41ec-9472-18fda0b16f5b · outbound

This paper cites Contrastive Region Guidance: Improving Grounding in Vision-Language Models without Training.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Contrastive Region Guidance: Improving Grounding in Vision-Language Models without Training

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.560452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.560452Z digest=sha256:d94f9ef3bcba2a40d900cd02108e05c5be2ef75ea4b3f3cd3deefb3614af73f4

Observation 407d1acc-13bc-4ac8-89f7-d4b7fceafdf1 · outbound

This paper cites Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.564572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.564572Z digest=sha256:0ae519c82bcf60e4281f6fe9c4c7e2dda54e447a7ed180c9af17b43c37bd4d6a

Observation 8066381b-7bc6-4b24-8c16-dff82742db3d · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.569125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.569125Z digest=sha256:b376f5b14b61725c82834c938e2ea37de87394c50f921dbd6b19ad3aec6499e3

Observation fb3bd056-b885-4e09-8458-b7ca1a314824 · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.573795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.573795Z digest=sha256:45a9502b5bdf61607c01cfee1ffb9ed381d73fe7066b453d6782c63a457abed0

Observation eb07f118-bb7e-425d-9a30-e11a15fd5a70 · outbound

This paper cites An Explanation of In-context Learning as Implicit Bayesian Inference.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision An Explanation of In-context Learning as Implicit Bayesian Inference

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.579540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.579540Z digest=sha256:bbcc9d01d8fba194b6a9e3573922a2a427429e44e78adb7395bc6510f7d73d3b

Observation 8d81d9d4-a9d7-4acd-8e69-4539c64f4c12 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 98

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:36.719148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:34:35.584821Z digest=sha256:56e2d9a337d6c08ba193ad26f0b1d1742e978747534b11508ac0478b403e2875

Observation f3ccbcf4-4a0d-4a35-bdf9-792dde0fd88b · outbound

This paper cites Qwen2 Technical Report.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Qwen2 Technical Report

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.589733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.589733Z digest=sha256:db0c7756e30f85d441523f7fb99474f03c2f2c5d28045c9c322a5398294ede10

Observation 77549c59-4f78-4c74-b2bc-8523a32e7d24 · outbound

This paper cites ReAct: Synergizing Reasoning and Acting in Language Models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision ReAct: Synergizing Reasoning and Acting in Language Models

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.594979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.594979Z digest=sha256:9071a07183f99adcf24bb0cb96afe97afdfec96b3d43af4dc7f96d8749f1b769

Pith citing papers

Observation e39f026f-64ca-46ac-9e31-8b1b235e05e1 · inbound

Grounding Hierarchical Vision-Language-Action Models Through Explicit Language-Action Alignment cites this paper.

Grounding Hierarchical Vision-Language-Action Models Through Explicit Language-Action Alignment Feedback-Driven Vision-Language Alignment with Minimal Human Supervision

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:35:51.423869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T19:41:58.673348Z digest=sha256:74a24cd598d199a6ae56d3846f30b701df22102e8068e2d14545232a8753338a