Pith. sign in

Paper Citation Record · LEDGER

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision

As of 12 August 2026, this Paper Citation Record lists 100 of 118 outbound references and 1 inbound Pith citation observation for arXiv:2501.04568.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.04568 v2

Coverage vector

measured 100 of 118 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:34:35.594979Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T19:41:58.673348Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T22:35:51.417589Z

Reference resolution

100 of 118 outbound references displayed

  • verified exact1
  • verified fuzzy8
  • unresolved91
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 878090a4-d779-44d2-b6f8-1d0a6d23754b · outbound

This paper cites GPT-4 Technical Report.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.931476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.931476Z digest=sha256:8cad055b484ff17f03c51b4fbe566c9ac5014500c660625bce510e94841c29b8

Observation 32d9879f-8e91-4db7-90ee-899051b96cc1 · outbound

This paper cites Agrawal, K.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Agrawal, K

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.937048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.937048Z digest=sha256:a7b8a495491c1f671e4fdfdd6116449403da8f1570b47366bbd65a6cba207421

Observation 3694a681-b2d0-4114-8ced-70895428e77c · outbound

This paper cites Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.941481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.941481Z digest=sha256:61b322a57ed0508a8d0476e854dfbfe55b59f7dc34668c70c46b74130c8ba639

Observation 4ffd809b-61fd-4843-96ab-870d088a292e · outbound

This paper cites What learning algorithm is in-context learning? Investigations with linear models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision What learning algorithm is in-context learning? Investigations with linear models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.947407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.947407Z digest=sha256:9d8c3bc3ba5f64507246539f7ddf46fd186f03292a98b425bd87bf5f0022d280

Observation 7474ee78-909d-4701-9a8f-384dc7bab561 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.953338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.953338Z digest=sha256:5b214930a6c0a8c76468aefe55705dfd63fd1574307ab7508f0fed4bf137f212

Observation 71d04822-d2d4-4bc0-9c9a-8ad64dc95e1a · outbound

This paper cites Anthony, Z.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Anthony, Z

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.958760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.958760Z digest=sha256:6948aa1782a92f1418d6921eb052edd8109dca75fc33f3cb02a78a155774fe54

Observation cdab2c24-dd18-40d9-89c0-c0f18369e1fe · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Constitutional AI: Harmlessness from AI Feedback

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.964208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.964208Z digest=sha256:feb1d8892284d0d51b188304fee1a29932e56369ce90d994bda0e9ef740e860c

Observation e6885b9d-e499-418e-8676-15396940c355 · outbound

This paper cites Hallucination of Multimodal Large Language Models: A Survey.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Hallucination of Multimodal Large Language Models: A Survey

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.973845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.973845Z digest=sha256:68e0940253dd710991b0e859d3316488e93c577672eb3c44314ed927e5e2ae1c

Observation eb961fdf-bffd-453e-9911-eaa1254e266c · outbound

This paper cites Banerjee and A.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Banerjee and A

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.978603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.978603Z digest=sha256:145f7e9e2cf6944da768b5c6def201be03daf8b2438036c82bba509072ba6982

Observation 34624016-5e5f-4eb1-980f-e0ba1c339fe0 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.982785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.982785Z digest=sha256:8fc35f4e9ea9a796004800717cbb7babfe95ccf43e356d311e27c51a5e3f5f35

Observation a2d0aaeb-d4cd-4ca1-8fbe-5d464f3bd4b3 · outbound

This paper cites PaliGemma: A versatile 3B VLM for transfer.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision PaliGemma: A versatile 3B VLM for transfer

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.987003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.987003Z digest=sha256:06b9b83272ba7e2e0fdb84322d11d14b84663824188a04089891cc899d976444

Observation 9d087100-76dd-4a0e-950f-c2b6f169a9fd · outbound

This paper cites An Introduction to Vision-Language Modeling.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision An Introduction to Vision-Language Modeling

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.991443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.991443Z digest=sha256:b495bb4ec86d472befeadc2293b1818e4385e5c47a5dc6223879baa3c68e70c1

Observation c495a7b7-1008-4907-8641-d6b95aabe1bc · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:34.995902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:34.995902Z digest=sha256:244cba8534866a418d5e4d61ea83a232486d1da7b797a1e5fd6e8b1994794a7a

Observation 5bf2d01e-dbba-43ac-9d0e-333859fa2f8a · outbound

This paper cites Carion, F.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Carion, F

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.000515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.000515Z digest=sha256:5e960ec3b6d66ae73d4c4e953c39806883c8f7e6f2841ba61ff5fbb2403fee7f

Observation b5c61565-a26f-4a92-9daf-179399d0330c · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.004387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.004387Z digest=sha256:131067d412e747b7a5afb51b87de47f9cb4a38a86e9f872957a4ef5c6c54c342

Observation 89342bce-3e91-4aac-bf88-d82c344e4c1a · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.008817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.008817Z digest=sha256:0806193379752dc6547c349864e1053b21cc75beeec54415cce707b1883e19e4

Observation 850fcc1d-5643-4686-9f1d-0fddf665bf05 · outbound

This paper cites Chiang, Z.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Chiang, Z

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.013082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.013082Z digest=sha256:892f94f82d2bdbcb4a0e680eed8c46083a8b9b1a17bc87322a6f468ed096f61e

Observation 2046ea3d-8e12-45da-97ec-a0f27ca6eb92 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.017220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.017220Z digest=sha256:8efe930e024b4e439724551687cb1b35ea2ac89dd66b8777ce8c3b514c4d4149

Observation fc4d0547-f5f6-462a-a702-3fa95d0eca58 · outbound

This paper cites Collerton, J.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Collerton, J

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.021685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.021685Z digest=sha256:7794b292f2709c2842e7fae4db88c28024d2e9400b3c828530aae29ab221b346

Observation 56f71cab-a7d4-463f-8650-3f89730ac78a · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.025848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.025848Z digest=sha256:6e92f38acdf9de15ecd364d10ada47737b4a7d52c65d94666611c3d87335d90d

Observation 9adc9177-a2b2-424f-8f65-f755a66aa877 · outbound

This paper cites Evaluating Large-Vocabulary Object Detectors: The Devil is in the Details.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Evaluating Large-Vocabulary Object Detectors: The Devil is in the Details

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.030102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.030102Z digest=sha256:34805c4d5d6a90ddf05e9e3858143f2392bea69acab89fd23ad5cf24c676f43a

Observation a2f69265-b911-483d-a4de-259bd3085b10 · outbound

This paper cites RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.034916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.034916Z digest=sha256:c4ebe0ec0016e4be6b17e2c3751f2e7598e1aeca1169e7eefcce227f5197b0fb

Observation db55ea79-7b36-47a2-b88d-100efad54b0a · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.039873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.039873Z digest=sha256:54f835e827db8930b5f9da81ce6a3b53a41d08c390a81027b993ddea9f53c566

Observation a2c326b4-a744-4bb0-bdc7-95995fa1fafa · outbound

This paper cites Favero, L.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Favero, L

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.044221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.044221Z digest=sha256:d37c3bce6f64ff5e783df4f3fc43caf0f5df0443c6c230001c579e2e9c25ea1a

Observation 5052f0c0-c76e-46bb-bea3-dac339e76a4f · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.048986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.048986Z digest=sha256:54b11a1236f0464bc7aeb042d1199dfb206bc94c34290bce7036de0bafdf1095

Observation 5cae555b-9c03-4fa3-ae0f-af6141fd5145 · outbound

This paper cites Girshick, J.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Girshick, J

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.053903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.053903Z digest=sha256:d20185819517dd2fd2e0b09d1112842629194ab91099ab2afc227590649647d6

Observation b32a54f5-af38-41f5-a890-2daa442ea00c · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.058226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.058226Z digest=sha256:15f9eb358b058b311a04da06d6c8aa3f8a0d8e63cf5c6146ca981665c140c7e0

Observation 698ca582-b5e9-483e-950e-fd79b10bf7c0 · outbound

This paper cites Aligning Language Models with Preferences through f-divergence Minimization.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Aligning Language Models with Preferences through f-divergence Minimization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.062540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.062540Z digest=sha256:c6295fcf74939447273e2a7c931b712c2d657b325fc742bf89e1c9845296d190

Observation f577cb62-6a42-4cc3-a61f-f0f2f6702c37 · outbound

This paper cites Reinforced Self-Training (ReST) for Language Modeling.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Reinforced Self-Training (ReST) for Language Modeling

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.067108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.067108Z digest=sha256:60e0369574a2d6db73d172eda27521f948d49060966b35a20deaf73fc8c67c93

Observation 68106966-7b72-4cba-a9b4-f035876b65f1 · outbound

This paper cites Gupta, P.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Gupta, P

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.071728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.071728Z digest=sha256:0b3b93380377365fcf64089098c79518d239b4dde0ae86cbf23bb777f9013c08

Observation 2eca5a40-89a6-41a3-854f-d06155598b7b · outbound

This paper cites Hattie and H.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Hattie and H

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.075865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.075865Z digest=sha256:490a60e760a5761534129230f478afc567903bd50ad5fa5f2aa5d1a805ae6046

Observation 1e843c24-4400-42a9-8c2d-29f2ad650c82 · outbound

This paper cites Foundation Model for Advancing Healthcare: Challenges, Opportunities, and Future Directions.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Foundation Model for Advancing Healthcare: Challenges, Opportunities, and Future Directions

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.080759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.080759Z digest=sha256:78f95ea0a5c141466e4856ee502ed5cffed5d438622ba52fdd74d4ba2dcfeb92

Observation 3b03dd56-acf9-45ee-a62e-ba5f1eb5c64b · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.085552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.085552Z digest=sha256:f7c41cc543405ca043f091c715931eaee3e078b32161d99448df301d409aed65

Observation b6fd9dc4-654f-4009-9412-cb8348171412 · outbound

This paper cites Meta-Learning in Neural Networks: A Survey.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Meta-Learning in Neural Networks: A Survey

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.090793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.090793Z digest=sha256:5f1149e415ad3d988c1ccbade80375b3d8e450f2eece00b95722a5282c7baca5

Observation 8cb71eeb-2f44-4129-8371-3e4fec8a6962 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision LoRA: Low-Rank Adaptation of Large Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.095790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.095790Z digest=sha256:07c63e7aaea7bdec79082aed3b22a6e32ce3375e313af252ea65e93a3e97c0c7

Observation 86a9a661-7f35-4a24-97f5-3305cf2a3f51 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.100708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.100708Z digest=sha256:b3659749cda8a5cb632ee364eeed429aad7b75cb0b3ef283529fca7854413532

Observation 23b50291-d88b-466a-8124-7a0bce5b3ec8 · outbound

This paper cites Mistral 7B.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Mistral 7B

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.104820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.104820Z digest=sha256:fade6971348435de8a26c3713fc5be211ecb6ae0c0bb88750d54ed50aee0b917

Observation 037b3905-0924-4610-a01c-dcac07746564 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.109247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.109247Z digest=sha256:6ce63b7afea252981b27ceab462c857b58e828ab1b8f598dd24213b132e22851

Observation 610cf59a-1b5b-4248-8419-0feacd8cdb6a · outbound

This paper cites Self-MoE: Towards Compositional Large Language Models with Self-Specialized Experts.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Self-MoE: Towards Compositional Large Language Models with Self-Specialized Experts

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.113712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.113712Z digest=sha256:402bdc3b82051a8b47e4cb182d590a177e64715bfbd1114efe15bbdc8b9ed8be

Observation 2f54bac0-1e87-4bd7-88f0-3e37e75f886f · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.118025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.118025Z digest=sha256:ed3aacb9270a21c7dab091b4ce01caeec3b0287f888d34b85d28ff4dcdb1e7b6

Observation 9eb64d4d-e5bd-4d12-8e1d-ca38b30f3a2e · outbound

This paper cites Kazemzadeh, V.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Kazemzadeh, V

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.122652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.122652Z digest=sha256:95ef4556c415b649b4301939fd95555c0ba070749dd20ae15424f07f4fc756aa

Observation 3b02740c-0921-4a76-9d19-b0fcbcad3f8d · outbound

This paper cites Kiefer and L.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Kiefer and L

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.126990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.126990Z digest=sha256:3fd50f0537afee6b7d7ea05fe80ae1687e78ed1de5f36b92949606332c516396

Observation d327302b-3b75-4d4d-8caa-a6c84789e86c · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.131178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.131178Z digest=sha256:e581f8ec2e13bba2ce9d202f7e314355c85831104cbf41a887f80c5352933276

Observation 1b8325a1-a876-4e4a-96ef-2575662909ff · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.135371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.135371Z digest=sha256:87b98727b02a2ec2e4d44abba1bf780889b57a06f3ceed12baebfb30723f9dd6

Observation ee277cb2-d287-40fa-b6ce-2d5dba39fffc · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision LLaVA-OneVision: Easy Visual Task Transfer

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.139694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.139694Z digest=sha256:84209766352a7077ce98988cc52b732d587915c514780852def307144d918d55

Observation cb2fafe6-72c7-4b7b-8f58-a838f84a00eb · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.144811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.144811Z digest=sha256:c282d4d7542c746dc321fe3f77ec1884aee4b7973b7d58ce874a0b67ad72da25

Observation bc133f54-8fdc-4192-8e62-de5ed743dd83 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.149220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.149220Z digest=sha256:3f4ad9cbdc65fd01f36b560908943b196b8a6e04cd973df4c1b4cc0ef00594e0

Observation 5631aa0a-b0c0-4041-b00e-ce2faefa1066 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.153357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.153357Z digest=sha256:2e2e7bf46ce379da0e3a42b6b9c06c5ea7fe96dc34c23e95e498de7d494912c1

Observation 78737c36-eb42-4e24-9a99-722fe21f42f0 · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Evaluating Object Hallucination in Large Vision-Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.159313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.159313Z digest=sha256:99c1c35a19ccd55ff284ed166952809ad2ef8a96857cbf00e1b91a17ffd916a0

Observation 5ba40cf8-7280-4d1b-a358-e40855b3a362 · outbound

This paper cites Meta-SGD: Learning to Learn Quickly for Few-Shot Learning.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Meta-SGD: Learning to Learn Quickly for Few-Shot Learning

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.164496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.164496Z digest=sha256:de8440934ea2df27e837dde4a0c647e9d2cddaa5f41e614086d0f5f4cbcc6cbd

Observation 867056cb-d36c-432d-a46a-5e5117686c13 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.169519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.169519Z digest=sha256:a3926d43147b55543cda549de5b5a47614ebd7f2dfb1114baf64ffb7ae64d86f

Observation 39b8c9e1-c9b8-4868-a147-d4d1fd770d66 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.174740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.174740Z digest=sha256:b2d95df7787316c71e2a3812bbe09e20cbae0fc2b77070dab695527776be0bdc

Observation 17b11415-edfa-4406-9077-25d3f6030d6a · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.178963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.178963Z digest=sha256:b50ccb6775a9451d3612d7d6e5d646a1d9f980415d30f2257e7db1cf97fe515f

Observation 0c16ae80-772e-4bc9-ac9e-e8efef9a8723 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.183587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.183587Z digest=sha256:14c230e8d4cad08a1abf95ddc415c73f6115710a2e29fb5e0b8c3e35abbcf421

Observation b45e967d-0689-43c1-951d-b377fa5b91ba · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.189707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.189707Z digest=sha256:cdd3a3815f8227760b90705b240b61186425bb4d177bfea78c525ed443db8cab

Observation ea07be76-999d-477b-8c8a-fc21358cd12d · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.195100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.195100Z digest=sha256:ca124b25bfc6c5436af389189f267d4b6f59d985dc95f3a6eb8178a5286fd545

Observation 61207065-261c-4e15-85d2-d48cc2b14958 · outbound

This paper cites Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.372992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.372992Z digest=sha256:45ddfe56e2fe18f578bf8050340c7883897226bc777d957870f278e4183da43b

Observation 845a080a-1088-4a7c-b38e-b51a555704dc · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.378684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.378684Z digest=sha256:9542bd8b95e384d7d9ca1fda3ef1026c2ad51444352bb24db09091c162bed925

Observation 91a31114-c52b-481b-9360-3f8f2a83cffa · outbound

This paper cites Madaan, N.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Madaan, N

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.383422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.383422Z digest=sha256:fac236a5d367b621e4f51d72ece64ffa15e48f63c76696032a072aec16bfbe7f

Observation d3da44bd-ce66-4737-8f13-cb77737f3385 · outbound

This paper cites Oquab, L.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Oquab, L

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.388417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.388417Z digest=sha256:7c4e34db13d930d9c104328bf0cbde7208759a9c9de84f36d3f8504fb0969ac8

Observation 83b172f2-c784-43df-af09-ab8179ac6f8e · outbound

This paper cites Training language models to follow instructions with human feedback.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Training language models to follow instructions with human feedback

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.393161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.393161Z digest=sha256:a5a5b2ee46229553740b42317681dd70a11152dcd4bec1006842500e4ecbf2c6

Observation f129c070-7340-42e9-9204-5b344f73b326 · outbound

This paper cites Papineni, S.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Papineni, S

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.398610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.398610Z digest=sha256:7bb2bbe334293aed1a78484a2f3635e6f09b7f5158b3bc2db8c8f1955a946a7a

Observation 46d47fc1-be05-47af-b5c2-de0f89acb80e · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:37.027583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:34:35.404501Z digest=sha256:ad24083278bf5c5cf705e5b67b0691e02b4e2adbd0f4bf089f61d1ecec3dbd2f

Observation 61e5d573-b8a4-4fba-b07a-c8bb7e0464d1 · outbound

This paper cites Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.410160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.410160Z digest=sha256:9a5a8e1734ee0710a7163299632f88491d724a435bcc583c3340c5a9236d9db7

Observation 57d8e92c-08d4-493a-aade-2acd2b7ec86e · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:37.008257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:34:35.417009Z digest=sha256:e03c98310f9419f86a8c9d5566c99a4367ced142c72da6d6584bf95e903c9137

Observation 0f4293f2-33f2-494f-9a2c-110b6374ac53 · outbound

This paper cites Peters and S.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Peters and S

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.421983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.421983Z digest=sha256:83d1d5b9498a565c8738942061560ff66d8a56f11987d6d94886f4981a477127

Observation f11008db-3ecd-42dc-b569-bac88f915e96 · outbound

This paper cites From Concept to Manufacturing: Evaluating Vision-Language Models for Engineering Design.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision From Concept to Manufacturing: Evaluating Vision-Language Models for Engineering Design

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.427447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.427447Z digest=sha256:6dd3e5297671002e89211145f7b8db59295375f8a149b2d8b4e97c485c8b29dc

Observation 9fe61691-2c1a-4d29-99a5-9929d98c2684 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.435022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.435022Z digest=sha256:1bc6e08b2277930742f0088f40faf344ae965859ad3325cc0f88dc621f014807

Observation a41a9d37-9ca0-461b-9249-475131ffde94 · outbound

This paper cites Radford, J.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Radford, J

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.440302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.440302Z digest=sha256:6e6a7536670abaf979c969f0005131788950e9cda554508e531ab1fed49a77ae

Observation be003b50-e0a8-4d17-8689-fcba51ad3b71 · outbound

This paper cites Rafailov, A.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Rafailov, A

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.956900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:34:35.445372Z digest=sha256:c1301d567687e4e0b0df67257ec5c58d1c971d23ddb44af6a5c1c4015e6be71c

Observation 1ab5e6f0-e331-4e55-beb0-0b0d50c531e6 · outbound

This paper cites Vision language models are blind: Failing to translate detailed visual features into words.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Vision language models are blind: Failing to translate detailed visual features into words

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.450573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.450573Z digest=sha256:06d8664811de8183a2083470a66276f776f006d24ac0ae9c7119c519c9b3157e

Observation 126f1b1d-4ab7-4e86-8581-e6c2e14c3c57 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:36.940756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:34:35.455262Z digest=sha256:97a7c0c0be77a3290c1ac5977562ac466be67f1afa9628877184ad0f799465c8

Observation e5621796-6727-4ed6-b694-2dfa5474df98 · outbound

This paper cites Tapered Off-Policy REINFORCE: Stable and efficient reinforcement learning for LLMs.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Tapered Off-Policy REINFORCE: Stable and efficient reinforcement learning for LLMs

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.460399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.460399Z digest=sha256:71bce945bf7bddf60771e92974989e837053b7ca4672b0a25d588556471e0a3f

Observation 10f49899-a4d2-4c1b-a778-eec86572a4d2 · outbound

This paper cites Saikh, T.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Saikh, T

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.925038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:34:35.465293Z digest=sha256:bfec4e42df3e7132004c59247f8faa6e855e7fa909926051771ea49cbebd0563

Observation c953b626-e876-4aa4-940e-8fb7a6d1ba7a · outbound

This paper cites debiaSAE: Benchmarking and Mitigating Vision-Language Model Bias.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision debiaSAE: Benchmarking and Mitigating Vision-Language Model Bias

Reference 75

Resolution
verified exact
local_arxiv, observed 2026-08-10T21:34:36.163169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:34:35.469776Z digest=sha256:e7006f59304374359444f12c1c77044839b53456292ceb341c31e50d0ef770ec

Observation 28e199e8-3641-487a-a636-1492a94fbc00 · outbound

This paper cites Schaul and J.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Schaul and J

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.907427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:34:35.475072Z digest=sha256:0e8d720ed2b5f30126413dc73fcb4e7c7f8197858c1e1319d41214557162f0f4

Observation 2ae740d4-afdc-4fcc-b674-b3b7deaec2e7 · outbound

This paper cites Schmidhuber.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Schmidhuber

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.889969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:34:35.479819Z digest=sha256:97160cd7ef5fa47b66a21910121ab4448906a14c9385e5f1dc0e838582e4120c

Observation e215bbec-0b46-4a8b-8181-8509df3e1371 · outbound

This paper cites Schmidhuber.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Schmidhuber

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.875814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:34:35.485458Z digest=sha256:26627c579a742ebcf8c9279f427e36e425675a716c3c1b5b6a0cd93c4b8ec0da

Observation fcabd321-b44f-4286-88ca-3da27a5b362f · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.490872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.490872Z digest=sha256:42ff3182524c3b07342ca3b95a87daefb41e562ec44673193b0d273879283849

Observation 46d187af-f7c2-4f15-ab81-a909bd26285e · outbound

This paper cites Shinn, F.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Shinn, F

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.847748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:34:35.497061Z digest=sha256:7386eea67f87c1e60c852f8af2e1af32d99d0c08d170faefaa1b15ca446736e1

Observation 2e1b7217-daaf-4198-8a5a-83d67092de2e · outbound

This paper cites Silver, A.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Silver, A

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.826926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:34:35.501869Z digest=sha256:8c983dc66043fbceb9615c7c2a49ea70fbfffdc6909cad8b84541805d2ee61b2

Observation 30721518-8d62-404a-a00c-90492671ca67 · outbound

This paper cites Silver, J.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Silver, J

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.506314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.506314Z digest=sha256:61ba8a909981766ebb857d25198e7a0a7554b8d58f10dd02d1a2b67f957f84c7

Observation 622fd121-1d72-4c7a-8940-df7fe757edb6 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 83

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:36.794520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:34:35.514137Z digest=sha256:4c131351ac917da186d2d9b4ab691ee5a2afdb9311c826e85469bcd6847834fb

Observation de6afcbe-64cb-4519-85b2-2a193a16f7bf · outbound

This paper cites LAB: Large-Scale Alignment for ChatBots.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision LAB: Large-Scale Alignment for ChatBots

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.518320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.518320Z digest=sha256:71316a9fa156545f3eabc6a4a42fc7ba758b634fe0285762a1ea41dc30b08859

Observation 5aa47013-f822-4ba3-bd82-29e57d74f91e · outbound

This paper cites Aligning Large Multimodal Models with Factually Augmented RLHF.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Aligning Large Multimodal Models with Factually Augmented RLHF

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.523168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.523168Z digest=sha256:f700d4e902d570636ff01b932bd1c4bd84552f41aea06cb962ecf5474050e9a2

Observation b88f8223-f487-4038-8072-35647fa28bd9 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 86

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:36.778848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:34:35.528275Z digest=sha256:136c67ed6ce4d2761d2267d90ac5d1cf91d9bff45e0904cda593683c1d550a28

Observation 38292c0a-2459-4ddd-80c4-20e34317a08d · outbound

This paper cites Tenenbaum and E.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Tenenbaum and E

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:34:36.764526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:34:35.532809Z digest=sha256:edb39b31e0556d50a1b2218ded4416a245ffcfcab01a2e33fa59be643fe29af9

Observation 49f63fe2-1ca2-4fc9-8902-aafa98facb7d · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.538223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.538223Z digest=sha256:f53fc2a0854aa04787dbba5a443853e21e075cd0ed1058106050d4f7c7a61e5f

Observation 07da72ab-9485-4eff-b109-686cc87aff44 · outbound

This paper cites Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.543022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.543022Z digest=sha256:6077e09aa892b8babceb81461407a086d95fa47fb42d5f5260d9b4a66ad595d7

Observation 84fd5bb2-6076-4f32-ba69-019c6427ff41 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision LLaMA: Open and Efficient Foundation Language Models

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.547860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.547860Z digest=sha256:c06f834136a9b7ea3ec55e6317f9b097153f9aa4f96471d7a3ed81fbc3a86de3

Observation 9c8b10f1-ac13-4b87-921f-85d627c45a9b · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 91

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:36.742039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:34:35.552315Z digest=sha256:227393c3f30807a4a6de8ee76cc2ecfe031b009371a015c8912998de635d51ec

Observation df5ac0f1-faca-4b31-aae2-791b4ddaeca2 · outbound

This paper cites Vedantam, C.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Vedantam, C

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.556464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.556464Z digest=sha256:3f61a0cceb338772989b3fa566ea4a2ed0746d72c1075d3b9ceb5f2666badb65

Observation ca17a4fe-520b-41ec-9472-18fda0b16f5b · outbound

This paper cites Contrastive Region Guidance: Improving Grounding in Vision-Language Models without Training.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Contrastive Region Guidance: Improving Grounding in Vision-Language Models without Training

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.560452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.560452Z digest=sha256:82e8da90ef868880eb404608c602d5061ff43d34cf7ed05d5527ba5f6e05dba4

Observation 407d1acc-13bc-4ac8-89f7-d4b7fceafdf1 · outbound

This paper cites Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.564572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.564572Z digest=sha256:ea3e09a412799f3d2b46c2ebe10cee96b65cb0d505b5227bebb9c48082641244

Observation 8066381b-7bc6-4b24-8c16-dff82742db3d · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.569125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.569125Z digest=sha256:d7499e3f5f94bf8b2846ad3194756b5ff5a1b3614a004f1465907ab62325a584

Observation fb3bd056-b885-4e09-8458-b7ca1a314824 · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.573795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.573795Z digest=sha256:ad49cac140756cba64ecd5825abbc290450114de27dd0409cd8a9aa6c7ebbf82

Observation eb07f118-bb7e-425d-9a30-e11a15fd5a70 · outbound

This paper cites An Explanation of In-context Learning as Implicit Bayesian Inference.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision An Explanation of In-context Learning as Implicit Bayesian Inference

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.579540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.579540Z digest=sha256:ede42fe6faf369c848a231c055cdcc6ec12c3903b463b77fd4537c5498e53759

Observation 8d81d9d4-a9d7-4acd-8e69-4539c64f4c12 · outbound

This paper cites an unresolved cited work.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Unresolved cited work

Reference 98

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:34:36.719148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:34:35.584821Z digest=sha256:7bef7f7728012d9b51bb5540ab02c2d0d576d4030bea0c8a13d76f25badfe149

Observation f3ccbcf4-4a0d-4a35-bdf9-792dde0fd88b · outbound

This paper cites Qwen2 Technical Report.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision Qwen2 Technical Report

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.589733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.589733Z digest=sha256:68bf0aa4ceb6032e06249efce038db8c3b55dc786f0683301e240c6190372b02

Observation 77549c59-4f78-4c74-b2bc-8523a32e7d24 · outbound

This paper cites ReAct: Synergizing Reasoning and Acting in Language Models.

Feedback-Driven Vision-Language Alignment with Minimal Human Supervision ReAct: Synergizing Reasoning and Acting in Language Models

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-10T21:34:35.594979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:34:35.594979Z digest=sha256:b4196ea642ca948195680724f9ef33f54971b91b88ec77243bef38d07fcba24f

Pith citing papers

Observation e39f026f-64ca-46ac-9e31-8b1b235e05e1 · inbound

Grounding Hierarchical Vision-Language-Action Models Through Explicit Language-Action Alignment cites this paper.

Grounding Hierarchical Vision-Language-Action Models Through Explicit Language-Action Alignment Feedback-Driven Vision-Language Alignment with Minimal Human Supervision

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:35:51.423869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T19:41:58.673348Z digest=sha256:6931ed93fe1257524575308f4124f4a46707fc6e5d7710372873a77d7e30dcd5