Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T21:15:45.548297Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 1 inbound Pith citation observation for arXiv:2412.04903.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T21:15:45.548297Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-16T15:05:21.907878Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-16T15:08:02.074896Z
62 of 62 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c20a20e2-9235-4bb3-8daf-3390873d5016 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfa5f7cc-b9ea-44ce-b6da-5432b3cd586a · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6075f1ff-8257-44bd-b27f-71af6e04cae8 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Introducing our multimodal models, 2023
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4122eb2-7ede-4595-85b8-ff6065f0956c · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Continual LLaVA: Continual Instruction Tuning in Large Vision-Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7531db02-af9a-4a18-b63e-6703c14ef1b4 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation ShareGPT4V: Improving Large Multi-Modal Models with Better Captions
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bfe13d4-debb-4035-b2ac-03525b209723 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation PCA-Bench: Evaluating Multimodal Large Language Models in Perception-Cognition-Action Chain
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 064bd64b-b1a0-49fd-9e69-17caf8363ac1 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Microsoft COCO Captions: Data Collection and Evaluation Server
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afab7f37-2681-4a98-b2e9-bb74ac27b656 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 900be8f2-8b35-4dfd-9cba-ace4b33dd605 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d0e3942-1d76-4462-b1b4-731178172ab8 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Enhancing Large Vision Language Models with Self-Training on Image Comprehension
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 278c5dde-25c1-4f08-abba-8d176d3e2e4a · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation What Makes for Good Visual Instructions? Synthesizing Complex Visual Reasoning Instructions for Visual Instruction Tuning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 399e962d-1baf-4ad7-9d1d-5d50926b6b24 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Vlmevalkit: An open-source toolkit for evaluating large multi-modality models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation be302331-927a-466f-8a70-b0f5d69dee51 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 869f7c6d-08f2-4dd3-8008-560deda00da0 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation HallusionBench: An Advanced Diagnostic Suite for Entangled Language Hallucination and Visual Illusion in Large Vision-Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47b76b57-8d5f-4f02-a2b6-6ebd0eaad57c · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Efficient Multimodal Learning from Data-centric Perspective
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14d1d624-dfba-4e20-94fb-6fb692d2dfdf · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation LoRA: Low-Rank Adaptation of Large Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 173b546f-be86-4f94-a87c-7f56937a687a · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0ca5f28-10e4-461c-91b1-241142baa4e9 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8238795-5298-42c2-962a-696a6a35e620 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 752bf7ef-9dcf-48ea-a232-c45f7efea877 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Silkie: Preference Distillation for Large Visual Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2837536-35ff-452f-816f-8af47781f5a6 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation M$^3$IT: A Large-Scale Dataset towards Multi-Modal Multilingual Instruction Tuning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50dcb1f3-f1c9-4d1a-8f0d-c8ef8a333b65 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Red Teaming Visual Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a7c7d87-f8d3-4c01-9b31-ecfd53584c39 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Evaluating Object Hallucination in Large Vision-Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff5a77d4-9051-48a1-b6fe-b71348d67325 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a89266fe-bff0-40d3-badf-f235576c5d77 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Improved Baselines with Visual Instruction Tuning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 969ef877-2a1a-4ec9-b1f6-bb1c97f6f50a · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Visual instruction tuning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2418008-9e24-4709-b8b6-79ba61956036 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation A Survey on Hallucination in Large Vision-Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 689856cc-c6a7-4f5a-9617-d0f9a1bcdc0d · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Learn to explain: Multimodal reasoning via thought chains for science question answering
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9bc3d518-bdfe-4e85-a664-19494be0f3b7 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Mathvista: Evaluating mathemat- ical reasoning of foundation models in visual contexts
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c91929f-c41c-443e-bb96-11fdea8b1140 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation SimPO: Simple Preference Optimization with a Reference-Free Reward
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bcec186-4edc-48ce-a99a-56c5c1b823e5 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation DINOv2: Learning Robust Visual Features without Supervision
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9ec3eef-ea7b-4fd0-b9a4-b9e280f06b30 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Training language models to follow instructions with human feedback
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 539e37b4-d317-4ea6-8ea7-4cc0bc67d028 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Learning transferable visual models from natural language supervi- sion
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46b8b64a-4161-4b9b-b325-50fc5e150257 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Direct preference optimization: Your language model is secretly a reward model
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d66c0994-55d1-4555-9588-607c3c1e22e7 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Proximal Policy Optimization Algorithms
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a179c69-c0b6-4d63-8011-d0a23245e2af · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Aligning Large Multimodal Models with Factually Augmented RLHF
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9d311ea-0459-4edd-9201-2f7766722622 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Gemini: A Family of Highly Capable Multimodal Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec515631-2aea-4330-9deb-950b9b20bb3e · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Eyes wide shut? exploring the 10 visual shortcomings of multimodal llms
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a8c18b14-d571-498e-94b0-8c673eaac303 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b781e7cd-e719-4fb8-bdf5-c8ba61f9eb03 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5dd3bf9-8510-42d6-8dd5-0c174b40ebe7 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Emu3: Next-Token Prediction is All You Need
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90bd0078-f692-4549-91bd-85fda6748089 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Chain-of-thought prompting elicits reasoning in large lan- guage models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 673ab646-95d2-4282-8bbc-bc16281423d8 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation LLaVA-Critic: Learning to Evaluate Multimodal Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f319dd1b-8512-41cd-9fe4-7f8eedb838b4 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Vigor: Improving visual ground- ing of large vision language models with fine-grained reward modeling
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e3e43e3-f35b-4de7-a02d-cb4cad39ea32 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1039a7d5-e991-4703-94d3-df29571e1b96 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Rlhf-v: Towards trustworthy mllms via behavior alignment from fine-grained correctional hu- man feedback
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d9c0d487-ee6f-4416-a2bc-58167ce97727 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Rlaif-v: Aligning mllms through open-source ai feedback for super gpt-4v trustworthiness
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3255b99-0504-4be1-867f-244dfc44a424 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Self-Rewarding Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c91f07c-3a13-41d5-aa0a-d4d3fd452b2e · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 538b0ddb-5dbf-4130-ab77-8f763dc35550 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation LLaVAR: Enhanced Visual Instruction Tuning for Text-Rich Image Understanding
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eda6088e-7fec-415d-ba2a-21aaecb2a30b · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation SVIT: Scaling up Visual Instruction Tuning
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75cd0b7e-36f6-420e-b4de-136dfbbaaccc · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Aligning Modalities in Vision Large Language Models via Preference Fine-tuning
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b241dee-3bf3-494f-83bf-bdb857a076f1 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Calibrated Self-Rewarding Vision Language Models
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ebe7466-a4a3-4a02-92d7-4503ba250d5b · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ade1ccc-ebd7-4f56-9f90-2287aa58bc8e · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation GPT-4o and our Critic model produce similar scores for responses, but they fail to identify the flaws in bad responses from the baseline LLA V A model
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 37566070-fac6-4220-bf83-d2c6b060ebaf · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation As shown in Table 8, most of the experiment is con- ducted with prompts in rating style, apart from the ablation study presented in Section 5.3
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9973f9c9-347f-40a6-8ac0-b9fa44c59baf · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation The training details are shown in Table 2
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 95650ce3-5617-4935-b03a-8d964506304d · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Using annotated preference data, one round of preference learning is conducted on LLaV A1.5
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e4972ee4-a8ae-4bcb-b70a-11b9cb3e5809 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Unresolved cited work
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9661aef2-010b-4c59-8e7b-598599019fca · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Unresolved cited work
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e55212cd-03c3-4adb-81a8-741fcfaca837 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation Here, we will show some examples between EACO and baseline LLaV A-v1.6- Mistral-7B in Table 9 and 10
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d6e4cda5-e15e-4dcd-8cd7-63f1716d6fc7 · outbound
EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation score:⟨total points⟩
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 822144ed-1b63-4e98-b397-83b8525ce130 · inbound
PALM: Progress-Aware Policy Learning via Affordance Reasoning for Long-Horizon Robotic Manipulation EACO: Enhancing Alignment in Multimodal LLMs via Critical Observation
Reference 115
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.