Pith. sign in

Paper Citation Record · LEDGER

RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2312.00849.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.00849 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:23:12.812692Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T23:04:01.934982Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 78c1bb63-6a42-468e-bdac-d91f81be24a7 · inbound

A Survey on Multimodal Large Language Models cites this paper.

A Survey on Multimodal Large Language Models RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback

Reference 116

Resolution
verified exact
arxiv_id, observed 2026-05-16T02:56:41.811143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T02:56:41.658658Z digest=sha256:e562812d14f18ee1eba45ee39437c9351625f692794501ca8b35a33c6fc09142

Observation 6f71323c-8d43-426b-b00b-c80b5c7e6655 · inbound

A Survey on Hallucination in Large Vision-Language Models cites this paper.

A Survey on Hallucination in Large Vision-Language Models RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:10:10.255593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T22:10:10.186950Z digest=sha256:2b70e90bb42a87c6680c61a0f92384ca594b607d2772062858a4395c3049b1a5

Observation 5db403f7-ea85-4282-b405-ea59fb98629b · inbound

Aligning Modalities in Vision Large Language Models via Preference Fine-tuning cites this paper.

Aligning Modalities in Vision Large Language Models via Preference Fine-tuning RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback

Reference 182

Resolution
verified exact
arxiv_id, observed 2026-05-17T10:58:53.490470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-17T10:58:53.215887Z digest=sha256:6039c673f1d5d9e12c4e9045c37b2b746379f2d27ed068c3b19ecf2913584fb2

Observation 9670d7be-b32b-4a89-8f63-b9aefcd79c9f · inbound

Hallucination of Multimodal Large Language Models: A Survey cites this paper.

Hallucination of Multimodal Large Language Models: A Survey RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback

Reference 198

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:33:33.317308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T12:33:32.631346Z digest=sha256:eed33ee42cf07d64d9a23f178dca51d5e3742ae6e99e219805c749aa8b631212

Observation d74eff4e-2d17-4523-8377-20f63bfd1fa6 · inbound

Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs cites this paper.

Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback

Reference 141

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:05:03.818126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T00:05:03.547664Z digest=sha256:cc8a9c4a8b2ea5074a4f3e1605cfd58035a195505f2c91a2ff83f125fd0d24a9

Observation 6892cf36-0521-42a5-ad70-9b14fc28cb78 · inbound

MetaMorph: Multimodal Understanding and Generation via Instruction Tuning cites this paper.

MetaMorph: Multimodal Understanding and Generation via Instruction Tuning RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback

Reference 159

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T07:51:13.118349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-17T07:51:12.953777Z digest=sha256:dcf3e9cfe729fa1dbb595a6c90db7b24af64408c55fb8f54452aace30c7c674b

Observation ec1ab486-8a66-4f9a-95ce-5a0ed0c2bd63 · inbound

D-Fusion: Direct Preference Optimization for Aligning Diffusion Models with Visually Consistent Samples cites this paper.

D-Fusion: Direct Preference Optimization for Aligning Diffusion Models with Visually Consistent Samples RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T13:23:12.812692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:23:12.812692Z digest=sha256:b039364949b1cf5ddf0ad4684f55987be9b12213a407126fa25d1a8da7a3e4d2

Observation 088feea2-f743-48af-9c46-217b475dd4ff · inbound

BlueLM-2.5-3B Technical Report cites this paper.

BlueLM-2.5-3B Technical Report RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:57.396552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:20:57.396552Z digest=sha256:44b5414995cc04a644f6e411fc399392cde702d06212e57c9ed0bdfe7ce4fc10

Observation 8f808cba-a471-4b70-aeb0-656436b2e9a7 · inbound

MCA-LLaVA: Manhattan Causal Attention for Reducing Hallucination in Large Vision-Language Models cites this paper.

MCA-LLaVA: Manhattan Causal Attention for Reducing Hallucination in Large Vision-Language Models RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-06T18:09:11.211296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:09:11.211296Z digest=sha256:a20b3e4bc8d48e27c1f2b4639d5c25f3bebcf17aaac66d1303496dec8d013894

Observation 57d57324-eb6c-4f45-af97-3a5a3dadc5b7 · inbound

Magic-MM-Embedding: Towards Visual-Token-Efficient Universal Multimodal Embedding with MLLMs cites this paper.

Magic-MM-Embedding: Towards Visual-Token-Efficient Universal Multimodal Embedding with MLLMs RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-03T04:20:58.564725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:20:58.564725Z digest=sha256:66c41e284bc7c859d72a71ccac1b6cd32482d347986157d98e501eadd0cac6ab

Observation a25d4354-5a11-41fc-9292-2c418054ab8a · inbound

CAST: Mitigating Object Hallucination in Large Vision-Language Models via Caption-Guided Visual Attention Steering cites this paper.

CAST: Mitigating Object Hallucination in Large Vision-Language Models via Caption-Guided Visual Attention Steering RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback

Reference 89

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T06:50:40.184164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-08T18:05:38.705480Z digest=sha256:e67d71c150c81864597a6b3cc416eca0ff167e7ceb9e151ac74d5b51488e8005

Observation a4888037-77a7-4eb7-b006-2866462455bb · inbound

Dual-Pathway Circuits of Object Hallucination in Vision-Language Models cites this paper.

Dual-Pathway Circuits of Object Hallucination in Vision-Language Models RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:19:23.918280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T19:18:41.987533Z digest=sha256:ceeb377915e86c240559f37781168e4b2e7e8e1da6da14d370726b650684f080

Observation 2b6eed57-b283-4a2a-bc1b-9a6fd41dd7f1 · inbound

Toward Native Multimodal Modeling: A Roadmap cites this paper.

Toward Native Multimodal Modeling: A Roadmap RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback

Reference 140

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:04:01.936915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T22:58:38.610609Z digest=sha256:0ed0f8f97497d2303f837c9354f9ad94897996a5fe4d7eff3dbe40912b470908

Observation d5573700-ee88-47f1-9cd3-67984f938448 · inbound

Rethinking Visual Neglect: Steering via Context-Preference for MLLM Hallucination Mitigation cites this paper.

Rethinking Visual Neglect: Steering via Context-Preference for MLLM Hallucination Mitigation RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:33:28.044262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-29T13:29:33.614965Z digest=sha256:ebc4b355da169cb102afb4e8e7286b16332848338276f1942d3dd9183162f2fb