Pith. sign in

Paper Citation Record · LEDGER

Understanding Alignment in Multimodal LLMs: A Comprehensive Study

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2407.02477.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.02477 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:13:40.330428Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T01:07:30.306320Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5e3ee981-f43b-485c-9ede-81c4fea061cf · inbound

Efficient Self-Improvement in Multimodal Large Language Models: A Model-Level Judge-Free Approach cites this paper.

Efficient Self-Improvement in Multimodal Large Language Models: A Model-Level Judge-Free Approach Understanding Alignment in Multimodal LLMs: A Comprehensive Study

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T12:42:49.168659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:42:49.168659Z digest=sha256:dcfae03699d903481be3ee0b3816e393699e70dbc0b1837ae3a2c1f99f6f64c8

Observation 6bfcede6-f38f-46a9-b53d-00e47c9960a9 · inbound

Beyond Human Data: Aligning Multimodal Large Language Models by Iterative Self-Evolution cites this paper.

Beyond Human Data: Aligning Multimodal Large Language Models by Iterative Self-Evolution Understanding Alignment in Multimodal LLMs: A Comprehensive Study

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T11:18:34.251721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:18:34.251721Z digest=sha256:dbbda4dc2c41c79b0c18fe17870d95eaf20a50d55ee38ed504e32d68c5f504af

Observation e2663749-bb91-47aa-9c7e-0f672f63ae3b · inbound

Align Anything: Training All-Modality Models to Follow Instructions with Language Feedback cites this paper.

Align Anything: Training All-Modality Models to Follow Instructions with Language Feedback Understanding Alignment in Multimodal LLMs: A Comprehensive Study

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T11:09:12.888377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:09:12.888377Z digest=sha256:6100ffc89396388e1eef5632e13d7a6dcde20fe5fbd0a42d9db945de7fe4c8fa

Observation b8703a87-72fa-44b0-a36f-1fdf8b04c253 · inbound

PerPO: Perceptual Preference Optimization via Discriminative Rewarding cites this paper.

PerPO: Perceptual Preference Optimization via Discriminative Rewarding Understanding Alignment in Multimodal LLMs: A Comprehensive Study

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T06:01:13.181293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T06:01:13.181293Z digest=sha256:8f43ad1b4bf42a806f952c569c6b165e9fb301aa358b8a4e3e277da1021131de

Observation a1c198b8-abb0-44c3-b207-214fe00f3088 · inbound

Seed1.5-VL Technical Report cites this paper.

Seed1.5-VL Technical Report Understanding Alignment in Multimodal LLMs: A Comprehensive Study

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:26:06.032367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-11T05:26:04.960844Z digest=sha256:992ea0c5c224c6a94c1be52bbc4bae282b52c17752437a75206d7a17cf7829b4

Observation a2e78cdb-1c50-4de0-aa61-8865a41e9cfc · inbound

TUNA: Comprehensive Fine-grained Temporal Understanding Evaluation on Dense Dynamic Videos cites this paper.

TUNA: Comprehensive Fine-grained Temporal Understanding Evaluation on Dense Dynamic Videos Understanding Alignment in Multimodal LLMs: A Comprehensive Study

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:02:59.199626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:02:59.199626Z digest=sha256:a97df345ef672ea9a6b075f489ed6857a4d54595a59997101c5c918fc6250bda

Observation f1ba0cee-9f31-4e8b-a0fc-4456da44a26f · inbound

Controlling Multimodal LLMs via Reward-guided Decoding cites this paper.

Controlling Multimodal LLMs via Reward-guided Decoding Understanding Alignment in Multimodal LLMs: A Comprehensive Study

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T19:52:50.159792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:52:50.159792Z digest=sha256:f3c920e24615a737fe53e0990fb002b06fdd3a51aa564b5de5ada7791eb626f1

Observation d92ef524-ed8b-4f9d-9da7-df4dd275bc81 · inbound

STORM: End-to-End Referring Multi-Object Tracking in Videos cites this paper.

STORM: End-to-End Referring Multi-Object Tracking in Videos Understanding Alignment in Multimodal LLMs: A Comprehensive Study

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T08:56:00.429658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-10T16:25:31.777907Z digest=sha256:062e9f42dbd9dfba06d13451c5bee71b3e7a9b22cc129a259df2a82e239d1d65

Observation ffbdc1ec-ecb5-437c-adf6-63ef1d32d0f6 · inbound

Vocabulary Hijacking in LVLMs: Unveiling Critical Attention Heads by Excluding Inert Tokens to Mitigate Hallucination cites this paper.

Vocabulary Hijacking in LVLMs: Unveiling Critical Attention Heads by Excluding Inert Tokens to Mitigate Hallucination Understanding Alignment in Multimodal LLMs: A Comprehensive Study

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:16:24.424276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-12T05:15:34.156717Z digest=sha256:fced22567f26dfd98923e5bf2ed14f0c039730c5c6cf109c416cf82117baec62

Observation d090b960-9e66-4b7d-8041-4da8f0f7ca45 · inbound

Personalization Meets Safety:Mechanisms,Risks,and Mitigations in Personalized LLMs cites this paper.

Personalization Meets Safety:Mechanisms,Risks,and Mitigations in Personalized LLMs Understanding Alignment in Multimodal LLMs: A Comprehensive Study

Reference 259

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:07:30.307746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-27T16:49:14.243931Z digest=sha256:872e0fac9246a53c6c30609ad891fbe28d3c4f83248661dce013e6ae010cc480

Observation 38e9c4e0-8eb8-43a0-a74a-20a97fa64a51 · inbound

Groc-PO: Grounded Context Preference Optimization for Truthful Multimodal LLMs cites this paper.

Groc-PO: Grounded Context Preference Optimization for Truthful Multimodal LLMs Understanding Alignment in Multimodal LLMs: A Comprehensive Study

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T04:15:35.851690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:15:35.851690Z digest=sha256:684e88727756f87c0d91c0bf25f41b731ac191d2550b0b699a8623a70072bb5c

Observation c9117fb5-96ab-4fb8-9283-475de13e7b86 · inbound

Toward a Theory of Value in AI Alignment cites this paper.

Toward a Theory of Value in AI Alignment Understanding Alignment in Multimodal LLMs: A Comprehensive Study

Reference 196

Resolution
unresolved
no resolver link, observed 2026-08-14T04:13:40.330428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:13:40.330428Z digest=sha256:2a9f7ba0615bbc7d4822f5fba95b4641429b67adeceaa27988cd2445c6c5e089