Pith. sign in

Paper Citation Record · LEDGER

Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2411.14432.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.14432 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T18:05:36.802960Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:18:57.819941Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d6f69275-87bb-439a-a51b-a089d5d6d31d · inbound

Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models cites this paper.

Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models

Reference 162

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:41:23.531557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T08:40:40.910461Z digest=sha256:b35e8d5fe3bbb7889a164b9ee9497f6ae3800d7908e9862ce0481888364177e0

Observation 8a505bfb-9f47-4ae9-b5b0-807212b8c0d2 · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models

Reference 91

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.752430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:5ec94bec88a860b63921666fd74d5788cc39f9ecbf988b77b45e47b58d4f097b

Observation 0b8427e1-700d-4a23-a967-7abf67faf2d8 · inbound

R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization cites this paper.

R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-16T15:04:22.821657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T15:04:22.690503Z digest=sha256:1b71c725059834131685b42f3a70dd6ddc4d9095d8a21fab3c3a8b0e7ab88953

Observation 8461ec5f-050a-4ead-b5ed-13ef51f6d045 · inbound

Grounded Reinforcement Learning for Visual Reasoning cites this paper.

Grounded Reinforcement Learning for Visual Reasoning Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:05:52.139807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T01:05:18.801388Z digest=sha256:6df84555d446eaf45c32cd60cb25cd76cab2b92f6696e37f1d2a2c02abe77352

Observation 5b7b1725-dd49-47a6-83dc-b953cace3bf9 · inbound

Vision-EKIPL: External Knowledge-Infused Policy Learning for Visual Reasoning cites this paper.

Vision-EKIPL: External Knowledge-Infused Policy Learning for Visual Reasoning Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-19T10:37:15.034292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T10:34:48.849524Z digest=sha256:6f3e4b1bc44323f90d658641ce8961754894a6b6bd89b2f5447ada5dee472d61

Observation 932f6aae-56d4-4782-a543-4a0fdd953aae · inbound

High-Resolution Visual Reasoning via Multi-Turn Grounding-Based Reinforcement Learning cites this paper.

High-Resolution Visual Reasoning via Multi-Turn Grounding-Based Reinforcement Learning Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-19T06:12:07.105593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T06:10:57.219445Z digest=sha256:d3793ea6de63d972ccb94558eb1ed0aae88bc9424d76059a1f5183ca3f99d6fa

Observation 10790b76-89d7-4bca-917b-6c638921155e · inbound

GenTune: Toward Traceable Prompts to Improve Controllability of Image Refinement in Environment Design cites this paper.

GenTune: Toward Traceable Prompts to Improve Controllability of Image Refinement in Environment Design Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T18:05:36.802960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:05:36.802960Z digest=sha256:cd37e8a837e7e757d5c53a1482187267bb110785776ba9e2c871eace8f791b69

Observation d62c896b-9f3a-4713-8d08-1799de8b65e5 · inbound

Aligning Large Vision-Language Models by Deep Reinforcement Learning and Direct Preference Optimization cites this paper.

Aligning Large Vision-Language Models by Deep Reinforcement Learning and Direct Preference Optimization Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T23:11:32.319995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:11:32.319995Z digest=sha256:a72b8aab9bd59f9acffe5eac9f38b1b8d179aa451aa29210f1e0f772e0d1a1c2

Observation 412a07ed-253d-4bc4-b971-b7bcb89e1096 · inbound

InPhyRe Discovers: Large Multimodal Models Struggle in Inductive Physical Reasoning cites this paper.

InPhyRe Discovers: Large Multimodal Models Struggle in Inductive Physical Reasoning Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T17:46:56.202091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:46:56.202091Z digest=sha256:43145350cd1e1cbbdac4fbec9df3927de429b99db9dbcd1685cce4029a084c84

Observation 845b5c78-7297-445e-8109-9ba35228f073 · inbound

Chat-Scene++: Exploiting Context-Rich Object Identification for 3D LLM cites this paper.

Chat-Scene++: Exploiting Context-Rich Object Identification for 3D LLM Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:32:59.478593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-14T21:31:00.764078Z digest=sha256:06377f02df5b7c4054b1daa419f81c48ad677b7095574ad3f162948dde220617

Observation 3b544053-1907-4c6a-89cb-7715de179913 · inbound

MathVis-Fine: Aligning Visual Supervision with Necessity via Progressive Dependency-Guided Training for Multimodal Mathematical Reasoning cites this paper.

MathVis-Fine: Aligning Visual Supervision with Necessity via Progressive Dependency-Guided Training for Multimodal Mathematical Reasoning Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models

Reference 58

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:18:57.823334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T01:23:40.564561Z digest=sha256:7cd931e965a187a23fd6fd4184bddd941ef453bc159f0b28f275ee8f68cf771d

Observation 70c9f362-4838-494c-b649-7b29b135f8fd · inbound

MIRROR: Learning from the Other View for Multi-Modal Reasoning cites this paper.

MIRROR: Learning from the Other View for Multi-Modal Reasoning Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T07:15:04.612610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T07:15:04.612610Z digest=sha256:4472c609c7d0294af0b35994efe8c4fc564c55e3572b57eacb1efbd3e06a34d3