Pith. sign in

Paper Citation Record · LEDGER

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering

As of 8 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 2 inbound Pith citation observations for arXiv:2607.04163.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.04163 v1

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-11T21:15:25.010442Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T22:30:28.401425Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

26 of 26 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4b7259a3-4a16-4092-b3d4-2f59411556bd · outbound

This paper cites Qwen Technical Report.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering Qwen Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:3923206633e00d6e3f575c827c40456a8acd6a1dec217ab8145228bab70242d7

Observation 10e41f4b-2d86-4135-a636-d290c04d5718 · outbound

This paper cites Token Merging: Your ViT But Faster.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering Token Merging: Your ViT But Faster

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:a524fca5b183d3462b29a2efd8d9888d89a3b13f9b634e2fe79d78cad4ba4281

Observation 4eaef454-e64d-4848-b478-eb1ab57e35ef · outbound

This paper cites Hallucinatory Image Tokens: A Training-free EAZY Approach on Detecting and Mitigating Object Hallucinations in LVLMs.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering Hallucinatory Image Tokens: A Training-free EAZY Approach on Detecting and Mitigating Object Hallucinations in LVLMs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:4dc564f47364a78b93f43b8572d0c7cd957286063a83fb5e7337be3c90b2d078

Observation d0b7326d-2383-42c9-9fb2-519686d2e957 · outbound

This paper cites An image is worth 1/2 tokens after layer 2: Plug-and-play inference acceleration for large vision- language models.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering An image is worth 1/2 tokens after layer 2: Plug-and-play inference acceleration for large vision- language models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:32ebf9a963549a18e2343af7b0f2a7160e56782f8d10d4fecf9164d6c403f81b

Observation d907c1e1-c2ba-4734-8e1c-cbe88d796015 · outbound

This paper cites DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:cc3ff542faf1431f3196cf1187644340cac1171d142953733679beb53637491c

Observation c0f44307-c052-4974-9a52-f26a5dcc9a25 · outbound

This paper cites Bert: Pre-training of deep bidirectional transformers for lan- guage understanding.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering Bert: Pre-training of deep bidirectional transformers for lan- guage understanding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:59985f9a462195ab63fb435e9aa2436aaf142d65dad86e1362b8ff996ce77e50

Observation 02d8d75a-719c-4320-9cea-51f7ea77b814 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:4e513e5ddeb2525cb7b4c98107d18acc24bc64b143c611be3691217a269c1b02

Observation 19c7353c-3bf3-4ee6-a47c-bd1674acfd41 · outbound

This paper cites STAR: Stage-Wise Attention-Guided Token Reduction for Efficient Large Vision-Language Models Inference.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering STAR: Stage-Wise Attention-Guided Token Reduction for Efficient Large Vision-Language Models Inference

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:3c38cf2e4fe7abfff48fafccfa540d0dc558c69bc07fc6e12eb9326f65095b8b

Observation a008a3cb-2cf6-46fd-9ab4-74ec8ed2dc2b · outbound

This paper cites FADE: Mitigating Hallucinations by Reducing Language-Prior Dominance in Large Vision-Language Models.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering FADE: Mitigating Hallucinations by Reducing Language-Prior Dominance in Large Vision-Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:d3e646aa643a7703d2761c802e51d9de6ad0cffec4d7e88170dc137540493981

Observation 44e06353-3e91-496d-b4e4-68ba6ab95e51 · outbound

This paper cites InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:fdce111cc17a597b70435ee05a88e071c4690f22d91acb0e81829fa11e643f97

Observation 560e52da-ad60-4c77-811e-890006c97e45 · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering Evaluating Object Hallucination in Large Vision-Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:1a31fa8707b37e1417705012875a0da5393109ba2a9dd7e3eb8120f711d86d2c

Observation 0cad0e59-4ada-494f-a6f0-dbbbdfd61b08 · outbound

This paper cites an unresolved cited work.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:16af589ede7a6c8a47a0ed1978fd599101ba832cc18aff6544999f496cbc4801

Observation 2fccedcd-8608-4de3-b7d7-cca237206623 · outbound

This paper cites Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:3e71b139346b62835d351bbeb52c0e836e8c733afc19ef0a20a123d36743f45e

Observation 0378b2c9-4a19-40b0-82d5-41e3ae4afd07 · outbound

This paper cites A Survey on Vision-Language-Action Models for Embodied AI.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering A Survey on Vision-Language-Action Models for Embodied AI

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:b5bf6eb1c2e45519c160487e123a1ced37ac2fe27cd5d8c8e380aad9911df17e

Observation 958f2fa3-56c7-466c-a839-649fd0ad9f5e · outbound

This paper cites A., and Kundu, S.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering A., and Kundu, S

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:70da783eb5e411f9a0c5ea2c32fe2e11fa298372b960a1b8f0e861dd231e3523

Observation 7bcb6fd0-3701-4cc2-a328-71e433d80097 · outbound

This paper cites J., and Yan, Y.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering J., and Yan, Y

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:5f736e2d72d718708e10e124d1e907aaf48e28dea9e3320e84c08b6b97ba7ec3

Observation 281d4bee-1cdd-495e-acc0-9192d1101549 · outbound

This paper cites Mitigating Hallucinations via Inter-Layer Consistency Aggregation in Large Vision-Language Models.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering Mitigating Hallucinations via Inter-Layer Consistency Aggregation in Large Vision-Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:f7658a00c225bcac887bb6255d8a3992490f86354d3464bacab1fcae854dc013

Observation 61c767f6-271d-412d-a0af-c84cc599ffe5 · outbound

This paper cites AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:a77786f444d73a125fb47b0ad5692eacc281d47ba5e6445f1added241a7490bf

Observation 1a135e4f-65fb-4554-bc10-866a59161538 · outbound

This paper cites Don't Miss the Forest for the Trees: Attentional Vision Calibration for Large Vision Language Models.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering Don't Miss the Forest for the Trees: Attentional Vision Calibration for Large Vision Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:fdcca4e24b012b328e42ef5e619960090225bed706cf648bb2afa2947b2aa732

Observation 3dbee225-cf8d-46f4-b824-5c33a77ee4f7 · outbound

This paper cites Hallucination is Inevitable: An Innate Limitation of Large Language Models.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering Hallucination is Inevitable: An Innate Limitation of Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:c0cfad67fd4033602ddac1437e54d4b11041b3108541defc8b023280b18a578e

Observation d222690d-2e51-4425-b6df-b17441f817d5 · outbound

This paper cites LLM Lies: Hallucinations are not Bugs, but Features as Adversarial Examples.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering LLM Lies: Hallucinations are not Bugs, but Features as Adversarial Examples

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:a688cad83831ed48aeaf109507c3a5ba46e7942b1b35fadbfbc5c02e2ec00953

Observation 1b35f242-db33-4035-8c48-f71bcdfba914 · outbound

This paper cites Not all errors are created equal: Ascot addresses late-stage fragility in efficient llm reasoning.arXiv Prepr.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering Not all errors are created equal: Ascot addresses late-stage fragility in efficient llm reasoning.arXiv Prepr

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:9ef582c1ddf5a3f98f4148b77091d95039fbac677135fcd35349d3a8fd472237

Observation c5c597f6-df0e-4cdf-af84-05820f8bfb73 · outbound

This paper cites Not all queries need deep thought: Coficot for adaptive coarse-to-fine stateful refinement.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering Not all queries need deep thought: Coficot for adaptive coarse-to-fine stateful refinement

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:2cd37895a50da6f05fbfbcf9ca9d29d4d392089f900be82bd248cf02d6ca2670

Observation 18157f53-17f8-435c-8620-248a523d780a · outbound

This paper cites SparseVLM: Visual Token Sparsification for Efficient Vision-Language Model Inference.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering SparseVLM: Visual Token Sparsification for Efficient Vision-Language Model Inference

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:793c16f871cd0082943a3233acfab322e6ee06dc4ca08477e6e325434dcb7860

Observation f236b25a-a134-43b3-a47d-85bbc32c7c0a · outbound

This paper cites InfMLLM: A Unified Framework for Visual-Language Tasks.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering InfMLLM: A Unified Framework for Visual-Language Tasks

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:dabd37d63eb3e08d94147fa501c059fe45a3ba583636433ccb2ba192f75c5058

Observation 2066eb10-ace4-4c3c-abf2-95c0f4d1aaf8 · outbound

This paper cites Analyzing and Mitigating Object Hallucination in Large Vision-Language Models.

SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering Analyzing and Mitigating Object Hallucination in Large Vision-Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-11T21:15:25.010442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:15:25.010442Z digest=sha256:a79958257d7d19828fdfdb5052b53708d5aa82ebdada929c3bce46c0cd0e6cda

Pith citing papers

Observation 4ac47558-83b9-4470-a5cc-20aa60767907 · inbound

SPARK: Susceptibility-Guided Profiling and Steering of Latent Reasoning States in Large Language Models cites this paper.

SPARK: Susceptibility-Guided Profiling and Steering of Latent Reasoning States in Large Language Models SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-14T12:50:33.254161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:50:33.254161Z digest=sha256:aec157c286c70e74600aa890afba6b15e833937bbaf876861d036bcb6230ab8d

Observation 38483ab5-3502-41d5-8e6b-e19e2bfebca4 · inbound

Better Starts, Better Ends: Bootstrapped Iterative Self-Reasoning Distillation for Compressed Reasoning cites this paper.

Better Starts, Better Ends: Bootstrapped Iterative Self-Reasoning Distillation for Compressed Reasoning SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual Token Engineering

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T22:30:28.401425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:30:28.401425Z digest=sha256:400702cb0ff7e200f2fb25df30d6852a1db6acae3bab770def9ed2ca1b445e4e