Pith. sign in

Paper Citation Record · LEDGER

LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 30 inbound Pith citation observations for arXiv:2312.17240.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.17240 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 30 of 30 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T17:37:43.652138Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T02:04:26.295057Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 187437e8-6175-4583-8aab-06a929c0177a · inbound

Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement cites this paper.

Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:31:43.561548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T12:31:43.494099Z digest=sha256:347e4def967440aef1266f2bd0d86f7886df5525ba78f037487d087941dc3a37

Observation 9ec78942-fa8f-4d43-9ca0-bbefc02526b5 · inbound

Mitigating Object Hallucinations via Sentence-Level Early Intervention cites this paper.

Mitigating Object Hallucinations via Sentence-Level Early Intervention LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-25T08:35:32.505195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T08:31:24.173135Z digest=sha256:47a43c433345256cc22f6bf83f1b44e1602e9cb0d9058481331dc5b9bc58a664

Observation 62f8d31d-e6ab-4856-8fba-4b6b33b1b816 · inbound

SCOPE: Speech-guided COllaborative PErception Framework for Surgical Scene Segmentation cites this paper.

SCOPE: Speech-guided COllaborative PErception Framework for Surgical Scene Segmentation LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T17:37:43.652138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:37:43.652138Z digest=sha256:adbdd5a618604c3d98e1301d5ce6ea24e3a09687199e8ee64a1754c37ff1b9d5

Observation b8e6c691-003d-4ddb-9931-4eabec1e20c5 · inbound

MediRound: Multi-Round Entity-Level Reasoning Segmentation in Medical Images cites this paper.

MediRound: Multi-Round Entity-Level Reasoning Segmentation in Medical Images LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-03T22:08:38.407309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:08:38.407309Z digest=sha256:bf82856eabfbf8db3278ac01fb0a2b6927a111b301977d000c2c5a002b2e901d

Observation a4c3445f-89dd-4afc-95ac-78bf5213f0a9 · inbound

Grounding Everything in Tokens for Multimodal Large Language Models cites this paper.

Grounding Everything in Tokens for Multimodal Large Language Models LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:31:21.918494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T23:31:05.422935Z digest=sha256:229769ab47f26e78afdbeb1b5228e9f4c3c93e0b8bafa6d7afa6a4f650d95f5b

Observation 42dec495-3e45-4b10-b056-02b1f948c5e3 · inbound

IBISAgent: Reinforcing Pixel-Level Visual Reasoning in MLLMs for Universal Biomedical Object Referring and Segmentation cites this paper.

IBISAgent: Reinforcing Pixel-Level Visual Reasoning in MLLMs for Universal Biomedical Object Referring and Segmentation LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-16T17:33:10.089758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T17:31:33.903063Z digest=sha256:bef882a63da95c9d2b45a536da96aad7262f5358d2c397a37464162d65c9e015

Observation 09cc0b98-6fe9-4af3-849a-c95f70d962c9 · inbound

StAR: Segment Anything Reasoner cites this paper.

StAR: Segment Anything Reasoner LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-02T18:14:38.024845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:14:38.024845Z digest=sha256:207fa6372e9a933e9a9b2c5233453dbf2446ef388d4fb5c12adeba00ba2a6138

Observation a7783ed9-6660-4ed5-a53b-1abea29bc48d · inbound

Speak, Segment, Track, Navigate: An Interactive System for Video-Guided Skull-Base Surgery cites this paper.

Speak, Segment, Track, Navigate: An Interactive System for Video-Guided Skull-Base Surgery LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:49:54.432320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T09:49:10.948061Z digest=sha256:933ac1ac40da4e012d8bdc168ce7f9c738bc496bb080bc69bde834bef0f84e25

Observation 4da97aa2-503b-4371-bb13-ba41ad31e781 · inbound

Tarot-SAM3: Training-free SAM3 for Any Referring Expression Segmentation cites this paper.

Tarot-SAM3: Training-free SAM3 for Any Referring Expression Segmentation LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T05:21:01.049928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T18:11:13.376684Z digest=sha256:b05d4f0fc2a22d458696236786040f16c874be5f45d832c7697e6a9f2cddf09f

Observation a5064ee8-1523-4465-bc78-854d3e1168a2 · inbound

WildDet3D: Scaling Promptable 3D Detection in the Wild cites this paper.

WildDet3D: Scaling Promptable 3D Detection in the Wild LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:26:00.559406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:38:13.336003Z digest=sha256:9a8fd35a4d0bf688ef60d288b21135824c3b379c28be3c72bf525770ba52a5a8

Observation 0fb09e23-dbbf-4626-9f6d-71729ae88fd2 · inbound

Online Reasoning Video Object Segmentation cites this paper.

Online Reasoning Video Object Segmentation LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 53

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T10:26:03.169679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T15:29:03.843441Z digest=sha256:71b072674bd517bef50329f06a19eb373b818fbc58ae25fd74db5f93a3b47424

Observation 4ae23ad6-39e9-4c56-82c5-b0cc1aa60114 · inbound

PixDLM: A Dual-Path Multimodal Language Model for UAV Reasoning Segmentation cites this paper.

PixDLM: A Dual-Path Multimodal Language Model for UAV Reasoning Segmentation LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:23:37.155620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T09:20:54.635375Z digest=sha256:c6b599339a618fc4d052be5bf2b67b9840c2063e1dc245e25c4469df1c3a13a0

Observation ff21ad11-394f-41f6-a46c-ea23bfa7bf56 · inbound

PixDLM: A Dual-Path Multimodal Language Model for UAV Reasoning Segmentation cites this paper.

PixDLM: A Dual-Path Multimodal Language Model for UAV Reasoning Segmentation LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T16:10:02.902806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:10:02.902806Z digest=sha256:83e8a90fbaea5565bd7f60d29c4603721b4319c094e94b9722287a5a8415b837

Observation bb6eae89-5126-4dbd-85e7-df2787c76d0d · inbound

From Failure to Feedback: Group Revision Unlocks Hard Cases in Object-Level Grounding cites this paper.

From Failure to Feedback: Group Revision Unlocks Hard Cases in Object-Level Grounding LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 85

Resolution
verified exact
arxiv_id, observed 2026-05-20T18:43:38.823338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T18:39:11.904941Z digest=sha256:3bbebfd94152c4d8815f31535736b262e5ce1c73e6b9284d280b06fb3b95d65e

Observation ecc43556-e6ee-4412-8332-5cae09ad7d80 · inbound

WOW-Seg: A Word-free Open World Segmentation Model cites this paper.

WOW-Seg: A Word-free Open World Segmentation Model LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-19T21:27:47.990120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T21:23:14.311122Z digest=sha256:02b2df05ab21400ec81e6cc1ec2f9df7550ef5365e885adbc8257b0fbf3a3379

Observation 5bbb6738-9c8a-4fe4-bfba-e452482681fd · inbound

Vision Harnessing Agent for Open Ad-hoc Segmentation cites this paper.

Vision Harnessing Agent for Open Ad-hoc Segmentation LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:53:04.398084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T05:52:40.429412Z digest=sha256:084fa3c4a794d7f7bceb50cb78d005bf4fd9d1a659d8f6ed6aa050a029bce206

Observation 5ffc6271-2490-4f0f-8c69-6b8fbb0dbb25 · inbound

B-GRTO: Bootstrapped Group Relative Tool Optimization for Referring Segmentation cites this paper.

B-GRTO: Bootstrapped Group Relative Tool Optimization for Referring Segmentation LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:30:20.640538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-25T04:27:15.911342Z digest=sha256:27b2122c794a1b81f7d38d22026483d7d444bdfc4688fbd3aa324184dfa17715

Observation 07be46d9-400d-43f4-9488-337d668c1e1b · inbound

B-GRTO: Bootstrapped Group Relative Tool Optimization for Referring Segmentation cites this paper.

B-GRTO: Bootstrapped Group Relative Tool Optimization for Referring Segmentation LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:04:52.970627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:00:40.812912Z digest=sha256:d6ba2a73226cc1b3dadb6ebd3d21c9e8609106f8930ea437fc6ecb5cac2ba7c4

Observation 4ff49ed0-1aca-4349-8b64-8134dcd75b98 · inbound

InstructSAM: Segment Any Instance with Any Instructions cites this paper.

InstructSAM: Segment Any Instance with Any Instructions LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:44:02.057177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-29T22:34:00.442420Z digest=sha256:b942cf751e3df6370a227bffeb2c1569d1cac86d8700c308d8d8c4f7198c6993

Observation 24910f6d-07cf-422e-b5f3-ece4397050b8 · inbound

An Open-Source Benchmark and Baseline for Multi-temporal Referring Segmentation cites this paper.

An Open-Source Benchmark and Baseline for Multi-temporal Referring Segmentation LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:46:14.171607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T17:40:16.265708Z digest=sha256:959e2a2b8a8f83fab5b99ab9b749c8522e150c6277666b69ce5ce2ae3ffddac8

Observation 19b7584a-bc32-4046-8aba-13d03e401982 · inbound

MedSIGHT: Towards Grounded Visual Comprehension in Medical Large Vision-Language Models cites this paper.

MedSIGHT: Towards Grounded Visual Comprehension in Medical Large Vision-Language Models LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:16:58.468492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T01:29:27.353291Z digest=sha256:c8000d1b6d3a41e5f9f1cdf3e0758fbfc0d5cb60ac193ab570202f90a3ef248f

Observation 1d4fe45b-f45e-43ee-b123-c5cadbbc5314 · inbound

Reason Twice: Segmentation via Candidate Discovery and Comparative Reasoning cites this paper.

Reason Twice: Segmentation via Candidate Discovery and Comparative Reasoning LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 84

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T00:47:30.643650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T17:01:13.745646Z digest=sha256:2c318ab7a71bc76a40437cbc2f12fc832b0ccd3c5bc11b1aa2af8bc39025cb53

Observation f3eaae0c-7e12-4fe9-b534-50f7d44475b9 · inbound

CABLE: Cloud-Assisted Bandwidth-efficient LMM-based Encoding for V2X Systems cites this paper.

CABLE: Cloud-Assisted Bandwidth-efficient LMM-based Encoding for V2X Systems LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:59:06.851192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-26T21:35:12.989339Z digest=sha256:315f21acf8e64aff1d2323b5fdf6ae07c7e918027d4ae6d2a2a1378b19d55fa1

Observation 0d660b49-7be2-4f8c-aab7-0ee0e76fd5de · inbound

Dynamic-dLLM: Dynamic Cache-Budget and Adaptive Parallel Decoding for Training-Free Acceleration of Diffusion LLM cites this paper.

Dynamic-dLLM: Dynamic Cache-Budget and Adaptive Parallel Decoding for Training-Free Acceleration of Diffusion LLM LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:33:28.303947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-29T13:27:45.796650Z digest=sha256:6a6222a82c800e0546524a908dc85dbb1e8f5d0dfe0e94ce555f21d0b45e3961

Observation 2d52c394-2ed8-4cbf-972e-56ff8d57e683 · inbound

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models cites this paper.

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:09:55.038018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-26T01:50:54.242508Z digest=sha256:6ad5fb87acf44f9429705f23881176a48dcb84f9576ff5f7547ed48d07dcd1b8

Observation c5298ee2-f63a-4f7a-aa66-fbad1f611bc4 · inbound

CCRC: A Change-Aware Captioning and Reasoning Chain for Image Change Captioning and Segmentation cites this paper.

CCRC: A Change-Aware Captioning and Reasoning Chain for Image Change Captioning and Segmentation LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-06-30T10:04:36.252137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T09:58:18.387486Z digest=sha256:826aff8d1d774fe79a0f6cdb09add9a93a0cd712942d4138e0382e14f1c78b27

Observation 578b0448-f727-4d2e-aeb4-fc2f4c601403 · inbound

InstanceControl: Controllable Complex Image Generation without Instance Labeling cites this paper.

InstanceControl: Controllable Complex Image Generation without Instance Labeling LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 55

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:15:45.101408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-01T05:37:41.030752Z digest=sha256:4142ec9a1241e0469210f2b0d285080655a482dab007a8eb487db2a3ed1cfdfa

Observation 9fe1f4f0-f427-4707-9a59-8a177d12232d · inbound

Vision as Unified Multimodal Generation cites this paper.

Vision as Unified Multimodal Generation LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 204

Resolution
verified exact
local_arxiv, observed 2026-07-08T02:04:26.296286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T01:54:30.649092Z digest=sha256:94d9b2823074e73e7f9562d3f0c84691944270f70494dbfdaecab55d658cf175

Observation ac68ccb2-f914-4178-9ab5-9416b617993c · inbound

Reasoning-Guided Part-Level Visual Grounding via Reinforcement Learning cites this paper.

Reasoning-Guided Part-Level Visual Grounding via Reinforcement Learning LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-01T23:38:50.776533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:38:50.776533Z digest=sha256:d0fcf964046b1e15600729daa1c366934412bdd714053966105f73c94a35546f

Observation b69baf3e-d8ac-470a-b82b-1b371b2783bc · inbound

ReferTrack: Referring Then Tracking for Embodied Visual Tracking cites this paper.

ReferTrack: Referring Then Tracking for Embodied Visual Tracking LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T11:00:26.207266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:00:26.207266Z digest=sha256:ff03adeb09ec22711f1134c33e173a4c60505c40c3fab5b96a7b2d263de3cec8