Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:03:03.566081Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 96 of 96 outbound references and 0 inbound Pith citation observations for arXiv:2507.09334.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:03:03.566081Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
96 of 96 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f48ccb24-04c1-4af9-ab85-115adccb60b5 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c97d3428-4d4b-4ca9-bfc9-53a764938a49 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 451f4ec4-0adb-4693-b0e8-cf0c196b7a00 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ed70302-e3d8-48dd-b8b9-60d04bc948aa · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Token Merging: Your ViT But Faster
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b25b53c-588f-4af3-9966-29ef4776d0ad · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4cbb8fe-1c56-4b2f-8079-d6820a086803 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 772f2383-4452-4096-8cc5-fa90cab7567e · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e65442a-8ef8-44e2-86c1-fa56cf9150ee · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Efficient Large Multi-modal Models via Visual Context Compression
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b05cc5d5-aee2-4d4d-885f-6acb86921161 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 015c0e41-aed1-4605-bda8-030c0e18aaa1 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07554a33-1893-4957-9184-ecb0773e65cf · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c7f5ec5-5aab-427a-bf75-90febf63009d · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 428ab502-aaaf-46c8-9e4d-2c368183afa6 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Grounded 3D-LLM with Referent Tokens
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58363efd-d264-40bf-8056-4fe131cdcc46 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7854f2f5-0054-4eab-884b-f903fa972108 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b81c9a9c-0bd5-4dd9-aa3b-85195db93ebe · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa1e8111-6433-4fed-b898-564526b8756d · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc7c0d4e-de84-4322-8dfb-1d57500b0b96 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding MobileVLM V2: Faster and Stronger Baseline for Vision Language Model
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c344f0a6-47bd-4900-b813-e89346807cdc · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9261afb5-5d0a-4769-93d5-a5c5a9b0cb68 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68ae865f-2237-4e64-b034-6bb324ef0489 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18c0dbd6-b05f-4c15-a4df-c38a6f616566 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Point-Bind & Point-LLM: Aligning Point Cloud with Multi-modality for 3D Understanding, Generation, and Instruction Following
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11b29fae-b0ba-4098-9893-4791cfeeef98 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding ImageBind-LLM: Multi-modality Instruction Tuning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dac39ce8-5bc7-49d5-bca1-aad175bc498d · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12dbf14c-b7ad-45d6-a642-dda8bf02b171 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3b9fa8c-0657-4b68-ac8d-f777b35998c1 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee98cb67-0a3a-478b-9812-a7ad333c6f60 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3eb360f5-3190-4106-8bdc-1133acb30210 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d929c75e-a43d-4e5b-a955-1e35386283a2 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 14f17569-c587-46af-bf64-25ab994e7398 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Chat-Scene: Bridging 3D Scene and Large Language Models with Object Identifiers
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d6927c8-a06d-4f62-82f4-2feec4d9382f · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding An Embodied Generalist Agent in 3D World
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31965077-7e67-48fe-aaa2-ced25d9f7d75 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e7bd4f78-18c4-45ac-8ee2-59ab9ed91133 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 943b7901-d968-4229-9a6e-ea74284ac571 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4d279c18-dd29-4371-a6dd-b9afe731d86f · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding What Kind of Visual Tokens Do We Need? Training-free Visual Token Pruning for Multi-modal Large Language Models from the Perspective of Graph
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 263f6650-c2bf-406b-bdf2-e5da196521b9 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a944f717-4782-49d0-8b83-340b38a70338 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Robin3D: Improving 3D Large Language Model via Robust Instruction Tuning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7346c2f8-6b98-4f16-8a59-a2848d296f0d · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58936284-d05d-4d50-9943-35cc6971894e · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding RedundancyLens: Revealing and Exploiting Visual Token Processing Redundancy for Efficient Decoder-Only MLLMs
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c424190-ddff-486c-a6ef-16aaf8b4e4ef · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding TokenPacker: Efficient Visual Projector for Multimodal LLM
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d46be39-c1eb-43b0-a664-74208764c257 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba871690-ba73-4293-b9cb-ef2c964ae505 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Not All Patches are What You Need: Expediting Vision Transformers via Token Reorganizations
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b21d098-658d-4f82-8300-bd55089c689f · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Boosting Multimodal Large Language Models with Visual Tokens Withdrawal for Rapid Inference
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5ab953b-5a41-4c58-abc4-31cbd716dbe0 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4209fa98-41a5-4b90-8146-f2305a0afdd0 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding A Survey on Text-guided 3D Visual Grounding: Elements, Recent Advances, and Future Directions
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63d6651f-e8da-4f2f-b415-e3cce06f15ba · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding A Survey of Attacks on Large Vision-Language Models: Resources, Advances, and Future Trends
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aaad7d22-b811-41a8-bde0-57f758a96a4a · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c4dfc515-b507-46d5-bb4a-5c2bee356090 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e495855-3c2f-4e86-871d-c49d21bf92f7 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Decoupled Weight Decay Regularization
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c05b8ad3-7462-4d43-a348-e4e2897b91a0 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding SQA3D: Situated Question Answering in 3D Scenes
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc8d66a7-de7c-48f2-ab20-092c5b8e0252 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c6747194-ffa5-4fad-a468-8e298098a2ef · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding DeepStack: Deeply Stacking Visual Tokens is Surprisingly Simple and Effective for LMMs
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4fb18f8-b5c6-409b-bd09-a7820b23a8e8 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding DINOv2: Learning Robust Visual Features without Supervision
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa459827-2401-489d-9602-c3cc74856acb · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f00de20a-bd16-44f9-a5d5-e95b21d78b99 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation beb01344-99fc-4f67-8d51-9643a7b40d2b · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 10c1911f-5be5-43f9-a8e1-23036f06c7d8 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85fc4b70-9b34-4241-9f9d-26410c2934b0 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3d5e37a-1175-423d-9e07-11cfb0186ccc · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1d2efe2b-79bc-4dd7-8171-3abf4037646f · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 826bb000-010a-4fe9-af72-5ff6515c331f · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding CrossGET: Cross-Guided Ensemble of Tokens for Accelerating Vision-Language Transformers
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f625af16-3a32-43a3-b587-62a26cd8d262 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Advances in neural information processing systems 34 (2021), 13937–13949
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e66426c-04bb-4b6e-ba8d-17fa51174131 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding LLaMA: Open and Efficient Foundation Language Models
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd9a1140-07e0-493f-9aba-660ef3bda262 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cd510ac-b256-40bb-8011-9c7fe44dc19e · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7e98b14-bf4d-4ba0-b598-48ecc44be977 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12439054-5c02-4e68-95df-1195896b9604 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 31a64dc9-6dae-47d2-8735-d4d256d72e78 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c0768b6-6793-4731-a6e9-a2dbe0c99a8a · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6db3fe3b-5dfe-4b00-85e9-0b3513d74123 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b402ff32-e7ce-435a-9dfc-b27fa2980217 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bf08dcb9-4566-433e-b52b-f72d9f749cab · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding [CLS] Token Tells Everything Needed for Training-free Efficient MLLMs
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 710f3385-21c6-4bd4-b355-9f7a54843c1a · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding DeCo: Decoupling Token Compression from Semantic Abstraction in Multimodal Large Language Models
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d18b4bf-9bd0-4340-9b6c-470f3b9efc82 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Fit and Prune: Fast and Training-free Visual Token Pruning for Multi-modal Large Language Models
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 093a4a1a-8679-4475-b912-ffb8858f64e3 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding In Findings of the Association for Computational Linguistics: NAACL
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 787ddb93-6028-4bca-8479-89bbd6852216 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Chat-3D: Data-efficiently Tuning Large Language Model for Universal Dialogue of 3D Scenes
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1643fde7-2768-4b46-bc1c-8a6a60bbdb1c · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding 3DGraphLLM: Combining Semantic Graphs and Large Language Models for 3D Scene Understanding
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69cbbbb6-fb80-4d49-9b2c-da94686edd61 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4583d938-0a81-440b-aa6e-02f54282bb68 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 35ef04ae-baf7-481c-b4bf-7ee6ea162144 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Accelerating Multimodal Large Language Models by Searching Optimal Vision Token Reduction
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5403498-287f-4bc9-8c20-0ef179aba5b5 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 84c13542-3f1b-49c4-b010-e4b1694cf24c · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding VoCo-LLaMA: Towards Vision Compression with Large Language Models
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b558d08-a2d2-449c-80fc-1505b8cbf49d · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b344fd13-b7d1-46a2-af72-80b526c10748 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding AIM: Adaptive Inference of Multi-Modal LLMs via Token Merging and Pruning
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebf472a7-2538-4504-9b01-4bad846c8f82 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 098af109-893c-4b64-a642-5475658e6011 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding SparseVLM: Visual Token Sparsification for Efficient Vision-Language Model Inference
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 688744cc-7d1a-49a3-a462-de351c6c7125 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Unresolved cited work
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bded14ac-604b-402d-8a64-1dfa3143d598 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding A Stitch in Time Saves Nine: Small VLM is a Precise Guidance for Accelerating Large VLMs
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87499f39-b287-4cd0-b263-0a6eadf54e95 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Dynamic Diffusion Transformer
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a42d758-339c-4ce3-842c-429e342f8015 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Dynamic Tuning Towards Parameter and Inference Efficiency for ViT Adaptation
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5276eb9e-4006-4c13-852f-11657fbe3d35 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding Uni3D: Exploring Unified 3D Representation at Scale
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88aba4f5-63ef-40eb-bcbc-6f0cab7a4d4c · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6237571-f163-467e-89ff-16f506e8ad34 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding ST$^3$: Accelerating Multimodal Large Language Model by Spatial-Temporal Visual Token Trimming
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae7110bb-e2d5-42b2-8559-7285a281dae2 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding IEEE Transactions on Pattern Analysis and Machine Intelligence 44, 11 (2021), 7436–7456
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65fb1d5d-d32b-4f50-9781-e7455cb29b16 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79870942-2f3b-4997-919b-9adaf1f622a7 · outbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding GPT4Scene: Understand 3D Scenes from Videos with Vision-Language Models
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.