Pith. sign in

Paper Citation Record · LEDGER

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching

As of 15 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:2608.04568.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04568 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:57:45.777325Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact2
  • verified fuzzy32
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 01067106-b163-41d0-b469-ab78d1f6b35a · outbound

This paper cites A survey on text-guided 3-d visual grounding: Elements, recent advances, and future directions,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching A survey on text-guided 3-d visual grounding: Elements, recent advances, and future directions,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:52.297649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:41.806149Z digest=sha256:79f9a973afd4a8dce246d20a83a933c329095d9c1ae9dfd5e8e6f91c9d22bff7

Observation 8c160fc7-fdd4-449f-bcc2-c14f12dac151 · outbound

This paper cites Mtrag: Multi-target referring and grounding via hybrid semantic-spatial integration,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Mtrag: Multi-target referring and grounding via hybrid semantic-spatial integration,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:52.126953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:41.932074Z digest=sha256:a29f98fc83b8c1e119cf23c9462332c1e5546320113dd8102d7d16eb7900eddc

Observation 6b7981d5-e9dd-428c-ba4a-716c2bcd9b85 · outbound

This paper cites Visual grounding in 2d and 3d: A unified perspective and survey,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Visual grounding in 2d and 3d: A unified perspective and survey,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:51.962238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:42.051937Z digest=sha256:63c5236917565187691a81cc2d3b081f6e10f51c135e1054dee906eb802590dd

Observation 153f2c58-1f1d-49db-a43e-f291a1817147 · outbound

This paper cites Embodied intelligence: A synergy of morphology, action, perception and learning,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Embodied intelligence: A synergy of morphology, action, perception and learning,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T21:57:42.197985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:57:42.197985Z digest=sha256:5cc7a6aece07663c68f8d7d3a4dbe89a8abeccba29707d6b034d01a49b506eb6

Observation 1adf9cc9-a459-4094-a301-5e1402c64481 · outbound

This paper cites Scanrefer: 3d object localiza- tion in rgb-d scans using natural language,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Scanrefer: 3d object localiza- tion in rgb-d scans using natural language,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:51.734573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:42.337529Z digest=sha256:4f823677352eb9d84d9023c5a4274ed43596f53af332de39a979093f4edfc4b9

Observation c645b4e4-6b7f-41a9-9e42-0a4eeb723348 · outbound

This paper cites Refer-it-in- rgbd: A bottom-up approach for 3d visual grounding in rgbd images,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Refer-it-in- rgbd: A bottom-up approach for 3d visual grounding in rgbd images,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:51.603127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:42.424629Z digest=sha256:549212f00903e984659c8e2675e366436f731c9353d7dadb02fb22ad87acd81b

Observation a030a884-56c8-4458-b1f4-06f753100db2 · outbound

This paper cites Referit3d: Neural listeners for fine-grained 3d object identification in real- world scenes,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Referit3d: Neural listeners for fine-grained 3d object identification in real- world scenes,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:51.415723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:42.538307Z digest=sha256:536d07d3dc56c1c5a6ced7e057afe547525fa875a4f200506a653c41aab4031b

Observation 3098abe7-50b0-46a2-9f35-a1aa20e97dbb · outbound

This paper cites Talk2car: Taking control of your self-driving car,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Talk2car: Taking control of your self-driving car,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:51.211165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:42.597982Z digest=sha256:e64986bf9974a3092eb99dfc3e8e8a62d1e5b3b04a63f18bb1668dff2d5f64cb

Observation ce6e5409-1676-414a-8ddc-7a2563c80b4c · outbound

This paper cites Talk2radar: Bridging natural language with 4d mmwave radar for 3d referring expression comprehension,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Talk2radar: Bridging natural language with 4d mmwave radar for 3d referring expression comprehension,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:51.042148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:42.646863Z digest=sha256:43ae1322969df0dd411e1de3a7107b870429429e3c70ca05e3af67f938333660

Observation a7889498-77b9-4f6f-88d8-0862c502ad58 · outbound

This paper cites Mono3dvg: 3d visual grounding in monocular images,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Mono3dvg: 3d visual grounding in monocular images,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:50.885811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:42.731019Z digest=sha256:c6c4076e5d554c6c090e779275d61701a2341f89430b7c8ba21876c564eff62d

Observation 081d450a-d864-4afe-a8fb-fd56462a9f23 · outbound

This paper cites Enhanced vision- language models for diverse sensor understanding: Cost-efficient opti- mization and benchmarking,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Enhanced vision- language models for diverse sensor understanding: Cost-efficient opti- mization and benchmarking,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:50.718754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:42.824598Z digest=sha256:176e39caba327849864f38a64748e91ec27a127d1747bef615c57797fb4a2ae8

Observation 8f9abcd8-f852-4a81-a728-458a0bd885d0 · outbound

This paper cites Seeground: See and ground for zero-shot open-vocabulary 3d visual grounding,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Seeground: See and ground for zero-shot open-vocabulary 3d visual grounding,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:50.500173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:42.936734Z digest=sha256:029e0bee1cd16f9642f84116f9f7206cb21d8c2fe8fb4b87f3215a202a9d0ebe

Observation ffd6a5ba-470f-4f69-b54e-7afedb5fe9b4 · outbound

This paper cites Enhance 3d visual grounding through lidar and radar point clouds fusion for autonomous driving,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Enhance 3d visual grounding through lidar and radar point clouds fusion for autonomous driving,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:50.300520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:43.090292Z digest=sha256:bdc103d11531cee2329c9fd5f1f91e48bdea278dc8273525356315e31142aaff

Observation c9d24e2b-f1bd-4930-8899-6eae09bbf57d · outbound

This paper cites Mmdrive: Interactive scene understanding beyond vision with multi-representational fusion,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Mmdrive: Interactive scene understanding beyond vision with multi-representational fusion,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:50.106770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:43.260049Z digest=sha256:9a2f854a4b40c12bd42045961a4cfdf8ec715c62a634a518dd06dff743b232c8

Observation 8dab6a81-c298-48e4-a129-10c04d7fd705 · outbound

This paper cites Bevfusion: Multi-task multi-sensor fusion with unified bird’s-eye view representation,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Bevfusion: Multi-task multi-sensor fusion with unified bird’s-eye view representation,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T21:57:43.433073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:57:43.433073Z digest=sha256:c64f5520aaa5c3f54cca6d88d47e3e08357747f31233907ae1fd4940408e1b43

Observation 430d9186-b5a8-41be-9c02-d001907c94a0 · outbound

This paper cites Futr3d: A unified sensor fusion framework for 3d detection,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Futr3d: A unified sensor fusion framework for 3d detection,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:49.848623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:43.620088Z digest=sha256:4a0687ac3b768f7af566d3da991be9b454bd34027904afdd7fa44b44418c97e6

Observation 10021b73-a009-4b65-80a5-1ffc4ce6f04d · outbound

This paper cites Gated multimodal units for information fusion,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Gated multimodal units for information fusion,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:49.617681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:43.735594Z digest=sha256:7757e752a2225a600568dc120fb462c1935e00de86be3a10fb689abfe1d717fa

Observation 4caeb6ea-20d2-4eeb-a2e9-fa8a1095004d · outbound

This paper cites Squeeze-and-excitation networks,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Squeeze-and-excitation networks,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T21:57:43.903160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:57:43.903160Z digest=sha256:ebfd6cb559a9da9d6f4e6c575cef1e5c5dfac485209889e7feb63ba89a318c89

Observation 7ef8d712-fe90-4fb4-bad4-22cdfa023ccf · outbound

This paper cites Talk2Radar: Bridging natural language with 4d mmwave radar for 3d referring expression comprehension,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Talk2Radar: Bridging natural language with 4d mmwave radar for 3d referring expression comprehension,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:49.418306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:44.083264Z digest=sha256:a87f094b84c576ef6352b9144d1a624091d00e632bd41d3cb9dbd406074c035b

Observation 0da53d71-2db6-4c49-8f78-3fd0b2d0b550 · outbound

This paper cites Detr3d: 3d object detection from multi-view images via 3d-to-2d queries,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Detr3d: 3d object detection from multi-view images via 3d-to-2d queries,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:49.239172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:44.158775Z digest=sha256:11b8e22edb32a503a5cc39e2bd367601908c3fa0a21c40d4c5b68b65f27a4c75

Observation 19e36284-b4f2-4edf-be4b-0a07eb1a4a3a · outbound

This paper cites Rolic: A robust lidar-camera fusion frame- work for 3d object detection,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Rolic: A robust lidar-camera fusion frame- work for 3d object detection,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:49.050566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:44.236709Z digest=sha256:fd9f1b6a7a3bfe464830464301a86ecf0aea60295cfffa09862acbafa59e43ef

Observation f43309fb-58fd-420b-8aa0-6e6f8041fa2c · outbound

This paper cites Omnihd-scenes: A next-generation multimodal dataset for autonomous driving,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Omnihd-scenes: A next-generation multimodal dataset for autonomous driving,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:48.839493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:44.310855Z digest=sha256:dfbe0c6595b9d5b5545063fc3cf1fd0f268df4cdf8b4b5bd9adb2b4f990f0a6c

Observation 711d8ef9-099f-4126-a324-92144389d187 · outbound

This paper cites Availability-aware sensor fusion via unified canonical space,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Availability-aware sensor fusion via unified canonical space,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:48.633471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:44.383781Z digest=sha256:1abc1212cdb808651a09b6d0a4d8fa74834644bc209c39f61f30c78999c172e9

Observation 6d84c426-bbb7-460f-8c2f-650b9d47b97d · outbound

This paper cites Samfusion: Sensor-adaptive multimodal fusion for 3d object detection in adverse weather,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Samfusion: Sensor-adaptive multimodal fusion for 3d object detection in adverse weather,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:48.445002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:44.451338Z digest=sha256:e0bb98424672e882c64dba49c01e8308fd630021e64f886a7c3432a117941912

Observation 08f9f94c-e228-48d5-b829-a3e274275eec · outbound

This paper cites Boosting faithful multi-modal llms via complementary visual grounding,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Boosting faithful multi-modal llms via complementary visual grounding,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:48.225299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:44.522616Z digest=sha256:0a463650122e2bc036d8d0e4fb58e6b2b0948ead161188ad64fe2814e63e8134

Observation f697fe4b-0e12-45be-b114-786878442a14 · outbound

This paper cites Talk to parallel lidars: A human-lidar interaction method based on 3d visual grounding,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Talk to parallel lidars: A human-lidar interaction method based on 3d visual grounding,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:48.029113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:44.602154Z digest=sha256:1044e1ddbad506b009bc9b49dfc5ce51100df18f13f75e6c91758390fb2c7f93

Observation 8a598bca-b01d-4add-800f-2f6eba6a5c36 · outbound

This paper cites LidaRefer: Context-aware Outdoor 3D Visual Grounding for Autonomous Driving.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching LidaRefer: Context-aware Outdoor 3D Visual Grounding for Autonomous Driving

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:57:46.151105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:44.663636Z digest=sha256:5c3efc47649c882809523739aba3413335c1fb2370c410b6bc5e51f2e176b272

Observation 8c1817b4-5048-4e09-afcd-ad3ad672c5b8 · outbound

This paper cites Multi-sensor fusion technology for 3d object detection in autonomous driving: A review,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Multi-sensor fusion technology for 3d object detection in autonomous driving: A review,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:47.818460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:44.746664Z digest=sha256:ea056f3678054651d227e08d4bced2d2a79a9988bc136ab98a131cb5e70a633d

Observation 9ae4e3dd-2fed-4593-bc0c-3c1967914179 · outbound

This paper cites Language-Guided 3D Object Detection in Point Cloud for Autonomous Driving.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Language-Guided 3D Object Detection in Point Cloud for Autonomous Driving

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-06T21:57:45.975290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:44.812965Z digest=sha256:8b3c2dd30cc0476de904bf63a752525b85b6513a3b9fee6d03fd61b9def542ca

Observation 31166c1b-4236-4518-a8fd-73ec82de9a31 · outbound

This paper cites Gpt-4 enhanced multimodal grounding for autonomous driving: Leveraging cross-modal attention with large language models,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Gpt-4 enhanced multimodal grounding for autonomous driving: Leveraging cross-modal attention with large language models,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:47.607611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:44.914171Z digest=sha256:23a584d2c95574a690400d44817587b1a8f221229acebe700edbc12bd6d061fc

Observation 6cda9f3c-6ead-4fd3-ada3-78caa908f4ba · outbound

This paper cites NuGrounding: A Multi-View 3D Visual Grounding Framework in Autonomous Driving.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching NuGrounding: A Multi-View 3D Visual Grounding Framework in Autonomous Driving

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T21:57:44.978031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:57:44.978031Z digest=sha256:2e8875bb1dc00b0e8cb257f649b701dc1ca139e9c6d40940a7051fa19d748881

Observation 174dfef7-f6eb-4c34-9be9-a3e11ae3c7dd · outbound

This paper cites VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:57:45.060750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:57:45.060750Z digest=sha256:79dfe51f05369fd9bbe08c8cdc43baf03e1ffe9e20576302db9b8ab85316ac0c

Observation ac39e5f9-14d5-4b0f-a38e-38f1ab008acd · outbound

This paper cites End-to-end object detection with transformers,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching End-to-end object detection with transformers,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:47.394196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:45.126474Z digest=sha256:14af44928a3cd5fa5f953fc789e1c97e35e756787b046e2a7568fc4f47497761

Observation fe5d6e02-ea1f-46fa-b8ba-8776989519fd · outbound

This paper cites Pointpillars: Fast encoders for object detection from point clouds,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Pointpillars: Fast encoders for object detection from point clouds,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:47.220879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:45.196964Z digest=sha256:2551991a7b1efc5729726df0b87bf7a92489d0714af099ced2012e75da694ab2

Observation 73ae91af-72f6-4a1c-9d74-c0c2f963f61b · outbound

This paper cites Second: Sparsely embedded convolutional detection,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Second: Sparsely embedded convolutional detection,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T21:57:45.286413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:57:45.286413Z digest=sha256:515a8174632b50132ef7775d020c367cb8544784ea1bbeb02c3ebf583554b8ad

Observation 47bf3d36-e1bd-41af-b1e2-7909855cc90d · outbound

This paper cites Generalized intersection over union: A metric and a loss for bounding box regression,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Generalized intersection over union: A metric and a loss for bounding box regression,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:47.006238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:45.358362Z digest=sha256:3af433da522418d03c5420153a04ea82f42209071894a6837fc410b505d7ab72

Observation 8d0ec622-11a1-4a3d-b278-4c37ec316e0a · outbound

This paper cites The hungarian method for the assignment problem,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching The hungarian method for the assignment problem,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T21:57:45.428732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:57:45.428732Z digest=sha256:c1dfd03071230a6e67965a4f193c450048e0812b80e4cd01bd464f1710ef3415

Observation 005210b2-148b-465c-9970-139967108447 · outbound

This paper cites Multi-class road user detection with 3+ 1d radar in the view-of-delft dataset,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Multi-class road user detection with 3+ 1d radar in the view-of-delft dataset,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T21:57:45.507621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:57:45.507621Z digest=sha256:fe73af1ff3e4a8dc42e25b812bb8b02c713687f7d192ff756609dbc483ab4860

Observation 074f90fb-9f35-4a8d-ae62-89c0305746fa · outbound

This paper cites A transformer-based framework for visual grounding on 3d point clouds,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching A transformer-based framework for visual grounding on 3d point clouds,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:46.850307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:45.582303Z digest=sha256:aa40caa4d8013cc4795230473682493554718caaf6c6357fc7aec21031ca1d2b

Observation b314e50b-e1c6-4ae4-9390-8ae4e8b6cc5f · outbound

This paper cites Eda: Explicit text-decoupling and dense alignment for 3d visual grounding,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Eda: Explicit text-decoupling and dense alignment for 3d visual grounding,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:46.677343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:45.643557Z digest=sha256:ed7de06fd25d150a763f3d0843d86c6b90e56461efa60fa7dec4c72df73d8b25

Observation 9d0c7b5f-a8e2-43e9-a9ae-dbef2b272fc4 · outbound

This paper cites Ges3vig: Incorporating pointing gestures into language-based 3d visual grounding for embodied reference understanding,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Ges3vig: Incorporating pointing gestures into language-based 3d visual grounding for embodied reference understanding,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:46.492125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:45.712106Z digest=sha256:17ea4bfc35f7ceeea6bb54a415544c899dbfb4033146020663a309d92d534a1c

Observation 0a8d75a1-d905-48fc-a0bb-c5a63b6b1524 · outbound

This paper cites Text-guided sparse voxel pruning for efficient 3d visual grounding,.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching Text-guided sparse voxel pruning for efficient 3d visual grounding,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:57:46.339849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T21:57:45.777325Z digest=sha256:3c14633ade52d2b1089ee24386742bac853b42001887e204aa46251972cab8fe

Pith citing papers

No inbound Pith citation observations are available.