Pith. sign in

Paper Citation Record · LEDGER

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding

As of 7 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 0 inbound Pith citation observations for arXiv:2506.09634.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.09634 v1

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:47:08.060954Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

52 of 52 outbound references displayed

  • verified exact1
  • verified fuzzy24
  • unresolved26
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 59d1df71-1389-4d58-8a2c-ee09b32c4a33 · outbound

This paper cites Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.620921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.620921Z digest=sha256:d5834da560e6dbe4eba10bfec99d556ecfeb84f1c4bc58ff4894e4f3c3235240

Observation 656efc58-2110-4998-a288-a64b06889b1a · outbound

This paper cites DCFormer: Efficient 3D Vision-Language Modeling with Decomposed Convolutions.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding DCFormer: Efficient 3D Vision-Language Modeling with Decomposed Convolutions

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.624830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.624830Z digest=sha256:8b00628560d84e89e93b00d09ec1540fbdc4330c6a27f71533ccec94ebb945bc

Observation 7a954a7c-d5a8-4442-ba12-c67648decd15 · outbound

This paper cites M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.628602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.628602Z digest=sha256:4e910300bd66690fb63ae453c3d9ebb8e9f7dc13af8aa7766dc391db7a7aa757

Observation 1c5b0f29-76a7-4997-9239-afd5e0597e8c · outbound

This paper cites Shah, Andrew Johnston, Robert D.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Shah, Andrew Johnston, Robert D

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.631850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.631850Z digest=sha256:b39a9c9258389e1ebb897d98e541be08b940cb805cea498ff019217bcdd0a3f0

Observation ddf966f4-e600-4c6d-aa4b-73395c2e83df · outbound

This paper cites Bruno, Eric A.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Bruno, Eric A

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:14.534106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.634653Z digest=sha256:f65c5706307179d0aa8f990afa259d524358d98e5052929ae04b82838e31ed70

Observation 4b26adec-b0b4-4955-ab61-8013d1320908 · outbound

This paper cites 3D-CT-GPT: Generating 3D Radiology Reports through Integration of Large Vision-Language Models.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding 3D-CT-GPT: Generating 3D Radiology Reports through Integration of Large Vision-Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.637441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.637441Z digest=sha256:8e2d54174ebe709c0ce16ea8548a642b66179eac78e1fa976c18732bf39d2ef7

Observation 48dcebf1-4bf3-43a9-822f-079ce5e67f00 · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.641054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.641054Z digest=sha256:5758a37d7d491ae740b338c529eb85610fd4dec59eaa556efe1b5975f80cf613

Observation 95305907-eca2-414a-9d4f-20ead735841f · outbound

This paper cites Dragonfly: Multi-Resolution Zoom-In Encoding Enhances Vision-Language Models.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Dragonfly: Multi-Resolution Zoom-In Encoding Enhances Vision-Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.644114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.644114Z digest=sha256:149890d2d040da5d1bbb49a0f07af1f069ff50a2bf5c3f280d64b0fe6964f22c

Observation a76d515b-13ab-4ff0-8da7-1aa827f86195 · outbound

This paper cites MedBLIP: Bootstrapping Language-Image Pre-training from 3D Medical Images and Texts.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding MedBLIP: Bootstrapping Language-Image Pre-training from 3D Medical Images and Texts

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.647075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.647075Z digest=sha256:384488ff2edfc6b5efcea5b360f097b9b71208fa138d50c21375687f6b9032b4

Observation 37b70fec-5257-4c93-851d-29b00c830b61 · outbound

This paper cites BIMCV-R: A Landmark Dataset for 3D CT Text-Image Retrieval.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding BIMCV-R: A Landmark Dataset for 3D CT Text-Image Retrieval

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:14.321988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.650006Z digest=sha256:170e09006317f2886546e6feb25d9c52fa4c0344d129670a33efaee756b39b8e

Observation 3a2baa62-f388-436a-95e3-6391794fb94d · outbound

This paper cites Dia-LLaMA: Towards Large Language Model-driven CT Report Generation.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Dia-LLaMA: Towards Large Language Model-driven CT Report Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.652832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.652832Z digest=sha256:23d4f0b18cb9c6c271927432083d4f256f141b67e9918c1f4831b7c0ae9670fc

Observation 594844b1-255a-40dc-ad4d-807471d347b9 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:14.100340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.655589Z digest=sha256:740d699e0d5cfacd029764f8914de5abff0a0b8acb6ea6e234dbf12467f9f741

Observation 124937a7-342c-49b7-8b75-0f59990804b1 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:13.968497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.658248Z digest=sha256:2d51bd9a9efb7d4272bcc15bb653c4a7d01de4265aa409e05e52504d2b74b60b

Observation 2b779508-9103-42b8-a0cb-631d79a054b6 · outbound

This paper cites Developing Generalist Foundation Models from a Multimodal Dataset for 3D Computed Tomography.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Developing Generalist Foundation Models from a Multimodal Dataset for 3D Computed Tomography

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.660922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.660922Z digest=sha256:8cc0f4023c6e7bd6038fd6b9bd749269a49d693f11c51a37a92b115712728893

Observation 4feccbb4-ac79-4e5d-ad6f-7d192f85fe18 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding LoRA: Low-Rank Adaptation of Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.663591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.663591Z digest=sha256:92dfde49a4b065137ed80005dacb4477a8b756f571322b96c59306e55d7b83b5

Observation f5fff607-8b6c-4f1d-923e-6379ef9deebd · outbound

This paper cites Ball, Norah Borus, Andrew Huang, Bhavik N.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Ball, Norah Borus, Andrew Huang, Bhavik N

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:13.708650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.666325Z digest=sha256:fab7c45d794d14feda48ede8625135b0aa62ef8e3ce4a6d52a07e0fa42c8a6fb

Observation b356897e-0e06-49ab-90ae-43215cb14aa4 · outbound

This paper cites Lungren, und Serena Yeung.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Lungren, und Serena Yeung

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:13.488370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.669216Z digest=sha256:12bfc43aafecccd593b94ffb3e67b5515bfe4d89af61fe4051633ba2f3f543e1

Observation 556ea59f-01cf-47ae-9322-af38c4d2464b · outbound

This paper cites an unresolved cited work.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:13.321335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.672081Z digest=sha256:5a17a41b554eadf04209b124da4575857d02bf5955003e8df54b15f6fdaeaaad

Observation 79a40d5c-1666-4444-a0c4-39568d4a1496 · outbound

This paper cites MDAgents: An Adaptive Collaboration of LLMs for Medical Decision-Making.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding MDAgents: An Adaptive Collaboration of LLMs for Medical Decision-Making

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:13.159421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.675122Z digest=sha256:5763baa874e00d7a9243685d731becb0832291f90092d0c3a70222e476bc7854

Observation 3cbc623f-fd10-44ae-abf3-1f67173aa724 · outbound

This paper cites E3D-GPT: Enhanced 3D Visual Foundation for Medical Vision-Language Model.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding E3D-GPT: Enhanced 3D Visual Foundation for Medical Vision-Language Model

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.677788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.677788Z digest=sha256:3647574c8e3d10675ef3e5671d78000a028fe7aa400c966d2ff151df2e69d472

Observation f398a14c-3cd1-4287-a0f3-2c58b6662927 · outbound

This paper cites Kevin Zhou.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Kevin Zhou

Reference 21

Resolution
verified exact
raw_fallback, observed 2026-08-07T04:47:08.567433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.680780Z digest=sha256:f1a14c56e9c41626a54375be89223bdc0c6e5cbaeec74f98e50985da2d031f2d

Observation a5794760-a898-4bd0-beee-2b7d456c3087 · outbound

This paper cites METEOR: An Automatic Metric for MT Evaluation with High Levels of Correlation with Human Judgments.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding METEOR: An Automatic Metric for MT Evaluation with High Levels of Correlation with Human Judgments

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:12.937482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.683384Z digest=sha256:14e078db441bf38810809e25d0e19558ce7fbb58379495357e65099f024eba49

Observation e444ef61-6a4c-40c4-b933-5116bbda84a8 · outbound

This paper cites Towards a holistic framework for multimodal LLM in 3D brain CT radiology report generation.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Towards a holistic framework for multimodal LLM in 3D brain CT radiology report generation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:12.755672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.686018Z digest=sha256:28dd93df6431fc0ebfc74d35d57221ed67441488c253e01f1eeebc704efda35e

Observation 7e10a83b-decf-4894-b774-0bef881cfbcb · outbound

This paper cites LLaV A-Med: Training a Large Language-and- Vision Assistant for Biomedicine in One Day.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding LLaV A-Med: Training a Large Language-and- Vision Assistant for Biomedicine in One Day

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:12.584478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.688760Z digest=sha256:085615ae4fee9ba0572b56f055d0485ed7a5686a6277b2bd7930627e8eba7f52

Observation ff9013ea-85a7-4230-833d-a2291a85dc49 · outbound

This paper cites an unresolved cited work.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:12.340999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.691469Z digest=sha256:b765230cae631dd1d4b2d182c77630a8d57436cd77810166374d80d53d2bfee2

Observation 1e1d1347-ba9f-44c9-99d4-540f68fcd32a · outbound

This paper cites Dynamic Graph Enhanced Contrastive Learning for Chest X-ray Report Generation.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Dynamic Graph Enhanced Contrastive Learning for Chest X-ray Report Generation

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:12.149866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.694107Z digest=sha256:9e5e2d8e588bb065fa7816781172175953d0d65f06e16af13ef726bf05a51171

Observation 4adaa931-5207-495e-bbd1-7f0ca55cb3db · outbound

This paper cites TokenPacker: Efficient Visual Projector for Multimodal LLM.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding TokenPacker: Efficient Visual Projector for Multimodal LLM

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.697677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.697677Z digest=sha256:36a0776f71187efcd95872270ed6fbf82e867670462b4dcc9ff50c3b057de29a

Observation d752036f-f3da-41d7-bf24-79bda8a8573c · outbound

This paper cites Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.700642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.700642Z digest=sha256:6ef8495e64e44b621378bef810b0a8b325550823cf9efd323fcf9d66e9bca3e3

Observation 3f76dae1-5a53-4bd6-b0d0-5d850f82c64a · outbound

This paper cites Macro-and micro- anatomical, histological and computed tomography scan characterization of the nasopalatine canal.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Macro-and micro- anatomical, histological and computed tomography scan characterization of the nasopalatine canal

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:12.020007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.703490Z digest=sha256:9efb985bea563bcead40ca52437b00109c598e59ec8127b6d65bcdf24c669776

Observation 910fd7f9-d15c-495d-9346-30f4a4f64400 · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Rouge: A package for automatic evaluation of summaries

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:11.842518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.706327Z digest=sha256:26e1dbfbc4790ef6bdd1b26b5faea8a8a106d79e4f21b15b97fd9ebd4d344034

Observation 40c0597a-cc1a-444e-999d-9166c23816f1 · outbound

This paper cites MG-3D: Multi-Grained Knowledge-Enhanced 3D Medical Vision-Language Pre-training.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding MG-3D: Multi-Grained Knowledge-Enhanced 3D Medical Vision-Language Pre-training

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.708862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.708862Z digest=sha256:b78b72932360960cb647f28465925f829e652969eca7f04a73eeeee24a788a14

Observation 2a42e4d0-3fa2-499c-a749-a9664e804bd4 · outbound

This paper cites Bleu: a Method for Auto- matic Evaluation of Machine Translation.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Bleu: a Method for Auto- matic Evaluation of Machine Translation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:11.676720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.711866Z digest=sha256:f12cc48fc9094ce717dc2635b32b17ce3eeef0a842cbab071686b6c006fdaad4

Observation d16d19a9-7886-4572-a920-c0f099914bdc · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Learning Transferable Visual Models From Natural Language Supervision

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:11.474165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.714448Z digest=sha256:47ffeab9be02a51d2b4fef95c5f9ef4526da6b8ec9d04c924933a44034811196

Observation e89def40-1a80-42a2-b779-0d5302249e3b · outbound

This paper cites Salvolini, E.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Salvolini, E

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:11.349751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.717125Z digest=sha256:af1c60f343fa66fa42f1ba86b6d336da1832932696df7659e651db551375ec22

Observation d7c55625-fda1-40c0-b6bb-b382ad468434 · outbound

This paper cites Time Is Money: Considerations for Measuring the Radiological Reading Time.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Time Is Money: Considerations for Measuring the Radiological Reading Time

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:11.173127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.719912Z digest=sha256:9548e3bf4314e4dea358bb5be23dd5f97daa88f78f689164a2ba7e2a6fcfd61e

Observation 5b7dceb2-6ae9-4f9d-9767-12fea05b9215 · outbound

This paper cites Med-2E3: A 2D-Enhanced 3D Medical Multimodal Large Language Model.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Med-2E3: A 2D-Enhanced 3D Medical Multimodal Large Language Model

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.722774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.722774Z digest=sha256:47a8b477ed8894a1cfd417c1ab048ec2617108b8ffe923424aa3756dc2e36471

Observation 7480eb91-463e-4c1b-b889-80e3340626b9 · outbound

This paper cites an unresolved cited work.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:11.001306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.725473Z digest=sha256:5199b80095fdd79f42bb17978fa1db7468235ae1b860a76f209d2071b9a978e3

Observation 7e9939ad-ea6f-4414-a99a-344e21c0a209 · outbound

This paper cites an unresolved cited work.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:10.861999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.728638Z digest=sha256:9f941c98f4101dc128e012c3b19e4196925e2d069c0afa2d4197b400fcc269da

Observation d3d44794-c411-4187-8c28-c7d1cc90b1b5 · outbound

This paper cites Towards Generalist Biomedical AI.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Towards Generalist Biomedical AI

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.731105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.731105Z digest=sha256:232dce76ea07edefc1e172c737e9a2fdc3d7e5ce94c982580279725011d10e1b

Observation 1f57a25c-8f75-4339-8799-4a0e942798c7 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Representation Learning with Contrastive Predictive Coding

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.733934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.733934Z digest=sha256:328ef0737dd6a250c7318ee1c10a34a19cb3079c739ea0266734cb0f8dc1fd38

Observation fa98463d-1155-416d-867b-0f2e6da219b4 · outbound

This paper cites Cross-modal prototype driven network for radiology report generation.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Cross-modal prototype driven network for radiology report generation

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:10.670325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.736992Z digest=sha256:09545dac8374961be4166a069a1ec2e51a2746fb4aa7c6f83664580b1226a057

Observation fd024453-c5ee-4cd6-b23a-8214a5f63df9 · outbound

This paper cites MMCLIP: Cross-modal Attention Masked Modelling for Medical Language-Image Pre-Training.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding MMCLIP: Cross-modal Attention Masked Modelling for Medical Language-Image Pre-Training

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.741096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.741096Z digest=sha256:29e7f1df85a64c5dfbd754ca00665f5e0b68f04b8468979e1dd64c675a46cb84

Observation 98a3ec93-657c-4a92-bf45-7c1b99b0c292 · outbound

This paper cites Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.745722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.745722Z digest=sha256:34b9746a8c48d6cacde388e7f9f29702bae43ec60165f22d4815c2083f6f9f97

Observation 883a7492-0f61-419c-868a-29c40e50901f · outbound

This paper cites Zou, und Huaxiu Yao.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Zou, und Huaxiu Yao

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:10.399927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.752474Z digest=sha256:d32a193d8212d53b13a2a3195fab8e354e568f2f6f302024e811acf0986b70a5

Observation 0d50dc99-07f8-4d3a-a0b8-dbf602292b1d · outbound

This paper cites Med3DVLM: An Efficient Vision- Language Model for 3D Medical Image Analysis.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Med3DVLM: An Efficient Vision- Language Model for 3D Medical Image Analysis

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.756129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.756129Z digest=sha256:c3cfb7d8a5402b26e1b7332943a6530396ca447258b591e2451bac384cd0173f

Observation f3bc874f-1ec7-4a1a-a361-98eefc53a0bb · outbound

This paper cites Sigmoid Loss for Language Image Pre-Training.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Sigmoid Loss for Language Image Pre-Training

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:10.209432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.760736Z digest=sha256:8034430145ca66599109aab8ffaa997fd899827bdb15e8b6f956e5765591ee46

Observation 3eb5ce39-5de0-4804-8dca-ccda4f254aa2 · outbound

This paper cites Lungren, Tristan Naumann, Sheng Wang, und Hoifung Poon.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Lungren, Tristan Naumann, Sheng Wang, und Hoifung Poon

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:09.979845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.770261Z digest=sha256:b83d646052afb7a56ed04d59ef3196e6de7fb4b25722114a7fb6eb7dae4edff2

Observation a8424c63-97ce-46e9-8a0a-4604076e458d · outbound

This paper cites Weinberger, und Yoav Artzi.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Weinberger, und Yoav Artzi

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:09.799164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.784587Z digest=sha256:2a926bc75ee9dbc74b7c15ff82050b72f315a7f2308392aa4a0a11e8068511c2

Observation 12f782b5-3cf3-4c2f-9a47-c9d7f591e4ac · outbound

This paper cites MEPNet: Medical Entity-Balanced Prompting Network for Brain CT Report Generation.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding MEPNet: Medical Entity-Balanced Prompting Network for Brain CT Report Generation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:09.557804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.805586Z digest=sha256:de21bb3d0888db3ad0a12054d6f666220e71c523cfc12c538eee2fc63ccdb189

Observation 3f99ebed-d651-47b5-87b0-1f2f8c00f6b6 · outbound

This paper cites RadGenome-Chest CT: A Grounded Vision-Language Dataset for Chest CT Analysis.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding RadGenome-Chest CT: A Grounded Vision-Language Dataset for Chest CT Analysis

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.828501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.828501Z digest=sha256:fd3fabeca903c690bf51941deb59c494a3553f6d4d3e570af22286f6f5038b93

Observation dd49e689-3feb-418a-a7b8-e50e15fce348 · outbound

This paper cites There are emphysematous changes in both lungs.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding There are emphysematous changes in both lungs

Reference 51

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T04:47:09.276205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:07.941763Z digest=sha256:4f6015cf4b4880b6756798c7727f4f745b98dfce474d1f1dc1e1842326c6251a

Observation 16f3d46d-e57f-4dec-a54b-23b8802df80d · outbound

This paper cites Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:08.987645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:47:08.060954Z digest=sha256:4fcfa02c19417ff835d2be4a63af3400a4946345c4ac59e65565f0bf77b90c1a

Pith citing papers

No inbound Pith citation observations are available.