Pith. sign in

Paper Citation Record · LEDGER

Exploring The Visual Feature Space for Multimodal Neural Decoding

As of 23 August 2026, this Paper Citation Record lists 76 of 76 outbound references and 1 inbound Pith citation observation for arXiv:2505.15755.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15755 v1

Coverage vector

measured 76 of 76 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:16:44.559080Z

measured 77 of 77 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-18T21:13:41.362773Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T21:16:51.304458Z

Reference resolution

76 of 76 outbound references displayed

  • verified exact0
  • verified fuzzy64
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 388e3681-7b31-489e-a6a3-610181a59373 · outbound

This paper cites GPT-4 Technical Report.

Exploring The Visual Feature Space for Multimodal Neural Decoding GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:39.121046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:39.121046Z digest=sha256:3d358602748329851bd1ebc255c0b1fb6e88b32c1d21c7f708684bf48d428f7a

Observation 0a1f4fdf-84b8-45fe-88f8-68b67d0d0d6b · outbound

This paper cites A massive 7t fmri dataset to bridge cognitive neuroscience and artificial intelligence.Nature neuroscience, 25(1):116–126, 2022.

Exploring The Visual Feature Space for Multimodal Neural Decoding A massive 7t fmri dataset to bridge cognitive neuroscience and artificial intelligence.Nature neuroscience, 25(1):116–126, 2022

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:54.955379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:39.175887Z digest=sha256:606316a1e735b022934f57cbdcfb383b70dbd3873261dbebed2f3ce2639715ab

Observation b27846bb-acfb-49d7-907e-56fed05c6f3e · outbound

This paper cites Spice: Semantic propositional image caption evaluation.

Exploring The Visual Feature Space for Multimodal Neural Decoding Spice: Semantic propositional image caption evaluation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:54.736347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:39.215937Z digest=sha256:66f186d800242f479193db9751e94ced8faf68b5a704ea18393cd48041fb781e

Observation 14313ba6-2bac-416d-801c-64b0818819fd · outbound

This paper cites Leo: Boosting mixture of vision encoders for multimodal large language models.arXiv preprint arXiv:2501.06986, 2025.

Exploring The Visual Feature Space for Multimodal Neural Decoding Leo: Boosting mixture of vision encoders for multimodal large language models.arXiv preprint arXiv:2501.06986, 2025

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:39.318179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:39.318179Z digest=sha256:687caccaffd576192b3b490c782f755b28a52b03ddba5a9045aebfb632627f05

Observation 5c742277-b718-42fd-8e3c-5f465a930335 · outbound

This paper cites Layer Normalization.

Exploring The Visual Feature Space for Multimodal Neural Decoding Layer Normalization

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:39.427338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:39.427338Z digest=sha256:497bc826e4e0bcce30606f66d9c1dce2ede17c2858ce955b9660eb4ebed18aaf

Observation a33319b0-cd3d-4002-b85a-5eee5ada4081 · outbound

This paper cites Meteor: An automatic metric for mt evaluation with improved correlation with human judgments.

Exploring The Visual Feature Space for Multimodal Neural Decoding Meteor: An automatic metric for mt evaluation with improved correlation with human judgments

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:54.522685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:39.527524Z digest=sha256:83f6da7d4072c45e865b01c4ce3663476666b7104a65e3588fe87661e9677cc0

Observation 0728e3c3-fea1-4186-9749-da08669ceddc · outbound

This paper cites Token merging: Your vit but faster.

Exploring The Visual Feature Space for Multimodal Neural Decoding Token merging: Your vit but faster

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:54.394220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:39.643607Z digest=sha256:4c7818dd5d9e161e3cfe91bef8c899779fa54355885b6249793bd322eb0ca911

Observation 86606b4c-4ec8-45d0-a58d-675569d137c4 · outbound

This paper cites Matryoshka multimodal models.

Exploring The Visual Feature Space for Multimodal Neural Decoding Matryoshka multimodal models

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:54.286249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:39.726607Z digest=sha256:f307172e8cebcf272704fbab83a613dcf2177873925dc8ef55a49e8806ecdeb6

Observation 1f8a7a84-6517-4757-a1a2-7c86839b3320 · outbound

This paper cites Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic.

Exploring The Visual Feature Space for Multimodal Neural Decoding Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:39.782659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:39.782659Z digest=sha256:38821b9c9215181211035277650d15eed7385af6b7f298032adf7b951c7bc6eb

Observation d6e973b4-5e26-406e-85ca-985572024c10 · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks.

Exploring The Visual Feature Space for Multimodal Neural Decoding Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:54.104757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:39.879597Z digest=sha256:1b000cab19e154c1d8b8c84fb8a935fd400d89f7f6aa0a4c338be3114277a5d4

Observation 332b05ac-8d32-4eff-9e6b-8312a4a097bc · outbound

This paper cites Unifying Specialized Visual Encoders for Video Language Models.

Exploring The Visual Feature Space for Multimodal Neural Decoding Unifying Specialized Visual Encoders for Video Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:39.968294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:39.968294Z digest=sha256:3511ad34ee947f2a8cbd8677f584e1bad0f1138807b0afc062edfb8b9b8ec86d

Observation a0833825-e0ea-4391-af46-3cd30e241cc9 · outbound

This paper cites Stimulus-selective properties of inferior temporal neurons in the macaque.Journal of Neuroscience, 4(8):2051–2062, 1984.

Exploring The Visual Feature Space for Multimodal Neural Decoding Stimulus-selective properties of inferior temporal neurons in the macaque.Journal of Neuroscience, 4(8):2051–2062, 1984

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:53.997039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:40.111400Z digest=sha256:af938691c03bbb94ce8f67b11a7cbfa318889376d187aee04221a83341a9c872

Observation 60695a16-2667-4633-927a-546143c50a99 · outbound

This paper cites Benchmarking and Improving Detail Image Caption.

Exploring The Visual Feature Space for Multimodal Neural Decoding Benchmarking and Improving Detail Image Caption

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:40.200691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:40.200691Z digest=sha256:00f6136d743773e41ec6961a17c90f279a7a90c2b7c10f99eda699caa798eeea

Observation de2d61fa-de9a-46af-82fc-18461c90b123 · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale.

Exploring The Visual Feature Space for Multimodal Neural Decoding An image is worth 16x16 words: Transformers for image recognition at scale

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:53.883272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:40.310420Z digest=sha256:1a075208fbf7d09972c7e998526430abacb8290dda2aabc1e3ccb470ed7d4294

Observation 58b7263b-a11f-41c7-9296-c36648d6e9b5 · outbound

This paper cites EV A: Exploring the limits of masked visual representation learning at scale.

Exploring The Visual Feature Space for Multimodal Neural Decoding EV A: Exploring the limits of masked visual representation learning at scale

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:53.659796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:40.362527Z digest=sha256:cfc885f3f019b0b8f6f11721f3224d0c7adde7178925346b7d8dc07a38ec821e

Observation c15c161f-5483-4191-bd26-441478c74ed0 · outbound

This paper cites EV A-02: A visual representation for neon genesis.Image and Vision Computing, 2024.

Exploring The Visual Feature Space for Multimodal Neural Decoding EV A-02: A visual representation for neon genesis.Image and Vision Computing, 2024

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:53.451823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:40.473232Z digest=sha256:070803ee1979be748895b77ea9d5b7c5f0cc31c26514e721edefce8a8e131086

Observation fea478de-afc1-4a5c-9c97-f23bdc2b1ed1 · outbound

This paper cites Brain Captioning: Decoding human brain activity into images and text.

Exploring The Visual Feature Space for Multimodal Neural Decoding Brain Captioning: Decoding human brain activity into images and text

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:40.581833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:40.581833Z digest=sha256:ab2577937f2ae59ef7dab9f34a80b63afd7d1cddd92e92b44da91248c78cba3e

Observation 1deedd2f-8eb3-4647-b7ef-6e1d7af3aa23 · outbound

This paper cites Onellm: One framework to align all modalities with language.

Exploring The Visual Feature Space for Multimodal Neural Decoding Onellm: One framework to align all modalities with language

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:53.223410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:40.684760Z digest=sha256:6e5365dbd67c6865434e49fd28ec5fd68e4b20357883e4f35d055529f3bde4d5

Observation 28a720fd-e4cf-43f3-a748-8861f3d6a1bd · outbound

This paper cites Deep residual learning for image recognition.

Exploring The Visual Feature Space for Multimodal Neural Decoding Deep residual learning for image recognition

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:52.954101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:40.762389Z digest=sha256:f29147da266c471be0d5d5f42bad7859d5256885a65005d132c4b280a8b37624

Observation 657b446d-7c8b-4a5f-bbca-47141b7bcd84 · outbound

This paper cites Masked autoencoders are scalable vision learners.

Exploring The Visual Feature Space for Multimodal Neural Decoding Masked autoencoders are scalable vision learners

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:52.705212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:40.795863Z digest=sha256:3b6bdfb8328c5c850459665bf778a203b705083897eccae14983f6a9e40a6e31

Observation 0a6a28e7-76a1-476d-b8b1-1be9f7ad0fe5 · outbound

This paper cites Clipscore: A reference-free evaluation metric for image captioning.

Exploring The Visual Feature Space for Multimodal Neural Decoding Clipscore: A reference-free evaluation metric for image captioning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:52.528355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:40.851544Z digest=sha256:11b0b0fc105704745a686f19bbd5c4c6561bc81be7a1b113b5dc2dd7862990c6

Observation 7ea27c20-0ff8-4ee5-bc91-8666b9c86b52 · outbound

This paper cites Distilling the knowledge in a neural network.

Exploring The Visual Feature Space for Multimodal Neural Decoding Distilling the knowledge in a neural network

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:52.402038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:40.899861Z digest=sha256:4348591a92ef24a99e6c3a2937df58be07822d2e0fffe2ccdcef5423f85d327b

Observation 1a4e0909-d1e7-49d2-a539-f44c4fb74cb9 · outbound

This paper cites Denoising diffusion probabilistic models.

Exploring The Visual Feature Space for Multimodal Neural Decoding Denoising diffusion probabilistic models

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:52.246292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:40.953896Z digest=sha256:b9cc628feab5431b59f2d2fd35aa0a716940730d20fdd5de285235862cb6a544

Observation ea347c78-4b73-441a-90e2-ebcee0799b6f · outbound

This paper cites GQA: A new dataset for real-world visual reasoning and compositional question answering.

Exploring The Visual Feature Space for Multimodal Neural Decoding GQA: A new dataset for real-world visual reasoning and compositional question answering

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:51.999624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:41.020468Z digest=sha256:d05583fd4f2da24d95c82987b1085945fefe80ea0926ec94254450dbf548fa8e

Observation 0a65854e-8224-4471-aa53-7cbc70ec7c55 · outbound

This paper cites Perceiver: General perception with iterative attention.

Exploring The Visual Feature Space for Multimodal Neural Decoding Perceiver: General perception with iterative attention

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:51.874722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:41.084629Z digest=sha256:02ca348b9f4145871eb28485766beffb5bc8953d69cfbbed21464041132c93c9

Observation 33355c04-f6a1-4421-9d5e-1064fb44ed41 · outbound

This paper cites The fusiform face area: a module in human extrastriate cortex specialized for face perception.Journal of neuroscience, 17(11):4302–4311, 1997.

Exploring The Visual Feature Space for Multimodal Neural Decoding The fusiform face area: a module in human extrastriate cortex specialized for face perception.Journal of neuroscience, 17(11):4302–4311, 1997

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:51.778661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:41.161774Z digest=sha256:8da866f523bacf8632b5c5f837dfc2f1af05fa64a7c997fbc70ed5aaab199865

Observation 78b0fd9c-d8b8-4c01-8440-a4cc7e73a300 · outbound

This paper cites Brave: Broadening the visual encoding of vision-language models.

Exploring The Visual Feature Space for Multimodal Neural Decoding Brave: Broadening the visual encoding of vision-language models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:51.547441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:41.210687Z digest=sha256:756dd66599b25cfe576b7152ab385e71eebbfb0e26db2883f455334dca39cc11

Observation f2a9e73d-621c-4847-ba89-521d0a7bac30 · outbound

This paper cites Analyzing and improving the image quality of StyleGAN.

Exploring The Visual Feature Space for Multimodal Neural Decoding Analyzing and improving the image quality of StyleGAN

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:51.328737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:41.280838Z digest=sha256:c285bd731d31a98e61708efc2cddb0a6fc77340b0d913507b9bd5b959fdfe166

Observation abd7f974-9a22-4d08-9b4c-e0cfb807daf3 · outbound

This paper cites Token fusion: Bridging the gap between token pruning and token merging.

Exploring The Visual Feature Space for Multimodal Neural Decoding Token fusion: Bridging the gap between token pruning and token merging

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:51.162382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:41.332896Z digest=sha256:2c62255da09e187e3095b00cde1c341cab349134114a319dcc0007c6a2f77f3d

Observation b03de65f-0afe-4c33-a1c0-77c5b69afcde · outbound

This paper cites Segment anything.

Exploring The Visual Feature Space for Multimodal Neural Decoding Segment anything

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:51.036563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:41.393450Z digest=sha256:d3d9287734fefbc47e09b6204fd20468e71c73b488407d1220ce9eff05fb3e73

Observation 0da5b2b8-1c17-4620-87cb-e31e8ba1f682 · outbound

This paper cites Matryoshka representation learning.

Exploring The Visual Feature Space for Multimodal Neural Decoding Matryoshka representation learning

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:50.806787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:41.457482Z digest=sha256:2e3b8f7e7ce852f42e2940632cfe24c27ab1e1a5bebe8e48b4635348a17e0483

Observation f7ac94bc-f61a-4437-a658-224935fe01de · outbound

This paper cites Pix2Struct: Screenshot parsing as pretraining for visual language understanding.

Exploring The Visual Feature Space for Multimodal Neural Decoding Pix2Struct: Screenshot parsing as pretraining for visual language understanding

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:50.578210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:41.544049Z digest=sha256:9e270f2f194d06b65bdaee33f63179443fee9e5a01f4ddbd1e07d7dc7960d567

Observation 94bf7d81-00de-45fa-a74d-5aea7949d114 · outbound

This paper cites Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation.

Exploring The Visual Feature Space for Multimodal Neural Decoding Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:50.415669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:41.642846Z digest=sha256:aa104acfeb8869ff657272c6843b29bb02967964f379ca5f636dc9f269874b69

Observation c9852cf7-11af-4ce2-b4bb-dee6783c6c76 · outbound

This paper cites Autoregressive image generation without vector quantization.

Exploring The Visual Feature Space for Multimodal Neural Decoding Autoregressive image generation without vector quantization

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:50.165291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:41.728624Z digest=sha256:d82e0e6ce40bbfd2cb874fc64e33975c83b8c3f2cbd58f416cf6c3d1d2907a3a

Observation 7ce839d9-13c5-4484-a914-15706ae8ce6c · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

Exploring The Visual Feature Space for Multimodal Neural Decoding Evaluating Object Hallucination in Large Vision-Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:41.823139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:41.823139Z digest=sha256:1bcad2651bf68bf44d84a1a21f28228ff37778653a0a3a9d3d98e11ff4d21a51

Observation 53704016-9590-4dcd-b3ac-05b22c00250b · outbound

This paper cites Factual: A benchmark for faithful and consistent textual scene graph parsing.

Exploring The Visual Feature Space for Multimodal Neural Decoding Factual: A benchmark for faithful and consistent textual scene graph parsing

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:49.958546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:41.936538Z digest=sha256:2d714cc5589c8f15c519922c7d74e7784131c48f51a6af7deee08e7cc54fa720

Observation 60a740ce-7852-4f43-b11b-72dbc513bf21 · outbound

This paper cites Mind Reader: Reconstructing complex images from brain activities.NeurIPS, 35: 29624–29636, 2022.

Exploring The Visual Feature Space for Multimodal Neural Decoding Mind Reader: Reconstructing complex images from brain activities.NeurIPS, 35: 29624–29636, 2022

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:49.854513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:42.047073Z digest=sha256:18bb743a370f3086b0bdea57794dea9de6b7565e50bf47dc707999d0ddbe8b55

Observation 63c620e6-b456-497a-8be6-6f161dcb017a · outbound

This paper cites Lawrence Zitnick.

Exploring The Visual Feature Space for Multimodal Neural Decoding Lawrence Zitnick

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:49.634326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:42.094523Z digest=sha256:f5ef532a2019b9536e5831d3ffef4e48138de81a5b1e2a6da30bdfd77732c963

Observation 894dd367-bf7a-4100-a4bd-02bd42a32cbe · outbound

This paper cites Visual instruction tuning.

Exploring The Visual Feature Space for Multimodal Neural Decoding Visual instruction tuning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:49.445157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:42.131165Z digest=sha256:7327b58b3741460df83f1bbd2025b66b9b2dc6d72411524938ca067485caec72

Observation 498ccef0-237d-4434-bf16-409cbf5d9197 · outbound

This paper cites Nltk: The natural language toolkit.

Exploring The Visual Feature Space for Multimodal Neural Decoding Nltk: The natural language toolkit

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:49.173873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:42.176007Z digest=sha256:2bb0e406dd41f75dd038d699edbe6c9bcf807e85377be737e4beaeb287efd07b

Observation 3c72c227-f608-4376-a7f5-2f04dba91cf2 · outbound

This paper cites Decoupled weight decay regularization.

Exploring The Visual Feature Space for Multimodal Neural Decoding Decoupled weight decay regularization

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:42.240170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:42.240170Z digest=sha256:089fbefb4919e155c89b7e974a68e53f22d4a309ac364ff081c8ee1efa8c57c8

Observation dbcbf07b-08ef-4466-8677-9adbaf041ebe · outbound

This paper cites Benchmarking large vision-language models via directed scene graph for comprehensive image captioning.

Exploring The Visual Feature Space for Multimodal Neural Decoding Benchmarking large vision-language models via directed scene graph for comprehensive image captioning

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:49.000460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:42.287542Z digest=sha256:e553c0b45a21619d052df146fb6cd5412c8c2da3ca97f58e20f16da1b4834ff2

Observation 585b1d00-8f20-4c8a-9123-ffe8d5215fd8 · outbound

This paper cites UMAP: Uniform Manifold Approximation and Projection for Dimension Reduction.

Exploring The Visual Feature Space for Multimodal Neural Decoding UMAP: Uniform Manifold Approximation and Projection for Dimension Reduction

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:42.359834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:42.359834Z digest=sha256:7411a36ca4da2f9668c7c6627f07c2f1f021a3309696d60ebf8cad23e587b5fd

Observation fa808e6d-fc21-4d96-b83b-55a9268f9515 · outbound

This paper cites Improved denoising diffusion probabilistic models.

Exploring The Visual Feature Space for Multimodal Neural Decoding Improved denoising diffusion probabilistic models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.859241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:42.446975Z digest=sha256:27d32b605025839ae78d8655c998f3082669f798404db600779db907fbbfd234

Observation af3dd298-a5aa-4998-bde5-1b1149381372 · outbound

This paper cites DINOv2: Learning robust visual features without supervision.TMLR, 2023.

Exploring The Visual Feature Space for Multimodal Neural Decoding DINOv2: Learning robust visual features without supervision.TMLR, 2023

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.657448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:42.531754Z digest=sha256:309487a80967a18d7b6dcee9f32908f0ab1efe70e1c746b6d0d595052afe3620

Observation 47499c3e-1aa8-42df-951f-d00b9bc44d63 · outbound

This paper cites Brain-Diffuser: Natural scene reconstruction from fMRI signals using generative latent diffusion.

Exploring The Visual Feature Space for Multimodal Neural Decoding Brain-Diffuser: Natural scene reconstruction from fMRI signals using generative latent diffusion

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.461309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:42.619184Z digest=sha256:61d0872f5aeddb8e29f4be1a88c4cf0bd352aebc9f4af5481d594cf5a5161fa1

Observation 3e2f230f-b7cb-48d1-9e2b-32aa31467837 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

Exploring The Visual Feature Space for Multimodal Neural Decoding Bleu: a method for automatic evaluation of machine translation

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.369074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:42.726780Z digest=sha256:1e7d49daa65a161965c13df2647bb4d2827163d591b50abf7f2ea7ee264aad76

Observation cbaa93ff-d50a-4fa8-ab4e-ce9c6395b1be · outbound

This paper cites Differential sensitivity of human visual cortex to faces, letterstrings, and textures: a functional magnetic resonance imaging study.Journal of neuroscience, 16(16):5205–5215, 1996.

Exploring The Visual Feature Space for Multimodal Neural Decoding Differential sensitivity of human visual cortex to faces, letterstrings, and textures: a functional magnetic resonance imaging study.Journal of neuroscience, 16(16):5205–5215, 1996

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.201025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:42.798335Z digest=sha256:02df273cd8eb6cc0fe5636631c49bf1004d7c5549794494080d370fed3ff6903

Observation f6683897-d20c-482e-94c5-7b52b60d4b3c · outbound

This paper cites Learning transferable visual models from natural language supervision.

Exploring The Visual Feature Space for Multimodal Neural Decoding Learning transferable visual models from natural language supervision

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:48.102969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:42.865111Z digest=sha256:52637bf4e3e42a11683f987c25c7b0bc6225afd90aca50391ea2982f6f43e93d

Observation a658534c-f50a-4e1a-bf67-015ce6026761 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.JMLR, 21(140):1–67, 2020.

Exploring The Visual Feature Space for Multimodal Neural Decoding Exploring the limits of transfer learning with a unified text-to-text transformer.JMLR, 21(140):1–67, 2020

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.811135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:42.983593Z digest=sha256:501421a6d7fbc2a0f52f2e6cfa8d37310717f24ee0ed36236418be474dce02cd

Observation 7943ce1a-d2a0-42d9-87bf-22234e1fd9e8 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Exploring The Visual Feature Space for Multimodal Neural Decoding High-resolution image synthesis with latent diffusion models

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.660888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:43.008703Z digest=sha256:8d7ce62a52380fa1df82d98f8b6c1411024672f9a980859ae14c59c720f4e31b

Observation 4236eac0-0419-46be-b8cd-cb1036ad93a2 · outbound

This paper cites LAION-5B: An open large-scale dataset for training next generation image-text models.

Exploring The Visual Feature Space for Multimodal Neural Decoding LAION-5B: An open large-scale dataset for training next generation image-text models

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.526897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:43.035521Z digest=sha256:fcb066d32329dec9cc7aef50d913f3768bb7949ddd75d678265f1a1768a9f4a4

Observation d0d58c55-7050-4e21-8fb9-9281d6da150d · outbound

This paper cites Reconstructing the mind’s eye: fmri-to-image with contrastive learning and diffusion priors.

Exploring The Visual Feature Space for Multimodal Neural Decoding Reconstructing the mind’s eye: fmri-to-image with contrastive learning and diffusion priors

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.375067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:43.097221Z digest=sha256:0d1ce7a514f96b2f30a301b6412c59e509951bbff6c199dc540a26f7e607900a

Observation 71d82b78-cb1f-4096-9994-5d29fd536196 · outbound

This paper cites Mindeye2: Shared-subject models enable fmri-to-image with 1 hour of data.

Exploring The Visual Feature Space for Multimodal Neural Decoding Mindeye2: Shared-subject models enable fmri-to-image with 1 hour of data

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.262328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:43.196098Z digest=sha256:b3a30bc21b9c92788fbdc845ed2dc808b2f443ea0d90b6c4d68878ac819f17ee

Observation e6f3ed65-fbc0-4e0d-90e2-37bba47c8aae · outbound

This paper cites Neuro-vision to language: Enhancing brain recording-based visual reconstruction and language interaction.

Exploring The Visual Feature Space for Multimodal Neural Decoding Neuro-vision to language: Enhancing brain recording-based visual reconstruction and language interaction

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.133289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:43.267167Z digest=sha256:59cb62a4dff5c97ee6dac0876b406ab062d116f83c3e142510876b67e40c1426

Observation c76b1743-8c99-41b5-bdd3-ace2936fc17e · outbound

This paper cites Mome: Mixture of multimodal experts for generalist multimodal large language models.

Exploring The Visual Feature Space for Multimodal Neural Decoding Mome: Mixture of multimodal experts for generalist multimodal large language models

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:47.047238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:43.348922Z digest=sha256:29e44c006b53d3fb5e3a0d6f59bce050e9b355862e45ba8bc54ec07f4fbf6d3b

Observation 17e9cb20-1596-490e-be4c-24f3b6c49fe5 · outbound

This paper cites Eagle: Exploring the design space for multimodal llms with mixture of encoders.

Exploring The Visual Feature Space for Multimodal Neural Decoding Eagle: Exploring the design space for multimodal llms with mixture of encoders

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.941690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:43.398601Z digest=sha256:5b04876728bc528b7c21056df3a63331d602b0451ab3bb2a81f29cce414e487c

Observation a9fb3157-714f-4946-85d1-1bc7f0fb998a · outbound

This paper cites Super-convergence: Very fast training of neural networks using large learning rates.

Exploring The Visual Feature Space for Multimodal Neural Decoding Super-convergence: Very fast training of neural networks using large learning rates

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.813615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:43.435758Z digest=sha256:028c8e8d6fd141bd9ee360b4b51a3a1af6fadcc9e5201b3300e153df2a2cbbb2

Observation 7f87395c-76b9-4ff2-ba18-e00b903d2b30 · outbound

This paper cites Denoising diffusion implicit models.

Exploring The Visual Feature Space for Multimodal Neural Decoding Denoising diffusion implicit models

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.706861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:43.476022Z digest=sha256:29b11793a72b25f59e13328f8f42093b7566b79d68d72f25d5e936b35cbe2430

Observation f756171b-d78d-4936-a86f-343c0b9ef436 · outbound

This paper cites Generative modeling by estimating gradients of the data distribution.

Exploring The Visual Feature Space for Multimodal Neural Decoding Generative modeling by estimating gradients of the data distribution

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.569992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:43.537605Z digest=sha256:3b2694db9fc2b8468f3c4c971219f0ca66601f63b991cf35c14aa44885915612

Observation 3890e0d1-e7bf-46b4-80cb-ee014125da9a · outbound

This paper cites Improving visual image reconstruction from human brain activity using latent diffusion models via multiple decoded inputs.

Exploring The Visual Feature Space for Multimodal Neural Decoding Improving visual image reconstruction from human brain activity using latent diffusion models via multiple decoded inputs

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:43.626381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:43.626381Z digest=sha256:4a2b34901a88ba37e65e4c09d135db493322136ef06813feab7713d9b6161708

Observation 3ecf22bc-6ccc-46dd-831d-616cddc5a3d4 · outbound

This paper cites Eyes wide shut? exploring the visual shortcomings of multimodal llms.

Exploring The Visual Feature Space for Multimodal Neural Decoding Eyes wide shut? exploring the visual shortcomings of multimodal llms

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.408440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:43.723593Z digest=sha256:fb32d8d074e9fd1401e6b28fcfa1c14840872cd31313791e75593858f24b0690

Observation 04b15a58-02f8-4e7d-bb52-0695f8b08729 · outbound

This paper cites Cider: Consensus-based image description evaluation.

Exploring The Visual Feature Space for Multimodal Neural Decoding Cider: Consensus-based image description evaluation

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.307705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:43.797248Z digest=sha256:668701a21b02be86470b4fd79287ae9e88dfac6675f5809c0e750eff4dab4761

Observation 97b93f9e-9851-4077-9911-88fe79ea0f6e · outbound

This paper cites Reconstructive visual instruction tuning.

Exploring The Visual Feature Space for Multimodal Neural Decoding Reconstructive visual instruction tuning

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.181743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:43.896942Z digest=sha256:c1aa0d4aa534ee1a161c86c92ce48e912007853ca354332225beabb03acd27b5

Observation 7aa6e714-839d-4e38-bdf8-262679d9b4a6 · outbound

This paper cites Mindbridge: A cross-subject brain decoding framework.

Exploring The Visual Feature Space for Multimodal Neural Decoding Mindbridge: A cross-subject brain decoding framework

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:46.063111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:43.939911Z digest=sha256:c222b6fa49a942ffddac3568520c03f6fac6cf2402dd03cdb837721cb37e4c27

Observation ecdfed39-eb49-472d-bc3e-6a231b43675a · outbound

This paper cites ConvNeXt V2: Co-designing and scaling convnets with masked autoencoders.

Exploring The Visual Feature Space for Multimodal Neural Decoding ConvNeXt V2: Co-designing and scaling convnets with masked autoencoders

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:45.937209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:43.974482Z digest=sha256:d6aa0d39c6368b51c371410cdad43eafe00d013f32afd825debef3e7f90440c0

Observation ec1cb38e-45e2-40d0-bdca-85671ec11111 · outbound

This paper cites Umbrae: Unified multimodal brain decoding.

Exploring The Visual Feature Space for Multimodal Neural Decoding Umbrae: Unified multimodal brain decoding

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:45.770874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:44.048688Z digest=sha256:e2e61179f6b70c9bb48348d0e57c43899e53407ea22d2d32f954c6e4c8f85f9e

Observation 0c38ce74-1042-4916-b061-326441070714 · outbound

This paper cites Dream: Visual decoding from reversing human visual system.

Exploring The Visual Feature Space for Multimodal Neural Decoding Dream: Visual decoding from reversing human visual system

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:45.661906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:44.114827Z digest=sha256:c199a83dce508930face75e4a23a6e4e8e82f8d11758b31645aff1a1365c7019

Observation b030d902-dbd5-4944-b644-9f6302edeeb5 · outbound

This paper cites Mevox: Multi-task vision experts for brain captioning.

Exploring The Visual Feature Space for Multimodal Neural Decoding Mevox: Multi-task vision experts for brain captioning

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:45.591386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:44.155145Z digest=sha256:9b2647ba81b626a578919734ad39ac07a6683131af0acdf62a38570519341ba1

Observation ea090e99-9cb1-46ea-ac05-0f03c753473d · outbound

This paper cites Versatile diffusion: Text, images and variations all in one diffusion model.

Exploring The Visual Feature Space for Multimodal Neural Decoding Versatile diffusion: Text, images and variations all in one diffusion model

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:45.450042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:44.199109Z digest=sha256:17ac3739b9cedf638d2df2f0a7ade8f1f30c783972b1acd51135daad74110231

Observation dc5a8203-d38c-44dd-b926-b14620841690 · outbound

This paper cites Dense connector for mllms.

Exploring The Visual Feature Space for Multimodal Neural Decoding Dense connector for mllms

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:45.331633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:44.239920Z digest=sha256:5466b5ecba1f2403f301fc8958d8a3a551c7cbd8700fd235a1528ec6b5446478

Observation 19bf626b-5eb4-4490-a700-f1ec87a71bfa · outbound

This paper cites Sigmoid loss for language image pre-training.

Exploring The Visual Feature Space for Multimodal Neural Decoding Sigmoid loss for language image pre-training

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:45.259364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:44.302629Z digest=sha256:e4112d93bc81c0f53df6304f18f3cb70064e9399f989aaf676270a56c50d5449

Observation c59b7996-ecb1-49e5-8385-cb8d999ee82a · outbound

This paper cites Languagebind: Extending video-language pretraining to n-modality by language-based semantic alignment.

Exploring The Visual Feature Space for Multimodal Neural Decoding Languagebind: Extending video-language pretraining to n-modality by language-based semantic alignment

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:45.147295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:44.430939Z digest=sha256:867a80e2fab391abd320a97f479ae7bd21d3606b55366097e69d51dfe9e74dbf

Observation cb6ffba9-e0ca-4262-b05d-32df0f7f2165 · outbound

This paper cites Detrs with collaborative hybrid assignments training.

Exploring The Visual Feature Space for Multimodal Neural Decoding Detrs with collaborative hybrid assignments training

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:45.070960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:44.504615Z digest=sha256:cb2adfd78ece205175ea8021907c256992dc2ec1903b63dbd7b7642bb7547b99

Observation f9f7e3e5-9712-4045-885b-a0eb277e0cc7 · outbound

This paper cites Mova: Adapting mixture of vision experts to multimodal context.

Exploring The Visual Feature Space for Multimodal Neural Decoding Mova: Adapting mixture of vision experts to multimodal context

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:44.986961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:44.559080Z digest=sha256:6aee33aa3957264fab1829daa2c3497d4800e3c9b2adbd7df4f11ceae6f9adeb

Observation 603cf714-87c3-46a0-9af4-38d6b48f5976 · outbound

This paper cites an unresolved cited work.

Exploring The Visual Feature Space for Multimodal Neural Decoding Unresolved cited work

Reference 2021

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:16:47.936612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:16:42.935474Z digest=sha256:ca2c9df5e43dccccb342905e4e527823a6e92f5a722f0a2be980626f4439fd0f

Pith citing papers

Observation 5c1aa832-789d-4219-b8f2-ca7a3a3890a2 · inbound

BRAIN: Bias-Mitigation Continual Learning Approach to Vision-Brain Understanding cites this paper.

BRAIN: Bias-Mitigation Continual Learning Approach to Vision-Brain Understanding Exploring The Visual Feature Space for Multimodal Neural Decoding

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-18T21:16:51.306848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-18T21:13:41.362773Z digest=sha256:a13abbe948b9ab59ea58521c17a8a837a49bef83ddd15e40b9262d7915ac3781