Pith. sign in

Paper Citation Record · LEDGER

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks

As of 17 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 1 inbound Pith citation observation for arXiv:2508.07803.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.07803 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:52:25.017468Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-11T23:54:32.839280Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

47 of 47 outbound references displayed

  • verified exact2
  • verified fuzzy38
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 168e4721-7e28-416b-8eea-e6e21ab8983a · outbound

This paper cites A deep learning framework for infrared and visible image fusion without strict registration.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks A deep learning framework for infrared and visible image fusion without strict registration

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:33.914320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:20.843522Z digest=sha256:d9f5f8b09520505a281dcb03b7e20fdb3bb670217d371aefd354000532cc07e9

Observation a5777d18-1a03-47f1-8b16-353c70a1a5b0 · outbound

This paper cites All-weather multi-modality image fusion: Unified framework and 100k benchmark.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks All-weather multi-modality image fusion: Unified framework and 100k benchmark

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:20.944258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:20.944258Z digest=sha256:a7d1d8a8e9509b2cd813944dc5d0a872e1a2d4dbe5a8165cf39b3102fc285bba

Observation 50ddb353-3e72-4768-adcb-9ec1833028c3 · outbound

This paper cites Infrared and visible image fusion based on domain transform filtering and sparse representation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Infrared and visible image fusion based on domain transform filtering and sparse representation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:33.758344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.052336Z digest=sha256:f1121084156649876ba66e0b2b5849734ecfdb6bb760fe3cc38c93454f6bc8a2

Observation f646ed0c-ad4d-4340-860a-01556952b60f · outbound

This paper cites Infrared and visible image fusion: From data compatibility to task adaption.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Infrared and visible image fusion: From data compatibility to task adaption

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:33.536389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.160857Z digest=sha256:828a645ccfd2140620e42072df6f84cd77da579f0c625bf8e6f4f9713a9af69c

Observation 9e8b1584-9e0b-4dbf-8c83-bc5edf9765eb · outbound

This paper cites Simultaneous tri-modal medical image fusion and super-resolution using conditional diffusion model.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Simultaneous tri-modal medical image fusion and super-resolution using conditional diffusion model

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:33.295541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.259991Z digest=sha256:4b1dce2af2d855a3304dcee3618ba938c066635b72ca05500b664103dfe94d40

Observation f2d9ac42-ac80-4ce0-bf67-5b308133d718 · outbound

This paper cites Fast r-cnn.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Fast r-cnn

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:33.066319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.481919Z digest=sha256:d4da4e4afc70cba1a7a8931f749ec251e3f5afe6ab6f1d1072a4a75b7fc7080a

Observation f8ae765a-2b45-451b-a981-b47b86160e2c · outbound

This paper cites Microsoft coco: Common objects in context.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Microsoft coco: Common objects in context

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:32.837681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.572846Z digest=sha256:fba3f2c21a230f506c95b9ba8bf8d6f4438a3f51af9bb2fd1fa07d9a7f6d67ff

Observation 95698043-7484-408d-aca9-888a1a635f15 · outbound

This paper cites Visualizing data using t-sne.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Visualizing data using t-sne

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:21.713153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:21.713153Z digest=sha256:272a9acdf88234de19812066e8962bc7643cdbda3fadffa1855582b63ac71647

Observation 3dc53547-8931-4a95-a8f5-f46167226cb4 · outbound

This paper cites Tsjnet: A multi-modality target and semantic awareness joint-driven image fusion network.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Tsjnet: A multi-modality target and semantic awareness joint-driven image fusion network

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:21.795023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:21.795023Z digest=sha256:5d4b41da68f76e123024762f56adaa4e42ab3d45170c8d0f4c0b199b7b066103

Observation 5957c9c7-fda6-4c8a-8295-3b4d8b2673e0 · outbound

This paper cites Target-aware dual adversarial learning and a multi-scenario multi-modality benchmark to fuse infrared and visible for object detection.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Target-aware dual adversarial learning and a multi-scenario multi-modality benchmark to fuse infrared and visible for object detection

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:32.653248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.874881Z digest=sha256:a67a817b64096229e92be9c508a2bdf2a78260ecfc1af6acfcfea0d24cd570d5

Observation 0baafcb8-965e-4a87-9a73-8d1f3a5b69ed · outbound

This paper cites Fs-diff: Semantic guidance and clarity-aware simultaneous multimodal image fusion and super-resolution.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Fs-diff: Semantic guidance and clarity-aware simultaneous multimodal image fusion and super-resolution

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:32.377319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:21.983819Z digest=sha256:bc05250ed5ba73246c9facc38d351c47674e00fc4fbd9d7723188a0522e03832

Observation f2a1e272-e385-4dd3-bc79-ab46406dff3e · outbound

This paper cites A task-guided, implicitly-searched and metainitialized deep model for image fusion.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks A task-guided, implicitly-searched and metainitialized deep model for image fusion

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:32.242375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.083478Z digest=sha256:3e23cbe75062f1fd010464339e70463fb62a34c73938a3902a07c8ce6787e1dd

Observation 330debc3-4de8-4489-87c9-cdf7d6008795 · outbound

This paper cites Image fusion in the loop of high-level vision tasks: A semantic-aware real-time infrared and visible image fusion network.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Image fusion in the loop of high-level vision tasks: A semantic-aware real-time infrared and visible image fusion network

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:32.072413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.190602Z digest=sha256:f26af3be4144b6abe623d360ba9977f9175e5e2b855b37e834667f876067fcf8

Observation e1cbd90c-fd38-4701-8161-0ce3eb1aa078 · outbound

This paper cites Mrfs: Mutually reinforcing image fusion and segmentation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Mrfs: Mutually reinforcing image fusion and segmentation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:31.755632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.300125Z digest=sha256:c20eb85e2225e12cde3963911f0e5ee0aff9c68c015ef205fbb78bbdf0096641

Observation 3fa7d815-f3a3-414c-9f10-8e61c88de446 · outbound

This paper cites Image-to-image translation: Methods and applications.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Image-to-image translation: Methods and applications

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:31.569164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.415638Z digest=sha256:e76ca1672583f1a45515500e2edecf26c12aa40e895ca1ac2b51aae105b4148b

Observation 628f808a-9b94-4aa0-9ef2-77bf7ac7885a · outbound

This paper cites Source-free open compound domain adaptation in semantic segmentation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Source-free open compound domain adaptation in semantic segmentation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:31.357536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.499013Z digest=sha256:ab8fb011ca2dea69b3fe4d72ba014aad7f2f9b17dcd43216946b7d8d45cc1b30

Observation 0d2eabd2-5c76-49ee-ac91-80f5a3b7dafb · outbound

This paper cites Infragan: A gan architecture to transfer visible images to infrared domain.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Infragan: A gan architecture to transfer visible images to infrared domain

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:31.076644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.571106Z digest=sha256:e1b03c6cb5b0ec5e23d06079926067dfb722f28be464bc0484d5462e0ff9fea8

Observation 37b2628c-cfeb-4652-a9a3-197fa0478a0b · outbound

This paper cites Cnn-based thermal infrared person detection by domain adaptation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Cnn-based thermal infrared person detection by domain adaptation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:30.887222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.633845Z digest=sha256:165f74c5b245956ccf1cb32b75d83b2db53798f661082dfdef4723bebc5bb303

Observation 56bfd79f-6e98-45f7-974b-c32740ef96a2 · outbound

This paper cites Hallucidet: hallucinating rgb modality for person detection through privileged information.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Hallucidet: hallucinating rgb modality for person detection through privileged information

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:30.699864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.737979Z digest=sha256:eddfbf97840bb61f7aacf54888eb58a79a6004100b635927c9b728448ba6a6ef

Observation 7a470abe-68f6-4c2d-ac65-ed82df8276a5 · outbound

This paper cites Modality translation for object detection adaptation without forgetting prior knowledge.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Modality translation for object detection adaptation without forgetting prior knowledge

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:30.456942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.812104Z digest=sha256:350d4cfb02e3bd3f891401a3defc19dd081ef96b750d9751da73f93fed72d911

Observation 42e047fa-8044-48e9-8f00-d3fc651b7201 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:22.912267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:22.912267Z digest=sha256:ba9c53f8a068258bd3049043b30bcc49b1aa80803ccebcaebbadc4332e881bd6

Observation a5d5d2c5-e078-4562-b432-69a715a0aae2 · outbound

This paper cites Vmamba: Visual state space model.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Vmamba: Visual state space model

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:30.269904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:22.998893Z digest=sha256:0e9da80ba0dcef80e9158cdb634f8163326d0c3509b393cdcbedd8fc4ef23488

Observation e0502ec9-18c2-4fc5-8539-79198965130f · outbound

This paper cites Mambair: A simple baseline for image restoration with state-space model.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Mambair: A simple baseline for image restoration with state-space model

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:30.002559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.077544Z digest=sha256:9d23a87fc4b1792d3a15523d46173a9365fba142b00cca9cf09d608c2b0828d6

Observation 324bcfbb-487b-47e7-ae1e-64013fd9f8a8 · outbound

This paper cites Vision mamba: efficient visual representation learning with bidirectional state space model.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Vision mamba: efficient visual representation learning with bidirectional state space model

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:29.682263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.182088Z digest=sha256:b39c6daad0d025c54ddf844da5707b5a2896b9cdf7872379aae303f1d6dd4813

Observation 6e766b80-804f-4162-839a-63acd56fd705 · outbound

This paper cites GLU Variants Improve Transformer.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks GLU Variants Improve Transformer

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:23.234051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:23.234051Z digest=sha256:8b90697b58f5af83748d9c58dbb18f720abd86d045e9f97a1353aa9ed5b49bbc

Observation 0a9be89a-6367-4f9a-82d8-b1f73f844dfc · outbound

This paper cites Faster r-cnn: Towards real-time object detection with region proposal networks.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Faster r-cnn: Towards real-time object detection with region proposal networks

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:23.332065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:23.332065Z digest=sha256:7fe4c5a7666ef424221501df43d99fa59ff680bb16f93a7f3175789c1bb35fae

Observation d099b7d5-f44e-47f0-8b07-4acc0cfe7479 · outbound

This paper cites Piafusion: A progressive infrared and visible image fusion network based on illumination aware.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Piafusion: A progressive infrared and visible image fusion network based on illumination aware

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:29.362717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.431911Z digest=sha256:fdae317812467c9904072ccaac4ab08d7b07b903dc524cc27b02ace20bbaf6cc

Observation 5fa405c6-2f88-4b12-aa1f-b3ab263b726c · outbound

This paper cites Every SAM Drop Counts: Embracing Semantic Priors for Multi-Modality Image Fusion and Beyond.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Every SAM Drop Counts: Embracing Semantic Priors for Multi-Modality Image Fusion and Beyond

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:52:25.307217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.523944Z digest=sha256:096a149cab153414dd844d5b72ccac11ef5c4958c6a3d2a52709fd29fd1d9208

Observation c6a96941-b04a-45d1-b68f-934966f34a9e · outbound

This paper cites DCEvo: Discriminative Cross-Dimensional Evolutionary Learning for Infrared and Visible Image Fusion.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks DCEvo: Discriminative Cross-Dimensional Evolutionary Learning for Infrared and Visible Image Fusion

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:52:25.159340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.586355Z digest=sha256:ef9e38050135b97275e468d19ec604724d0414edaf1d169d3584b3d8cab4cd5b

Observation a4137315-f119-4e1f-854a-992dc18a4587 · outbound

This paper cites Where elegance meets precision: towards a compact, automatic, and flexible framework for multi-modality image fusion and applications.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Where elegance meets precision: towards a compact, automatic, and flexible framework for multi-modality image fusion and applications

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:29.079716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.647965Z digest=sha256:fce866a8968ae45dd8d0e1e611a7c6f0f1167893dffffef11bdc45612835b6ef

Observation 5dc11091-632f-42ca-ac3b-9c475a758d04 · outbound

This paper cites Equivariant multi-modality image fusion.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Equivariant multi-modality image fusion

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:28.797720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.710572Z digest=sha256:919f15382c9768689b3f44426d14fd4aded75400898159a020d8d8b27000a3eb

Observation 255a2d7b-c502-452d-b0b0-523e20e56bb5 · outbound

This paper cites Cddfuse: Correlation-driven dual-branch feature decomposition for multi-modality image fusion.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Cddfuse: Correlation-driven dual-branch feature decomposition for multi-modality image fusion

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:28.529331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.786134Z digest=sha256:ffce778e323c33755ad44035894d7e2cffc1aa157c1b933efb5d30b3246b6c44

Observation 6273535e-be8e-4d07-85b2-fb4cdc44f161 · outbound

This paper cites Probing synergistic high-order interaction in infrared and visible image fusion.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Probing synergistic high-order interaction in infrared and visible image fusion

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:28.297204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.842113Z digest=sha256:d53cfb4329f8a390a1c9b3f73a21c79c2f8052b17909b959f8fbd13baa385f2c

Observation b702cb31-6238-41f4-b971-534b2127bb34 · outbound

This paper cites Coconet: Coupled contrastive learning network with multi-level feature ensemble for multi-modality image fusion.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Coconet: Coupled contrastive learning network with multi-level feature ensemble for multi-modality image fusion

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:28.094732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.934211Z digest=sha256:fc564e9970eeebbe9cea7fc046e656a7b10605b2edafd0f62c3607abc53b0665

Observation 572f25c5-9d32-4158-82de-22b71dfa0ef8 · outbound

This paper cites Contrastive learning for unpaired image-to-image translation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Contrastive learning for unpaired image-to-image translation

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:27.905015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:23.996337Z digest=sha256:f7367e587e32d29f25106c8dfdbde2ccb5653856434021a86e20f961c1390ad2

Observation e7609665-980b-4198-b359-3000148efe3b · outbound

This paper cites The pascal visual object classes challenge: A retrospective.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks The pascal visual object classes challenge: A retrospective

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:27.731790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.109184Z digest=sha256:34f8e40993e058d8e9151912dbd075a73115e26ec91b05696823142b880e4fce

Observation b6d9f432-c6e3-46b5-8ed4-e3975836ac9b · outbound

This paper cites Assessment of image fusion procedures using entropy, image quality, and multispectral classification.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Assessment of image fusion procedures using entropy, image quality, and multispectral classification

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:27.510761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.196100Z digest=sha256:cb1dcbb04b68b30458f6b38b909e82e01ba51fe3a20d5c2dca84cabde20e188d

Observation cc129a2d-01ec-4c29-800e-51208e30d266 · outbound

This paper cites Detail preserved fusion of visible and infrared images using regional saliency extraction and multi-scale image decomposition.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Detail preserved fusion of visible and infrared images using regional saliency extraction and multi-scale image decomposition

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:27.340977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.285735Z digest=sha256:dd0beaba977ec74239092950b1c8659d8960da754b757f7a3d90f2e12c0d6cbe

Observation b8565c7b-b79f-498e-b190-0ad0f5caacb8 · outbound

This paper cites Fsim: A feature similarity index for image quality assessment.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Fsim: A feature similarity index for image quality assessment

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:27.130819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.378057Z digest=sha256:8d748317e48bfafc9effae4bb13142811c0ac58694b4a4193a5b4b0376f403a7

Observation b5b73b88-d3eb-4bbd-91b5-d3110176138c · outbound

This paper cites Image fusion metric based on mutual information and tsallis entropy.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Image fusion metric based on mutual information and tsallis entropy

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:26.953724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.458774Z digest=sha256:ff63f45e447ccbde46813fae666d062b8dfe6a76d2c1f805499f93b213ff7ead

Observation b1156e03-b3a0-4d7c-a857-8c796da41b4c · outbound

This paper cites Image quality measures and their performance.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Image quality measures and their performance

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:26.804297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.550186Z digest=sha256:80a23302e8e52fe8c30084f06868ebc3601ab219148b0216a84bc487c8d55877

Observation 5595f56a-d325-46f5-a2cc-d8ba3b43b6f7 · outbound

This paper cites Contrastive learning for unpaired image-to-image translation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Contrastive learning for unpaired image-to-image translation

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:26.584059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.609982Z digest=sha256:fb7a2412de885c7870703b22a15ea39bb46a94b7c7ba39d17d23d040c7400cf5

Observation 7d03129e-0337-429b-a8e3-b68856ba74bb · outbound

This paper cites Doubao vision pro 32k.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Doubao vision pro 32k

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:26.427719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.695539Z digest=sha256:c275b91eed4bffd42a5f00fe08315767b4c0210062fdc0b4cdb3b5371a34ee10

Observation 0a875a7b-ce21-4c50-a7c2-56eff4607016 · outbound

This paper cites Fastinst: A simple query-based model for real-time instance segmentation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Fastinst: A simple query-based model for real-time instance segmentation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:26.250516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.779892Z digest=sha256:cefc04cb5481e50f579e3011c33807ad55113e50ea4d838e5ac502f96424abbb

Observation 85e2643c-66e2-4d01-9075-53beed274f0e · outbound

This paper cites Cascade r-cnn: Delving into high quality object detection.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Cascade r-cnn: Delving into high quality object detection

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:26.089616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.838186Z digest=sha256:e5ed152c6a341bac3e0c769f2c4d764d6a4a8b5261a5ba2da3c1053da586299e

Observation e28765f1-bead-406f-80c3-2e89ef03ec38 · outbound

This paper cites Mask dino: Towards a unified transformer-based framework for object detection and segmentation.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks Mask dino: Towards a unified transformer-based framework for object detection and segmentation

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:52:25.921270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-05T21:52:24.927151Z digest=sha256:f8212451fc2d2fcccbaaf864297337f69f15f10595a26ac80973d133ec32093b

Observation 694e4564-e847-47f0-a216-ccae6b2ecc79 · outbound

This paper cites YOLOv11: An Overview of the Key Architectural Enhancements.

MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks YOLOv11: An Overview of the Key Architectural Enhancements

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T21:52:25.017468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:52:25.017468Z digest=sha256:a176f8525c637df046eb5824fb868b101a0d511071a529189357c084e1903809

Pith citing papers

Observation 8186a886-802d-47e0-bb27-ebe0460eea66 · inbound

InfraNet: Quality-Aware RGB Guidance for Efficient Infrared Object Detection cites this paper.

InfraNet: Quality-Aware RGB Guidance for Efficient Infrared Object Detection MambaTrans: Multimodal Fusion Image Translation via Large Language Model Priors for Downstream Visual Tasks

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-11T23:54:32.839280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:54:32.839280Z digest=sha256:e969c59f36e0a438de64470b5cddc17acc2b7710fe159a487c98ec098877e275