Pith. sign in

Paper Citation Record · LEDGER

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer

As of 10 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 0 inbound Pith citation observations for arXiv:2506.12982.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.12982 v1

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:42:03.479929Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

34 of 34 outbound references displayed

  • verified exact1
  • verified fuzzy25
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 121603b3-c3c8-4fa1-91cb-0d361e9a426d · outbound

This paper cites Computing receptive fields of convolutional neural networks.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Computing receptive fields of convolutional neural networks

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.536866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:00.047124Z digest=sha256:7423d5255fe72434a0f0429e99e5af2cb25b7abd0b3886fbefbde7719d3136e9

Observation d9c37fc8-f615-41b0-b61b-32b42d6841be · outbound

This paper cites Advances in medical image analysis with vision transformers: a comprehensive review.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Advances in medical image analysis with vision transformers: a comprehensive review

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.530417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:00.133965Z digest=sha256:3009791d52ca0cc3e996f85131175e0643247989e3f4c3c6936f3c8adf2f1d0e

Observation 298bc71d-fd4c-457e-bb0e-de4202cf6ffd · outbound

This paper cites Med-former: A transformer based architecture for medical image classification.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Med-former: A transformer based architecture for medical image classification

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.523287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:00.178727Z digest=sha256:3fec8f2d00746bc41119edc70a0d702fefda58e6d04d5c3dcd8a02444906cc28

Observation e314b164-6d93-4c17-bac4-6063c24cea80 · outbound

This paper cites Coatnet: Marrying convolution and attention for all data sizes.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Coatnet: Marrying convolution and attention for all data sizes

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.516390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:00.235061Z digest=sha256:028acfe76cfb1d62b2557f94a34b101bbf2ad24deaed131901ab7df88c92c0b9

Observation 50520a72-5c57-44a4-8a2c-aed2f291d68b · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:00.295713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:00.295713Z digest=sha256:11d77b014462249882053a056c605bdbd5d002fd8a5054b75ca7d5bca7237754

Observation b3af8723-7c72-4972-8629-18843815f874 · outbound

This paper cites Convit: Improving vision transformers with soft convolutional inductive biases.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Convit: Improving vision transformers with soft convolutional inductive biases

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.509453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:00.354803Z digest=sha256:ad22e83c931b054dac6ed4d09e7f25d468d58e9ecd28f62a18500c2d41e1131c

Observation d4639b79-7dd1-4def-b9ba-5baafb7d5221 · outbound

This paper cites Rmt: Retentive networks meet vision transformers.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Rmt: Retentive networks meet vision transformers

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.502640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:00.411487Z digest=sha256:91dda23383828e06209b32ff964b14972dded17931642858b376857df6e98ed8

Observation 1f0487c9-c931-4c7b-b903-10e3e60a3165 · outbound

This paper cites Cmt: Convolutional neural networks meet vision transformers.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Cmt: Convolutional neural networks meet vision transformers

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.496119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:00.507518Z digest=sha256:def9f5c3810036359dad9dd305f3e5a98c5622ce35b33fa8a794c56ceca412b1

Observation d4222645-5575-48b0-9e1a-89fe1025b94e · outbound

This paper cites Higt: Hierarchical interaction graph-transformer for whole slide image analysis.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Higt: Hierarchical interaction graph-transformer for whole slide image analysis

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.489171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:00.562683Z digest=sha256:6679713f1ba9f841af4f75993f729fa57a3c571f2d680f5f534996d51032f396

Observation 5f5091d8-991a-4355-94ed-01c3315b7ef6 · outbound

This paper cites Deep residual learning for image recognition.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Deep residual learning for image recognition

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:00.613900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:00.613900Z digest=sha256:774a7a4df922ebee0e32270152d07aae63d7bb7f18491743164adb354d7c3b6e

Observation def84805-d051-4d8a-bd84-ccae8d84e67e · outbound

This paper cites Conv2former: A simple transformer-style convnet for visual recognition.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Conv2former: A simple transformer-style convnet for visual recognition

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.478390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:00.668212Z digest=sha256:9c086a7fbc095c341efe316c8ed6569abb080f0656d52a83a965522469584803

Observation 92195be2-b7e0-409c-8cb0-cc7e922c4c70 · outbound

This paper cites Benchmarking self-supervised learning on diverse pathology datasets.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Benchmarking self-supervised learning on diverse pathology datasets

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:00.721756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:00.721756Z digest=sha256:46aee2d8b4845da1db313d5fdd5d831b4986e415b4bed06f258f2db5c7a7d18a

Observation 9b464d33-8b4f-4436-ac47-0df59a554708 · outbound

This paper cites Vision Transformer for Small-Size Datasets.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Vision Transformer for Small-Size Datasets

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:00.774613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:00.774613Z digest=sha256:0be3022377aa8104a8e394b680e78b7008ee989973c10e22079c2eae3fbb5a01

Observation 3053d9d7-c1e7-44c5-a0a8-de337fa5567b · outbound

This paper cites Mvitv2: Improved multiscale vision transformers for classification and detection.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Mvitv2: Improved multiscale vision transformers for classification and detection

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.468463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:00.855810Z digest=sha256:e1dba0215e23156a8511ad053a5bb77c3364b21be6573dee25a7a46117577f62

Observation 52397976-09f5-4ae8-9038-7b709136d20c · outbound

This paper cites LocalViT: Analyzing Locality in Vision Transformers.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer LocalViT: Analyzing Locality in Vision Transformers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:00.924074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:00.924074Z digest=sha256:6bb1d31d8b1c1f1c0524d924f2c65a0401910c5c1e3edba857cedee41fbab092

Observation 5f28b332-df95-4833-b04a-638ef511a154 · outbound

This paper cites Scale-aware modulation meet transformer.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Scale-aware modulation meet transformer

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.461660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:00.975737Z digest=sha256:b0cb84e3677c7267f9b35ba71aaaf3bc663aefedfb20602ef09299227c0f1330

Observation cdd0b2fb-d693-44db-8011-068cc29c180d · outbound

This paper cites Exploiting geometric features via hierarchical graph pyramid transformer for cancer diagnosis using histopathological images.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Exploiting geometric features via hierarchical graph pyramid transformer for cancer diagnosis using histopathological images

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.455126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:01.050750Z digest=sha256:b7887a27693ea662957ab7528b4aaac0d064156d061e254dce5477ac1dae7268

Observation 5e1bf034-b649-4439-931d-92c2e631d6c2 · outbound

This paper cites Efficient training of visual transformers with small datasets.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Efficient training of visual transformers with small datasets

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.448424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:01.186246Z digest=sha256:c75d739312f134b66d4b4b122f395e9a518156b14622c98347939418189365d6

Observation 3f4df035-9a75-439e-a79c-e617f38d61ce · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Swin transformer: Hierarchical vision transformer using shifted windows

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.441813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:01.293237Z digest=sha256:348936dc0cc579121f4ed0cc45f70161107acf41f342bc286227cbfeddd3bc61

Observation eeb6cd80-34e1-48e8-8777-2d5faebc73f0 · outbound

This paper cites Hybrid ladder transformers with efficient parallel-cross attention for medical image segmentation.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Hybrid ladder transformers with efficient parallel-cross attention for medical image segmentation

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.434969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:01.341331Z digest=sha256:455ae46c8eb21a59aaacd7289301171e443ab6e0f2c9706f84d6ca33914f6a86

Observation 65557212-abf7-4534-b817-70631fa746e1 · outbound

This paper cites Medvit: a robust vision transformer for generalized medical image classification.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Medvit: a robust vision transformer for generalized medical image classification

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.428180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:01.430781Z digest=sha256:d69a3a2765017fa5a7162eec2f5e567a6e18a124be39411d49ac36f1a72ff8d9

Observation 18ce1c21-9bc2-4adc-b51f-0a62698bd0cb · outbound

This paper cites Cell-detr: Efficient cell detection and classification in wsis with transformers.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Cell-detr: Efficient cell detection and classification in wsis with transformers

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.421183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:01.662919Z digest=sha256:996101c015b8b5e3f6e4fc0c49c85167fd9a9cc9f5daaf16d7e521b72ba082f4

Observation 00624969-f0c4-4b55-8cff-7f266306305d · outbound

This paper cites Do vision transformers see like convolutional neural networks? Advances in neural information processing systems, 34: 0 12116--12128, 2021.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Do vision transformers see like convolutional neural networks? Advances in neural information processing systems, 34: 0 12116--12128, 2021

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.305020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:01.795141Z digest=sha256:46e9447a2fac31b801649f51fcfe7261b66a33bc6f1e61b7639fac2cbd6f6d5b

Observation 78f39b2a-4334-4bc3-8079-eb6aa7f762a9 · outbound

This paper cites Transformers in medical imaging: A survey.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Transformers in medical imaging: A survey

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:05.209668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:01.938556Z digest=sha256:36583128eda6566062f7ad02abab2baaee492e0ae9e117343aed5b025b284d8e

Observation 7469ffd3-ae0c-49b7-9e3e-835abbd6deea · outbound

This paper cites Transmil: Transformer based correlated multiple instance learning for whole slide image classification.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Transmil: Transformer based correlated multiple instance learning for whole slide image classification

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:04.989612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:02.070143Z digest=sha256:bc33202c8f455b8e5aac066e2bc89278dddce8d1f823e569db07ee8938bae4f0

Observation 92a00660-ca49-474d-9aae-997f94d1973c · outbound

This paper cites Training data-efficient image transformers & distillation through attention.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Training data-efficient image transformers & distillation through attention

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:04.786575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:02.235797Z digest=sha256:c3b1adea4640ac432f5bcdca3e927751a9e54d4ec74dfaaca41fc7c82a9ed7d6

Observation 0c0b9d61-4cf6-4e23-a466-f9b00e5cca5a · outbound

This paper cites Uctransnet: rethinking the skip connections in u-net from a channel-wise perspective with transformer.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Uctransnet: rethinking the skip connections in u-net from a channel-wise perspective with transformer

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:04.591573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:02.352240Z digest=sha256:a4480912cf98304178d552ab2ca6e5271a97dbae6ab98884f3fb9bc49e665721

Observation 121df421-3e5a-4ac0-afa0-34baf9cc7704 · outbound

This paper cites The cancer genome atlas pan-cancer analysis project.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer The cancer genome atlas pan-cancer analysis project

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:02.553062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:02.553062Z digest=sha256:f4be1f4654f8897b8fd8ef142a770d1d0d76d7b52312c81fca6ee22e195bbddf

Observation 2ea0001f-c4b9-4962-9c5b-63b9e7cdde7c · outbound

This paper cites Cvt: Introducing convolutions to vision transformers.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Cvt: Introducing convolutions to vision transformers

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:04.415585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:02.701230Z digest=sha256:55465ed7fa778711bd881fd1e02dcf68c1ebe8d37ac717f35542049e1aa1f1df

Observation 5b54213b-59f5-44c4-a77a-9f180444fed0 · outbound

This paper cites Co-scale conv-attentional image transformers.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Co-scale conv-attentional image transformers

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:04.170358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:02.815883Z digest=sha256:02510eefe09a671d7c4ca5ab384f8f2ca09caa76a08464a5958c8f91e2ddc3bc

Observation 22ee1fca-c355-4cad-817c-890a6cf07767 · outbound

This paper cites Incorporating convolution designs into visual transformers.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Incorporating convolution designs into visual transformers

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:03.960318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:02.981271Z digest=sha256:400b7d8da38991c7805e216b5c154a94befb4a5b34c049c70fcf53e684bcb91a

Observation 7267ffd3-39ec-4805-a027-f49bd04c39d4 · outbound

This paper cites CLASS-M: Adaptive stain separation-based contrastive learning with pseudo-labeling for histopathological image classification.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer CLASS-M: Adaptive stain separation-based contrastive learning with pseudo-labeling for histopathological image classification

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:42:03.788724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:42:03.137920Z digest=sha256:689c77db8fe0330221e1ede9e5de91b8068ac01222f3f3913ecbc1419ce8b9a1

Observation 002b6ea9-b74b-47ae-ab6c-537edb0442c0 · outbound

This paper cites Crossformer: Transformer utilizing cross-dimension dependency for multivariate time series forecasting.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Crossformer: Transformer utilizing cross-dimension dependency for multivariate time series forecasting

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:03.322168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:03.322168Z digest=sha256:71f2f399d427038b76623dbb958605e1f95a54503605b307ab685286a9027adc

Observation edb0aeb1-3b24-468a-9ba1-c8724a95bdc9 · outbound

This paper cites Deformable DETR: Deformable Transformers for End-to-End Object Detection.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer Deformable DETR: Deformable Transformers for End-to-End Object Detection

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:03.479929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:03.479929Z digest=sha256:318ae8e7934a2e8afddd7bbf0d0f25a46fa4b126ea83270fe557864a8af67329

Pith citing papers

No inbound Pith citation observations are available.