Pith. sign in

Paper Citation Record · LEDGER

EMOv2: Pushing 5M Vision Model Frontier

As of 12 August 2026, this Paper Citation Record lists 100 of 117 outbound references and 1 inbound Pith citation observation for arXiv:2412.06674.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.06674 v1

Coverage vector

measured 100 of 117 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T19:33:09.159780Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T16:30:19.100173Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T16:30:19.157894Z

Reference resolution

100 of 117 outbound references displayed

  • verified exact2
  • verified fuzzy35
  • unresolved63
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8e84cf61-7c8c-40c1-a18e-8b065172ec78 · outbound

This paper cites Rethinking vision transformers for mobilenet size and speed,.

EMOv2: Pushing 5M Vision Model Frontier Rethinking vision transformers for mobilenet size and speed,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.654918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.654918Z digest=sha256:deb0be5a140cc2681631865c0da1278d10992212b19383092774c70420e02159

Observation 73934214-ad02-419b-9ea6-953354eb7b0d · outbound

This paper cites Edgenext: efficiently amalgamated cnn-transformer architecture for mobile vision applications,.

EMOv2: Pushing 5M Vision Model Frontier Edgenext: efficiently amalgamated cnn-transformer architecture for mobile vision applications,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.661571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.661571Z digest=sha256:4ac0f3140136634dea3e4a08072183208b11dd3b23cb5fc877ee9e2fc8f8fd28

Observation 411ff54c-7915-4fda-8451-f04aa6e20fc8 · outbound

This paper cites TinySAM: Pushing the Envelope for Efficient Segment Anything Model.

EMOv2: Pushing 5M Vision Model Frontier TinySAM: Pushing the Envelope for Efficient Segment Anything Model

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.666958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.666958Z digest=sha256:1b07dbd6405e33df145fd0a46fc46101871610a492adfe724af28ff10bdace94

Observation 7e58ad30-8e85-455e-8d61-0653dc7ac474 · outbound

This paper cites EdgeSAM: Prompt-In-the-Loop Distillation for SAM.

EMOv2: Pushing 5M Vision Model Frontier EdgeSAM: Prompt-In-the-Loop Distillation for SAM

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.673378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.673378Z digest=sha256:05c6055e328036774097b99c1a5bf53c491da23d32873e41341e7fe813242c35

Observation 4626c645-2d66-4f12-8eb1-3f575496e272 · outbound

This paper cites RMP-SAM: Towards Real-Time Multi-Purpose Segment Anything.

EMOv2: Pushing 5M Vision Model Frontier RMP-SAM: Towards Real-Time Multi-Purpose Segment Anything

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-08-11T19:33:09.930388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:08.678897Z digest=sha256:00e50af1b5b579ee5aa62b04c2e3d08e4f118bd840916defd4d4be2e9e74ea79

Observation f4a81fa6-25d5-402f-a38b-0421641e9c6a · outbound

This paper cites Semantic flow for fast and accurate scene parsing,.

EMOv2: Pushing 5M Vision Model Frontier Semantic flow for fast and accurate scene parsing,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.684783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.684783Z digest=sha256:a586d91ad3f57d2fc9600fbeb9f460c48da10d2c0499d50c212795d549f96f1e

Observation be0314c3-7ad2-4c6c-bef4-ff62ad9c4b8f · outbound

This paper cites RTMO: Towards high-performance one-stage real-time multi-person pose estimation,.

EMOv2: Pushing 5M Vision Model Frontier RTMO: Towards high-performance one-stage real-time multi-person pose estimation,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.691436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.691436Z digest=sha256:afb1beecb0775531ef1361401e9e03c6109e30ba415d85495038c39eb1f43e9e

Observation c7ad3ff9-e6ec-404f-a405-d69773bf8967 · outbound

This paper cites MobileNets: Efficient Convolutional Neural Networks for Mobile Vision Applications.

EMOv2: Pushing 5M Vision Model Frontier MobileNets: Efficient Convolutional Neural Networks for Mobile Vision Applications

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.696530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.696530Z digest=sha256:d5f82a82a8f8e14c6dc6f4c95b1ad607097379406c9cf183446afff51c36da44

Observation c3b56807-9f23-4947-a900-903807f3d1ea · outbound

This paper cites Mobilenetv2: Inverted residuals and linear bottlenecks,.

EMOv2: Pushing 5M Vision Model Frontier Mobilenetv2: Inverted residuals and linear bottlenecks,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.701456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.701456Z digest=sha256:38ed7f19224db321a63ee72685f474ae40a0d696a40c8c6a17927c781d5be26d

Observation 1ffe787f-beb9-42a5-a433-f44eaf2ffbdf · outbound

This paper cites Searching for mobilenetv3,.

EMOv2: Pushing 5M Vision Model Frontier Searching for mobilenetv3,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.706363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.706363Z digest=sha256:06dcac6d1c8c6d9669f5ef13dd9cc33432495b974cd61f5c0cc3801da73a0ee3

Observation c715fc29-f051-455c-a572-bde4120ccf4c · outbound

This paper cites Ghostnet: More features from cheap operations,.

EMOv2: Pushing 5M Vision Model Frontier Ghostnet: More features from cheap operations,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.711971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.711971Z digest=sha256:16765d22b6b316b6dceb56d9d877c4871fbe86719dbac597a0d0980d08baf27a

Observation 8ac428f7-739d-4c7f-8c21-da27321d5c9d · outbound

This paper cites Efficientnet: Rethinking model scaling for convolu- tional neural networks,.

EMOv2: Pushing 5M Vision Model Frontier Efficientnet: Rethinking model scaling for convolu- tional neural networks,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.717647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.717647Z digest=sha256:f05d463f43faf6773f8ea324e54cdeb0a3805350294f71ea8230e8412663a14c

Observation abbfe0d3-ec84-4b43-bed4-8b2cb3d7107b · outbound

This paper cites Rethinking mobile block for efficient attention- based models,.

EMOv2: Pushing 5M Vision Model Frontier Rethinking mobile block for efficient attention- based models,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.723263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.723263Z digest=sha256:ea4632b80c08263f9807060abb3e629a0e8b1c01179c1c0f9951912a46612823

Observation 7ab1bf7e-2601-4303-8c09-96d2f1a3c2f2 · outbound

This paper cites Separable self-attention for mobile vision transformers,.

EMOv2: Pushing 5M Vision Model Frontier Separable self-attention for mobile vision transformers,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.727666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.727666Z digest=sha256:0aa4d17492a2e77ed1b508f3d9fac787a47498716818c8ceef2a864ca13fc80e

Observation 0990dfa0-8335-417b-9287-aff0f9e29755 · outbound

This paper cites The need for speed in ai,.

EMOv2: Pushing 5M Vision Model Frontier The need for speed in ai,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.731853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.731853Z digest=sha256:7bb3d5f4b01ab4d8a2ac667783152a25552ee5017a2e9cb3487e894639d84e2e

Observation 36848301-9378-4db7-abfe-9730811e6041 · outbound

This paper cites Morgan Kaufmann, 1994.

EMOv2: Pushing 5M Vision Model Frontier Morgan Kaufmann, 1994

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.736200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.736200Z digest=sha256:10282927871bbb4b9ef6bd163563b2465fd90bda7cb4e733ff5752c625270f40

Observation 70f6f7a5-ae5c-43b7-9ad6-fb4e2ae1207a · outbound

This paper cites Mobilevit: Light-weight, general-purpose, and mobile-friendly vision transformer,.

EMOv2: Pushing 5M Vision Model Frontier Mobilevit: Light-weight, general-purpose, and mobile-friendly vision transformer,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.741940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.741940Z digest=sha256:fed08e9a4a51f5645a1212ce69388a1dccf90e640d5e0f56f5b1983f374b15de

Observation e00b2a56-e6d6-4f7d-aa3b-3f0cdebdf639 · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale,.

EMOv2: Pushing 5M Vision Model Frontier An image is worth 16x16 words: Transformers for image recognition at scale,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.747913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.747913Z digest=sha256:bca8d457e19773cb4cb56bd5135cbef4205cc48d9d0ac92185156278ae171025

Observation 3158886e-6a80-4b9c-b34d-ba2450d64983 · outbound

This paper cites Pyramid vision transformer: A versatile backbone for dense prediction without convolutions,.

EMOv2: Pushing 5M Vision Model Frontier Pyramid vision transformer: A versatile backbone for dense prediction without convolutions,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.752576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.752576Z digest=sha256:cca9ea4b31e96ea846ebf7d64e1b3473f4b00cdf3bc322ceb4de0066f82acda8

Observation c4a2ca76-9dc6-49f7-b599-2c84ec0deab5 · outbound

This paper cites Pvt v2: Improved baselines with pyramid vision transformer,.

EMOv2: Pushing 5M Vision Model Frontier Pvt v2: Improved baselines with pyramid vision transformer,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.757577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.757577Z digest=sha256:7a1019bc1cda803d41dffc8779547a79d68bd4d95534a5026b9675ab8257ef29

Observation a8be071b-6bba-4e0e-a390-252fdd28640a · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows,.

EMOv2: Pushing 5M Vision Model Frontier Swin transformer: Hierarchical vision transformer using shifted windows,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.761844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.761844Z digest=sha256:bf6ebc6ba3936d07f2a86b5a9771a6fe35d198e9846de82ec10914ac2646b197

Observation acd20323-67f1-4521-8e2b-b175589b0b70 · outbound

This paper cites Swin transformer v2: Scaling up capacity and resolution,.

EMOv2: Pushing 5M Vision Model Frontier Swin transformer v2: Scaling up capacity and resolution,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.766281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.766281Z digest=sha256:95739f49edc341aa1c1c34a4e72af55c23e513203d653c07fb59ecf1d6c9ccd6

Observation f755c53c-fc13-49d4-8fa3-ed7c03243643 · outbound

This paper cites Analogous to evolutionary algorithm: Designing a unified sequence model,.

EMOv2: Pushing 5M Vision Model Frontier Analogous to evolutionary algorithm: Designing a unified sequence model,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.770568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.770568Z digest=sha256:8b68f126a3363aee730d8454480106c1154d52a236323cee9677f14e213f38cb

Observation ff6ec03f-1536-42c6-9abe-77e2a5d307c0 · outbound

This paper cites Eatformer: improving vision transformer inspired by evolutionary algorithm,.

EMOv2: Pushing 5M Vision Model Frontier Eatformer: improving vision transformer inspired by evolutionary algorithm,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.774881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.774881Z digest=sha256:f6860172adfba106b88994e6decdb6d432efa895f959a52d4ac6b45f75e6e166

Observation f65a7298-ff7b-40a5-8b96-24dabe04e7b0 · outbound

This paper cites Transformer-based visual segmentation: A survey,.

EMOv2: Pushing 5M Vision Model Frontier Transformer-based visual segmentation: A survey,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.779095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.779095Z digest=sha256:f133b2041babe3cd5e7b6291caf0e14fc7facd55dbfe25c0e9b43ea36929756f

Observation 1d98f55a-1abc-47b8-be2e-842544e36605 · outbound

This paper cites Involution: Inverting the inherence of convolution for visual recognition,.

EMOv2: Pushing 5M Vision Model Frontier Involution: Inverting the inherence of convolution for visual recognition,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.784290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.784290Z digest=sha256:c91e21663bf548817a05f499f00460040b2dd6a8a2d8d63c4ba91c2cda478636

Observation fd42233a-f80b-49f1-b593-17445f5468d1 · outbound

This paper cites Reformer: The efficient transformer,.

EMOv2: Pushing 5M Vision Model Frontier Reformer: The efficient transformer,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.789159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.789159Z digest=sha256:8650b1bf79fb7f25679a113cb4d876067b7df545d48628896abb57f1348e62b8

Observation 34e7f7c7-731d-4326-ba20-9dff7667a4df · outbound

This paper cites Rethinking attention with performers,.

EMOv2: Pushing 5M Vision Model Frontier Rethinking attention with performers,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.793762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.793762Z digest=sha256:0bdf2cb90b4d0de8b7b88352daa56364197e828096d4d0258eafcd715dbefd52

Observation 27bd2904-f201-4682-921f-3896a3aad547 · outbound

This paper cites Cvt: Introducing convolutions to vision transformers,.

EMOv2: Pushing 5M Vision Model Frontier Cvt: Introducing convolutions to vision transformers,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.799084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.799084Z digest=sha256:c3e0fc1f96b007e7596ba3ed9b18d5ce548d5ad5c26fd966076f266925aa7b4a

Observation e5eb6ccc-8c74-492c-b035-4c28345ed52e · outbound

This paper cites Next-ViT: Next Generation Vision Transformer for Efficient Deployment in Realistic Industrial Scenarios.

EMOv2: Pushing 5M Vision Model Frontier Next-ViT: Next Generation Vision Transformer for Efficient Deployment in Realistic Industrial Scenarios

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.804743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.804743Z digest=sha256:0b4e72cae968f585bf7956d2cc72239e06b226b23d7c80d3ad1b87e691fe121e

Observation afe24a19-3941-4870-95cd-d62da4f76f0f · outbound

This paper cites Delight: Deep and light-weight transformer,.

EMOv2: Pushing 5M Vision Model Frontier Delight: Deep and light-weight transformer,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.810047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.810047Z digest=sha256:bd9930119fd06baa3fee85c1715304b364eef822e2a71d2e0ff95f092db443a2

Observation 992e5d3b-2184-4fd1-94de-0a93dfc5beb2 · outbound

This paper cites MobileViTv3: Mobile-Friendly Vision Transformer with Simple and Effective Fusion of Local, Global and Input Features.

EMOv2: Pushing 5M Vision Model Frontier MobileViTv3: Mobile-Friendly Vision Transformer with Simple and Effective Fusion of Local, Global and Input Features

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.814498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.814498Z digest=sha256:fcb08892a43625c1edc043bca34f5f5094a718925913cea7562620338904312d

Observation e520eb0b-6d36-4c8f-86c3-bcb393c5ea46 · outbound

This paper cites Mobile-former: Bridging mobilenet and transformer,.

EMOv2: Pushing 5M Vision Model Frontier Mobile-former: Bridging mobilenet and transformer,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.820780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.820780Z digest=sha256:465dd534786d25f1144bbabf308d379550de138de38fda78c0d5658418e066bf

Observation 03f0c357-36fa-4f03-ada7-af4d35b2fa63 · outbound

This paper cites Efficientformer: Vision transformers at mobilenet speed,.

EMOv2: Pushing 5M Vision Model Frontier Efficientformer: Vision transformers at mobilenet speed,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.827247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.827247Z digest=sha256:de6fa3b7a94ab9ae3d2a2bfeefa021958aa337d2cf41b969e83e46a436ca8505

Observation 58c51e32-7a46-4087-80c5-8b3e216cc763 · outbound

This paper cites Attention is all you need,.

EMOv2: Pushing 5M Vision Model Frontier Attention is all you need,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.833265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.833265Z digest=sha256:5b7810452ebd8891c2104735251d01b2ee072479cf9043931296c96b180e3c42

Observation 392d599a-51f6-4ae9-b3fa-f39d0fa02526 · outbound

This paper cites Focal loss for dense object detection,.

EMOv2: Pushing 5M Vision Model Frontier Focal loss for dense object detection,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.839656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.839656Z digest=sha256:ffbbe36c81bbb4353905ae6e1897faaaae06d37787e73fbba01548680d61314b

Observation 55a4544f-fa93-43c1-8266-15eec9444fc9 · outbound

This paper cites SqueezeNet: AlexNet-level accuracy with 50x fewer parameters and <0.5MB model size.

EMOv2: Pushing 5M Vision Model Frontier SqueezeNet: AlexNet-level accuracy with 50x fewer parameters and <0.5MB model size

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.845432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.845432Z digest=sha256:64e33217593c1647e3513c998bae3b6578d4921eb67a212b5ec4578ef057199d

Observation 8a1d4f5b-bc96-480f-aa03-d4e603c7c3ad · outbound

This paper cites Rethinking the inception architecture for computer vision,.

EMOv2: Pushing 5M Vision Model Frontier Rethinking the inception architecture for computer vision,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.851325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.851325Z digest=sha256:3cd0f47837cc7639e9536e72c7db6c3270132ae27889aec90217d2aeeb8f7305

Observation 486bcff3-5a35-4f06-9161-707ae2c96535 · outbound

This paper cites Sfnet: Faster, accurate, and domain agnostic semantic segmentation via semantic flow,.

EMOv2: Pushing 5M Vision Model Frontier Sfnet: Faster, accurate, and domain agnostic semantic segmentation via semantic flow,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.858186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.858186Z digest=sha256:be41df7a5affea188624fda1e9b0cfb4e6528af315d474f26396330300ff4003

Observation 8ff64f92-79ed-4bd1-8890-cc59de2db743 · outbound

This paper cites RepViT: Revisiting Mobile CNN From ViT Perspective.

EMOv2: Pushing 5M Vision Model Frontier RepViT: Revisiting Mobile CNN From ViT Perspective

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.863249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.863249Z digest=sha256:68bfdd61aea02e16fa72231dfa396e26da5a06c1e70b11f736cb36b429ad2012

Observation 9a640a60-d755-402b-9f9c-619af5ab2cac · outbound

This paper cites GhostNetV3: Exploring the Training Strategies for Compact Models.

EMOv2: Pushing 5M Vision Model Frontier GhostNetV3: Exploring the Training Strategies for Compact Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.868843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.868843Z digest=sha256:0020bbe7c1ebba1f89b916863d0f1c546690b68d4a0507a8570ddcf69c360ad5

Observation 4e6f82c6-2a2c-41ae-ad0e-3b70f686dfae · outbound

This paper cites MobileNetV4 -- Universal Models for the Mobile Ecosystem.

EMOv2: Pushing 5M Vision Model Frontier MobileNetV4 -- Universal Models for the Mobile Ecosystem

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.873961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.873961Z digest=sha256:418ce6acd34babbc2a303531a8fb055460bd6d35b9ccf59831d0cc0568cd6c36

Observation c377e4c1-abe1-44de-8b56-044a72afed09 · outbound

This paper cites Training data-efficient image transformers & distillation through attention,.

EMOv2: Pushing 5M Vision Model Frontier Training data-efficient image transformers & distillation through attention,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.879036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.879036Z digest=sha256:b813814662cd94e335553f94534d1cb6c5487d12220659c9e7192b61c86a2331

Observation 1955f161-b3e5-4173-9c88-c9e2ac5f9627 · outbound

This paper cites Deep residual learning for image recognition,.

EMOv2: Pushing 5M Vision Model Frontier Deep residual learning for image recognition,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.883318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.883318Z digest=sha256:3def40e5d422624fed53d8f1bde885999074c86f8500cfff8607e17e3a307018

Observation 32ea3e88-d033-46fa-975a-85b102a9e5fb · outbound

This paper cites Visual Attention Methods in Deep Learning: An In-Depth Survey.

EMOv2: Pushing 5M Vision Model Frontier Visual Attention Methods in Deep Learning: An In-Depth Survey

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.887397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.887397Z digest=sha256:cd38118905907fc30fe412b0a83001984b58c087b6e63a31d5dd905d81df42aa

Observation 743ef52e-aec7-42fd-ac69-d156a532d765 · outbound

This paper cites Recent Advances in Vision Transformer: A Survey and Outlook of Recent Work.

EMOv2: Pushing 5M Vision Model Frontier Recent Advances in Vision Transformer: A Survey and Outlook of Recent Work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.892005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.892005Z digest=sha256:248b259cceb2c381bc9ae1a7ff149a8208ed756c4d1ad76273ffe80972404eb9

Observation cbc42525-401c-4862-83c6-7dea658736a9 · outbound

This paper cites Incorporating convolution designs into visual transformers,.

EMOv2: Pushing 5M Vision Model Frontier Incorporating convolution designs into visual transformers,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.896971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.896971Z digest=sha256:6a02c56092458f859edc3fa3c96ef11428ae8f259481bcdd1c2ee06296ebf38c

Observation 2f21b8c0-dc0a-4499-9607-3048338d7061 · outbound

This paper cites Conditional positional encodings for vision transformers,.

EMOv2: Pushing 5M Vision Model Frontier Conditional positional encodings for vision transformers,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.901253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.901253Z digest=sha256:a400429df7b7c8e46f6a051e36a57adbb4e3b9c51ddd865d94a9a0e5eb0330ad

Observation b0aaec46-3205-4273-81c5-603398431aca · outbound

This paper cites Uniformer: Unified transformer for efficient spatial-temporal representation learning,.

EMOv2: Pushing 5M Vision Model Frontier Uniformer: Unified transformer for efficient spatial-temporal representation learning,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.905345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.905345Z digest=sha256:6ed5088c74bb95f1c3d2a6a98254854a3945221e7056c5ce3f1c150b18b9bc42

Observation efbbadd6-3bd9-46c9-8526-31080f65c0c0 · outbound

This paper cites Moganet: Multi-order gated aggregation network,.

EMOv2: Pushing 5M Vision Model Frontier Moganet: Multi-order gated aggregation network,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:11.116756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:08.909672Z digest=sha256:111e9a5e4b795b2205959c3042eb6d42a0a46a09afbfdc7bb40a9eefd9c3d9c9

Observation 1397ff16-5ff2-4ee2-a7c3-2b64b661da38 · outbound

This paper cites Shvit: Single-head vision transformer with memory efficient macro design,.

EMOv2: Pushing 5M Vision Model Frontier Shvit: Single-head vision transformer with memory efficient macro design,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:11.096081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:08.914141Z digest=sha256:fc744132cbbc0601fd0fa1121825639ed7274f8a44e79c7b52aba51bdd102a43

Observation 47747279-bde4-4292-a45e-1dcd761b2e6d · outbound

This paper cites Metaformer is actually what you need for vision,.

EMOv2: Pushing 5M Vision Model Frontier Metaformer is actually what you need for vision,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:11.076066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:08.919166Z digest=sha256:a8162b3f7cb17e46924a8f41387329f7eaee202874029a8d3ee76db7bef161e7

Observation 43aac6a7-d236-42b7-b7e6-7525849696e0 · outbound

This paper cites LightViT: Towards Light-Weight Convolution-Free Vision Transformers.

EMOv2: Pushing 5M Vision Model Frontier LightViT: Towards Light-Weight Convolution-Free Vision Transformers

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.925512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.925512Z digest=sha256:d68b21b72ab5fe19206220dd714748dd2c8d0766ed112657858328220d4d187e

Observation 27f40d56-3242-48a8-8afc-9e3be376ecb0 · outbound

This paper cites Rest: An efficient transformer for visual recognition,.

EMOv2: Pushing 5M Vision Model Frontier Rest: An efficient transformer for visual recognition,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:11.048115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:08.931911Z digest=sha256:2e02ecc28104c42bf57881deaab2d6834b75a591cae98173507f500f8f59c66e

Observation c917cc1b-f326-4ddd-928b-77f635b36e8b · outbound

This paper cites Edgevits: Competing light-weight cnns on mobile devices with vision transformers,.

EMOv2: Pushing 5M Vision Model Frontier Edgevits: Competing light-weight cnns on mobile devices with vision transformers,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:11.025015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:08.936398Z digest=sha256:c94b3e982234a358824c3736d67f7b65e53605bc083ca6d13a7ab694b567b8c2

Observation e2d96dbe-2661-4e47-baf4-d231bb7d05f0 · outbound

This paper cites Res2net: A new multi-scale backbone architecture,.

EMOv2: Pushing 5M Vision Model Frontier Res2net: A new multi-scale backbone architecture,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:11.005459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:08.941031Z digest=sha256:387edd692129152c605676e96cb91c079f11e588200387e404da0d02ba47d8bf

Observation a1f46c91-27f5-495c-99de-d53d18a39504 · outbound

This paper cites Xcit: Cross- covariance image transformers,.

EMOv2: Pushing 5M Vision Model Frontier Xcit: Cross- covariance image transformers,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.986956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:08.945517Z digest=sha256:c257618ba60d2c49c6f9941b99cfd0542ffd08e9ba9d7b613ed95cbbcbf8f777

Observation a92d1563-4dab-4dfd-97dd-790e49f0ba0c · outbound

This paper cites ViG: Linear-complexity Visual Sequence Learning with Gated Linear Attention.

EMOv2: Pushing 5M Vision Model Frontier ViG: Linear-complexity Visual Sequence Learning with Gated Linear Attention

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-08-11T19:33:09.678654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:08.950493Z digest=sha256:8693047016e4f4261bc5ce1ac3435309d9548fbeb6de10e3526eaee6f188cdfd

Observation a4a37822-38fa-4b5a-acb1-e31952461ee9 · outbound

This paper cites PointRWKV: Efficient RWKV-Like Model for Hierarchical Point Cloud Learning.

EMOv2: Pushing 5M Vision Model Frontier PointRWKV: Efficient RWKV-Like Model for Hierarchical Point Cloud Learning

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.955174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.955174Z digest=sha256:5ae105f20acec02d9831ec1ff4fc054bbf67eaf63d619d3371ab45ccd92cbd7e

Observation a479897f-bf53-4400-a6e2-ada955ad23ca · outbound

This paper cites Vision-RWKV: Efficient and Scalable Visual Perception with RWKV-Like Architectures.

EMOv2: Pushing 5M Vision Model Frontier Vision-RWKV: Efficient and Scalable Visual Perception with RWKV-Like Architectures

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.960085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.960085Z digest=sha256:eb401f1bb0d394031075c40160e35341e791abe76059f7cbb1828944dc73cbc4

Observation 17fad12d-aff0-499a-b5e7-59f97270620c · outbound

This paper cites Mamba or RWKV: Exploring High-Quality and High-Efficiency Segment Anything Model.

EMOv2: Pushing 5M Vision Model Frontier Mamba or RWKV: Exploring High-Quality and High-Efficiency Segment Anything Model

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.964841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.964841Z digest=sha256:c65d8ced735cd7c70b3b1e005319ea740e76f06877c54fbe90c055dac2085d6d

Observation 846bddbb-113b-45fe-8c27-736174fb5682 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

EMOv2: Pushing 5M Vision Model Frontier Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.969473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.969473Z digest=sha256:4e100f5b65b7af057a4e311d432eda5c8f445e2521faeb8c64239e9398325e22

Observation 17db84ff-8a86-4532-9a30-958bf9e1c72a · outbound

This paper cites RWKV: Reinventing RNNs for the Transformer Era.

EMOv2: Pushing 5M Vision Model Frontier RWKV: Reinventing RNNs for the Transformer Era

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.974566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.974566Z digest=sha256:0a23b41f86575b9e2ff9e1519f6e9173790485076d24dd815ff1bb320e6dcb56

Observation cecb269c-59f2-46ad-b5ee-234fec159eaa · outbound

This paper cites Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model.

EMOv2: Pushing 5M Vision Model Frontier Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.979919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.979919Z digest=sha256:2dc7ad653207f833afa1359cf14812dfac785b2ede1043ee314aa78c0d285b9d

Observation f113bcfd-9f82-42a5-a765-b1fedf2ac45c · outbound

This paper cites EfficientVMamba: Atrous Selective Scan for Light Weight Visual Mamba.

EMOv2: Pushing 5M Vision Model Frontier EfficientVMamba: Atrous Selective Scan for Light Weight Visual Mamba

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.985119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.985119Z digest=sha256:50fdaa7c9c1401401c477d6fde95d1f9c45f6596201bd2290845685663326e76

Observation ca7ddef0-e575-4101-a4b6-35eb9fb6ce84 · outbound

This paper cites MobileMamba: Lightweight Multi-Receptive Visual Mamba Network.

EMOv2: Pushing 5M Vision Model Frontier MobileMamba: Lightweight Multi-Receptive Visual Mamba Network

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:08.990933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:08.990933Z digest=sha256:da8c537e7ded924dd5c9e3832cdab8f1a4a1e46b2dbbb055c0166a056cedb15c

Observation d5be2b96-5ac2-42ea-bd25-a18e90e94b8a · outbound

This paper cites Scalable diffusion models with transformers,.

EMOv2: Pushing 5M Vision Model Frontier Scalable diffusion models with transformers,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.969001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:08.996411Z digest=sha256:15ff683cf30604a61a8ad90499a5c81cf2ee456464d39777d21841fd49bbbc88

Observation dd7ee496-e397-4116-8610-07581aeae030 · outbound

This paper cites Focal attention for long-range interactions in vision transformers,.

EMOv2: Pushing 5M Vision Model Frontier Focal attention for long-range interactions in vision transformers,

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.950398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.002489Z digest=sha256:1d529b93b34a51c5961b5f8e860e1e817c08a12897407ecca56b73c5ffb6180d

Observation e62db3f7-e045-4f42-b810-cd7277157d9b · outbound

This paper cites Cswin transformer: A general vision transformer backbone with cross-shaped windows,.

EMOv2: Pushing 5M Vision Model Frontier Cswin transformer: A general vision transformer backbone with cross-shaped windows,

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.928923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.008272Z digest=sha256:137f3c9dd27b6adf1e8c580739646cf37f20c8cb824ba0e5bce93dfdaecdc804

Observation 7ad91af7-83be-49a6-9a73-bc19de038b30 · outbound

This paper cites Inception transformer,.

EMOv2: Pushing 5M Vision Model Frontier Inception transformer,

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.902466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.013061Z digest=sha256:1762232ff9541fbb0479746cf62aae3b78cc8c220080081ac9e42397c0e67a20

Observation 58e8c158-a0f4-46c0-b058-998897949334 · outbound

This paper cites Pay attention to mlps,.

EMOv2: Pushing 5M Vision Model Frontier Pay attention to mlps,

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.879202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.017652Z digest=sha256:b497abe5559e531ecda2189023b02becfb40d30882dbbee0d3c71fd172b55bf6

Observation f5339a8e-65cc-45fd-bba3-85cb20cbf269 · outbound

This paper cites Mlp-mixer: An all-mlp architecture for vision,.

EMOv2: Pushing 5M Vision Model Frontier Mlp-mixer: An all-mlp architecture for vision,

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.856023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.022746Z digest=sha256:7f0f2ca5719d59e6a2d628827d3100efdd3ba3e51d2eaba91476414ae53beb08

Observation edde3b16-5899-4a8c-b443-f4f29f6b4e98 · outbound

This paper cites Resmlp: Feed- forward networks for image classification with data-efficient training,.

EMOv2: Pushing 5M Vision Model Frontier Resmlp: Feed- forward networks for image classification with data-efficient training,

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.839369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.026899Z digest=sha256:0dd6dc5beb4cfe7cfcf6896619b7a1bb8e06a731f7be27e15c681092637d6211

Observation fe8409ad-4abb-4bb7-9f3f-eab73da92ada · outbound

This paper cites Shufflenet v2: Practical guidelines for efficient cnn architecture design,.

EMOv2: Pushing 5M Vision Model Frontier Shufflenet v2: Practical guidelines for efficient cnn architecture design,

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.807090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.031454Z digest=sha256:cae2aea0828df332bfea36a006ebdaace636b5c0bdd70dda4d2a2a3561636ddc

Observation 3b036c57-67c3-49b0-88df-ee8f5064aa7b · outbound

This paper cites Moat: Alternating mobile convolution and attention brings strong vision models,.

EMOv2: Pushing 5M Vision Model Frontier Moat: Alternating mobile convolution and attention brings strong vision models,

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.780993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.036522Z digest=sha256:88dab7d08b6f93d09d6bccdc74361147be71484c09b0dbf307f73c7f90628ddf

Observation f207e548-e190-410e-88ce-ff782ccfacbf · outbound

This paper cites Batch normalization: Accelerating deep network training by reducing internal covariate shift,.

EMOv2: Pushing 5M Vision Model Frontier Batch normalization: Accelerating deep network training by reducing internal covariate shift,

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.751839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.042235Z digest=sha256:d500b8f7d143f5be03f83cd6295d664f36c8f9b596d27e4b26bd5a2c48df2da1

Observation d97791ea-07fd-46d4-ad78-f68d35816897 · outbound

This paper cites Gaussian Error Linear Units (GELUs).

EMOv2: Pushing 5M Vision Model Frontier Gaussian Error Linear Units (GELUs)

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:09.047105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:09.047105Z digest=sha256:85b6baac88db55b913604319cdc5b4b1cb2b0e151e765a2c63d20a908941662a

Observation 894c76f9-82e0-46f1-bab7-574b3f42d27c · outbound

This paper cites Layer Normalization.

EMOv2: Pushing 5M Vision Model Frontier Layer Normalization

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:09.052064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:09.052064Z digest=sha256:79d58c8e7f486648205e3f6d9ea509c75b0810d7c48a873f9ef207a7c4951c60

Observation e8a62178-94ae-470f-9605-8e38e2c5a708 · outbound

This paper cites Imagenet: A large-scale hierarchical image database,.

EMOv2: Pushing 5M Vision Model Frontier Imagenet: A large-scale hierarchical image database,

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.731458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.056904Z digest=sha256:7bd5a454297b79a0a1ce1e8455f3db48d64ff96b1284a3d1505f977bb38ce384

Observation 57fc92c9-d1a3-4d3d-aad8-2b07d4b3bfb5 · outbound

This paper cites Decoupled weight decay regularization,.

EMOv2: Pushing 5M Vision Model Frontier Decoupled weight decay regularization,

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.708528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.061008Z digest=sha256:641b01250cd53d4ee359f79898ea91cfcc3cadadfd9e8934cfb4ffd99d387aa1

Observation ad719448-0ba3-43b0-a4e6-4fc178cfbcaf · outbound

This paper cites SGDR: Stochastic gradient descent with warm restarts,.

EMOv2: Pushing 5M Vision Model Frontier SGDR: Stochastic gradient descent with warm restarts,

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.684108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.065269Z digest=sha256:c740322521d337b25b47550453ff437a28afa7c2f61c04b1a7071278ade5b536

Observation 8684aa9d-264d-40d0-bd48-c941080f277d · outbound

This paper cites Rethinking the inception architecture for computer vision,.

EMOv2: Pushing 5M Vision Model Frontier Rethinking the inception architecture for computer vision,

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.657560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.069479Z digest=sha256:d5fd8b1a702b6a5bebcea9bb008f08dac6314485de6c790d2a732e56a0752298

Observation baa63e5f-30b8-4daa-9ff7-d511a5d314f1 · outbound

This paper cites Deep networks with stochastic depth,.

EMOv2: Pushing 5M Vision Model Frontier Deep networks with stochastic depth,

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.623924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.073594Z digest=sha256:6129b1a711912ed7b79bc79d71bc724f25a779b201b06203408cb83850ff30b8

Observation 042320f6-541c-466e-b4d8-dd8411d8701e · outbound

This paper cites Randaugment: Practical automated data augmentation with a reduced search space,.

EMOv2: Pushing 5M Vision Model Frontier Randaugment: Practical automated data augmentation with a reduced search space,

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.590744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.078375Z digest=sha256:9f70765a3b5f82cad452580142ab82e7c2333647306dd29a2907d98cbe713348

Observation b8afa672-506a-4e7b-b2cb-f96e1bfa5a93 · outbound

This paper cites Going deeper with image transformers,.

EMOv2: Pushing 5M Vision Model Frontier Going deeper with image transformers,

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.568617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.083189Z digest=sha256:c5e4af87b454c9e2b5feccb6b0b538fe187b26c49caa8517775b66f0c1baad14

Observation c9e714b6-8083-4929-bc1e-23725ee1c6e0 · outbound

This paper cites Dropout: a simple way to prevent neural networks from overfitting,.

EMOv2: Pushing 5M Vision Model Frontier Dropout: a simple way to prevent neural networks from overfitting,

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.541249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.088459Z digest=sha256:1ad7d545b4f20e33154444b4625e3990a8d3ea9c78c6fb785a4623fe7fb8c3e8

Observation 18d1dcc5-5954-40cf-945f-0ea0af80b9df · outbound

This paper cites mixup: Beyond empirical risk minimization,.

EMOv2: Pushing 5M Vision Model Frontier mixup: Beyond empirical risk minimization,

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.515981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.093649Z digest=sha256:fe659fa4448b91d003bfe14cf20a98e0631321ce82d956fb2c5d9500b5a8f993

Observation 33bf2199-9b74-4fc9-bfee-73954325550b · outbound

This paper cites Cutmix: Regularization strategy to train strong classifiers with localizable features,.

EMOv2: Pushing 5M Vision Model Frontier Cutmix: Regularization strategy to train strong classifiers with localizable features,

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.493164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.099815Z digest=sha256:f525d3c5c783040b8eb7fe7057f292aeb6f2605895a250b126674b3a6c319455

Observation 30b561b1-19a5-4bd4-932c-afd800402dda · outbound

This paper cites Random erasing data augmentation,.

EMOv2: Pushing 5M Vision Model Frontier Random erasing data augmentation,

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.471017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.105749Z digest=sha256:377a6444acae176cd88ea2c551bd4c0a295bf77c88d9b37aadb052950e0c8dfc

Observation 3119840b-3a8f-4a63-9f7b-63eaaee56d96 · outbound

This paper cites All tokens matter: Token labeling for training better vision transformers,.

EMOv2: Pushing 5M Vision Model Frontier All tokens matter: Token labeling for training better vision transformers,

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.442430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.111014Z digest=sha256:3dc1a055ee2f2d5811eab819030ebfc485e7b93180e33b8dcd7462ca576e7e3b

Observation 01d4d43a-718c-47c0-b2b4-9a604226b8d6 · outbound

This paper cites Pytorch image models,.

EMOv2: Pushing 5M Vision Model Frontier Pytorch image models,

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:09.116189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:09.116189Z digest=sha256:b4698ef2172ffd8bb0d0351e76cb16a931c76eeca624f4b6ee79030750422c20

Observation 43f888f2-8d32-43d2-a133-5c8e753c3104 · outbound

This paper cites Tresnet: High performance gpu-dedicated architecture,.

EMOv2: Pushing 5M Vision Model Frontier Tresnet: High performance gpu-dedicated architecture,

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.399510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.120756Z digest=sha256:df9674af5502b7d1a07252c09addd27371c390834e2c0f613561a820a420d4f2

Observation e360292b-ab38-4b64-92ab-7c9887d2741d · outbound

This paper cites Run, don’t walk: chasing higher flops for faster neural networks,.

EMOv2: Pushing 5M Vision Model Frontier Run, don’t walk: chasing higher flops for faster neural networks,

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.376257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.125356Z digest=sha256:7fce1930df2fb2b8adca682aa6e6c51d0a8035ad5a2aec0d9aa3c47d06a12c47

Observation db85b25a-659b-4da1-8527-28afa9548087 · outbound

This paper cites MoCoViT: Mobile Convolutional Vision Transformer.

EMOv2: Pushing 5M Vision Model Frontier MoCoViT: Mobile Convolutional Vision Transformer

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:09.129596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:09.129596Z digest=sha256:a02895c76fd0d23ed022c14b5fc305f713c097af74cce4e5ed73fc2a3db44725

Observation daac5484-93a7-4379-86b7-776dbb425025 · outbound

This paper cites Efficientvit: Memory efficient vision transformer with cascaded group attention,.

EMOv2: Pushing 5M Vision Model Frontier Efficientvit: Memory efficient vision transformer with cascaded group attention,

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.354143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.134748Z digest=sha256:0067e22369af9fe04a0c5b3c536d71a785b14723b9aebc253bb1887c14e1b476

Observation 24faf258-60a5-4ef8-927a-48bd188e380d · outbound

This paper cites Mpvit: Multi-path vision transformer for dense prediction,.

EMOv2: Pushing 5M Vision Model Frontier Mpvit: Multi-path vision transformer for dense prediction,

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.317228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.139328Z digest=sha256:a1d4c03895f021dea46e80fb9e51d501a9d7e22d49a52a132533cd03c720bf7f

Observation 063dfb65-5f9b-401b-9606-be7d48e94af5 · outbound

This paper cites Multi-Scale VMamba: Hierarchy in Hierarchy Visual State Space Model.

EMOv2: Pushing 5M Vision Model Frontier Multi-Scale VMamba: Hierarchy in Hierarchy Visual State Space Model

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:09.144083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:09.144083Z digest=sha256:cc8dd6c84f2e62350ebb395434684d4a12403252059103cddbf87fcdfaf21c55

Observation f4e01e62-c4f0-4e4d-b604-710d53478543 · outbound

This paper cites MambaOut: Do We Really Need Mamba for Vision?.

EMOv2: Pushing 5M Vision Model Frontier MambaOut: Do We Really Need Mamba for Vision?

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-11T19:33:09.149397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:33:09.149397Z digest=sha256:bfb19edd96839b22d85cb03fbdc0c89bce381b5b023e32bd65d96e0ccb1333f1

Observation 61065d6f-849c-4e43-9cbf-9c3fc1bc013b · outbound

This paper cites Microsoft coco: Common objects in context,.

EMOv2: Pushing 5M Vision Model Frontier Microsoft coco: Common objects in context,

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.294337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.154553Z digest=sha256:02b8458a634fe3040816ad9aa118a0e69080b7a09be348811194e3faa937eae3

Observation 0131d5fc-e6ef-4ee5-94de-c487610a00b3 · outbound

This paper cites Mask r-cnn,.

EMOv2: Pushing 5M Vision Model Frontier Mask r-cnn,

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T19:33:10.277774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T19:33:09.159780Z digest=sha256:5098c9ff897f4e557513effd607948c025c36d28e4466712d3b46b305a9264fc

Pith citing papers

Observation fd592c2d-bfac-4ee6-98bc-6e25b4ba6797 · inbound

VQualA 2025 Challenge on Face Image Quality Assessment: Methods and Results cites this paper.

VQualA 2025 Challenge on Face Image Quality Assessment: Methods and Results EMOv2: Pushing 5M Vision Model Frontier

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:30:19.165100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T16:30:19.100173Z digest=sha256:2948a5c9443659f7778d51915adf68305237d02ccb6481cdfe5c77709cad4112