Pith. sign in

Paper Citation Record · LEDGER

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation

As of 11 August 2026, this Paper Citation Record lists 100 of 111 outbound references and 0 inbound Pith citation observations for arXiv:2607.25527.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.25527 v1

Coverage vector

measured 100 of 111 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T02:14:09.853239Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 111 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved100
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6d822cfa-5d27-4f91-a107-d3382ffde034 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation LLaMA: Open and Efficient Foundation Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.635442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.635442Z digest=sha256:a2abe7f0d8b09cd17b742a4747e90c0b1095df47709831a5d5f7c749d02298c8

Observation b7668705-e96e-44ad-b5bf-5908ca006fa0 · outbound

This paper cites 2018 , publisher=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation 2018 , publisher=

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.638491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.638491Z digest=sha256:a6fd77af8fb54c4c10e27f30786719d37ab322dca0a21dacaef49b943c25f729

Observation c5e68546-645f-4267-971d-1006b5f57955 · outbound

This paper cites OpenAI blog , volume=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation OpenAI blog , volume=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.640771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.640771Z digest=sha256:835c58a5be0aae6fb7901073113317d15814cd210faf7bd903c51bf82504588f

Observation 1190516d-1a05-4d98-8c67-dd2a61529c8d · outbound

This paper cites GPT-4 Technical Report.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation GPT-4 Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.643040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.643040Z digest=sha256:5d67bfa927016ba7ed6ebd31add21f970ce005aeddb3c594f00579c314c76bc9

Observation 917eec41-5b12-456a-a0d7-a0a1828968f7 · outbound

This paper cites GLM: General Language Model Pretraining with Autoregressive Blank Infilling.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation GLM: General Language Model Pretraining with Autoregressive Blank Infilling

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.645560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.645560Z digest=sha256:7d37efcd0ca433124c7d4f0e6c91497c642fc51d45599061924d7461d3725f01

Observation bdfef975-41eb-4719-856a-2a56887ebb08 · outbound

This paper cites Qwen3 Technical Report.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Qwen3 Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.648170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.648170Z digest=sha256:19478077d2ea54c00dcb0b59d071efe9d7c02259d98aa64cca884a1daa92db8d

Observation 01db86fc-20f5-4feb-b9d3-ac4ecdf7f5bb · outbound

This paper cites Advances in neural information processing systems , volume=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Advances in neural information processing systems , volume=

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.650835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.650835Z digest=sha256:856c88b3cf4bf71b1481cb2f64f2db142cbaadf4647d2c8a3b02ddd22fb83cb2

Observation 59fe1502-9dad-4090-a301-104a6dd0583d · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.653137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.653137Z digest=sha256:1bb0d03b60869df3aa2ac2f5451dff4c7e072decea098bf8896f352884816129

Observation 5f4f585e-2efb-41d1-8ae3-8b82277da990 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.655279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.655279Z digest=sha256:a0770d9f8bdb2cb8997696055b975149b7fad5cc9883b7458d17eb84b8963b12

Observation b4963aa3-b34e-4146-80ad-30b776d668b0 · outbound

This paper cites Qwen2.5-VL Technical Report.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Qwen2.5-VL Technical Report

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.657370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.657370Z digest=sha256:7be124564ac003cb8ea685b21e21ae3584e23d3f744c9d43bd1e73f234cc9a79

Observation dfb57b7f-6ec7-44f0-bb3c-61f9bb92d819 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.659839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.659839Z digest=sha256:f72b6fdd5803bf903d91eb7af0619af20f284df2bf6ebfb65239ef7e897a0fb6

Observation e9236777-6ac9-48e9-a2a6-e246e03ed9dc · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.662243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.662243Z digest=sha256:686605f62408752d6b44f1770fb3ba612071c1acb190477464f19dd6b002aee5

Observation d57e628c-7ca7-420a-8eff-85740b4f658c · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.664298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.664298Z digest=sha256:11d6907015960d731b5d04cf120362bca53398c1d788040eb9aadae19364d4f7

Observation cce9ad22-26cd-415c-afe4-39ddaa34608d · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.666550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.666550Z digest=sha256:464ee8be938439da4b6bbe6b820f35a6f67d9b08c1ce1fedd8bd179dced63f3f

Observation 6f888a35-1279-4bec-b4ad-412d461792df · outbound

This paper cites SmolVLM: Redefining small and efficient multimodal models.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation SmolVLM: Redefining small and efficient multimodal models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.668753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.668753Z digest=sha256:5ef6039b4d50a66d5fd6fcf2b0bbff91dcc7566f609d34cb64cf8138608f86cb

Observation 8ab7e486-4f61-4e3f-bd85-d547503facae · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.670883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.670883Z digest=sha256:03443d03fcad3eec7312f6a2dac1e565edf2e504020510e697a2acdd2f6dd889

Observation 3f2abd4a-d34b-48c6-9f78-7247f82f084c · outbound

This paper cites HART: Efficient Visual Generation with Hybrid Autoregressive Transformer.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation HART: Efficient Visual Generation with Hybrid Autoregressive Transformer

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.672981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.672981Z digest=sha256:19c33360d6bbc6b74c46dbf2015ad9d6fdf2c131f88c442649f8e00cbba717c0

Observation a510151f-7821-4a04-b832-67c7e74bdf58 · outbound

This paper cites Advances in neural information processing systems , volume=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Advances in neural information processing systems , volume=

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.675146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.675146Z digest=sha256:c72dd70a36ef49efa7a98c504b5b1eca2417c8cbdbaf41ce7348fd9b401bda1b

Observation 7b7b6e72-cbaf-434c-8104-8b61b00877cb · outbound

This paper cites Proceedings of machine learning and systems , volume=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of machine learning and systems , volume=

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.677127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.677127Z digest=sha256:0ba360a677faa650eef4c0d02b2c4a9bb2a13cd9bd442528fa9fdbfe1677e7af

Observation 390f89a3-3abe-4bf9-92ed-d9f8115a6173 · outbound

This paper cites Proceedings of the 29th symposium on operating systems principles , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the 29th symposium on operating systems principles , pages=

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.679232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.679232Z digest=sha256:be51c81551c6537358e228775d4250b9a234d7a56a28f18e5866d4db466b7a37

Observation a99611c5-82a4-41ab-b4af-91a3bd64e73b · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.681151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.681151Z digest=sha256:5cecf9c761ce4f826707d4daa5a2f8d8d0a5e58815b1a67cc1b646d76d978369

Observation 3572b3f5-d612-404b-9db0-2e1c41d08a09 · outbound

This paper cites VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.683080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.683080Z digest=sha256:dfa571a95ae4ff15e0b0ea48b224f35e38cb76eb3c4ca99e2fbac3d63454717a

Observation c01a7ccf-c41a-4012-907d-1d97ffd1e34d · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Emu3: Next-Token Prediction is All You Need

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.685408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.685408Z digest=sha256:4004e2bec995f66165ec1ba21633e4cf51b98b2ef84bb9ace4324193e13c3b89

Observation 2ef15d77-cc04-4d18-80d0-a35ec502ee27 · outbound

This paper cites Liquid: Language Models are Scalable and Unified Multi-modal Generators.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Liquid: Language Models are Scalable and Unified Multi-modal Generators

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.687868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.687868Z digest=sha256:d12583ac3f18e7c96509c624b94104a044d808a0fecdcecda31287c48192fabd

Observation 5376db81-4679-4749-aa0f-5ca47ef27479 · outbound

This paper cites Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.690130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.690130Z digest=sha256:9dd933a361b178b0829a1167e9cee0aba597ee7333b9ea6ca195a92844babc64

Observation 9dec31d3-6546-4bb5-ba74-89a94bc7150c · outbound

This paper cites Advances in neural information processing systems , volume=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Advances in neural information processing systems , volume=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.692456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.692456Z digest=sha256:09edbe4919d145d0952016b9b5a6aa60a88f7430ec85e19e2f51e994c6d72ef9

Observation 09fc30c7-4ae0-49fe-a82f-ab221e803c7e · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.694503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.694503Z digest=sha256:f23c8e26d5cb5561dfe0713e8a6025ce54da512cf7cd89626152005936c5b0ea

Observation 4811978a-ff29-47f9-8f60-8cd71c995fc4 · outbound

This paper cites arXiv preprint arXiv:2502.20321 , year=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation arXiv preprint arXiv:2502.20321 , year=

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.696645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.696645Z digest=sha256:9a8a9e7122f58d203912bd00687c2317347f3f3c326937d679ea9fe995aa67e4

Observation b9238f48-ba38-4d02-8eed-10159ccc1c92 · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.699068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.699068Z digest=sha256:c1c4dda9c76230d7652ccc0e2b06b113abbd917b410dae8878d78635c3f5ade9

Observation 825332d0-8eb1-4b92-90d7-2ab865f844d6 · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.701000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.701000Z digest=sha256:95266aa9ad20be5af9d238d3921eb5bb89514216079512c144322e7f39d19493

Observation d79baf54-52ee-4497-8cb1-01b3e01de288 · outbound

This paper cites International conference on machine learning , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation International conference on machine learning , pages=

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.703243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.703243Z digest=sha256:d2ed4b2340711e2e4612562932431bb90c05bf7b0335bb61ea42021a2b2c1783

Observation 5c7c9a39-e79e-47da-9f4e-4c60089271f5 · outbound

This paper cites International conference on machine learning , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation International conference on machine learning , pages=

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.705509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.705509Z digest=sha256:41ef51c0fcb0de0d8683edd0c6c7c1bf51c4cb05acad189b898714d4a4def13a

Observation d574293c-b8ee-4008-a4cd-46ea9be80b7f · outbound

This paper cites International conference on machine learning , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation International conference on machine learning , pages=

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.707578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.707578Z digest=sha256:4f05da45f7cc93cee671a575345c1060b3877283cdfc5f699f2fcf763e69c6f5

Observation 4be5d177-d5cd-4d79-aa4d-e7343b38925d · outbound

This paper cites CoCa: Contrastive Captioners are Image-Text Foundation Models.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation CoCa: Contrastive Captioners are Image-Text Foundation Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.709753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.709753Z digest=sha256:d9e5e20a8d21115166c7395e3f91faac265a4d02b542bb844a63f690c20a1249

Observation cd52193b-2f7b-45d8-89dd-1614479348c3 · outbound

This paper cites FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.712121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.712121Z digest=sha256:df902fcd18e4dd0b4b17dbe55b75b0f67295b242fd6e38fdadb4094d473b1455

Observation 51904937-ec7c-4dab-b955-c80c59c6ae8a · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.714236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.714236Z digest=sha256:30804289947472c5cea3097b1778c70005179459ce849ca252077022e3200686

Observation eb1b247c-ea41-4e4e-897d-d37491af8b77 · outbound

This paper cites Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.716304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.716304Z digest=sha256:9b4b47d7cfeefc9f58f8bbcc77c2c8780c6ff381285b4c138670d6be078d7511

Observation 360da8aa-56a0-45c1-b85a-9c0d09b937dd · outbound

This paper cites arXiv preprint arXiv:2503.10772 , year=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation arXiv preprint arXiv:2503.10772 , year=

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.718770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.718770Z digest=sha256:936c3e42a1475acebfd1191bb0eb79b44c8eaf5317ea5e817435c7ce90ae9309

Observation b22f4a52-4331-4435-b897-eb0f9c897d3e · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.720886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.720886Z digest=sha256:0ce6a0cb978ddaee7306f4867c75c6d41110dbcaa4d1f76c40d36911e65c4535

Observation 2b968a28-79dc-4ea6-8983-bc194ac38dc9 · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.722926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.722926Z digest=sha256:1d7b56d5f9713d7ffb10efe943d4fae86ed499d9364819ff14ce583bd2151a2c

Observation a64a2dd6-6cd8-4ee7-b93b-1f0f9773476d · outbound

This paper cites Show-o2: Improved Native Unified Multimodal Models.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Show-o2: Improved Native Unified Multimodal Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.725080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.725080Z digest=sha256:a26f06ad44cdf4a2aeac3f4e679f0b50b0b2e9067b847ce25e9031a69224bbae

Observation e20ba4f8-be2c-49fd-a6e7-27aa64fc3fa0 · outbound

This paper cites MetaMorph: Multimodal Understanding and Generation via Instruction Tuning.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation MetaMorph: Multimodal Understanding and Generation via Instruction Tuning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.727225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.727225Z digest=sha256:6d76a0ee7b93a4727c37d5714be10daab482f4c7ddc5d1089b28b4e2dfeed48b

Observation 9891d8fc-a98f-462e-bb44-7c0e6fce057c · outbound

This paper cites Transfer between Modalities with MetaQueries.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Transfer between Modalities with MetaQueries

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.729444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.729444Z digest=sha256:90e6fb7e5cd9a2cef6e69c03a0ec74f621972dca05fb10b8dd13fa37811ac12f

Observation 62dca7ef-37bf-4491-a82c-44c6e7582eac · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.731512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.731512Z digest=sha256:0ebf2dc82ce568734de753f4235ca4653750d93f845a1cf60606c7c82680b765

Observation fbe590f5-d3f7-4fde-81cd-e79630b097cd · outbound

This paper cites ImageFolder: Autoregressive Image Generation with Folded Tokens.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.733716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.733716Z digest=sha256:dcd328d57128a159c0a0ac68c16dab8a8b060eaba1127146e3cd995e06459bb4

Observation e646cae8-fff4-420a-97ea-8897560acc96 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.735764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.735764Z digest=sha256:7bd99a2826b75aa09dfb07f4f7bac8cff896189522f77a758364c396fee46487

Observation b3895f10-3282-42ed-948d-ac737497dd2c · outbound

This paper cites arXiv e-prints , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation arXiv e-prints , pages=

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.738444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.738444Z digest=sha256:706de4569fdc69dba3e968b09ed70be4d2bdc59e280f3b8b272e514e3ea29217

Observation a8be3992-5757-46b5-b975-1cf416216006 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.740667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.740667Z digest=sha256:8dcad0571984307e6749c209352fdd430fbf6023594f125e449fd92e6a737867

Observation 47e8f753-f8c2-4325-9060-04229bb2a91f · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.742611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.742611Z digest=sha256:23f3d71a2d7301422f0995f8f756843e11fee34d2ab7f3ed7b8f0d2ca23a6bf5

Observation 63d19b94-950c-48c9-8380-44c9883d215d · outbound

This paper cites Proceedings of the IEEE international conference on computer vision , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the IEEE international conference on computer vision , pages=

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.744773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.744773Z digest=sha256:dbdb5af68c0b104243bb0db5fffcd9d3d0d63b67f9f9b01e3349a0e031663c63

Observation 3042cbe0-f884-49ab-b783-7ae951a04795 · outbound

This paper cites Proceedings of the IEEE conference on computer vision and pattern recognition , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the IEEE conference on computer vision and pattern recognition , pages=

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.746753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.746753Z digest=sha256:1e2ee7a12292c60c567a2132bb9b5bb290425c02b802e0d1524ba6aaa134365a

Observation 9a51c0b7-2270-4d22-a939-06e769077240 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Advances in Neural Information Processing Systems , volume=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.748873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.748873Z digest=sha256:338e7e00238abbed261e1de4ea7678a7862857bc8c2335d0db7ce9dc6cb5ecfb

Observation 90382785-b2a9-419e-af69-49a30189fd95 · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Evaluating Object Hallucination in Large Vision-Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.750827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.750827Z digest=sha256:cc560031eaed6ba9f6c742f04857ae4c5f51300388bc77b9bba2e627f9fcf4d1

Observation 2294b431-5d8f-4bc8-88d9-2718a466ad2b · outbound

This paper cites Advances in neural information processing systems , volume=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Advances in neural information processing systems , volume=

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.753111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.753111Z digest=sha256:094891ed29f8717ca44d6c85ac3a2309bcdcd1b153f2ba0c00720f6ffb394cb2

Observation 53cbad54-0881-452b-8e7a-872d784756ef · outbound

This paper cites Proceedings of the national academy of sciences , volume=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the national academy of sciences , volume=

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.755218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.755218Z digest=sha256:47829d0fb395e145075f01d8ec8147cfa3ac3e23692dc299dad9e9d0ea665525

Observation f1bf7fe7-9353-4b1f-94aa-08881026b2bb · outbound

This paper cites European Conference on Computer Vision , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation European Conference on Computer Vision , pages=

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.757589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.757589Z digest=sha256:db6760997faf16409886f20efca55cb1d0af3f93f646b3a84b95af22a0db9956

Observation d5a2408c-9b0f-4f75-870d-63bfe10e646d · outbound

This paper cites Proceedings of the 44th international ACM SIGIR conference on research and development in information retrieval , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the 44th international ACM SIGIR conference on research and development in information retrieval , pages=

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.759988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.759988Z digest=sha256:575c9d05b92331c52b49cad04032e8f26014847c8368e5c9ea083628eaec7828

Observation ee0a4277-ee5e-4f3e-ba16-6697ec990925 · outbound

This paper cites ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.761937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.761937Z digest=sha256:b04b1295389eb3082d9d07f1398a7745fa9b9b49e61b37e139391efddbf9b56e

Observation 6a86ad0d-b695-40fb-a84b-960544069b0f · outbound

This paper cites Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.764129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.764129Z digest=sha256:e83a7ee4ebada968b55fb7ef07babc9a0a3e38599699c0cde06ba5bb8321a4ea

Observation 4fe96f0f-d08f-42f0-a117-300e8b751b81 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Advances in Neural Information Processing Systems , volume=

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.766313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.766313Z digest=sha256:64056b537aa6a17c4056082612b5d2d4b1341f97e9c613e41b63ed8d2d4fbbf4

Observation ed34f350-5f97-4804-82a7-15386ea342bc · outbound

This paper cites Advances in neural information processing systems , volume=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Advances in neural information processing systems , volume=

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.768102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.768102Z digest=sha256:43c246bd13d9107083af8f11b5d51b6009c7b828e351854087d29c267de6ecb4

Observation bc232d89-644f-46fc-a3e2-96cb6a617efd · outbound

This paper cites BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.770060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.770060Z digest=sha256:a0c331b3f8f8604f5218df7ad59f01473eb1a0a20e7f9ddc73851ab8aa355d02

Observation e6e8cc14-ba48-4b0f-8e93-6b5832611c76 · outbound

This paper cites Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.772449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.772449Z digest=sha256:cd80aa5fe1e3b35678b88857e41ce360dc5a584ab4532d4f3237fa05b8ea6ef1

Observation 3b58d059-40fe-451f-935a-30723005f784 · outbound

This paper cites Proceedings of ACL , year =.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of ACL , year =

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.774589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.774589Z digest=sha256:664cf786c9f51e34dc1f61e0bf4b0b57d765db49706ad5367562d1778408e3a3

Observation 241528e2-3e44-46f6-b779-f58ec4a13e7e · outbound

This paper cites DiffusionDB: A Large-scale Prompt Gallery Dataset for Text-to-Image Generative Models.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation DiffusionDB: A Large-scale Prompt Gallery Dataset for Text-to-Image Generative Models

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.776624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.776624Z digest=sha256:a0f94104e8bce670c77e33471bb1483312c5dbe8faa6395025193d4f03123c02

Observation 278f40d2-61ce-4b21-8027-3d91c2c439cb · outbound

This paper cites Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.778827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.778827Z digest=sha256:6679b703dd8a45ceb6aa47100766994c996d787407ff4a3a9129afb5d3080b37

Observation bc81e8e2-9b13-4857-809e-b78e8571c12f · outbound

This paper cites 5 technical report , author=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation 5 technical report , author=

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.780943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.780943Z digest=sha256:0439de482f49016296b7d521aa23bc5dc0356c59086c3b03b86aa96f370349ac

Observation 3cd63f28-6d7a-480a-a040-80310a9154b8 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.782823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.782823Z digest=sha256:8882ccd362a1e2823b8639e2ea86e0b1eb15690e6dc0666dd162970eb6f341a9

Observation 8feca1b1-55b9-4f77-832a-070719b648a4 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.784756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.784756Z digest=sha256:f62b0f23e9f3c146130fbb49daeb332376c592dfd35488f3fd10f7c5cacacd75

Observation 203173a3-893b-4629-8c2f-147226f30496 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.786607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.786607Z digest=sha256:8afda2c8325d73df4fd0d9780ac4d3f329331b461eee753926215a1173234d42

Observation 231c1717-8a8d-452d-8840-1cedb5534a6d · outbound

This paper cites MMaDA: Multimodal Large Diffusion Language Models.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation MMaDA: Multimodal Large Diffusion Language Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.788866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.788866Z digest=sha256:2576b675176032741b27bb580c3315ea2550f952feea1230e030e44deafaa3a1

Observation e794fb98-3746-4095-a919-f93c121b53c6 · outbound

This paper cites Harmonizing Visual Representations for Unified Multimodal Understanding and Generation.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Harmonizing Visual Representations for Unified Multimodal Understanding and Generation

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.791253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.791253Z digest=sha256:2eaa7b1deb51da6d62bf7dc51800995836dea09fb99726bdb6e494956e030de9

Observation c57f59f7-d39f-4680-ac51-a8ba07ad02bf · outbound

This paper cites UniFork: Exploring Modality Alignment for Unified Multimodal Understanding and Generation.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation UniFork: Exploring Modality Alignment for Unified Multimodal Understanding and Generation

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.793608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.793608Z digest=sha256:4005b33ac4c694c170e85a1109581555f65205a4112901c7a739f318d54b51cf

Observation 0cecab19-7a31-466c-968c-1ccc28564717 · outbound

This paper cites SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.795993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.795993Z digest=sha256:e424d7789642e5d3a3436c330b750a0c90c24787f5058f2bccb45fbdf90f281a

Observation 6b2073f1-f657-4d1f-9533-62ddf20c180d · outbound

This paper cites ILLUME+: Illuminating Unified MLLM with Dual Visual Tokenization and Diffusion Refinement.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation ILLUME+: Illuminating Unified MLLM with Dual Visual Tokenization and Diffusion Refinement

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.798269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.798269Z digest=sha256:9482e9c0d6bde0420c23c7aafe15f6c45081d6e32a7e318e77eece620b3d3d0b

Observation 787f9f45-44db-44f2-84a9-f155bee2870d · outbound

This paper cites arXiv preprint arXiv:2309.04669 , year=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation arXiv preprint arXiv:2309.04669 , year=

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.800492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.800492Z digest=sha256:cb7745b5e55ed961cb553c9419e1e03379c4df6060b92cb640c2668513b9ed40

Observation 37aa285a-3d1e-409e-b0f2-380700a2578c · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Advances in Neural Information Processing Systems , volume=

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.802506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.802506Z digest=sha256:10b24bc4e0eb42d0437b64a7560108063da86196cdf012fb30a52005f395e93a

Observation 4a159646-b2b4-4ac0-ad78-8e1b35836874 · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.804623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.804623Z digest=sha256:c9fee3b2a7ef2d13e9f154a6f76c111dc061542221ac54a3808525430c5e28eb

Observation a19dabe8-8a04-4838-95b8-47275033b060 · outbound

This paper cites World Model on Million-Length Video And Language With Blockwise RingAttention.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation World Model on Million-Length Video And Language With Blockwise RingAttention

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.806696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.806696Z digest=sha256:c2df76d1c31c12093f8309e85ead373df1dff4fc7fd1173dd11985b0cbe703a0

Observation 4a1beea4-5392-4b8f-9918-fb76ac006b76 · outbound

This paper cites arXiv preprint arXiv:2503.06764 , year=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation arXiv preprint arXiv:2503.06764 , year=

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.808991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.808991Z digest=sha256:66141bc1dc8278b5b3ec94cc2868321fd56f722f0c2328d81eeeb1b1d65eaa36

Observation 4112da53-a8be-491f-b95c-e48894b0801c · outbound

This paper cites TokLIP: Marry Visual Tokens to CLIP for Multimodal Comprehension and Generation.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation TokLIP: Marry Visual Tokens to CLIP for Multimodal Comprehension and Generation

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.811094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.811094Z digest=sha256:919301de745ef6108e49f404ffff4f8001da1f059c67f9404d40010011ddd5f7

Observation 13f9d1c7-3bd9-4927-a71c-8d41b2b8c2b1 · outbound

This paper cites Proceedings of the 1st International Workshop on Efficient Multimedia Computing under Limited , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the 1st International Workshop on Efficient Multimedia Computing under Limited , pages=

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.813254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.813254Z digest=sha256:9edd4b7ab2514a973ef27e3caac2d3f2bea9da821e93b1fbd409668e9d773b7c

Observation 1fee9ea4-365d-4765-a35f-c423d3ef6543 · outbound

This paper cites MobileVLM V2: Faster and Stronger Baseline for Vision Language Model.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation MobileVLM V2: Faster and Stronger Baseline for Vision Language Model

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.815337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.815337Z digest=sha256:6b5f5b4734f1c62ef448b0df8feb964b4fd6da6058df00d0bc62a01e2bdab662

Observation 33549019-2312-4c34-9195-99dedf88e1c9 · outbound

This paper cites MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.818128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.818128Z digest=sha256:6e69b83aa6206d2924aa90de431dcb32e6db18939145e0adaeb33f2fe5643dcd

Observation 0b00afa8-409c-4857-87be-53d685216ede · outbound

This paper cites Advances in neural information processing systems , volume=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Advances in neural information processing systems , volume=

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.820421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.820421Z digest=sha256:bd36cabfffd4254517b8c393dbbf003bf37db891ece06120720a379f56654106

Observation c99377bf-fb3e-474a-95bd-ad2a3ed90afc · outbound

This paper cites URL https://huggingface.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation URL https://huggingface

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.822478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.822478Z digest=sha256:a3364c56e1e75b8b4769aef352e79dd62001588ddf26c2c03cf7f88338c8ac30

Observation 039e9fa4-aebc-4199-beee-ffa5abe1efba · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.824594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.824594Z digest=sha256:09499631efd9d9d9c8319f4a8960979e4d8acc2ae7d94540cfed1b6bbb74b3b4

Observation 71086539-402b-466c-92cf-78b7f6da92d6 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.826810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.826810Z digest=sha256:79d32b17ae9ed7bc71310663a5cc557893bc08715082069e1fb96e5391018d60

Observation 76c74b5a-c4b8-46c7-926c-96013cb82fe5 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.828909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.828909Z digest=sha256:3337d2b6dbf2caecccd305b1d421a32876a26267ec97ab2be34689859a5b9777

Observation fa765995-4e6a-4e7e-a51d-9a80d5c5647f · outbound

This paper cites PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.831191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.831191Z digest=sha256:bccbfa0c43f7d36262d831d7a769c598e70e54e38b76ccce8d886d8cb3150b81

Observation a9489052-6c21-45e5-bd2a-197da020d181 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.833278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.833278Z digest=sha256:a4ababf9b818299f401a52a904cd14c283ff3a5336675cfc67019fc228dfbddf

Observation fe323864-5202-45f9-b029-48da66092156 · outbound

This paper cites Computer Science.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Computer Science

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.835632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.835632Z digest=sha256:08ab4952e68d9f6685926306caee7d9db18779c9b2743966813fdf0ebb03288d

Observation 9a95067b-f562-4ca4-b9fa-91b0e3db2062 · outbound

This paper cites Forty-first international conference on machine learning , year=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Forty-first international conference on machine learning , year=

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.837682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.837682Z digest=sha256:e307a997b086e35e2447ab699aaafca4103a6e23bc766df27d9c9f965854c93e

Observation e7f67f8a-bf49-47ce-9df6-2ebe5a186c8b · outbound

This paper cites Classifier-Free Diffusion Guidance.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Classifier-Free Diffusion Guidance

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.839746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.839746Z digest=sha256:3bcf365b42c3499038d1713c50fbfa02b7f485a33d16bc27279a4758539e1efb

Observation 8e271c92-4cec-4a90-8e63-fdf8bca508e0 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.841977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.841977Z digest=sha256:676b3aea81ce463d69cc34fe5180a0d3528c4a45103cc99f16b0d6384f39a184

Observation 8e949925-4cef-40f3-bf3f-8217bdd7cbdd · outbound

This paper cites Proceedings of the IEEE/CVF international conference on computer vision , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the IEEE/CVF international conference on computer vision , pages=

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.844769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.844769Z digest=sha256:cd13c7f47abbffc31ff771b96eef76693493821ba38dacee5224d3afd6942f07

Observation 4a9202ba-68e4-49f2-a0eb-4557b954c661 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.847148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.847148Z digest=sha256:a23f54ea2a8eea02ab080c973b45813a3dcd91997e12e7c350c561422058786b

Observation d8d3bd69-52a7-444e-ba80-c18ca5de2e0b · outbound

This paper cites Forty-first International Conference on Machine Learning , year=.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Forty-first International Conference on Machine Learning , year=

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.849151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.849151Z digest=sha256:4b50bd71b7a65b08bfb8c4776be3d113f730d00046debddeda043ead94413d73

Observation f4daf8f9-2b65-4284-95f0-a1d6d5f0226c · outbound

This paper cites 2024 , month =.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation 2024 , month =

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.851223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.851223Z digest=sha256:f7e807888a0379b33351a6d773e47abcd5f2c524984b7cb8e0ef2eabeeaa7c4f

Observation 132aaca5-1c1c-41aa-86f7-28e0cdfde46d · outbound

This paper cites Flow Matching for Generative Modeling.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Flow Matching for Generative Modeling

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.853239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.853239Z digest=sha256:082660fa83fa163fdb2b7dcbf4e17942fdfb86cdea5e95a39f27f2d4f5218443

Pith citing papers

No inbound Pith citation observations are available.