Pith. sign in

Paper Citation Record · LEDGER

Emu: Generative Pretraining in Multimodality

As of 25 July 2026, this Paper Citation Record lists 22 of 22 outbound references and 31 inbound Pith citation observations for arXiv:2307.05222.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.05222 v2

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-16T20:22:11.395162Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-07-24T06:31:00.690269+00:00

measured 31 of 31 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T03:31:19.309532Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-04T10:09:44.721463Z

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy20
  • unresolved0
  • parse uncertain1
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3bac87b3-5a9b-4ec8-9d6d-1fee103dcddc · outbound

This paper cites {question}.

Emu: Generative Pretraining in Multimodality {question}

Reference 1

Resolution
malformed identifier
raw_fallback, observed 2026-05-16T20:22:11.411505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:25309f368d05066ed65c8577bee252924545842d8999cef7f51f257050850810

Observation eaa0e76b-78f9-4eac-88b3-e9f495b71797 · outbound

This paper cites Make sure to check the weather forecast before your visit and pack appropriate clothing and gear.

Emu: Generative Pretraining in Multimodality Make sure to check the weather forecast before your visit and pack appropriate clothing and gear

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.414586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:0a3e2c0f197d5fe451598023ab53e060d684f2924cd1d0760d2fb562c87e1403

Observation 940fe0bf-a2e5-4e09-8b48-eb5004057409 · outbound

This paper cites Make sure to stay on designated trails and keep your distance from any wildlife you encounter.

Emu: Generative Pretraining in Multimodality Make sure to stay on designated trails and keep your distance from any wildlife you encounter

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.417392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:d4b166ca6166dbc7df1bcf46d54f38ebc98c9cccf64668ed0d02a95487c9b0c8

Observation a2db8400-24ac-4246-afa8-9ec047f00bee · outbound

This paper cites Make sure to check with local authorities before swimming or boating in the lake to ensure it is safe to do so.

Emu: Generative Pretraining in Multimodality Make sure to check with local authorities before swimming or boating in the lake to ensure it is safe to do so

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.420323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:7a28267c3dba0082a847b6bf593db16d9aefc49430b25ae43fe11ac3386468e3

Observation 3b6c6c0a-0424-4f7d-8aad-ef732049151f · outbound

This paper cites Make sure to familiarize yourself with the lake's layout and any potential hazards before venturing out on the water.

Emu: Generative Pretraining in Multimodality Make sure to familiarize yourself with the lake's layout and any potential hazards before venturing out on the water

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.422995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:3523413de255d43b24148c8795224b989bbe8e2e1c2ff1a09f4ad2f0ae8af1b1

Observation 6bee841b-9a46-4747-b37e-e33ec3f14a42 · outbound

This paper cites By taking these precautions, you can ensure a safe and enjoyable visit to the lake.

Emu: Generative Pretraining in Multimodality By taking these precautions, you can ensure a safe and enjoyable visit to the lake

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.425595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:e4dddc1d98370191fc4b8997131076b3fa4f3f7f42179f373e7badb0b425bc67

Observation 37b4e904-0359-43e1-9465-ae5eabae8e71 · outbound

This paper cites an unresolved cited work.

Emu: Generative Pretraining in Multimodality Unresolved cited work

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.427964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:53497253723d3d7547003aecf47dbc974281d6f76d0540768ca9f4623107171e

Observation 6115fa07-497c-4697-99be-2a317ce8448a · outbound

This paper cites Make sure to include items such as bandages, antiseptic wipes, and pain relievers.

Emu: Generative Pretraining in Multimodality Make sure to include items such as bandages, antiseptic wipes, and pain relievers

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.430380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:f9c0271253c4047aa05976c6ddcf97211f6346cbea3f179103ccc5636e42caec

Observation 30fd068e-cea1-42f9-92fc-0672b5b35fb1 · outbound

This paper cites an unresolved cited work.

Emu: Generative Pretraining in Multimodality Unresolved cited work

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.432641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:e5708266603aff146078af4d8ac356c0255bd0adf1c6f4d08ac5382d55d53f8b

Observation 037d8bb0-9ac2-4e7d-8fa9-c2acef75d5c8 · outbound

This paper cites Impression, Sunrise.

Emu: Generative Pretraining in Multimodality Impression, Sunrise

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.435019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:8879c80bebabdb336ae882796fff4709d6650ff7c6ca6a3e72289e242a6dd7a7

Observation a0687970-3070-46df-b764-9815c5070560 · outbound

This paper cites The Mysterious Affair at Styles.

Emu: Generative Pretraining in Multimodality The Mysterious Affair at Styles

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.437324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:7d419d6b70d78b217c454474bd08fba443e6cc250ca156969a634339d26ed370

Observation 4c2b2c19-caa4-4f4e-aaae-22ca6136cfec · outbound

This paper cites The Secret Adversary.

Emu: Generative Pretraining in Multimodality The Secret Adversary

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.439602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:47f43c646c577e736974eaf2c5de415acbd8303d5f198ccba817d74464f97eba

Observation 01a46233-7974-43ac-ae25-a5966d7df925 · outbound

This paper cites The Murder on the Links.

Emu: Generative Pretraining in Multimodality The Murder on the Links

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.441742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:6772aa8157f7a585d6826a6922a7c7837bed79fc784a652cfb350a9f3239795e

Observation ba8b2382-1921-478e-a0be-803c800be50c · outbound

This paper cites The Man in the Brown Suit.

Emu: Generative Pretraining in Multimodality The Man in the Brown Suit

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.443970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:beeba98524e254a4c64589e956391f98ec41377b511a9125eeb168fdab243b8a

Observation 4140cbb6-d135-4f23-967e-ea335de27b75 · outbound

This paper cites The Secret of Chimneys.

Emu: Generative Pretraining in Multimodality The Secret of Chimneys

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.446380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:0ba4a374f2b20920411f070ec337188b64a756b7539b5bb697fb452f45789fa1

Observation 7c5891d9-78c2-4c24-b914-51ded3e293eb · outbound

This paper cites The Murder of Roger Ackroyd.

Emu: Generative Pretraining in Multimodality The Murder of Roger Ackroyd

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.448535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:c342ac7f3c8f9c33ebace70ba9069cc5c98edc245f920924957ec45f1ec6c3d7

Observation f346b73b-e48c-4a71-b2bf-27ee41c5572a · outbound

This paper cites The Big Four.

Emu: Generative Pretraining in Multimodality The Big Four

Reference 17

Resolution
parse uncertain
raw_fallback, observed 2026-05-16T20:22:11.450657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:62e63029787b8284dbd21b33bbb2561c7cef94eaaaee888873a8c3b48e727caf

Observation d76b426a-4cbc-4c1a-8263-1f9f3726ceb2 · outbound

This paper cites The Murder at the Vicarage.

Emu: Generative Pretraining in Multimodality The Murder at the Vicarage

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.452739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:69701db8b1b9c9c5e3a03ca8a50b26cf0d393c28715333fd2067e7d8e32a946b

Observation e554165e-f686-4b75-a1df-a6d1caae0a98 · outbound

This paper cites an unresolved cited work.

Emu: Generative Pretraining in Multimodality Unresolved cited work

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.454802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:d6f6fdb064ad9b0fbf05670538f0c7b50e91c5d410955aa11be437b79cf1329b

Observation dd15778c-93c9-4f48-9969-3738a75419f2 · outbound

This paper cites an unresolved cited work.

Emu: Generative Pretraining in Multimodality Unresolved cited work

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.456721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:923c3aa139381e978828124fa2f13c8650ab67eec641b928cebf80f461c68101

Observation 3872f2f3-1312-465f-a006-92393f3503d4 · outbound

This paper cites an unresolved cited work.

Emu: Generative Pretraining in Multimodality Unresolved cited work

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.458916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:9312f205b667a4ee8f4fe23b042f3b0d7f368611a70c5f6d3a2f5a31839680c0

Observation 825beea7-d9ba-437b-b0b7-e1ec9c592bf7 · outbound

This paper cites A play that was adapted for a movie and later became a TV mini-series.

Emu: Generative Pretraining in Multimodality A play that was adapted for a movie and later became a TV mini-series

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T20:22:11.461390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T20:22:11.395162Z digest=sha256:20f9c065975b402699e3bc90fe53fa1a3e87c67a778af4f4b0ab4c73a38c5baf

Pith citing papers

Observation e286c21f-c588-4593-b40d-6307d6af5bc6 · inbound

A Survey on Multimodal Large Language Models cites this paper.

A Survey on Multimodal Large Language Models Emu: Generative Pretraining in Multimodality

Reference 175

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:22:11.462164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T02:56:41.658658Z digest=sha256:c57dd87d7d523e2402f11929d8db43ce48a0e2c50cef654cf30a4a5894e08695

Observation dea1c44d-70f9-48e1-9744-6e55e2ab21ef · inbound

SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension cites this paper.

SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension Emu: Generative Pretraining in Multimodality

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:22:11.462164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-12T16:59:50.495335Z digest=sha256:d0556a9f451a89bb994b41454d6484f80ac2757b63f39b5b7b3af9b43583cc97

Observation d1b73c67-50be-45a3-9491-94e911eddeed · inbound

Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models cites this paper.

Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models Emu: Generative Pretraining in Multimodality

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:22:11.462164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-12T18:57:28.666194Z digest=sha256:d4011dca1595c34a0d4eadccbf5419f513a229b539054c69d1a0eb4e6f859f75

Observation d95fb4b5-7215-4c16-b11b-75ea0c07de63 · inbound

InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks cites this paper.

InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks Emu: Generative Pretraining in Multimodality

Reference 132

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:22:11.462164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-13T22:46:09.693156Z digest=sha256:44afc441f1113b659dcd36d846575f402052b547953509517667e08782c87b68

Observation de3df9a3-5e22-4eb5-a759-a2d67358f386 · inbound

NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation cites this paper.

NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation Emu: Generative Pretraining in Multimodality

Reference 93

Resolution
verified exact
local_arxiv, observed 2026-05-18T04:55:20.521578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-18T04:55:20.362512Z digest=sha256:abb7e5de0d112218158c2c132e71dadf62c0763507f1bdba90bf6acb73875de6

Observation 0fd7862f-db26-45ee-bef4-ba405e8211a5 · inbound

DeepSeek-VL: Towards Real-World Vision-Language Understanding cites this paper.

DeepSeek-VL: Towards Real-World Vision-Language Understanding Emu: Generative Pretraining in Multimodality

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:22:11.462164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-11T17:58:54.177359Z digest=sha256:d2db8259508ad4f2f18ae40845141c33879cc4a068b95a2471258d2d233c949c

Observation 0f60332d-6007-4aa8-bf55-6e8e3f17346b · inbound

Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models cites this paper.

Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models Emu: Generative Pretraining in Multimodality

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-05-17T07:44:47.439143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-17T07:44:47.355960Z digest=sha256:5cffbc61d8a770465d9f18471363069fa8eaa89e67434da85e095d3c9f139f42

Observation 74006cb6-215a-437c-9783-f0cca1e050b1 · inbound

SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation cites this paper.

SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation Emu: Generative Pretraining in Multimodality

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:22:11.462164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-15T22:48:36.010306Z digest=sha256:f4fdf7d48c0b7cb37c7b15ab0e36b994f0c963be253dfd6be4e861c6ca9f1211

Observation 16261b77-81e4-4ecc-9dfe-7a50450556a2 · inbound

Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation cites this paper.

Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation Emu: Generative Pretraining in Multimodality

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:22:11.462164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-11T22:09:16.622717Z digest=sha256:0474613fb7e89848dd13de02b1191442e5c548e787fcfaaebdb0bf45ae7bd6fb

Observation 7320952a-a2b1-4fec-9046-2177212f5d5e · inbound

Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation cites this paper.

Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation Emu: Generative Pretraining in Multimodality

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:22:11.462164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-15T22:09:16.001309Z digest=sha256:11f68007bb162e130c40a9ff0fabd167b03ec29d51cb23b1e7f6651f29d9a2a9

Observation 59904974-7e3f-44f0-8cea-fb0fec9f2a6e · inbound

Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling cites this paper.

Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling Emu: Generative Pretraining in Multimodality

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:22:11.462164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-11T08:14:52.890145Z digest=sha256:75d807d0dcca338ad0182298594691329a58f53c5b880f6ed93289d19cb64ffc

Observation 9e530a6d-6202-4f97-988c-26e35bf5ccc0 · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey Emu: Generative Pretraining in Multimodality

Reference 203

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:22:11.462164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:feffbd7bc3286f6667305766d3e62803f31ec62ffa980aaeaabe853ca5880f9d

Observation 42a084c4-5234-420d-a97d-71ba78dd32ea · inbound

Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation cites this paper.

Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation Emu: Generative Pretraining in Multimodality

Reference 73

Resolution
verified exact
local_arxiv, observed 2026-05-17T07:24:04.634180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-17T07:24:04.460276Z digest=sha256:9529b207d46c543fbc6ae68c9b9b26bc1ad51e027e2588767cc2e5f4cc6c41ba

Observation 4303dca9-f30a-408c-a5cc-4df08b3fda1f · inbound

Emerging Properties in Unified Multimodal Pretraining cites this paper.

Emerging Properties in Unified Multimodal Pretraining Emu: Generative Pretraining in Multimodality

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:22:11.462164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-10T16:23:41.854132Z digest=sha256:0f43dad8c6ace11ad63d7607f5f552437d743db610dfab5ceeca8c226342d026

Observation 24ae7dfa-e2f3-4a95-857e-278095069e6f · inbound

MMaDA: Multimodal Large Diffusion Language Models cites this paper.

MMaDA: Multimodal Large Diffusion Language Models Emu: Generative Pretraining in Multimodality

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:22:11.462164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-15T14:50:59.661153Z digest=sha256:8abd103053017ffc90c0bcebaad4a3b1d864f95016b5bc42d6e20aff8caf8562

Observation 6ffdbd29-d8de-4838-8190-daf5889c4a80 · inbound

Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model cites this paper.

Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model Emu: Generative Pretraining in Multimodality

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-19T13:02:18.320357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-19T12:59:31.454155Z digest=sha256:e7d5cc3cf367894d907f18a91ec14da7883614391f59d45cb9a918ca763ae7a5

Observation 654a4f5b-817b-4a81-9bb5-383d6de8755d · inbound

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems cites this paper.

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems Emu: Generative Pretraining in Multimodality

Reference 165

Resolution
metadata mismatch
local_arxiv, observed 2026-05-19T11:52:16.293180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-19T11:49:36.574471Z digest=sha256:123bfd98df45d683ba35baf13a967e9b53f5d1b16eade26703a3764a93d6bff3

Observation f87ced91-b975-4f81-b9f5-ff6c44103be5 · inbound

Beyond Patches: Global-aware Autoregressive Model for Multimodal Few-Shot Font Generation cites this paper.

Beyond Patches: Global-aware Autoregressive Model for Multimodal Few-Shot Font Generation Emu: Generative Pretraining in Multimodality

Reference 53

Resolution
verified exact
local_arxiv, observed 2026-05-21T17:00:24.187048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-21T16:57:09.843691Z digest=sha256:7b6acd8a53f64ee30ef2fe05e7f3c6d562ca72e437499e6793f9b3aadee5b988

Observation c6b5cc46-2869-4679-a03a-b10f9ad36b3d · inbound

Mind the Gap No More: Achieving Zero-Gap Multimodal Integration via One Tokenizer cites this paper.

Mind the Gap No More: Achieving Zero-Gap Multimodal Integration via One Tokenizer Emu: Generative Pretraining in Multimodality

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:22:11.462164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T12:48:53.670343Z digest=sha256:d5fd4b1861d7b3e7b0ff377dab70c02c01c4393c6f4c19033ddbbb14e7df0f7f

Observation 4f99a9fe-14e3-459c-bd7e-33a2a93876ea · inbound

UniICL: Systematizing Unified Multimodal In-context Learning through a Capability-Oriented Taxonomy cites this paper.

UniICL: Systematizing Unified Multimodal In-context Learning through a Capability-Oriented Taxonomy Emu: Generative Pretraining in Multimodality

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-13T18:42:33.178622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T18:42:33.178622Z digest=sha256:69c14f4fbfc39486945cd76a84299f30ebf4d15f6631e84b16bd91a61d8a1b40

Observation 84da22eb-0467-4b51-be24-57b6f21bc375 · inbound

Symbiotic-MoE: Unlocking the Synergy between Generation and Understanding cites this paper.

Symbiotic-MoE: Unlocking the Synergy between Generation and Understanding Emu: Generative Pretraining in Multimodality

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T20:22:11.462164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-10T18:24:17.872120Z digest=sha256:115f845dd49b38e6280de8b6f15776724e48383f15d6766ee7105d8d86b4936a

Observation 89a03a03-ce3b-46b1-9057-c09691efb59e · inbound

Memory-Efficient Transfer Learning with Fading Side Networks via Masked Dual Path Distillation cites this paper.

Memory-Efficient Transfer Learning with Fading Side Networks via Masked Dual Path Distillation Emu: Generative Pretraining in Multimodality

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:22:11.462164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-10T16:45:36.306400Z digest=sha256:343fa52e797f405e97f659f000377ee465706d6921126adaef42cb88985fad44

Observation a19f30e2-243e-4626-b52b-1a40a891422f · inbound

CheXmix: Unified Generative Pretraining for Vision Language Models in Medical Imaging cites this paper.

CheXmix: Unified Generative Pretraining for Vision Language Models in Medical Imaging Emu: Generative Pretraining in Multimodality

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:22:11.462164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-08T12:22:11.995334Z digest=sha256:7e703ecb2d080face2594d64689b7db4971255e307a0581f9220b61d5a4c4665

Observation 46d9b798-54c5-499a-97d3-fde29ed17993 · inbound

MP-ISMoE: Mixed-Precision Interactive Side Mixture-of-Experts for Efficient Transfer Learning cites this paper.

MP-ISMoE: Mixed-Precision Interactive Side Mixture-of-Experts for Efficient Transfer Learning Emu: Generative Pretraining in Multimodality

Reference 61

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T20:22:11.462164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=arxiv_source observed=2026-05-10T17:19:59.247074Z digest=sha256:aaade6c72b96cc0c9fed925e8b56429162491b9b99af219d629a000dbd332320

Observation b4443323-7f65-4b0c-85f9-41d026a1fb23 · inbound

Bernini: Latent Semantic Planning for Video Diffusion cites this paper.

Bernini: Latent Semantic Planning for Video Diffusion Emu: Generative Pretraining in Multimodality

Reference 65

Resolution
verified exact
local_arxiv, observed 2026-05-22T06:41:10.580946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-22T06:39:47.124605Z digest=sha256:94e901a48f25e9955cc15a5833b14207df8f339410198c64a0489da9d662134e

Observation bbf20112-1250-498e-87ed-c77f699baee4 · inbound

Polaris: Scaling Up Instruction-Guided Image Generation Towards Millions of Personalized Style Needs cites this paper.

Polaris: Scaling Up Instruction-Guided Image Generation Towards Millions of Personalized Style Needs Emu: Generative Pretraining in Multimodality

Reference 101

Resolution
verified exact
local_arxiv, observed 2026-07-01T22:46:18.697488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=arxiv_source observed=2026-06-28T15:07:03.439928Z digest=sha256:83d785dd84b4208d8d02e6bf68ce06c6e627ffe34b8065f5ba899366ba6eb96f

Observation 482b012e-291c-4814-be94-c02e4f2ca184 · inbound

Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation cites this paper.

Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation Emu: Generative Pretraining in Multimodality

Reference 111

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T17:18:43.980987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=arxiv_source observed=2026-06-27T04:19:26.332718Z digest=sha256:6c489f42783b3e67087f86fd58ec3f59ca74f92e8157f07be218a473128fc978

Observation eb768116-a103-4349-a1fc-fe6a24ddc000 · inbound

SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models cites this paper.

SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models Emu: Generative Pretraining in Multimodality

Reference 42

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T10:09:44.722485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-06-26T09:08:25.661515Z digest=sha256:ce842688268fe65d9788e9100f3f1ef7ebe4f0d074d226e2f1485dfd6f08ff49

Observation bd59a6f5-5a7b-4056-bb49-4ec979367383 · inbound

SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models cites this paper.

SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models Emu: Generative Pretraining in Multimodality

Reference 45

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T23:19:02.577001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-07-03T23:15:09.253879Z digest=sha256:b6dac1abda4d6459d7363c3b8d77ac0442283d50fd3e94f1856d7820022e4c7c

Observation 56cfd251-0366-4b6c-b83d-f2dc6f088f2b · inbound

MentalThink: Shaping Thoughts in Mental SVG World cites this paper.

MentalThink: Shaping Thoughts in Mental SVG World Emu: Generative Pretraining in Multimodality

Reference 105

Resolution
unresolved
no resolver link, observed 2026-07-12T01:50:59.184754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T01:50:59.184754Z digest=sha256:730d19d279778711a8db1b32769cd7e92e95b7bd7ba5f32c3e50f4b236426e48

Observation d91a875a-13db-47d5-a550-bfce8cab3c7f · inbound

Qwen-Audio-VAE Technical Report cites this paper.

Qwen-Audio-VAE Technical Report Emu: Generative Pretraining in Multimodality

Reference 126

Resolution
unresolved
no resolver link, observed 2026-07-14T03:31:19.309532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T03:31:19.309532Z digest=sha256:a25e623c2f2280983b8f8433a1908406e4065f15d48490d8f84c8727740a1252