Pith. sign in

Paper Citation Record · LEDGER

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs

As of 18 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 0 inbound Pith citation observations for arXiv:2607.17140.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.17140 v1

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T18:57:00.056228Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

51 of 51 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved51
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c0b86eb7-43f1-4fe7-b6c4-54b9b751d3e1 · outbound

This paper cites Ofa: Unifying architectures, tasks, and modalities through a simple sequence- to-sequence learning framework.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Ofa: Unifying architectures, tasks, and modalities through a simple sequence- to-sequence learning framework

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:56.754103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:56.754103Z digest=sha256:e1b2039cbaa599fcd80db0054f11e09ae9f4a0d73f5ee2456cfc71137a546e44

Observation 0e640c8d-0cf4-4bab-b8ac-d8b3d8042df1 · outbound

This paper cites Unified-io 2: Scaling autoregressive multimodal models with vision language audio and action.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Unified-io 2: Scaling autoregressive multimodal models with vision language audio and action

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:56.820114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:56.820114Z digest=sha256:adfcfa2ce82bdcd0adf73c7718f9f3c1bbade15c94e435020905678a27d885f3

Observation 9e8a90fe-cb3f-4083-afd7-176badcdeab0 · outbound

This paper cites Dreamllm: Synergistic multimodal comprehension and creation.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Dreamllm: Synergistic multimodal comprehension and creation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:56.883157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:56.883157Z digest=sha256:774f8826a23062206863890e69bdf678d98d5d5670b366bfc9656bb069511e4a

Observation c7cebcb5-1f8c-4e0a-809e-d6295edfff26 · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:56.938205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:56.938205Z digest=sha256:950785573bc778e1720ad227f6e1d6248224ab2c6833bfcd08bd4115bc84a8b7

Observation b3932ea1-c287-4589-968d-651343eb9b9c · outbound

This paper cites Show-o: One single transformer to unify multimodal understanding and generation.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Show-o: One single transformer to unify multimodal understanding and generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:56.993216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:56.993216Z digest=sha256:214c9653323b4ec0874f8e649b2dd43a829c4c60cb8c4a92eab72cdca2488e85

Observation e0242232-435b-4e31-bc4d-8e02974cf3fc · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Emu3: Next-Token Prediction is All You Need

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:57.047720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:57.047720Z digest=sha256:d49ff75f27970e724e11a81633f30c8daa4fa146683f426f66d241088b9bade2

Observation 082416f4-7371-4f61-ba75-fb98c2d2345d · outbound

This paper cites Janus: Decoupling visual encoding for unified multimodal understanding and generation.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Janus: Decoupling visual encoding for unified multimodal understanding and generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:57.156094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:57.156094Z digest=sha256:e81067540441950220376d16c001096fea42219cd1ffc9e6d4c654d9224234c2

Observation c4bc1a20-fea9-4de5-96c4-6c9985ed8245 · outbound

This paper cites Emerging Properties in Unified Multimodal Pretraining.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Emerging Properties in Unified Multimodal Pretraining

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:57.226990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:57.226990Z digest=sha256:e3c36c941da7706d3ed15d48e1e7764e7d994e63be359d1ef78170fa308316cf

Observation 1bb65791-50ff-4de3-bf65-7b61a260f14b · outbound

This paper cites SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:57.285851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:57.285851Z digest=sha256:d88fc0001d6641a8e1aa6e2f72132202b295d8f41763b39e3f75ff0619493d29

Observation 48aeea2a-6119-460e-846d-af6ef78e4955 · outbound

This paper cites Vila-u: a unified foundation model integrating visual understanding and generation.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Vila-u: a unified foundation model integrating visual understanding and generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:57.336956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:57.336956Z digest=sha256:b681a56678b219ac5073e35fa1c446afdb38558dc2193fb0c84e6d23e5c683d8

Observation 2796d1c4-7cce-4183-9e50-70fbbd0b02ad · outbound

This paper cites BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:57.402154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:57.402154Z digest=sha256:3015ecc7098b0b62a46127aaa0216f40edb7da0cac44c2ed2f31cd9364b0b270

Observation fe7b03a8-ce7d-440f-a555-f56e96957883 · outbound

This paper cites Internvl-u: Democratizing unified multimodal models for understanding, reasoning, generation and editing.arXiv preprint arXiv:2603.09877, 2026.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Internvl-u: Democratizing unified multimodal models for understanding, reasoning, generation and editing.arXiv preprint arXiv:2603.09877, 2026

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:57.457053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:57.457053Z digest=sha256:43e53c9923ffebea4c63d03f09f9a2e5aff692ebea5596c25a2d920542825f35

Observation c31cc9a7-9686-428e-853f-444bccabf723 · outbound

This paper cites Tuna: Taming unified visual representations for native unified multimodal models.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Tuna: Taming unified visual representations for native unified multimodal models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:57.513914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:57.513914Z digest=sha256:c6615c52f8911e5b936020405d5a80a4377114a1b7fecd61c2edaf8c74298087

Observation e72af9cb-8b76-4ba0-9cf0-57fbbfa3ebad · outbound

This paper cites Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:57.561814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:57.561814Z digest=sha256:e2b37171112cfa1693566f8a8086cf0be5087d8dcf9f4aa672a87abf7c909614

Observation 37a337e9-960c-47a3-aee4-a1388418953c · outbound

This paper cites SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:57.619206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:57.619206Z digest=sha256:7731c09a52ff51f4915db8f370db72b9c90047d0ec47a06add0d6315ced06bcc

Observation dc688b7a-b580-48ab-ad05-fdd95c9a605a · outbound

This paper cites Unified language-vision pretraining in llm with dynamic discrete visual tokenization.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Unified language-vision pretraining in llm with dynamic discrete visual tokenization

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:57.676565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:57.676565Z digest=sha256:50a976bda0ce0fda8b40b49f6226d068d1a5d9b0623c28669aa53f4bbd4b109d

Observation 88709946-dcf2-487f-870a-4537ffaccfc5 · outbound

This paper cites VL-GPT: A Generative Pre-trained Transformer for Vision and Language Understanding and Generation.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs VL-GPT: A Generative Pre-trained Transformer for Vision and Language Understanding and Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:57.735425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:57.735425Z digest=sha256:515a5338a7a3589e8b2893511e83708c38209ffe4b95846b1f44afc80e232da2

Observation 7636bf54-aed9-4c91-a203-bbafe8928509 · outbound

This paper cites Generative multimodal models are in-context learners.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Generative multimodal models are in-context learners

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:57.795533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:57.795533Z digest=sha256:a92daad8995e26d4d3ea104c13e2b2f98af9c6d8cc068fa734fce770f44f7d73

Observation 557ef632-529b-4302-b6ce-cca98222538b · outbound

This paper cites Show-o2: Improved native unified multimodal models.Advances in Neural Information Processing Systems, 38:47490–47518, 2026.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Show-o2: Improved native unified multimodal models.Advances in Neural Information Processing Systems, 38:47490–47518, 2026

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:57.849553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:57.849553Z digest=sha256:aa64ed164d770a8fc1f3c5c0bfd6861d4b5ccc6ee87e27025756667f6b2568f5

Observation 916c2745-7f74-4203-bdd6-497728d03727 · outbound

This paper cites Transfusion: Predict the next token and diffuse images with one multi-modal model.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Transfusion: Predict the next token and diffuse images with one multi-modal model

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:57.909523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:57.909523Z digest=sha256:22d3659e82c7002e8357e960b5d8a8a57c35f688295fefa6928d68829bfce3f4

Observation 8773cc45-8c0b-47ad-92ef-6872fbead13e · outbound

This paper cites Metamorph: Multimodal understanding and generation via instruction tuning.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Metamorph: Multimodal understanding and generation via instruction tuning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:57.966409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:57.966409Z digest=sha256:e637ef76b134e2205e60b65b1a5bcf18f75ccdd85737b086be22f5ab072bc5e8

Observation ac73d63b-6f13-43c0-935a-3b48a7d390b4 · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:58.022089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:58.022089Z digest=sha256:4d3114502e7dc791bde1ead8878f3599d756657aee0b3579501114067cb5bc03

Observation df3d68b0-15d6-4697-8bbe-9d943f510673 · outbound

This paper cites Onecat: Decoder-only auto-regressive model for unified un- derstanding and generation.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Onecat: Decoder-only auto-regressive model for unified un- derstanding and generation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:58.067777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:58.067777Z digest=sha256:b2f905e4f4723db79ffe59ab5ee6f66233c9596ef3db1cfc86801b70de811863

Observation 6f643404-2e2c-4b8c-9a15-12044ab1837d · outbound

This paper cites Unigen: Enhanced training & test-time strategies for unified multimodal understanding and generation.Advances in Neural Information Processing Systems, 38:152386–152415, 2026.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Unigen: Enhanced training & test-time strategies for unified multimodal understanding and generation.Advances in Neural Information Processing Systems, 38:152386–152415, 2026

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:58.129664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:58.129664Z digest=sha256:2f7f7ee640aa49d4d0b85f5c22560c0c3f252e3ab2686557da5d9bf963fcc118

Observation a2b02145-2d74-4af7-b2cf-57e6af75cced · outbound

This paper cites UniRL: Self-Improving Unified Multimodal Models via Supervised and Reinforcement Learning.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs UniRL: Self-Improving Unified Multimodal Models via Supervised and Reinforcement Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:58.185926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:58.185926Z digest=sha256:a2eee0c8716b0b91b7916f55dd9b4cf5f8d3c84191563b7eed959244d269e756

Observation f8acfc31-515b-42f5-812e-3ef4d24df194 · outbound

This paper cites LaViDa-R1: Advancing Reasoning for Unified Multimodal Diffusion Language Models.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs LaViDa-R1: Advancing Reasoning for Unified Multimodal Diffusion Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:58.246966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:58.246966Z digest=sha256:d5bb603bb7331d7949b54a39433b9ea8e1e60438eb4585053f960269f5ae1f1e

Observation 0b475d5a-72b9-4920-a703-6281b7369990 · outbound

This paper cites Towards unified multimodal interleaved generation via group relative policy optimization.Advances in Neural Information Processing Systems, 38:5332–5353, 2026.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Towards unified multimodal interleaved generation via group relative policy optimization.Advances in Neural Information Processing Systems, 38:5332–5353, 2026

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:58.307608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:58.307608Z digest=sha256:0a3bb90487a1e2c703512ebe9c6a7513c4235252a1dcb6b2f10b085358729094

Observation aed24d9f-ca30-45ad-93f7-1e994a8233a1 · outbound

This paper cites Lvrpo: Language-visual alignment with grpo for multimodal under- standing and generation.arXiv preprint arXiv:2603.27693, 2026.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Lvrpo: Language-visual alignment with grpo for multimodal under- standing and generation.arXiv preprint arXiv:2603.27693, 2026

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:58.365146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:58.365146Z digest=sha256:d64f70ac21310e45710e6f27b65b506da677ee5043bc979ae32260d8e8c1283d

Observation 3338d5c3-8fd9-495b-b2e4-510c87a01a14 · outbound

This paper cites Generation Enhances Understanding in Unified Multimodal Models via Multi-Representation Generation.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Generation Enhances Understanding in Unified Multimodal Models via Multi-Representation Generation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:58.412871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:58.412871Z digest=sha256:5dc481d279abf55b8c2afd40e9124c55c53aa196e38fa58a59c71f4c37fca42e

Observation e29cf26c-52fb-46fd-aefe-7c1e503fd3fa · outbound

This paper cites Unified multimodal models as auto-encoders.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Unified multimodal models as auto-encoders

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:58.483489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:58.483489Z digest=sha256:00d06ca0d2e021555eda9602096fe548eba01acffa80d5a445c9a3627776dd4f

Observation 002f0bdf-85c7-4bf2-af29-b94827f52057 · outbound

This paper cites Reconstruction Alignment Improves Unified Multimodal Models.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Reconstruction Alignment Improves Unified Multimodal Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:58.587330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:58.587330Z digest=sha256:e753cf55ed9215f445f5f9aeccbc7ff22098cdaa0a0061ca0ec8cc25701de146

Observation df5130ca-d6d9-42b0-9e32-21fe1555638e · outbound

This paper cites Learning to generate via understanding: Understanding-driven intrinsic rewarding for unified multimodal models.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Learning to generate via understanding: Understanding-driven intrinsic rewarding for unified multimodal models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:58.654007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:58.654007Z digest=sha256:35850c14f47770d21a89a6e01d86bc937be197fc6d6c7486ec9ccd4dcfaddfa3

Observation 34c57d94-5c67-40a6-93a0-01c71264e2b1 · outbound

This paper cites Steering Visual Generation in Unified Multimodal Models with Understanding Supervision.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Steering Visual Generation in Unified Multimodal Models with Understanding Supervision

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:58.714370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:58.714370Z digest=sha256:fadd9bc5c5916ed2e35e3c304f0571e75787b57ae65c19d442c91285824f146c

Observation 9d07a7b5-ea16-4598-be28-2ebb0316e786 · outbound

This paper cites Flow-grpo: Training flow matching models via online rl.Advances in neural information processing systems, 38:40783–40818, 2026.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Flow-grpo: Training flow matching models via online rl.Advances in neural information processing systems, 38:40783–40818, 2026

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:58.769990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:58.769990Z digest=sha256:6077d9c80f5553d08e0e42a3f88bd9363e4bd40d69828ff03ff6bb3514ba0dd3

Observation 9a09be2e-4f18-419a-82e3-e1b14bb4f540 · outbound

This paper cites MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:58.829878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:58.829878Z digest=sha256:ef968430bc38e10f3ef0054ac67beb12c6f08cf6d982228b6f877fdf2fc0af9b

Observation 24e643f5-3fe4-4626-9afc-9b1e180fccfa · outbound

This paper cites Editreward: A human- aligned reward model for instruction-guided image editing.arXiv preprint arXiv:2509.26346, 2025.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Editreward: A human- aligned reward model for instruction-guided image editing.arXiv preprint arXiv:2509.26346, 2025

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:58.888826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:58.888826Z digest=sha256:56b76761da5668aa421e362fda642a6178cd3a791804a474435b80704c2a227c

Observation afaa18da-9a25-48f7-a886-1de74b58cfc0 · outbound

This paper cites Thinkrl-edit: Thinking in reinforcement learning for reasoning-centric image editing.arXiv preprint arXiv:2601.03467, 2026.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Thinkrl-edit: Thinking in reinforcement learning for reasoning-centric image editing.arXiv preprint arXiv:2601.03467, 2026

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:59.025476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:59.025476Z digest=sha256:63b22bbaeed1b063e34b7db5c9cb72313c551f4883829189a839df7f58ce33f3

Observation 8b88c607-9014-439a-bb65-1c40d193fcaf · outbound

This paper cites Leveraging verifier-based reinforcement learning in image editing.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Leveraging verifier-based reinforcement learning in image editing

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:59.113488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:59.113488Z digest=sha256:aa299a12ed3a5c60371e57ee2ae9bb12f22bbbcef2b07884310b0c1b1bac9e67

Observation ac0d21e7-1637-40d7-9c6a-1fb5a21f301c · outbound

This paper cites Pico-banana-400k: A large-scale dataset for text-guided image editing.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Pico-banana-400k: A large-scale dataset for text-guided image editing

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:59.189410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:59.189410Z digest=sha256:132681c6232f43c73e3eb594fc5f018712d00a509dc66604c844c5e983e25189

Observation f9217aca-024b-42fa-a3c6-894c4464b699 · outbound

This paper cites Qwen3 Technical Report.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Qwen3 Technical Report

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:59.252101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:59.252101Z digest=sha256:ba3d8a17ae6c5801b00f7953df9c7a87bf790209dc494c9d92395f1e5f03250c

Observation 3b9daf6c-c2d2-4ee4-8041-debc9c1dd3d7 · outbound

This paper cites an unresolved cited work.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:59.315445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:59.315445Z digest=sha256:0495d532b25f395e4d58282e9847738d2b13719c4fa4bb678386d7e1e359913c

Observation 86aecf24-41e0-4e8c-8e45-acf9fa3bfe2a · outbound

This paper cites Editthinker: Unlocking iterative reasoning for any image editor.arXiv preprint arXiv:2512.05965, 2025.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Editthinker: Unlocking iterative reasoning for any image editor.arXiv preprint arXiv:2512.05965, 2025

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:59.374345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:59.374345Z digest=sha256:f2ccc409a9aea798eedff656b0c9e6cfc4ea27927215940872f864e339674dfd

Observation f28a57ae-bd12-4180-a370-5a35e94acbc8 · outbound

This paper cites Blink: Multimodal large language models can see but not perceive.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Blink: Multimodal large language models can see but not perceive

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:59.435425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:59.435425Z digest=sha256:d94f52d2c444845a2edb689e839e7cb5b4e2987d1997c91acb9b1a501dab6b80

Observation 294bc429-dc3d-434a-977b-fdae81065aa4 · outbound

This paper cites Cambrian-1: A fully open, vision-centric exploration of multimodal llms.Advances in Neural Information Processing Systems, 37:87310–87356, 2024.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Cambrian-1: A fully open, vision-centric exploration of multimodal llms.Advances in Neural Information Processing Systems, 37:87310–87356, 2024

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:59.546062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:59.546062Z digest=sha256:efdbcdd0642fde0708af52d5540e33d18309b6fb865a4113e3c5414030f39b19

Observation 8146b69e-0641-4ca9-8aa3-f5f140f83143 · outbound

This paper cites Thinkmorph: Emergent properties in multimodal interleaved chain-of-thought reasoning.arXiv preprint arXiv:2510.27492, 2025.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Thinkmorph: Emergent properties in multimodal interleaved chain-of-thought reasoning.arXiv preprint arXiv:2510.27492, 2025

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:59.624256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:59.624256Z digest=sha256:38d53b860bf2e93d4e3b5fab7b620e314f2fc464f4de227ee0fcc1d34226a14e

Observation ccd2384b-0de4-45f1-990f-269c862a6fdd · outbound

This paper cites MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:59.711456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:59.711456Z digest=sha256:23544029b6024c8f1c1f602555356694bc71816a3f6a0bb373d727e13e6ba71e

Observation ad24bab6-197a-4361-ade8-f905a11b8284 · outbound

This paper cites Eyes wide shut? exploring the visual shortcomings of multimodal llms.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Eyes wide shut? exploring the visual shortcomings of multimodal llms

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:59.778381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:59.778381Z digest=sha256:1bd6a2184d69c731c4bfe04f6780fb6c3ce49f9def8a681db340cf290236e081

Observation 04f19fb9-0956-4f9e-8ce8-d1f21bb21a8e · outbound

This paper cites Oddgridbench: Expos- ing the lack of fine-grained visual discrepancy sensitivity in multimodal large language models.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Oddgridbench: Expos- ing the lack of fine-grained visual discrepancy sensitivity in multimodal large language models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:59.831659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:59.831659Z digest=sha256:982a429d6eb68f909868a61bc95294c10479c5a22b286a4281142fc911c03e75

Observation 165cffc8-b8eb-4bb3-8abd-938d396c08b9 · outbound

This paper cites WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:59.905835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:59.905835Z digest=sha256:04067174f6da7f7b01d155458bcc26f4186c780a6279eee9037b8bbf7543c6dd

Observation 4c2447dd-3fb7-4a90-8516-749aa5b1e468 · outbound

This paper cites Imgedit: A unified image editing dataset and benchmark.Advances in Neural Information Processing Systems, 38, 2026.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Imgedit: A unified image editing dataset and benchmark.Advances in Neural Information Processing Systems, 38, 2026

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T18:56:59.980892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:56:59.980892Z digest=sha256:33ad4912331afe03bb8a949ac2097bc14204efa97335f3401da38608d8be5af2

Observation 56e0e08a-42c1-4168-afc6-d4f8ae3d9238 · outbound

This paper cites Envisioning beyond the pixels: Benchmarking reasoning- informed visual editing.Advances in Neural Information Processing Systems, 38, 2026.

STBridge: Shared-Target Alignment for Bridging Understanding and Generation in UMMs Envisioning beyond the pixels: Benchmarking reasoning- informed visual editing.Advances in Neural Information Processing Systems, 38, 2026

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-01T18:57:00.056228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:57:00.056228Z digest=sha256:a91529b1908ec40d81a1210f6c8df03425ba0e12076d863b36e16049cdf5f0e4

Pith citing papers

No inbound Pith citation observations are available.