Pith. sign in

Paper Citation Record · LEDGER

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents

As of 5 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 0 inbound Pith citation observations for arXiv:2607.23588.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.23588 v1

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-30T18:12:01.235057Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

52 of 52 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved52
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 21ea2c63-b769-4758-9ea7-433ab68aa066 · outbound

This paper cites Claude design.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Claude design

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:55.766093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:55.766093Z digest=sha256:76e4754d1560eb9d3d9c1108d75fd83ef4ee2ec5f67d385793e8935d91add474

Observation cd86e9c6-02f6-4636-8b18-ce4a899f0adf · outbound

This paper cites FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:55.821375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:55.821375Z digest=sha256:d49e9d3638a32ae34be7f3dc1e3a413bfa53f19819397f9e3f21789313bcbb77

Observation 6f371a22-9669-4d16-b0f1-0282e6f2c60d · outbound

This paper cites Instructpix2pix: Learning to follow image editing instructions.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Instructpix2pix: Learning to follow image editing instructions

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:55.940278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:55.940278Z digest=sha256:0969b33a913d9b9b26d590842908ba454b2e79dcdde408f5a183604c6f8f6d36

Observation 979b85a9-1103-41cc-9cdc-a0376b93a12b · outbound

This paper cites Seedream 5.0 Lite.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Seedream 5.0 Lite

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:56.063095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:56.063095Z digest=sha256:6e13f4a3143dac81ca8eb9ef039ed852614deee65bc784ff193bcf2d1f095c51

Observation 26d42d83-6d93-404c-a81e-263ce44c7f93 · outbound

This paper cites PhotoArtAgent: Intelligent Photo Retouching with Language Model-Based Artist Agents.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents PhotoArtAgent: Intelligent Photo Retouching with Language Model-Based Artist Agents

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:56.188903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:56.188903Z digest=sha256:4c2b49b5ece8853b4d269c703a468b7c454a6e99102eb7fbc42baf57e0873882

Observation 7ed4222a-a15a-417f-9dee-f4f4c4b073c7 · outbound

This paper cites Unify-agent: A unified multimodal agent for world-grounded image synthesis.arXiv preprint arXiv:2603.29620, 2026.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Unify-agent: A unified multimodal agent for world-grounded image synthesis.arXiv preprint arXiv:2603.29620, 2026

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:56.309493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:56.309493Z digest=sha256:dc3b472145f0ff2d362b6134c7e6528c59929408ed84110b746ec9f178e91e7c

Observation 1cb04005-c1fd-4aa0-af9d-d64afbbb2533 · outbound

This paper cites PosterCraft: Rethinking High-Quality Aesthetic Poster Generation in a Unified Framework.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents PosterCraft: Rethinking High-Quality Aesthetic Poster Generation in a Unified Framework

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:56.433952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:56.433952Z digest=sha256:a370804acf69e634981c21845b4a8dc43b99c089b1ee6c1ac441c55725968672

Observation 8542a092-b4c0-44b0-b3e9-aebd46e21b88 · outbound

This paper cites GenEvolve: Self-Evolving Image Generation Agents via Tool-Orchestrated Visual Experience Distillation.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents GenEvolve: Self-Evolving Image Generation Agents via Tool-Orchestrated Visual Experience Distillation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:56.550690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:56.550690Z digest=sha256:211c75f20dfa7692ddcbcd2c0b650758c40d5147fee7318516bd3c4db06ef547

Observation 37e624cd-d920-4eb3-b233-9aa7d5a5ad18 · outbound

This paper cites Emerging Properties in Unified Multimodal Pretraining.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Emerging Properties in Unified Multimodal Pretraining

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:56.665887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:56.665887Z digest=sha256:58f09c77ea58e904f80e271c18c0ca578859f99b543b23d054995a1dfdd97ee1

Observation 44f46cfe-0750-4c24-b165-f68bacaf0d47 · outbound

This paper cites Kimi-Audio Technical Report.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Kimi-Audio Technical Report

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:56.784415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:56.784415Z digest=sha256:c12781328bc7ecbb2fb2d94f4ad6cef5f9a1927556aeb682b3d613851fe90256

Observation 3a19b9ec-dd19-4685-9e33-8efe744cbeae · outbound

This paper cites MonetGPT: Solving puzzles enhances MLLMs’ image retouching skills.ACM Transactions on Graphics, 44(4):1–12, 2025.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents MonetGPT: Solving puzzles enhances MLLMs’ image retouching skills.ACM Transactions on Graphics, 44(4):1–12, 2025

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:56.904633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:56.904633Z digest=sha256:6e27b3dbb2db226b29665a5f78350678670c2faee575a238d8a085c7f2cde740

Observation a42c2564-ae86-47bc-b4d0-73937b5c0289 · outbound

This paper cites Gen-Searcher: Reinforcing Agentic Search for Image Generation.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Gen-Searcher: Reinforcing Agentic Search for Image Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:57.022662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:57.022662Z digest=sha256:efe7ee0e1c0613eb447b1afb0c78feebe063e1c21b708af21878880cbf84290a

Observation d99da952-6288-4f1a-a1fb-a78f0ec2f0e2 · outbound

This paper cites Guiding instruction-based image editing via multimodal large language models.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Guiding instruction-based image editing via multimodal large language models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:57.162269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:57.162269Z digest=sha256:8e31fcf9943de52644b0a8a3089f7aba28df0524da632391e7c7197d3e861e91

Observation 058d0f09-92b1-4fb7-a294-2207015ccb66 · outbound

This paper cites Advancing vision-language models in front-end development via data synthesis.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Advancing vision-language models in front-end development via data synthesis

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:57.328467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:57.328467Z digest=sha256:0feb0fdff9f0eea228d53e9d0e4b5388fda8736df2feb787acd05f2082e55523

Observation ae5280a9-57cb-4062-818e-b74362d592e6 · outbound

This paper cites Gemini 3.1 Pro Model Card.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Gemini 3.1 Pro Model Card

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:57.430533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:57.430533Z digest=sha256:3432a9e217bbd1b167d26f9b846e25cdc0b5eb3f32851f478c09bfc66e14e028

Observation c156b2f4-be16-4c3e-b887-a1da8e1b8344 · outbound

This paper cites Gemini 3.1 Flash Image Preview.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Gemini 3.1 Flash Image Preview

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:57.579067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:57.579067Z digest=sha256:501886b08a43f71e1812379c6cc063ff44f1e4ef74578aaf84357ccf139ad6fd

Observation d8c797ee-b6f2-46bc-a662-ed1b2b2fe421 · outbound

This paper cites Vision2Web: A Hierarchical Benchmark for Visual Website Development with Agent Verification.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Vision2Web: A Hierarchical Benchmark for Visual Website Development with Agent Verification

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:57.745081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:57.745081Z digest=sha256:0dc2d4382687978729f5adf07532bc0cc0b65d8b6a17c2fbf90df33d7deb7598

Observation 8a944490-009a-4f5f-a037-58f3d819745b · outbound

This paper cites ComfyGPT: A self-optimizing multi-agent system for comprehensive ComfyUI workflow generation.arXiv preprint arXiv:2503.17671, 2025.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents ComfyGPT: A self-optimizing multi-agent system for comprehensive ComfyUI workflow generation.arXiv preprint arXiv:2503.17671, 2025

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:57.911212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:57.911212Z digest=sha256:b44b7abfb8405d0e46e1ab5fb489408d198fc1ba3eccb5e530571327b72260f3

Observation db7a25dc-37d6-4961-94f6-a85ebf4cbc0e · outbound

This paper cites HQ-Edit: A High-Quality Dataset for Instruction-based Image Editing.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents HQ-Edit: A High-Quality Dataset for Instruction-based Image Editing

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:58.077397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:58.077397Z digest=sha256:614abc223e9f2e2c65290a702b5e5bb33bb4a4319f4ad3ad52b5e065480a556a

Observation 40985152-0e6e-4b6b-b772-80eace24448f · outbound

This paper cites Node-based editing for multimodal generation of text, audio, image, and video.arXiv preprint arXiv:2511.03227, 2025.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Node-based editing for multimodal generation of text, audio, image, and video.arXiv preprint arXiv:2511.03227, 2025

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:58.222414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:58.222414Z digest=sha256:121d17a669e60f4850b3b5679773744bb2028e73ab7e800ca50916e8986a66bc

Observation fbe66ef7-d5d0-4a28-a05d-28d380f4ebff · outbound

This paper cites StoryNodes: Human-AI co-creation for multimodal media generation using an agentic node-based interface.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents StoryNodes: Human-AI co-creation for multimodal media generation using an agentic node-based interface

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:58.307788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:58.307788Z digest=sha256:2db5dec35fad0230331605ba806caaba31ee018f1a9e11f2ccc496c00e33d6f0

Observation 03a9f316-3538-4abd-988c-a9f591c49e52 · outbound

This paper cites Stitch: A new way to design user interfaces using ai.https://blog.google/innovation-and-ai/ models-and-research/google-labs/stitch-ai-ui-design/, 2025.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Stitch: A new way to design user interfaces using ai.https://blog.google/innovation-and-ai/ models-and-research/google-labs/stitch-ai-ui-design/, 2025

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:58.367536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:58.367536Z digest=sha256:999a42a5c8810d35efb1ff1cc14fc35af252515f4a0ae334450596bb51c9d7a9

Observation aae6c768-67e6-4c26-a3af-86f203b8247a · outbound

This paper cites Unlocking the conversion of Web Screenshots into HTML Code with the WebSight Dataset.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Unlocking the conversion of Web Screenshots into HTML Code with the WebSight Dataset

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:58.456621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:58.456621Z digest=sha256:4af0feae10528f0356f229cb740747a0e57d550c3ba5cda88cc9f64030aa565e

Observation 8f64374c-0e29-456a-a649-1e6ef6df9098 · outbound

This paper cites GENEVA: GENErating and Visualizing branching narratives using LLMs.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents GENEVA: GENErating and Visualizing branching narratives using LLMs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:58.521121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:58.521121Z digest=sha256:f5e1c9853c18654ee88aa58d0fb34a3f99fbee91254fc89c5adaf062070f5f94

Observation 34531799-0305-48b4-b461-7dcbccfb4e84 · outbound

This paper cites Claw-Eval-Live: A Live Agent Benchmark for Evolving Real-World Workflows.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Claw-Eval-Live: A Live Agent Benchmark for Evolving Real-World Workflows

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:58.595774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:58.595774Z digest=sha256:1e0795d410e15e0c58cce369bbbdf139577356eac9b7f9a4a1be3f0c3ef2eb88

Observation 9aaff17b-9009-463c-9c75-e0be667817ad · outbound

This paper cites Libtv skills.https://github.com/libtv-labs/libtv-skills, 2025.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Libtv skills.https://github.com/libtv-labs/libtv-skills, 2025

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:58.688998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:58.688998Z digest=sha256:585dc004fbcbd244f8365ed9bda76b585821fcaa62ba2de3e108b355f552d475

Observation 4b154fe1-d545-4072-9131-2c0d5c10df58 · outbound

This paper cites JarvisArt: Liberating Human Artistic Creativity via an Intelligent Photo Retouching Agent.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents JarvisArt: Liberating Human Artistic Creativity via an Intelligent Photo Retouching Agent

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:58.736145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:58.736145Z digest=sha256:34cc89a44b671a9b6eedca7696b1d2dc8fcae3b4a7012dd3a176bd238da222bd

Observation a6ce469a-ed6b-4aad-a94d-ce4f2764537d · outbound

This paper cites Jarvisevo: Towards a self-evolving photo editing agent with synergistic editor-evaluator optimization.arXiv preprint arXiv:2511.23002, 2025.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Jarvisevo: Towards a self-evolving photo editing agent with synergistic editor-evaluator optimization.arXiv preprint arXiv:2511.23002, 2025

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:58.781908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:58.781908Z digest=sha256:7e5f6a1e568ea18efd580a8c72b74df5d2679bdd5613723f9bfdb76f5b14864a

Observation 539ba8ad-04d3-4935-b663-52fa5d49abb4 · outbound

This paper cites Step1X-Edit: A Practical Framework for General Image Editing.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Step1X-Edit: A Practical Framework for General Image Editing

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:58.832877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:58.832877Z digest=sha256:a8e2ab6bc29a85ff042a023dbb01dad01aa9aad99efad36215c0bc4811ac2e34

Observation 11d60bbc-73de-4819-81ed-5ef01e1afb68 · outbound

This paper cites Minimax hub.https://hub.minimax.io/, 2026.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Minimax hub.https://hub.minimax.io/, 2026

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:58.894339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:58.894339Z digest=sha256:503619a0272f8338ad6ffc06558ca3c48cc616c22fda0b798dee61d829463784

Observation e32f5fc7-b003-4762-aef0-4b8dfbee8309 · outbound

This paper cites Magentic-UI: Towards Human-in-the-loop Agentic Systems.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Magentic-UI: Towards Human-in-the-loop Agentic Systems

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:58.917187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:58.917187Z digest=sha256:ca2776dd7663a4ecd2f4cb93fd323435ede5655b480a1bf41a2fb235a69817b9

Observation 676fe693-835b-4eff-9247-03065bb364e3 · outbound

This paper cites GPT-5.5 System Card.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents GPT-5.5 System Card

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:59.022066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:59.022066Z digest=sha256:df37dda2b976b388e4ee849b2573705db08fbeb94a1cd8692bf9d3849271ebab

Observation d491df65-7c0b-4995-a504-8b9adcf8c54c · outbound

This paper cites GPT Image 2.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents GPT Image 2

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:59.213291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:59.213291Z digest=sha256:49cbbfcb339189bb9a2f3f9e48759e8dfe7d94697cae362ae492c0e72d8912b9

Observation 720708b6-5cb8-43d4-a462-76d119c287f4 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:59.404060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:59.404060Z digest=sha256:69cd8a84c13b12d8ef6436c6d9b5c6a989b0cd5bd29cdc202ad2225faa1162ed

Observation cd4e249d-0015-47ad-8ad3-ab3fb2d2f8ad · outbound

This paper cites Dynamic Storyboard Generation in an Engine-based Virtual Environment for Video Production.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Dynamic Storyboard Generation in an Engine-based Virtual Environment for Video Production

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:59.521115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:59.521115Z digest=sha256:be3771d79579b088ecb05dff304c03ba7f5544fcae90aaca4176fd27606f3812

Observation 60b2647d-3915-4a57-8b98-3a0c93205878 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents High-resolution image synthesis with latent diffusion models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:59.617080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:59.617080Z digest=sha256:0f6f6fc52663f41647c627af5a026ec337c94e49d82a70c98c6ff789059e0c96

Observation 49794ce4-c0b7-4e0a-9b08-0785327263e4 · outbound

This paper cites Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:59.621120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:59.621120Z digest=sha256:26dd68847981b671a714aa16dc7db3620ed7a7eb20b12176cda3a141ed3c41ca

Observation 6878d94b-bcd2-4992-a3e1-371bb99b03dc · outbound

This paper cites Seedream 4.0: Toward Next-generation Multimodal Image Generation.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Seedream 4.0: Toward Next-generation Multimodal Image Generation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:59.668552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:59.668552Z digest=sha256:6eb56f2b6da18d838d92a2ea04d82faf4da06858d88972e656c2d1d79b2116e4

Observation 8d09ca1d-8cea-451c-b3d1-ad53c360101e · outbound

This paper cites Design2code: Benchmarking multimodal code generation for automated front-end engineering.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Design2code: Benchmarking multimodal code generation for automated front-end engineering

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:59.809001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:59.809001Z digest=sha256:b7b88badffcf8669eddd3ab0015047c5d148c8e7cf61bc286e011bd414773cc3

Observation 6efc636a-e35c-4f9c-89ae-5438206cfdbb · outbound

This paper cites Tapnow ai creative platform documentation.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Tapnow ai creative platform documentation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-30T18:11:59.982896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:11:59.982896Z digest=sha256:0c09fe7d15646958c2306a2f4e9ef4910d238f20a698815ad9be3487a2e6c74e

Observation c157b6c6-1a0e-4025-947e-c8fdd252402b · outbound

This paper cites Seedance 2.0: Advancing Video Generation for World Complexity.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Seedance 2.0: Advancing Video Generation for World Complexity

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-30T18:12:00.088705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:12:00.088705Z digest=sha256:c3ee6d25ba36a5eabe5d1717b4c8c84de9812df72aa683f686a87d23eca8f384

Observation 6c8e2f5a-5045-4b3d-a715-fb1d5753f5d3 · outbound

This paper cites Audio-Omni: Extending Multi-modal Understanding to Versatile Audio Generation and Editing.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Audio-Omni: Extending Multi-modal Understanding to Versatile Audio Generation and Editing

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-30T18:12:00.209282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:12:00.209282Z digest=sha256:5b28c76ec15e4f9dd66ce4bd585d6ff7470a66d896e9e271bac9c40e722139ce

Observation 5b54d182-01d6-411e-a9c3-a08d3f6b1a63 · outbound

This paper cites Qwen-Image Technical Report.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Qwen-Image Technical Report

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-30T18:12:00.311201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:12:00.311201Z digest=sha256:6741a0e740852aafa3d1483e53ef827d4d8779b6c87ad8d569a464981bbe84f5

Observation 37cd124a-ae18-4090-aa2f-695e12154fd7 · outbound

This paper cites PromptChainer: Chaining large language model prompts through visual programming.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents PromptChainer: Chaining large language model prompts through visual programming

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-30T18:12:00.466983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:12:00.466983Z digest=sha256:eb9bbe91e2b5fbb6002cd0a750dc452b3952b72e07c2b3f029805b991bceab49

Observation 6237bb20-3860-4477-8552-2da13e6cf190 · outbound

This paper cites OmniGen: Unified Image Generation.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents OmniGen: Unified Image Generation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-30T18:12:00.567686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:12:00.567686Z digest=sha256:e6ff061cc5102ef4f9f2c694cb233bf361030d76c61d4b9a17684ff76d68a9b9

Observation 9b5ac9f8-6b9b-466d-9092-05d312e031be · outbound

This paper cites Show-o2: Improved Native Unified Multimodal Models.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Show-o2: Improved Native Unified Multimodal Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-30T18:12:00.656802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:12:00.656802Z digest=sha256:705446568db1b58e1e72c6afd84907dede205617476ffc554a4ca4324d98e6a6

Observation 698f80f4-da27-4cb7-8042-717c544608a2 · outbound

This paper cites ComfyUI-R1: Exploring Reasoning Models for Workflow Generation.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents ComfyUI-R1: Exploring Reasoning Models for Workflow Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-30T18:12:00.760876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:12:00.760876Z digest=sha256:b39986865a4cbb16af2329536265376ca7351b18fad480343ce99a36c1b6a7b9

Observation 7629497d-923e-4c8e-8544-ee70bfafe077 · outbound

This paper cites ComfyUI-Copilot: An intelligent assistant for automated workflow development.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents ComfyUI-Copilot: An intelligent assistant for automated workflow development

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-30T18:12:00.863447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:12:00.863447Z digest=sha256:4a0b62e5ce9fc0d2e6dfdeddfdbbf5ebf0f677f7e9e96e7ddc1263c1ac0181d6

Observation 80705135-ee9b-486f-874d-24b9e08188f2 · outbound

This paper cites ComfyBench: Benchmarking LLM-based agents in ComfyUI for autonomously designing collaborative AI systems.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents ComfyBench: Benchmarking LLM-based agents in ComfyUI for autonomously designing collaborative AI systems

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-30T18:12:00.966317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:12:00.966317Z digest=sha256:7ff68a4b1050173bb0d73d197f0818ef4543d6645b92b74414ae7a338dea4a29

Observation f0f2b316-81ac-4709-be89-890e94c9ae8d · outbound

This paper cites Scaling Autoregressive Models for Content-Rich Text-to-Image Generation.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Scaling Autoregressive Models for Content-Rich Text-to-Image Generation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-07-30T18:12:01.025913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:12:01.025913Z digest=sha256:92f1a3dfe4c7d1a2d9aff7530d22a1a55a7b427f11b3b714d2b64bc02b556e25

Observation 620c4fd8-93c1-4ec2-84a6-bc0ae413d493 · outbound

This paper cites Magicbrush: A manually annotated dataset for instruction-guided image editing.Advances in Neural Information Processing Systems, 36:31428–31449, 2023.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Magicbrush: A manually annotated dataset for instruction-guided image editing.Advances in Neural Information Processing Systems, 36:31428–31449, 2023

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-30T18:12:01.102433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:12:01.102433Z digest=sha256:00de4e924e8f3fd5333b18bb0033863d4b9c73f73ad019994831924387ae787e

Observation 75620615-899e-4220-bdda-5be7a7d02b66 · outbound

This paper cites Ultraedit: Instruction-based fine-grained image editing at scale.Advances in Neural Information Processing Systems, 37:3058–3093, 2024.

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Ultraedit: Instruction-based fine-grained image editing at scale.Advances in Neural Information Processing Systems, 37:3058–3093, 2024

Reference 52

Resolution
unresolved
no resolver link, observed 2026-07-30T18:12:01.235057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:12:01.235057Z digest=sha256:20d8a6eaf487452e322e7a7c21934f94231c0ad7a61b8162eeb93bed33030d0c

Pith citing papers

No inbound Pith citation observations are available.