Pith. sign in

Paper Citation Record · LEDGER

Visual prompt engineering for video models

As of 4 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2607.25537.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.25537 v1

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T02:13:12.723564Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

61 of 61 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved61
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5f090e9a-fe50-4d4c-94f5-f60cc954ff98 · outbound

This paper cites Demonstrate-Search-Predict: Composing retrieval and language models for knowledge-intensive NLP.

Visual prompt engineering for video models Demonstrate-Search-Predict: Composing retrieval and language models for knowledge-intensive NLP

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:11.448951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:11.448951Z digest=sha256:2939452a43e9a7039964f7a4b464bac4e92242b8de2bd78aa810cdc7db1049f8

Observation 09c09a26-2eb7-4188-a66e-f7914423a6b8 · outbound

This paper cites A Prompt Pattern Catalog to Enhance Prompt Engineering with ChatGPT.

Visual prompt engineering for video models A Prompt Pattern Catalog to Enhance Prompt Engineering with ChatGPT

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:11.462591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:11.462591Z digest=sha256:f4942757fb0f44a63e89a30808bda199d1a42d2e7b43a4ac79eaefe86f49ac61

Observation 89a7b524-47c3-4ab3-b752-0fecacec8c1f · outbound

This paper cites Prompt engineering with ChatGPT: a guide for academic writers.Annals of biomedical engineering, 51(12):2629–2633, 2023.

Visual prompt engineering for video models Prompt engineering with ChatGPT: a guide for academic writers.Annals of biomedical engineering, 51(12):2629–2633, 2023

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:11.479275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:11.479275Z digest=sha256:68668e4e1786a79d802198cb5cbaea36c8c01eb67a63fb30d1f0e09148fcfbbe

Observation d5eedd9b-796c-43da-bdf3-c67c584abc2f · outbound

This paper cites Prompt programming for large language models: Beyond the few-shot paradigm.

Visual prompt engineering for video models Prompt programming for large language models: Beyond the few-shot paradigm

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:11.490979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:11.490979Z digest=sha256:8dfe80a15a2af8eaf515848ab038dacca5c6ffcc6a1a98b9b82388df445b2f90

Observation 744a43c0-473a-478e-a94a-3098827c6b5e · outbound

This paper cites Dspy: compiling declarative language model calls into state-of-the-art pipelines.

Visual prompt engineering for video models Dspy: compiling declarative language model calls into state-of-the-art pipelines

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:11.507043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:11.507043Z digest=sha256:a309242b9b64cb94e68bd824209773f55aeba6186f6594970fbea4b57a75aa0d

Observation bee71271-569c-4350-b0be-ca4ba20dbfa9 · outbound

This paper cites TextGrad: Automatic "Differentiation" via Text.

Visual prompt engineering for video models TextGrad: Automatic "Differentiation" via Text

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:11.524453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:11.524453Z digest=sha256:03b5969accef06025cf81fca3309a016967aba4f799130f0073578bf2b4c82a0

Observation ebe7cd9b-39d2-4f59-9b08-de03a11b14a2 · outbound

This paper cites Prompt engineering in large language models.

Visual prompt engineering for video models Prompt engineering in large language models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:11.540171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:11.540171Z digest=sha256:caec46c6945e434bfce4e80ed15827110cdf31bc78a8f5860ce4f64925867173

Observation b0a38dd5-468c-4434-bfe6-c0d6a8719c8d · outbound

This paper cites Prompt engineering as an important emerging skill for medical professionals: tutorial.Journal of medical Internet research, 25:e50638, 2023.

Visual prompt engineering for video models Prompt engineering as an important emerging skill for medical professionals: tutorial.Journal of medical Internet research, 25:e50638, 2023

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:11.561565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:11.561565Z digest=sha256:b9780fc6b0ce06171c23b82b4c9966859948a39f01c04594c1f7e60ceba89a2f

Observation 0654ae9d-0209-4ca9-b583-f60db8edc336 · outbound

This paper cites Video models are zero-shot learners and reasoners.

Visual prompt engineering for video models Video models are zero-shot learners and reasoners

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:11.573564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:11.573564Z digest=sha256:2f4db488b27db6169814f9ab0030a7de4590ae2538c572473ed4fbf42bdef3da

Observation a8759541-caa2-4281-a86f-58fc7310dc89 · outbound

This paper cites Video as the New Language for Real-World Decision Making.

Visual prompt engineering for video models Video as the New Language for Real-World Decision Making

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:11.587979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:11.587979Z digest=sha256:2ff060a04d27086197e6e300ece28cb2a04236f00655d156c95ab72c2fc091b5

Observation d062959b-d2a0-4e0b-8cd5-74fecac58800 · outbound

This paper cites Rethinking visual intelligence: Insights from video pretraining.arXiv preprint arXiv:2510.24448, 2025.

Visual prompt engineering for video models Rethinking visual intelligence: Insights from video pretraining.arXiv preprint arXiv:2510.24448, 2025

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:11.632251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:11.632251Z digest=sha256:0e4ab089afdc954047c62b769ea64aea004779b056054e351cce875608ef5aa4

Observation f48ddc45-34d6-46fe-9b10-ff51575dd041 · outbound

This paper cites A very big video reasoning suite.arXiv preprint arXiv:2602.20159, 2026.

Visual prompt engineering for video models A very big video reasoning suite.arXiv preprint arXiv:2602.20159, 2026

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:11.677056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:11.677056Z digest=sha256:b3c5eaa680bdb0cfc4521d15db3cb869f342ccacc284becdca65778e8d4364e4

Observation b40ca782-8d73-490d-826a-4cd7ac835f78 · outbound

This paper cites MentisOculi: Revealing the Limits of Reasoning with Mental Imagery.

Visual prompt engineering for video models MentisOculi: Revealing the Limits of Reasoning with Mental Imagery

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:11.701130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:11.701130Z digest=sha256:a6cd9b43c2b6ffd3247f61ce6be2c302051b2cea7ce5dc22923d806b91b00da6

Observation 22e635c4-1d8e-417e-99d7-c1a8bc7657b1 · outbound

This paper cites Video models reason early: Exploiting plan commitment for maze solving.arXiv preprint arXiv:2603.30043, 2026.

Visual prompt engineering for video models Video models reason early: Exploiting plan commitment for maze solving.arXiv preprint arXiv:2603.30043, 2026

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:11.730294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:11.730294Z digest=sha256:1cbd6c4143c9ca6f57f3b15f8e744162b4bb4a88b64cf5c97a83fbad96387a14

Observation 1ddbc3ed-f6a2-4f2e-b577-cb97588e1911 · outbound

This paper cites Are video models ready as zero-shot reasoners? an empirical study with the mme-cof benchmark.

Visual prompt engineering for video models Are video models ready as zero-shot reasoners? an empirical study with the mme-cof benchmark

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:11.741999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:11.741999Z digest=sha256:29efb9872322363ede4ddbd09a081ccffa6f9b5a806101fe4dea3b51fffc5be2

Observation a1e35699-3f7b-4289-b025-7f4af4bed506 · outbound

This paper cites Demystifying Video Reasoning.

Visual prompt engineering for video models Demystifying Video Reasoning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:11.772187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:11.772187Z digest=sha256:9ff9533d63a6099c840e11e46e5bc98d5f0a026116ada343ae86fe2fe1c13b3c

Observation 80735478-d855-4326-a579-d2a52006468a · outbound

This paper cites Thinking in frames: How visual context and test-time scaling empower video reasoning.arXiv preprint arXiv:2601.21037, 2026.

Visual prompt engineering for video models Thinking in frames: How visual context and test-time scaling empower video reasoning.arXiv preprint arXiv:2601.21037, 2026

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:11.894883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:11.894883Z digest=sha256:aaa8ac3b6400db7880fa8062a59261b6836a42f85cabc56453887d9b90131f23

Observation 0668deb4-1aad-42da-b63c-07be808d368b · outbound

This paper cites VLMs are Good Teachers for Video Reasoning via Adaptive Test-Time Optimization.

Visual prompt engineering for video models VLMs are Good Teachers for Video Reasoning via Adaptive Test-Time Optimization

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:11.938873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:11.938873Z digest=sha256:f8da682f407c5e045b58bdad211007bd8971c71f7dc70cb605f4daaae3d3a937

Observation bd531f51-11e3-419b-9ecd-d8ddce809eba · outbound

This paper cites Thinking with video: Video generation as a promising multimodal reasoning paradigm.

Visual prompt engineering for video models Thinking with video: Video generation as a promising multimodal reasoning paradigm

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.013390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.013390Z digest=sha256:4109f62598681d68c68cc2a96837f1eb3cfdc7e8e0ed1efb93df61e7cea59346

Observation e3435844-0e9a-4b6a-b693-41c87750df24 · outbound

This paper cites Review of large vision models and visual prompt engineering.Meta-Radiology, 1(3):100047, 2023.

Visual prompt engineering for video models Review of large vision models and visual prompt engineering.Meta-Radiology, 1(3):100047, 2023

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.146963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.146963Z digest=sha256:8b79dd75ad1fc9bc7be96106dce5aac631162e8386b3b1b647345a05ece4f446

Observation b7cda6a1-eada-4c41-8416-e152951f0bc6 · outbound

This paper cites A Systematic Survey of Prompt Engineering on Vision-Language Foundation Models.

Visual prompt engineering for video models A Systematic Survey of Prompt Engineering on Vision-Language Foundation Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.198478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.198478Z digest=sha256:882338b1637a1aaa2b432a15b05ef02f85303feeea3dc993af666796f5c057d0

Observation 0cd51713-ca86-417a-abe8-98a1a2fa6d14 · outbound

This paper cites What does clip know about a red circle? visual prompt engineering for vlms.

Visual prompt engineering for video models What does clip know about a red circle? visual prompt engineering for vlms

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.253737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.253737Z digest=sha256:f9813763814ec251ee26b8caf9ccb6968c983bce31646a90d5242f19a0d5abc5

Observation 2d743149-ad87-428e-a26b-16dbbce48e60 · outbound

This paper cites Cpt: Colorful prompt tuning for pre-trained vision-language models.AI Open, 5:30–38, 2024.

Visual prompt engineering for video models Cpt: Colorful prompt tuning for pre-trained vision-language models.AI Open, 5:30–38, 2024

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.300836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.300836Z digest=sha256:268c7f2122627b63c18d32de51dcbcfaeb0082a28d507c704213f711fa61dd0e

Observation 085e645b-ba73-4dbc-98ac-2d97d35e2bfc · outbound

This paper cites Exploring Visual Prompts for Adapting Large-Scale Models.

Visual prompt engineering for video models Exploring Visual Prompts for Adapting Large-Scale Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.362932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.362932Z digest=sha256:f0cf1f4047865aca7659ba2a40eee4c9b117639dded25e4c5750967b20a94da8

Observation 3070168c-54c1-4a00-9076-42c423b9a614 · outbound

This paper cites Highlight: Learning visual prompts for vision-language models, 2024.

Visual prompt engineering for video models Highlight: Learning visual prompts for vision-language models, 2024

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.425127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.425127Z digest=sha256:003c1d0c110df6bf51a0547ddade1be67c1c90f9fa2cf84207cd23577d3e5950

Observation f20ac7b9-241a-433e-869a-2982d5afeb3c · outbound

This paper cites Visual prompt tuning.

Visual prompt engineering for video models Visual prompt tuning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.495284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.495284Z digest=sha256:5f3e92a50818a4ff348302d4b0d5182c540aac7852840c5914c9c8b64b679a00

Observation 54a50607-9700-44b8-85bc-541c0487cc3f · outbound

This paper cites Visual promptingviaimageinpainting.

Visual prompt engineering for video models Visual promptingviaimageinpainting

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.562064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.562064Z digest=sha256:69230c631c6d81b18a75094cafc2c3e28ee04b8301cb3d91474852bf187da24f

Observation 345bc233-e3ed-46bd-bfe9-e89dfc631b4f · outbound

This paper cites Images speak in images: A generalist painter for in-context visual learning.

Visual prompt engineering for video models Images speak in images: A generalist painter for in-context visual learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.575509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.575509Z digest=sha256:4c01adafcc0bb4ffdac3d244a3f80c9feb8e47db06128bd927d071d582875b12

Observation a64133c8-05fe-46d5-80fd-e478ce72cf3b · outbound

This paper cites Visualcloze: A universal image generation framework via visual in-context learning.

Visual prompt engineering for video models Visualcloze: A universal image generation framework via visual in-context learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.579748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.579748Z digest=sha256:2615ff99a234bcae930290754074bf36dd94077d88efbe842aa93b6c213be87b

Observation 4c4ec16f-038e-440b-8c90-77973ef24ef5 · outbound

This paper cites Draw-and-understand: Leveraging visual prompts to enable mllms to comprehend what you want.

Visual prompt engineering for video models Draw-and-understand: Leveraging visual prompts to enable mllms to comprehend what you want

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.583911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.583911Z digest=sha256:f47b2c682a5111e6497eb69195a71cd5b13644e634990b56ac357107e074b162

Observation a38f6da9-6952-4955-bd88-71f4f76b3eff · outbound

This paper cites Visual Physics Comprehension Test (VPCT) Dataset.https://huggingface.

Visual prompt engineering for video models Visual Physics Comprehension Test (VPCT) Dataset.https://huggingface

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.588062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.588062Z digest=sha256:e3fa79907e544fdb52943442705015db9c236809e683f501700fbbd7ca03a8bb

Observation 5a03b69f-f3fb-40bf-8a81-95d7e64f3319 · outbound

This paper cites Nano Banana 2: Gemini Image Generation Overview.https://gemini.google/ov erview/image-generation/, 2026.

Visual prompt engineering for video models Nano Banana 2: Gemini Image Generation Overview.https://gemini.google/ov erview/image-generation/, 2026

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.592298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.592298Z digest=sha256:c7383c1fe438f8d3d5997e0f95ede089724f3eec57f2b88a82806a472c9b1b74

Observation 61c54c8f-14b9-47a8-8e17-518fb2d2cc1c · outbound

This paper cites Gemini 3.1 Pro.

Visual prompt engineering for video models Gemini 3.1 Pro

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.596436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.596436Z digest=sha256:8a3561357c0ed8564be64e7804880095637d3e6d29aa14fd96be3657ae75933d

Observation 589565e5-19eb-431f-a5b4-421644604b78 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Visual prompt engineering for video models Wan: Open and Advanced Large-Scale Video Generative Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.600739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.600739Z digest=sha256:e6904972bd7bf09cd2ce44766c1b8e45035b74db727e8fa12f74fd5ff4e8c608

Observation ec29296a-05f9-4d7d-954a-5385078ca917 · outbound

This paper cites an unresolved cited work.

Visual prompt engineering for video models Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.605421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.605421Z digest=sha256:7ae0751aa71bff97c22dde5ff45e950aca804e3f1bd638e43b999d6ede632cd8

Observation a2921345-95d1-4bca-b538-6fe676ddc70f · outbound

This paper cites Omni Flash Model Card.https://deepmind.google/models/model-cards/g emini-omni-flash/, 2026.

Visual prompt engineering for video models Omni Flash Model Card.https://deepmind.google/models/model-cards/g emini-omni-flash/, 2026

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.609889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.609889Z digest=sha256:2de3b7bdc8de94269d0ffafb36053ef3705f3eea243b29b78c2e58a85e1ccce1

Observation 58327f99-3e6a-4d34-a34f-99a46d33de5c · outbound

This paper cites Interpreting and controlling model behavior via constitutions for atomic concept edits.

Visual prompt engineering for video models Interpreting and controlling model behavior via constitutions for atomic concept edits

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.614086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.614086Z digest=sha256:2ff6888fb4a3f5ded77a76d7af18b3c7cd855e7cbe777cca46a753f606bcf1f1

Observation 843dabc7-1838-4840-8831-6ec344419360 · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

Visual prompt engineering for video models Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.618377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.618377Z digest=sha256:380e82959e324c46fbe2d827dc5481edbc553a4951bf49924270f6348b75f6f5

Observation cc6a05d0-6c30-4c7f-8176-41394e174f2b · outbound

This paper cites Gemini developer API pricing.https://ai.google.dev/gemini-api/docs/pr icing#veo-3.1, 2026.

Visual prompt engineering for video models Gemini developer API pricing.https://ai.google.dev/gemini-api/docs/pr icing#veo-3.1, 2026

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.622700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.622700Z digest=sha256:fbce3e43e893935eda6f61ef654c86f3e2738bf71265afc95acf59bc1c7a3f1a

Observation f69b5237-8b0e-463e-8ae0-e3291e97c427 · outbound

This paper cites Performance vs.

Visual prompt engineering for video models Performance vs

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.626805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.626805Z digest=sha256:2765ffb7f8390ba4387ca4f7b1426cfcd90793adfc7b0034cff7c645886d8e60

Observation d65c511e-c7ab-4aa3-ac98-bd961710bf65 · outbound

This paper cites How can we know what language models know?Transactions of the Association for Computational Linguistics, 8:423–438, 2020.

Visual prompt engineering for video models How can we know what language models know?Transactions of the Association for Computational Linguistics, 8:423–438, 2020

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.631186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.631186Z digest=sha256:3e93b9c0c5e8b33f8045eddac70eee625e1a61e6195626a72df2027e91f3406f

Observation 9e66877a-2c83-43d4-9e57-4047f3e4fe62 · outbound

This paper cites Inducing relational knowledge from BERT.

Visual prompt engineering for video models Inducing relational knowledge from BERT

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.635363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.635363Z digest=sha256:95325157dee7094773bb8679e4b939f93d1b712391b0d2128079213af960673d

Observation 40892ac7-8b43-41a3-91b1-ca301b55024b · outbound

This paper cites Shortcut learning in deep neural networks.Nature Machine Intelligence, 2(11):665–673, 2020.

Visual prompt engineering for video models Shortcut learning in deep neural networks.Nature Machine Intelligence, 2(11):665–673, 2020

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.639407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.639407Z digest=sha256:fb71c98c087d91f89a20de89e8e4c5744d9c10ce2eea4a5933300237c758dd54

Observation f7810ad5-54e8-48c9-914a-5599cbf52dd5 · outbound

This paper cites Unmasking Clever Hans predictors and assessing what machines really learn.Nature communications, 10(1):1096, 2019.

Visual prompt engineering for video models Unmasking Clever Hans predictors and assessing what machines really learn.Nature communications, 10(1):1096, 2019

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.643831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.643831Z digest=sha256:1768b7efb378eb263d4a0279ff9ca0242e7cec07577fce2450da868144bce01c

Observation 0fba6822-e52a-409b-8bdf-7093defd00d9 · outbound

This paper cites Unbiased look at dataset bias.

Visual prompt engineering for video models Unbiased look at dataset bias

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.648237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.648237Z digest=sha256:ac9836312db79ca1c5152850a91b22c0aac07af0e4870842cfa42bed1769d74f

Observation 88035d6b-db89-4b45-b581-abb7c189b321 · outbound

This paper cites Cosmos 3: Omnimodal World Models for Physical AI.

Visual prompt engineering for video models Cosmos 3: Omnimodal World Models for Physical AI

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.652594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.652594Z digest=sha256:86b55bb287c5670b9ad8714e8eba28e854108a7be47db1fe8d19ea01ea9b5cec

Observation 558e47f8-d117-4434-bee0-2a3ca4ee0768 · outbound

This paper cites this vertical drop leads directly into the first bucket on the left.

Visual prompt engineering for video models this vertical drop leads directly into the first bucket on the left

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.656912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.656912Z digest=sha256:f4fe25e978d432c286358d6fb02213afb93b8e72d95a64784e9c0b6fac7ca54f

Observation 0ee5b61a-16dc-40fc-9423-fd5f1be4e28a · outbound

This paper cites an unresolved cited work.

Visual prompt engineering for video models Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.661963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.661963Z digest=sha256:59c8148636e6ae7ef6df986f50b095d1554106dc74529e7b4673551d0c80056f

Observation 1814a3af-73f5-4c0a-9383-453be937d88a · outbound

This paper cites an unresolved cited work.

Visual prompt engineering for video models Unresolved cited work

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.666424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.666424Z digest=sha256:2d2e046d1d19668c5d77173a6417c5bbef95c5e3d77d4b1f9005349f9dce649e

Observation f6e5c226-f8c0-4a60-9e7f-e307a3cfdf3c · outbound

This paper cites an unresolved cited work.

Visual prompt engineering for video models Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.670666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.670666Z digest=sha256:7913ea24281f54e0bc8092ac92ce4be68526990eb4a5a95715f21a8f8745ec0d

Observation 2df7766c-1d9d-4856-8a91-229bcfb18784 · outbound

This paper cites an unresolved cited work.

Visual prompt engineering for video models Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.674895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.674895Z digest=sha256:e78def786c3286ee280bcb3430434ae363392e84b3d8e49dd8f4d8941cb9c3ba

Observation 003ff0e1-6050-4106-8dc5-2cc570bcba7c · outbound

This paper cites an unresolved cited work.

Visual prompt engineering for video models Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.679494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.679494Z digest=sha256:d6f21ec2e6d38b352c0f8c45fcb573bf260142e62cb7f28a1bb4e3cb94fc962d

Observation 5e611fd3-90b4-45f2-a345-eaf0524a8e35 · outbound

This paper cites an unresolved cited work.

Visual prompt engineering for video models Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.683907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.683907Z digest=sha256:7b0854a93c517588d41c91cdace2aea3235c98b74157c65517f6d9cdb1f5af38

Observation f5104bf3-1644-4950-9952-94d42ad2e31b · outbound

This paper cites an unresolved cited work.

Visual prompt engineering for video models Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.688751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.688751Z digest=sha256:2915da9937aa7abe135b9e4e7b2be36138666bfaed60d274e7452c518f6bcbd0

Observation 09508719-b996-4a91-ab5c-ae012492def5 · outbound

This paper cites an unresolved cited work.

Visual prompt engineering for video models Unresolved cited work

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.692912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.692912Z digest=sha256:8d6b1de5823da6dd8f179e9f7952be4411aad6944cbe3aa488a666bc82889c1c

Observation f6225924-a716-48ff-9187-a2c7911213a9 · outbound

This paper cites Static shot, no zoom or pan.

Visual prompt engineering for video models Static shot, no zoom or pan

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.696949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.696949Z digest=sha256:8470f7920f4d42838a598a06a852268c1ccab45269d35d0251a2a8f0e7a60b10

Observation da033842-a0c7-4868-b0ee-0e7477f01da3 · outbound

This paper cites an unresolved cited work.

Visual prompt engineering for video models Unresolved cited work

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.702377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.702377Z digest=sha256:ef10f7bda9a7a27e9be2cb7a790571269b5666ab9f9632dfe4301071648c27ca

Observation 7139350d-05e2-440a-8f02-301b721cfce4 · outbound

This paper cites justification.

Visual prompt engineering for video models justification

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.710844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.710844Z digest=sha256:7f0a061abc28fa008125939fe41f5a6a238fefe4a57c0f881434a9eac7a50a24

Observation 08a7fe50-9548-444a-b797-d6fdceb30cd2 · outbound

This paper cites moves”: A list of strings representing the extracted moves in order, e.g., [“A N.

Visual prompt engineering for video models moves”: A list of strings representing the extracted moves in order, e.g., [“A N

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.715433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.715433Z digest=sha256:dc36ebc5782926ede0f0711a736f3979c062b31bcc57a690af81e1f28a616c03

Observation b1c04840-bd69-40ac-adae-20f12b9cc66e · outbound

This paper cites invalid_after_seconds.

Visual prompt engineering for video models invalid_after_seconds

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.719328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.719328Z digest=sha256:a4e5417f5d603fa5b063df772ec9fb7ed40e0bb6fe93d9ef0de5fca7d6703b6c

Observation 94721f54-1154-44b3-ac2c-bda806b17fb9 · outbound

This paper cites justification.

Visual prompt engineering for video models justification

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:12.723564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:12.723564Z digest=sha256:3a9c8fbc16c14f75a6d4db4c149ebe7397a92ffd91c1e57c4e09ccb4a88565bd

Pith citing papers

No inbound Pith citation observations are available.