Pith. sign in

Paper Citation Record · LEDGER

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing

As of 17 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 0 inbound Pith citation observations for arXiv:2607.25300.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.25300 v1

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T02:57:35.417802Z

measured 66 of 66 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

66 of 66 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved65
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ecced3ec-8956-43b8-8873-8915804c5eca · outbound

This paper cites Adopting self- supervised learning into unsupervised video summarization through restorative score.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Adopting self- supervised learning into unsupervised video summarization through restorative score

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.526010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.526010Z digest=sha256:95b198a535fa21fa602d21f2022743c70d4624f4f38ed0dc2d358e3f826ed947

Observation acda662c-6162-4503-9b3c-fd3458d436e4 · outbound

This paper cites Combining global and local attention with positional encoding for video summarization.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Combining global and local attention with positional encoding for video summarization

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.531817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.531817Z digest=sha256:9736c1646160330eb68db2e662f8b1139c9f922581d220e344e2e8c3b357664f

Observation 27d5d21a-0523-4123-a40c-ae1ba4b9053d · outbound

This paper cites Summarizing videos using con- centrated attention and considering the uniqueness and diver- sity of the video frames.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Summarizing videos using con- centrated attention and considering the uniqueness and diver- sity of the video frames

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.537215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.537215Z digest=sha256:d2eb8d3762b3326afd630d9a5a665ed43995a4fcfffbe7c87cf85953029401c8

Observation 59baaf17-8b2b-4b27-b110-c428ee0f6d7c · outbound

This paper cites Scaling up video summarization pretraining with large language models.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Scaling up video summarization pretraining with large language models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.542503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.542503Z digest=sha256:c7abc82a4f9efdf8221fe901bd0bc24a05a4921ffc478a596224185c3b2a9527

Observation 1de320e2-cd3a-4139-94e1-2109e9b99123 · outbound

This paper cites Qwen3-VL Technical Report.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Qwen3-VL Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.547372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.547372Z digest=sha256:c9c5726d1685673b5be861e6c46579397a4c16d1302f501456ea45b71c197cb1

Observation ff241b79-9bbc-432b-ad24-16efc41ab4d6 · outbound

This paper cites Blender studio films.https://studio.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Blender studio films.https://studio

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.553920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.553920Z digest=sha256:8140e06c8c404ce52061776c051b25ad05eb6c56268987de9dba05fffefa5356

Observation 093d8ffd-64fb-4612-9cf7-98c1eedbb5ec · outbound

This paper cites VSUMM: A mechanism designed to produce static video summaries and a novel evaluation method.Pattern Recog- nit.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing VSUMM: A mechanism designed to produce static video summaries and a novel evaluation method.Pattern Recog- nit

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.559812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.559812Z digest=sha256:60d716ec3ec482996902af0aacea612b1abf224705586a39cee231f3b24772a9

Observation 87394ed8-fd48-4735-977e-5a16af907239 · outbound

This paper cites Summarizing videos with attention.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Summarizing videos with attention

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.566410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.566410Z digest=sha256:e4eaef282d61ba3575542ad030737d473eff5bb2813d7bf956c39de26bb1ce38

Observation f517d32b-6f65-4f09-81b4-b8854a870cae · outbound

This paper cites Video-R1: Reinforcing video reasoning in MLLMs.NeurIPS, 38:99114–99137,.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Video-R1: Reinforcing video reasoning in MLLMs.NeurIPS, 38:99114–99137,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.573564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.573564Z digest=sha256:e964245c170e21da3885cf721a134edab364867523a060f7f23281c3692bcb1f

Observation 1526b89c-f3bf-481d-a0f6-4b3f32a69e5b · outbound

This paper cites Au- tomatic non-linear video editing transfer.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Au- tomatic non-linear video editing transfer

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.577949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.577949Z digest=sha256:2963438d48dd5b1a006c6c1e99e1efcafe3a1b58823b85c8cdb043afd764d362

Observation 4612b575-b863-4be5-8268-5df9bfe7066b · outbound

This paper cites Video-MME: The first-ever comprehensive evaluation benchmark of multi- modal LLMs in video analysis.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Video-MME: The first-ever comprehensive evaluation benchmark of multi- modal LLMs in video analysis

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.581992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.581992Z digest=sha256:dce4737c8a17282502c11f8a0d48557e024dddb3d027d8ea38bf00b3b54b7a23

Observation 1049e734-6252-40a7-b747-b7c66e8b30f5 · outbound

This paper cites Training-free language-guided video summarization via multi-grained saliency scoring.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Training-free language-guided video summarization via multi-grained saliency scoring

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.586663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.586663Z digest=sha256:fdc9c3831bcb483461a455c563109f381bf29239722fddd5c4183b7d061bf7a3

Observation 9aa6d92c-f5ab-46ce-84bf-d1d51241a196 · outbound

This paper cites Supervised video summarization via multiple feature sets with parallel attention.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Supervised video summarization via multiple feature sets with parallel attention

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.591243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.591243Z digest=sha256:50f118128e5f92fb2b8bc5f057f99318cb6f5c3249b8a192db0c8753f7e66afb

Observation 5a58f19a-82d1-447d-a574-17ed5d27a40b · outbound

This paper cites A Survey on LLM-as-a-Judge.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing A Survey on LLM-as-a-Judge

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.595868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.595868Z digest=sha256:95359cb49ed94836fef2f3d192eaf1c8a19b26bed787b5b1fd9f73060cfe4fe1

Observation c728aa61-d9b1-4197-b6a2-5da281883f28 · outbound

This paper cites VTG-LLM: Integrating timestamp knowledge into video LLMs for enhanced video temporal grounding.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing VTG-LLM: Integrating timestamp knowledge into video LLMs for enhanced video temporal grounding

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.600907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.600907Z digest=sha256:5257731224d809c0801b2378a0979930fe709557e3930d39a87f3f03aafd4a48

Observation 3ca9901d-096e-4061-ab4c-97a148843892 · outbound

This paper cites Creating summaries from user videos.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Creating summaries from user videos

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.606418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.606418Z digest=sha256:e5b25ce2897d601585848dbbdb037df3ebbe97592b0fd95836024c2e46860f27

Observation a7d90a2c-4405-41d9-aae1-0bf39a09d96b · outbound

This paper cites Align and attend: Multimodal summarization with dual contrastive losses.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Align and attend: Multimodal summarization with dual contrastive losses

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.611332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.611332Z digest=sha256:397011b3e3e0c52c409561a39a9daee0d190de2a362521e83c94271ed73ed9a2

Observation acfa2fc3-39d0-46a4-9ed5-1b8035c2a4de · outbound

This paper cites MovieNet: A holistic dataset for movie under- standing.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing MovieNet: A holistic dataset for movie under- standing

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.617191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.617191Z digest=sha256:79c7033494dfe0d3bd3f13d211c29e1692a17d5148b618944148674c9269858e

Observation 6fd57df8-6057-4080-ac17-65e9d06b8542 · outbound

This paper cites Joint video summarization and moment localization by cross-task sample transfer.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Joint video summarization and moment localization by cross-task sample transfer

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.623434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.623434Z digest=sha256:06987eaeb2316d141042a574c01459f52076df9bf2ff24826e6af574995c2d3f

Observation a192e377-fd68-4f53-8236-3dcd1e8de382 · outbound

This paper cites Discriminative feature learning for unsu- pervised video summarization.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Discriminative feature learning for unsu- pervised video summarization

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.633320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.633320Z digest=sha256:ca2ff34794164e757d404bc78732ad5ea701f97866f70cd74740d631dd483d5b

Observation e914b6eb-0492-4220-8877-28d176408329 · outbound

This paper cites Dense-captioning events in videos.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Dense-captioning events in videos

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.646692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.646692Z digest=sha256:44d89eccb2485b0ca90789d5f717f3d4f3fa83fe431d18d6e68138b6e7eb33f8

Observation 2da01de5-bbbf-4d77-8beb-e4ccfb1c4176 · outbound

This paper cites Computational video editing for dialogue-driven scenes.ACM Trans.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Computational video editing for dialogue-driven scenes.ACM Trans

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.659453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.659453Z digest=sha256:90ca579315f6c1c8132b85025a4f493d6c909d652c7044ddc767f5a5a7115d44

Observation f3f5c708-8ebf-4238-8d97-484837367b00 · outbound

This paper cites Video sum- marization with large language models.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Video sum- marization with large language models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.680677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.680677Z digest=sha256:e210b577980d2e7859d879953eab0a39fb2ec4b03ae900d5f46e09e65b4ff093

Observation 86900be1-3d86-456f-a173-501ca56e32a4 · outbound

This paper cites Detecting mo- ments and highlights in videos via natural language queries.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Detecting mo- ments and highlights in videos via natural language queries

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.700419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.700419Z digest=sha256:f103124983c92fe7d0f39c1c48e17c1ba7155b4dc722f60cb7d2fcdca5f12b37

Observation b63513ae-f924-4abf-ad7b-743101a434e0 · outbound

This paper cites Progressive video summarization via multimodal self- supervised learning.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Progressive video summarization via multimodal self- supervised learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.721473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.721473Z digest=sha256:bcbf0b8a42b9cb77f4f08c0f5fd734a58b83022a36272029c1d82cde106a6f48

Observation 062e7a7a-18d2-4467-adf9-01a0a304ad21 · outbound

This paper cites Progressive video summarization via multimodal self- supervised learning.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Progressive video summarization via multimodal self- supervised learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.747254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.747254Z digest=sha256:6c448737f51ba4787f1e6f39e12e3ea43cb13adadc08d5ed25f56105e97b7537

Observation da640c86-dd3c-4611-b5c5-b1c815b40c3a · outbound

This paper cites VideoChat-R1: Enhancing spatio-temporal percep- tion via reinforcement fine-tuning.NeurIPS, 2025.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing VideoChat-R1: Enhancing spatio-temporal percep- tion via reinforcement fine-tuning.NeurIPS, 2025

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.771469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.771469Z digest=sha256:3687d4a2f73e7578a0f125014636488bbf3d7d912f8ea0d9b1d821f24bc88ca3

Observation 41c3e89d-2252-4411-b6d1-10fcd170bf27 · outbound

This paper cites VideoXum: cross- modal visual and textural summarization of videos.IEEE Trans.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing VideoXum: cross- modal visual and textural summarization of videos.IEEE Trans

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.793614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.793614Z digest=sha256:702814d8dfad622b0a1e860fa9ee0d066af06894ebe9b5b72c3eb0ea1287f85b

Observation b70bf6a0-9822-46b6-8405-91157efb81a6 · outbound

This paper cites Visual instruction tuning.NeurIPS, 36:34892–34916, 2023.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Visual instruction tuning.NeurIPS, 36:34892–34916, 2023

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.819796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.819796Z digest=sha256:3cc0bc4a9fb5065710729a4cc1944ed1fd3a8a4ac5c7c204c303134bdd355e7f

Observation 4fcf045e-d7f8-4251-8ddc-1a187e0b7952 · outbound

This paper cites G-Eval: NLG evaluation using gpt-4 with better human alignment.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing G-Eval: NLG evaluation using gpt-4 with better human alignment

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.841722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.841722Z digest=sha256:b6102af977956ca2503f90c3cd1d1f4af036ace10ea81580b2d5b1b4ed68c2b1

Observation 49c0c03c-75b9-41ff-81c8-1c126a6ae26b · outbound

This paper cites an unresolved cited work.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:33.910460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:33.910460Z digest=sha256:752df2d35f8b9e915aa8d87da5b97eed4efc20031317d5db5c9567003c3cf6bf

Observation 1ebf3e82-6a7c-431d-8d72-8db30a0f133b · outbound

This paper cites Chrono: A simple blueprint for representing time in MLLMs.arXiv preprint arXiv:2406.18113, 2024.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Chrono: A simple blueprint for representing time in MLLMs.arXiv preprint arXiv:2406.18113, 2024

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:34.008992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:34.008992Z digest=sha256:9292c1224b6ef658c6d0601c437cc14b50b05bc1132640a50f115d7501143481

Observation ec5547e9-99dc-4c2c-b7f1-3d9f24c42fdf · outbound

This paper cites CLIP-It! language-guided video summarization.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing CLIP-It! language-guided video summarization

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:34.127674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:34.127674Z digest=sha256:74443b60134f860d545c844b224bf808f2cfbef280f14eb6fabcab2e9cd12161

Observation b48724bc-2013-4ce5-afd3-1dd42a55a6ee · outbound

This paper cites Rethinking the evaluation of video summaries.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Rethinking the evaluation of video summaries

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:34.285272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:34.285272Z digest=sha256:f9d3647b29838cdf72afa555d5fc5a6b7e512e5f3cf4e8ce0b74f967c37db99b

Observation f334b4cd-eb25-4f6f-bb75-206a797ed8ca · outbound

This paper cites Contrastive losses are natural criteria for unsu- pervised video summarization.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Contrastive losses are natural criteria for unsu- pervised video summarization

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:34.407233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:34.407233Z digest=sha256:89d698040e46cd8ababdbad538db90f2739e0533797a3a661b4ed45b980264cd

Observation 6b51240d-382e-4790-b7b0-646b837f855a · outbound

This paper cites Mea- sure Twice, Cut Once: A semantic-oriented approach to video temporal localization with video LLMs.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Mea- sure Twice, Cut Once: A semantic-oriented approach to video temporal localization with video LLMs

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:34.538157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:34.538157Z digest=sha256:f0722b606fa5fa72ccb7b38aa53b37eb35d28d4a7d31a9e2ba65dd5529f3f90c

Observation d5debe9c-38ca-407e-bc5a-c918a90f6f4f · outbound

This paper cites MovieCuts: A new dataset and benchmark for cut type recognition.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing MovieCuts: A new dataset and benchmark for cut type recognition

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:34.652076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:34.652076Z digest=sha256:7480fb82f588195fc097e58b511e072e4bdcf8de9bdea1c254a6b327f74b92c2

Observation 5922fd12-7f72-49b0-ae3d-b7d02d6b5cc7 · outbound

This paper cites Generative timelines for instructed visual assembly.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Generative timelines for instructed visual assembly

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:34.724308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:34.724308Z digest=sha256:a14ba505dcc16f1d0d1361e00f6e718c13a5bcdb4ce1843da2e6e6f56e7cb634

Observation 340a0024-cb23-46fc-ad60-2e35af0d907e · outbound

This paper cites MMSum: A dataset for multimodal summarization and thumbnail gen- eration of videos.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing MMSum: A dataset for multimodal summarization and thumbnail gen- eration of videos

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:34.884059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:34.884059Z digest=sha256:efcebb3811f72bb0c21396925b3047eeb5e1bcf29e29334a181f6e91bae45b2c

Observation 0449790b-45b0-41e3-9677-b1c772a3243a · outbound

This paper cites TimeChat: A time-sensitive multimodal large language model for long video understanding.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing TimeChat: A time-sensitive multimodal large language model for long video understanding

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:34.950203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:34.950203Z digest=sha256:c9c9f814aa00ad68cddf1e666db0e54d1a5ff14e3e638d36070c7c04814e13b0

Observation 41add252-88da-4500-82b0-0b2e80965a3d · outbound

This paper cites an unresolved cited work.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.052912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.052912Z digest=sha256:fbc67d81756315603d0877cfca29841b6c974204e2253faf8c5055d5a2c5a1d0

Observation 71a2a1fc-0ad2-49c5-b9d6-3f60a8bfdb18 · outbound

This paper cites Judging the judges: A system- atic study of position bias in LLM-as-a-Judge.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Judging the judges: A system- atic study of position bias in LLM-as-a-Judge

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.119099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.119099Z digest=sha256:204f96f9ce614867f3b72c79998287441001655e9ede5240e7a4c5db85b83548

Observation 05e84f1d-c50e-42e2-abff-77a382ff51e9 · outbound

This paper cites Generic event boundary de- tection: A benchmark for event segmentation.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Generic event boundary de- tection: A benchmark for event segmentation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.224975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.224975Z digest=sha256:23aae4bf33b812b35b1fc625a7eb50abb5d71efab9476e239174522d0586f3e8

Observation 129aa3b8-ae6f-44d0-b473-8fe1787e0a99 · outbound

This paper cites CSTA: Cnn- based spatiotemporal attention for video summarization.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing CSTA: Cnn- based spatiotemporal attention for video summarization

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.325730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.325730Z digest=sha256:13a0a7653c256004243331a33b386da5e0fcd2ae902fd4bc4db4978d4859e22b

Observation 79796f32-ba54-46e4-9a8b-1e47c22205b8 · outbound

This paper cites Csta: Cnn- based spatiotemporal attention for video summarization.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Csta: Cnn- based spatiotemporal attention for video summarization

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.330070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.330070Z digest=sha256:32515426fae74ffcefdcb288ea764ca19fd4b8d5ba01b1f2b4e42212e25f33a3

Observation e7cc1cd1-8c6a-4491-aa08-5d0f4bdeeb89 · outbound

This paper cites TVSum: Summarizing web videos using titles.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing TVSum: Summarizing web videos using titles

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.334102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.334102Z digest=sha256:efc6efd4e45bba431bc8dd19f5fd1a281d0061c46a3a864a1769a1af9714e8c2

Observation 4094f2cd-4abe-4731-b9d8-8afde550e466 · outbound

This paper cites Language-guided self-supervised video summarization using text semantic matching considering the diversity of the video.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Language-guided self-supervised video summarization using text semantic matching considering the diversity of the video

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.338185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.338185Z digest=sha256:b386bf0a39914bd2a2026add34bc07241a3e564a7a7a67924fb561adba4e253c

Observation 04c50e48-d688-4b89-b878-b6073eca71b3 · outbound

This paper cites Gemini: A family of highly capable multimodal models, 2025.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Gemini: A family of highly capable multimodal models, 2025

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.342337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.342337Z digest=sha256:f4664a701c4e82accfa1e28ded5cf7b1b7a85213c61f16b95656ef8c757d1c88

Observation 9aeab3b7-f291-467e-9636-b198e9faf1b7 · outbound

This paper cites QuickCut: An interactive tool for editing narrated video.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing QuickCut: An interactive tool for editing narrated video

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.347096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.347096Z digest=sha256:5f744b2cf2e91114d9641500c69ef657343451f093c807c9ac018ccdecdf96ae

Observation 91f8cce7-7d4a-472f-a3f5-f575e5e5fff6 · outbound

This paper cites Query Twice: Dual mixture attention meta learning for video summarization.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Query Twice: Dual mixture attention meta learning for video summarization

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.351267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.351267Z digest=sha256:f8b6bb7ed4751937338b1eb16bb744bc39701899430152db1042007061ac4ab4

Observation 3fb8ca8c-cf18-4fcf-839d-204a33781b98 · outbound

This paper cites Write-A-Video: Computational video montage from themed text.ACM Trans.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Write-A-Video: Computational video montage from themed text.ACM Trans

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.356284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.356284Z digest=sha256:d4a1cbb00ff90365c3895767b12b7d48067ecf9fbd4afa737700697b65b648ab

Observation c2aa9254-6cd7-4399-80e8-efc9746a1a6f · outbound

This paper cites InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.360849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.360849Z digest=sha256:72e77a431b1e83255643586b0d033fa89100ba15a41378ec587bdc5c8b42b823

Observation a1397759-3ae4-4575-a7bc-fe20cf4c8182 · outbound

This paper cites Time-R1: Post-training large vision language model for temporal video grounding.NeurIPS, 38:83330– 83364, 2026.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Time-R1: Post-training large vision language model for temporal video grounding.NeurIPS, 38:83330– 83364, 2026

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.364941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.364941Z digest=sha256:7a5a1e504fd7484687d48d0a86af4826a4a4dfcc497bc637fbb9d30dec3853a3

Observation a07e9033-6de5-465c-8aff-cee5b8e24bde · outbound

This paper cites LongVideoBench: a benchmark for long-context interleaved video-language understanding.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing LongVideoBench: a benchmark for long-context interleaved video-language understanding

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.368984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.368984Z digest=sha256:54b49d7f3316a9b2e6fe11255b35deea4f20d3324cb245e035bbbe6eae5d6999

Observation 2a774dfc-600c-434f-899c-2654861e571b · outbound

This paper cites Transcript to Video: Efficient clip sequencing from texts.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Transcript to Video: Efficient clip sequencing from texts

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.372928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.372928Z digest=sha256:011aa62a390d0ddfaff1528097e5218289dae96ea63f7eca45eb198b18e24864

Observation 000544e0-d069-4b55-a5de-b07b9c020bcb · outbound

This paper cites Vid2seq: Large-scale pretraining of a vi- sual language model for dense video captioning.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Vid2seq: Large-scale pretraining of a vi- sual language model for dense video captioning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.376928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.376928Z digest=sha256:dcd3833ecfa1bfd0c00d9f8cc943d28db48693b020a536678be1a5244dedb120

Observation d2dd359e-a90e-45f9-bfb7-3443b4f3b7f7 · outbound

This paper cites TimeLens: Rethinking video temporal grounding with multimodal LLMs.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing TimeLens: Rethinking video temporal grounding with multimodal LLMs

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.380934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.380934Z digest=sha256:763120807afef871bf185d77ed589971266241c23562eb7ff6a688d6e66c193c

Observation 355f80e7-92a6-489f-bca3-d1ee81bf4c9d · outbound

This paper cites GPT-4V(ision) as a Generalist Evaluator for Vision-Language Tasks.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing GPT-4V(ision) as a Generalist Evaluator for Vision-Language Tasks

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.385182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.385182Z digest=sha256:3e296f4ed104a451357b9cda6021e2f7b3cda54c9ccc6037c621f4aa90dfc33f

Observation 6d41d762-6f0e-471f-bb80-5be68a0be5f8 · outbound

This paper cites Re- constructive sequence-graph network for video summariza- tion.IEEE Trans.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Re- constructive sequence-graph network for video summariza- tion.IEEE Trans

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.389309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.389309Z digest=sha256:9dda235f5917cca40e44978b6bedddbe30cbcfb7f3818e4ed02d9f97139ca196

Observation 1c464df8-65b5-4080-a5f1-2c97c1e014db · outbound

This paper cites Xing, Hao Zhang, Joseph E.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Xing, Hao Zhang, Joseph E

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.393918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.393918Z digest=sha256:37dff507692d0dbbd042b51b3dfb87a9916228a1ca7794f5648e8a209b2e1a9b

Observation 38eaef64-af30-48cd-ab3a-81913c92698b · outbound

This paper cites Deep semantic and attentive network for unsupervised video summarization.ACM Trans.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Deep semantic and attentive network for unsupervised video summarization.ACM Trans

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.397619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.397619Z digest=sha256:002d2c66c0176aa39e140bdd04c5f9d5ecac816f7a37248cff3d9f2aa28b37d1

Observation 4d47acc8-b7f2-4d90-8e10-8ea365fe22cc · outbound

This paper cites MLVU: Benchmarking multi-task long video understanding.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing MLVU: Benchmarking multi-task long video understanding

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.401484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.401484Z digest=sha256:156c9dbd01e899dc4aecc9465cfd4c447b32f27cb5c5ea265229ff12d44581bc

Observation 975ca5e2-7685-46d7-9015-5f25ee07038b · outbound

This paper cites Deep reinforce- ment learning for unsupervised video summarization with diversity-representativeness reward.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Deep reinforce- ment learning for unsupervised video summarization with diversity-representativeness reward

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.405324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.405324Z digest=sha256:1c3bfec014873f02568ae1739f982ebc5ae60e7f0aa8f1da79fe4e6c8511e6fa

Observation a73ca798-edcd-4bac-baef-7dc7cf444de0 · outbound

This paper cites Edits” counts edits with at least one temporal reversal; “Cuts.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Edits” counts edits with at least one temporal reversal; “Cuts

Reference 64

Resolution
malformed identifier
no resolver link, observed 2026-08-01T02:57:35.408991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.408991Z digest=sha256:d82043ebb0df8f7c515ee2977ba7700d240f318f5b34b0a85cb01d5faefb3a9a

Observation ec6c8238-0f87-4c20-b4c3-11eabcdd9401 · outbound

This paper cites an unresolved cited work.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing Unresolved cited work

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.413753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.413753Z digest=sha256:14b34bf9701a0dc00abadcffebfa69ebf3f162d514e981505b2e27efb5a9f133

Observation 5f239f22-9ad7-4f3d-83af-9b65558bdbf3 · outbound

This paper cites 01:30:50.

MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing 01:30:50

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:35.417802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:35.417802Z digest=sha256:452b994bfb00013abd43aa626828ea5b221ad876efc317a4def2d771e3c61043

Pith citing papers

No inbound Pith citation observations are available.