Pith. sign in

Paper Citation Record · LEDGER

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting

As of 7 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 1 inbound Pith citation observation for arXiv:2507.16873.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.16873 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:17:58.243690Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-20T00:55:59.857014Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T00:57:54.247973Z

Reference resolution

60 of 60 outbound references displayed

  • verified exact1
  • verified fuzzy40
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 111e67ee-b96a-4a46-9964-b9ba7fe484d4 · outbound

This paper cites GPT-4 Technical Report.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.694812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.694812Z digest=sha256:d2f5058c8f61c35bdadf82058a84facfad578d31bdc83a84dda5867b692a2aaa

Observation 14b28a1e-c8e4-472a-b772-576670cb78ee · outbound

This paper cites Video summarization using deep neural networks: A survey.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Video summarization using deep neural networks: A survey

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.559165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.706308Z digest=sha256:776be3f343c5955ca5cf78460308c5613d737488ef6f34d65ba49231677c9250

Observation 4f93a273-ade6-4024-b885-94c9a4835eff · outbound

This paper cites Towards automated movie trailer generation.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Towards automated movie trailer generation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.526219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.715747Z digest=sha256:67736622b9e837b741f63fd4737cb791e156869001051e8eae05369904712f38

Observation f06a89c7-b8a4-444d-a8e3-b2d4a61b9262 · outbound

This paper cites Scaling up video summarization pretraining with large language models.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Scaling up video summarization pretraining with large language models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.494069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.732501Z digest=sha256:d10f0167156122c4c9786b3c00035c6e45e906ae3dc0caf6edce8ed96b2813f6

Observation 9d250bf7-c4b0-4f42-a71a-659e9ba1c1da · outbound

This paper cites Fast-DetectGPT: Efficient Zero-Shot Detection of Machine-Generated Text via Conditional Probability Curvature.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Fast-DetectGPT: Efficient Zero-Shot Detection of Machine-Generated Text via Conditional Probability Curvature

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.742908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.742908Z digest=sha256:e72fa662a3ad4711942ba6a87617ace238c590aa254a94c66ea21d218b090c3f

Observation c5e70f11-e328-4a11-9644-a2a5d94b2764 · outbound

This paper cites End-to-end object detection with transformers.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting End-to-end object detection with transformers

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.754625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.754625Z digest=sha256:1d9c925ecac4b31faa736403c131c6f9e29bf95bd5438994c38cddc4edb0fd83

Observation 971a1022-1cae-47db-bb32-8a8d388802f9 · outbound

This paper cites Personalized video summarization by multimodal video understanding.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Personalized video summarization by multimodal video understanding

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.446030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.772859Z digest=sha256:026ab675354779f5cda91ebba5c2f0399d4803754a6e86cb144842652ded77b4

Observation 2df36854-e189-4c3a-bb60-8d9fccec8895 · outbound

This paper cites Chatbot arena: An open platform for evaluating llms by human preference.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Chatbot arena: An open platform for evaluating llms by human preference

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.780042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.780042Z digest=sha256:048521c381c03e03d127c64003dffc4a4ea69bf3b9a35dd114e10d5f33114c6b

Observation 2f9c3627-2e56-4a5e-84c9-3cce8d9556ee · outbound

This paper cites Tall: Temporal activity localization via language query.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Tall: Temporal activity localization via language query

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.397974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.794841Z digest=sha256:56520552395cb5eb4c471abc78baeb056e238d1df67b5605f68b1469e3a4927f

Observation 6f3e2665-8d99-4161-b88d-a1bbf8e26051 · outbound

This paper cites Creating summaries from user videos.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Creating summaries from user videos

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.361373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.805901Z digest=sha256:2cce434d1d29e59fbeff56652710a15de85aa308cd52b2a6a5ec551da3b78768

Observation 2e82b1b4-5224-40cd-9210-3ccff6cd019a · outbound

This paper cites Video2gif: Automatic generation of animated gifs from video.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Video2gif: Automatic generation of animated gifs from video

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.314682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.813044Z digest=sha256:28d53af9999011e94ffbf7d14a0b989c14d9ac1c99041a48f45ad2e705cb2b64

Observation 6b864fc9-9403-4a05-a35b-e0262781c528 · outbound

This paper cites Shot2Story: A New Benchmark for Comprehensive Understanding of Multi-shot Videos.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Shot2Story: A New Benchmark for Comprehensive Understanding of Multi-shot Videos

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.824025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.824025Z digest=sha256:5e313eddede195c0a4c997a62499c65e6c4de6e98eb9b982ae0c651f37bd18c4

Observation 5b936faf-0f4c-404c-995e-7bdbeb775617 · outbound

This paper cites V2xum-llm: Cross-modal video summarization with temporal prompt instruction tuning.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting V2xum-llm: Cross-modal video summarization with temporal prompt instruction tuning

Reference 13

Resolution
verified exact
raw_fallback, observed 2026-08-06T15:17:58.535719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.833243Z digest=sha256:e39cce6a4cc650e54137ac4989be1c45c709f6417960e3b612fa634781cbcd60

Observation 178f61e5-5ca4-4ccb-bd6a-6b78c0c7cf7e · outbound

This paper cites Movienet: A holistic dataset for movie understanding.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Movienet: A holistic dataset for movie understanding

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.274641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.850713Z digest=sha256:372e8886ef59dc29a7ec5e608b0ead0064968b8b7c4fdf0252d11da81d21b31d

Observation 100bc01e-0ebb-4d22-a71c-1079ae8835d8 · outbound

This paper cites Video summarization with attention-based encoder--decoder networks.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Video summarization with attention-based encoder--decoder networks

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.233683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.866956Z digest=sha256:5af67095e4293ded4b41500dbd4aec2820d8eaae3f938c84d9e2ff4222285c8d

Observation b5e92f8f-3ef2-4c14-a09d-c8b297af2f97 · outbound

This paper cites Mdetr-modulated detection for end-to-end multi-modal understanding.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Mdetr-modulated detection for end-to-end multi-modal understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.878808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.878808Z digest=sha256:55a374da2ffe0ebc6ae4e49883d68f4ef60d92ef7eebe05827aaeaa2a9e34014

Observation 78658a8a-e7bf-41e2-9f1d-5f2293280a20 · outbound

This paper cites Self-attentive sequential recommendation.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Self-attentive sequential recommendation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.889332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.889332Z digest=sha256:cf3061480bcbe24f650f46817061db1fe856473f3656237d4161a6d198cadb4f

Observation 34c922f3-384c-43b9-822b-ac4b3f72e2e4 · outbound

This paper cites Tvr: A large-scale dataset for video-subtitle moment retrieval.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Tvr: A large-scale dataset for video-subtitle moment retrieval

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.099995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.894979Z digest=sha256:c28c46315c9a4ab38f4072567a4d34ba566ae276f3b2ffaf8bdf81ebd16379cd

Observation 92b3b6c7-adb3-4da3-8832-cbbe029447bf · outbound

This paper cites Detecting moments and highlights in videos via natural language queries.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Detecting moments and highlights in videos via natural language queries

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.900813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.900813Z digest=sha256:ca5f900ffaa2a44224bab4c577bb5e0f5bfe33e8900adb185c7c178350f06ad4

Observation 1186ada7-6db8-4df6-9898-174df4cae153 · outbound

This paper cites HERO: Hierarchical Encoder for Video+Language Omni-representation Pre-training.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting HERO: Hierarchical Encoder for Video+Language Omni-representation Pre-training

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.908906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.908906Z digest=sha256:22cbaeddc0394f987c9c0b80f3068263b4b286cc292492f5f9fda8e119cd88a7

Observation 2fffdc20-b80c-4b42-ad59-a56ad77212c1 · outbound

This paper cites Univtg: Towards unified video-language temporal grounding.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Univtg: Towards unified video-language temporal grounding

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.005787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.915019Z digest=sha256:8ef43023ebec0d756325738b6df1db998b74b04f531fc161e6bc4b40e4b25fcd

Observation d74afcc3-c1aa-4ae7-93cd-14992b383a2f · outbound

This paper cites Visual instruction tuning.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Visual instruction tuning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.923338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.923338Z digest=sha256:d2ba7af32044aac11c72c87940059bd0bb63741ce7a3debe68f2bdc029dd9bf2

Observation fbe27845-2ba9-4e8f-afa8-131644f5e416 · outbound

This paper cites Attentive moment retrieval in videos.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Attentive moment retrieval in videos

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.942980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.930475Z digest=sha256:70832ad4d135584e2565e16521720eee6451dc76f94325dc574ed651cf85539e

Observation cb57efb1-20af-4de0-9b32-4c036c771603 · outbound

This paper cites Multi-task deep visual-semantic embedding for video thumbnail selection.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Multi-task deep visual-semantic embedding for video thumbnail selection

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.916236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.938843Z digest=sha256:db21b41453536ade8d13aaabbe27ebe0d1cfa5a764b73bd01ddd0b1abc45c0f4

Observation 200f224b-e477-4969-b44a-eda34a757df2 · outbound

This paper cites Umt: Unified multi-modal transformers for joint video moment retrieval and highlight detection.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Umt: Unified multi-modal transformers for joint video moment retrieval and highlight detection

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.894062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.950702Z digest=sha256:2fea3c10ee97af43cb83a0491ad55b7661991667d576a60b008ab7af83122006

Observation 67d4abdc-b013-481e-acd5-af011be4bf6e · outbound

This paper cites Debug: A dense bottom-up grounding approach for natural language video localization.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Debug: A dense bottom-up grounding approach for natural language video localization

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.858277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.959213Z digest=sha256:e12897d6bf0118c43769161b13ffbfec5e981080b86145f49f029d0561e19842

Observation 14265216-7375-4778-b9c6-da0ad1b99330 · outbound

This paper cites Videoautoarena: An automated arena for evaluating large multimodal models in video analysis through user simulation.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Videoautoarena: An automated arena for evaluating large multimodal models in video analysis through user simulation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.813663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.968763Z digest=sha256:0cc0c021a4ac31093106577bd776c02734885a9bb7e2211cd50e80f2ab86ea0e

Observation 6370adc4-c383-4e9f-aa75-86741df5413c · outbound

This paper cites Howto100m: Learning a text-video embedding by watching hundred million narrated video clips.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Howto100m: Learning a text-video embedding by watching hundred million narrated video clips

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.978108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.978108Z digest=sha256:e34ab5e514af35e498208e57905a2d12d35b79856ec9973b0dd719efedc8610c

Observation 13ca2bd9-7318-45af-a15f-56f231f4b646 · outbound

This paper cites Detectgpt: Zero-shot machine-generated text detection using probability curvature.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Detectgpt: Zero-shot machine-generated text detection using probability curvature

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.747306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.986161Z digest=sha256:08d0b2d8a9b367e509fe48e588fb50da263012e0a2fac0d8df18f2b9c024fbb3

Observation 12fcf8e1-6451-4820-ac18-dff6ed7fb08f · outbound

This paper cites Query-dependent video representation for moment retrieval and highlight detection.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Query-dependent video representation for moment retrieval and highlight detection

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.706354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.997352Z digest=sha256:91c582362505bc980bcd435e2aee15fac2a446bdea5b8f75702197242a149660

Observation cf66e6ea-cd8a-4ed0-b290-2fcea3507307 · outbound

This paper cites Clip-it! language-guided video summarization.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Clip-it! language-guided video summarization

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.673686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.005873Z digest=sha256:a6b848ca0a6d1095ae21f06a17eba3ab98da459be4d18b59a8505217796f197b

Observation 69e54f2b-58c6-4089-b36f-483ebc85c181 · outbound

This paper cites Sumgraph: Video summarization via recursive graph modeling.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Sumgraph: Video summarization via recursive graph modeling

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.632181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.011349Z digest=sha256:291240ecc19071cc20827a4a503631b1dd0a460560b953626a459b1f8e19e601

Observation 0fe70fca-d3e8-468f-bda8-39f3cd0bbfde · outbound

This paper cites Mmsum: A dataset for multimodal summarization and thumbnail generation of videos.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Mmsum: A dataset for multimodal summarization and thumbnail generation of videos

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.587725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.017617Z digest=sha256:5fc76c7bc5aa34e875a82c84145e87b9fd7bc6cb0efcddfd697ca7fe8e40bbd1

Observation 612b5a05-1947-4417-b6d6-ed88046ad67c · outbound

This paper cites Learning transferable visual models from natural language supervision.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Learning transferable visual models from natural language supervision

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:58.026255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:58.026255Z digest=sha256:e9db457834a5fe1a2127ffb05241562d61cdf3f1f539a5667fde9a11b61602d9

Observation 12a5c193-1394-4a55-9fd9-3744ba6d020d · outbound

This paper cites Bpr: Bayesian personalized ranking from implicit feedback.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Bpr: Bayesian personalized ranking from implicit feedback

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.508037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.034667Z digest=sha256:65da54fd21e3284e5eda866ff282a5a8e1a9060a548cc4c7f8c0f289b91448bb

Observation fbca70f4-3f80-48e4-90c8-484a13ac1b35 · outbound

This paper cites Adaptive video highlight detection by learning from user history.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Adaptive video highlight detection by learning from user history

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.462359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.042738Z digest=sha256:745024a4621ceb84307697feed28a3fa0dcacd06eba38fbb68560657dc995a78

Observation 778060dd-f944-4ab2-b973-c03a99ae4853 · outbound

This paper cites Query-focused extractive video summarization.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Query-focused extractive video summarization

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.415123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.048754Z digest=sha256:43bd5392021d04f856628f25bc072b53c6fe4eb10bd00e1d4bb075d1e453bfc3

Observation 068a8527-2123-4f8e-b6f3-a414943f5124 · outbound

This paper cites Query-focused video summarization: Dataset, evaluation, and a memory network based approach.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Query-focused video summarization: Dataset, evaluation, and a memory network based approach

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.379996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.054036Z digest=sha256:1a1d4b24ceee8da2de9bbf3738966ed2359a5cd3d5c5350fa6402202cb3e3912

Observation 15bbd598-720d-4cd9-87df-175cf218596e · outbound

This paper cites Tvsum: Summarizing web videos using titles.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Tvsum: Summarizing web videos using titles

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.342002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.061056Z digest=sha256:adef0c7f35e6e6afec458c4228574c5d4ecbde3fd54f097a359e1dcc19c0600b

Observation fcf77a06-11a7-4cb3-8e6f-e4e37d409487 · outbound

This paper cites To click or not to click: Automatic selection of beautiful thumbnails from videos.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting To click or not to click: Automatic selection of beautiful thumbnails from videos

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.312544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.067085Z digest=sha256:0f25f7319493572c142e0d998694b18b27ceabb50e281d70d9b1e4ca7bcdcb2c

Observation a6087fab-8d80-46c5-90a8-546fbb73ef42 · outbound

This paper cites an unresolved cited work.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:17:59.284673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.076964Z digest=sha256:0e1c049ea40b624cea8d96edbb3f54679fc48ea92bfbfccb016abbb7d6f6c5c0

Observation 1600e304-033b-4d22-a5a0-ccfea5bc1e24 · outbound

This paper cites Tr-detr: Task-reciprocal transformer for joint moment retrieval and highlight detection.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Tr-detr: Task-reciprocal transformer for joint moment retrieval and highlight detection

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.250454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.087219Z digest=sha256:f5aed35b548f22b1f37ffa648909eac81057820d05a2c1baa2db806191ccce76

Observation 4fedb22b-42da-4fab-a4e1-2b9bc923309d · outbound

This paper cites Ranking domain-specific highlights by analyzing edited videos.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Ranking domain-specific highlights by analyzing edited videos

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.215237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.099601Z digest=sha256:785eb515934768afea99ee1678b52cbce73a7a40573a9d481dd07f06b8bfe4bb

Observation 87556897-de50-49d6-ba8a-993fb85164d3 · outbound

This paper cites Query-adaptive video summarization via quality-aware relevance estimation.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Query-adaptive video summarization via quality-aware relevance estimation

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.175888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.108373Z digest=sha256:f898c349acd1f178c4d7efb6b2206a866499e80cd612d9bfb041234fe32d120b

Observation 11843b59-aebc-42d7-a3f9-d711aec94c8c · outbound

This paper cites Videoagent: Long-form video understanding with large language model as agent.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Videoagent: Long-form video understanding with large language model as agent

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:58.113681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:58.113681Z digest=sha256:4d4d779d31d7fe16f3b84a40b9f6ad55ae7526ebeca41fa0b09ffa68523d0753

Observation 5b6a1701-91b5-4a9e-a64b-4b4a2456f62c · outbound

This paper cites Query-biased self-attentive network for query-focused video summarization.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Query-biased self-attentive network for query-focused video summarization

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.096946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.119440Z digest=sha256:3f76d7207a6eb337f5270fe2bb9d05c35c70dfc416925b3a66cd1dee36484246

Observation 4b5ea4cb-bc56-48d0-b451-09a8d639da2e · outbound

This paper cites Convolutional hierarchical attention network for query-focused video summarization.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Convolutional hierarchical attention network for query-focused video summarization

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.053131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.126430Z digest=sha256:e38a83ac6f5061618f9f7e93081d78b94ca2354f11922fdbf89b7b5976132b88

Observation fdfec61e-997d-4624-8fe8-0aaf263c8ae3 · outbound

This paper cites Bridging the gap: A unified video comprehension framework for moment retrieval and highlight detection.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Bridging the gap: A unified video comprehension framework for moment retrieval and highlight detection

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.017479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.135930Z digest=sha256:0e9fdc0a874a67272d57bc843013f5060cf81022353db983e48f34d45916b24c

Observation a63772b5-cfef-4cb1-8124-3312178cc852 · outbound

This paper cites Cross-category video highlight detection via set-based learning.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Cross-category video highlight detection via set-based learning

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:58.977655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.142535Z digest=sha256:94f47c979559c1b4bc1f448eec49a243ba29a7d5ed3d774a9e6d40d82fcecc89

Observation 150ced57-45b7-4bec-a253-c45bab977bf5 · outbound

This paper cites Mh-detr: Video moment and highlight detection with cross-modal transformer.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Mh-detr: Video moment and highlight detection with cross-modal transformer

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:58.936518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.147689Z digest=sha256:30e7746fada9a144d4caee65151866bfb2280f3c9a34aaf999a24d3ae1095fd3

Observation ed1d3f71-31c3-44a2-8ad1-67beec4c60e7 · outbound

This paper cites Highlight detection with pairwise deep ranking for first-person video summarization.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Highlight detection with pairwise deep ranking for first-person video summarization

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:58.909446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.153170Z digest=sha256:441bd24a82e2815bfe14e9b964078583940dfae4b0a30ddfc894a680bb9048eb

Observation 1326983a-f5e9-476d-af92-e90ef741f5d4 · outbound

This paper cites Semantic conditioned dynamic modulation for temporal sentence grounding in videos.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Semantic conditioned dynamic modulation for temporal sentence grounding in videos

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:58.872839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.158832Z digest=sha256:d19bd74c7795694218eeb2888986da5312021aae8be6e7c778cf0cde645badb9

Observation bf2b387c-6436-4f76-a3e4-5da413d9b1f2 · outbound

This paper cites Hierarchical video-moment retrieval and step-captioning.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Hierarchical video-moment retrieval and step-captioning

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:58.836105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.165948Z digest=sha256:b115de7d01f59473fd05479a48af7b15b7b3835f8852248c5c42b80e548b219d

Observation 04726dc6-2e68-4748-91af-cb19249712e3 · outbound

This paper cites Moment is important: Language-based video moment retrieval via adversarial learning.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Moment is important: Language-based video moment retrieval via adversarial learning

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:58.809922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.174479Z digest=sha256:89eb0e017b9340ec4199f41c00bace98fae9a172405fdae61c2701a7c70f521e

Observation 8d2e8ea1-3eb3-4cce-b704-39e7b0c783fe · outbound

This paper cites Span-based Localizing Network for Natural Language Video Localization.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Span-based Localizing Network for Natural Language Video Localization

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:58.184421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:58.184421Z digest=sha256:b44ac7293beb5d67bfc566cd6eb4eca223e5d18d8063621cd208b8fb2b764898

Observation 3a048f91-13ce-4e15-b481-71f3410b713f · outbound

This paper cites Towards automatic learning of procedures from web instructional videos.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Towards automatic learning of procedures from web instructional videos

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:58.779679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.193137Z digest=sha256:a5fb3fbfd99717074530a96416cb8b1413bfb5f1037cf37cf7faa96cac2c0f69

Observation 8eb3311f-c2fb-43f5-899b-7f6807e76280 · outbound

This paper cites write newline.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting write newline

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:58.209470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:58.209470Z digest=sha256:5158d622cd5eb687da7ffb9485f22db2b72812138b32676c4882d5245ed3a474

Observation a8d73be5-de15-45fd-bba3-b628a287f923 · outbound

This paper cites @esa (Ref.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting @esa (Ref

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:58.220624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:58.220624Z digest=sha256:90a4e7d153b75096512e9c01e37e236fe86a1858b1856f42580feca5afbbf96b

Observation 221f98c4-a519-4b45-af4a-53619fe51ada · outbound

This paper cites an unresolved cited work.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Unresolved cited work

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:58.235989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:58.235989Z digest=sha256:766a34e94f6f758015c5ddd7c76979cb7b684c8b445a99d21b80c96544df072a

Observation ffbfffb5-013f-445d-90c2-45a901683c77 · outbound

This paper cites an unresolved cited work.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Unresolved cited work

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:58.243690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:58.243690Z digest=sha256:3cc25b302a6765052945b8148b244f18e08a961239837233d834b02ce8ed2757

Pith citing papers

Observation 07909adf-c8b8-4550-8409-6fd6794033a1 · inbound

Will It Go Viral? Grounding Micro-Video Popularity Prediction on the Open Web cites this paper.

Will It Go Viral? Grounding Micro-Video Popularity Prediction on the Open Web HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-20T00:57:54.249989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T00:55:59.857014Z digest=sha256:759619a83aaa3f0e5a36ae33ea620ca1c7cb642ca91532002ca5b13c25e97da5