Pith. sign in

Paper Citation Record · LEDGER

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering

As of 18 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 0 inbound Pith citation observations for arXiv:2608.01660.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.01660 v1

Coverage vector

measured 75 of 75 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:08:38.986063Z

measured 75 of 75 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

75 of 75 outbound references displayed

  • verified exact1
  • verified fuzzy28
  • unresolved46
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3e07367e-ad2f-4286-8965-4b457e004f4d · outbound

This paper cites European Conference on Computer Vision , pages=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering European Conference on Computer Vision , pages=

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.715296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.137288Z digest=sha256:aca6d80308337d1130ef16d3c38194700eca2cce74054a4285446d1d15aa7d57

Observation f80d940b-fb5f-4d60-b764-3c00b82a9c30 · outbound

This paper cites 2023 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2023 , eprint=

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.236299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.236299Z digest=sha256:9d5ca2b6d2db43800f5366a0cdb29294d6e0cc0ac6a52564df3e50c16d26dfcc

Observation f1d5d609-1817-430d-b93d-2502aa81b6c0 · outbound

This paper cites 2025 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2025 , eprint=

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.699847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.240854Z digest=sha256:3e9536f5ada40bb8565a4e6b90153cf45e35c46f44cdc511072696de7d0cefbf

Observation 4c8f8d32-5713-481d-8617-51a6e8426dcb · outbound

This paper cites MovieChat+: Question-Aware Sparse Memory for Long Video Question Answering , year=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering MovieChat+: Question-Aware Sparse Memory for Long Video Question Answering , year=

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.688778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.244355Z digest=sha256:f031ff1189f260bfc533bd8dda09da07944bb08b74f3303548546ee2174c7b5d

Observation 30e998dd-7757-403b-a696-8cd1c87e8de0 · outbound

This paper cites 2026 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2026 , eprint=

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.556579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.248378Z digest=sha256:9052346d9270cb38929649c47b6474b2260ba52cacb4eb37006deca234502407

Observation f87b3c8b-b499-480b-986b-fe8454df5881 · outbound

This paper cites Proceedings of the 21st annual international ACM SIGIR conference on Research and development in information retrieval , pages=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the 21st annual international ACM SIGIR conference on Research and development in information retrieval , pages=

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.257284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.257284Z digest=sha256:834ea2a1383c46029412284adbeed963dccaa8741a64a10f13c2ba9e3c2dde44

Observation 731a362f-c9ea-41f9-a974-8dfd949df099 · outbound

This paper cites Classification Problem Solving.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Classification Problem Solving

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.369466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.369466Z digest=sha256:af558d872f9f2aafb961831bd35e9365856bb2deeaa6e944379e85b33ac97ff4

Observation 7fb517ed-fd23-488a-a795-59bd633fbe48 · outbound

This paper cites , title =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering , title =

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.450358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.450358Z digest=sha256:deb408d1019cba65accfb1420a8c1d38f5878c33431c68565d3f41008e42553e

Observation bcdd3945-3806-4e5c-8143-6df7a5e58a88 · outbound

This paper cites New Ways to Make Microcircuits Smaller---Duplicate Entry.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering New Ways to Make Microcircuits Smaller---Duplicate Entry

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.454244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.454244Z digest=sha256:25d2c8df3e7bb87b105eab8d9b3425e7511c1e0446de2c5ac38c28f03669fd9d

Observation 647134e6-4879-4caf-8f01-584f35045f15 · outbound

This paper cites Clancey and Glenn Rennels , abstract =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Clancey and Glenn Rennels , abstract =

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.469414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.469414Z digest=sha256:6c78ebf6bae76992c4d8520e10e0280e1ee189a2f3296a9a0aebd8c87cfb67ba

Observation 03e553fe-709a-4736-9ed4-5f3d6906e7b2 · outbound

This paper cites and Rennels, Glenn R.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering and Rennels, Glenn R

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.518806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.518806Z digest=sha256:b24fdcdbb04f15cce66562d2e58f8c10af2a687edaab7ab4ff71286562d10134

Observation 4eb3de7b-adc4-43bf-bfea-b75d9ec7ff45 · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.516542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.537659Z digest=sha256:30a592156c16f65c5d8deb7eeee97b71f01bc6d724dcd7320304deb89a4ca3e1

Observation 08d5c6c0-416a-457c-a5bd-ce61436bbdd4 · outbound

This paper cites Poligon: A System for Parallel Problem Solving.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Poligon: A System for Parallel Problem Solving

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.541546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.541546Z digest=sha256:5205da147fc0af75855029a352960d3cc322ddf162daa8ba61e528506b11365c

Observation 6398792e-5697-45d6-b6db-a6228abd5a73 · outbound

This paper cites Transfer of Rule-Based Expertise through a Tutorial Dialogue.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Transfer of Rule-Based Expertise through a Tutorial Dialogue

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.545441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.545441Z digest=sha256:e5b4f8152c0ae497b07a8ac9def32092363aa044d090b318a20734e2f8754214

Observation bbcc461a-59a8-48bd-bf56-dc6f6ae55936 · outbound

This paper cites The Engineering of Qualitative Models.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering The Engineering of Qualitative Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.549013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.549013Z digest=sha256:cbf261c43e96823b7ed8995486d463610dd1ba74c333f596d4d7caf651edeb2b

Observation 80a3d2a8-78a4-4a89-9fa7-e6cf7c4a5ebe · outbound

This paper cites 2023 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2023 , eprint=

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.552405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.552405Z digest=sha256:c05d98eb37a45ef858e33dce81c8e11975f45903fc80a72c4c8651f5d7dccb92

Observation 0b77ec0d-8f7f-48de-8bd0-2a1077c22345 · outbound

This paper cites Pluto: The 'Other' Red Planet.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Pluto: The 'Other' Red Planet

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.640023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.640023Z digest=sha256:19b11e2798aa024cd945fd0c4c066a6cf4ca12f05d0cde87d5d6f144fbb509e7

Observation 190f8748-9f09-47f7-81e2-b74a79508286 · outbound

This paper cites Proceedings of the 19th Conference of the European Chapter of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the 19th Conference of the European Chapter of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.428207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.756939Z digest=sha256:5dba98980f6e22d7d6a35707784bbbe0f96a60480cbfdd45df9154c5d331b847

Observation 8ece195d-8ce3-4ee6-82ad-c45dc7a5738f · outbound

This paper cites 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , year=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , year=

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.417185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.792374Z digest=sha256:250774da701321a334b4e29276f79296436979230a51f7df961dd64e99e734ce

Observation f6b9d767-4596-4ceb-9ac4-9f106a591760 · outbound

This paper cites Logic-in-Frames: Dynamic Keyframe Search via Visual Semantic-Logical Verification for Long Video Understanding , url =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Logic-in-Frames: Dynamic Keyframe Search via Visual Semantic-Logical Verification for Long Video Understanding , url =

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.287918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.799515Z digest=sha256:518d39babf273a1578e0cfe872b57a62b7948909bfc20a0d918e71d82aecc7b7

Observation 5ff192d5-0538-4ac6-bea1-732cd5b13155 · outbound

This paper cites 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , year=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , year=

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.277701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.802855Z digest=sha256:6a17feb8cb5fce66f21270a77d59f8e61c5d0d20d13842d0f8c1fa4d11265213

Observation bbccdefa-fb6a-4da3-af5c-978829d5e4b2 · outbound

This paper cites 2025 IEEE/CVF International Conference on Computer Vision (ICCV) , year=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2025 IEEE/CVF International Conference on Computer Vision (ICCV) , year=

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.174178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.806918Z digest=sha256:feb28c2293c9a07463851e89e6593299999354929ddfc7e08632208ee3e56adf

Observation 69f5769f-4d62-429f-b526-2fdffa94d299 · outbound

This paper cites European Conference on Computer Vision , pages=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering European Conference on Computer Vision , pages=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:37.810421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:37.810421Z digest=sha256:2d6a0098492716fd098e2c9fca776b80f44a98dfff97ba55bda21659c4bdfb1f

Observation bd063ae0-a4e3-4387-ba09-77952f165a2f · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , year=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , year=

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.055167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.900572Z digest=sha256:40a8cdf0fdbb1518466bccb538fcc3d4b474cfc247a3503c0b913370b2830cc2

Observation d3dc87d4-d1b2-40a7-b602-0dd1753bf505 · outbound

This paper cites 2026 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2026 , eprint=

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:41.043524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.949122Z digest=sha256:c80d70c71d17e15a09e5827bfa16e1362cac4752e6459558e5c0fdad2c4a666e

Observation 2b627819-3db4-45ef-b362-27c4c5a7b65a · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.926114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.953951Z digest=sha256:5207f2847a132a85310500f94d7858629a7afd8aea7bf5d44280b3a1c381a2f1

Observation cbe1ffc0-9a39-46c0-bac1-c287eff57ef1 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Findings , month =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Findings , month =

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.867359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.957665Z digest=sha256:ae9cfc7cb89a6130e717fb9da8e6cdf59c7a1c1ca94df7f1f4db476f65292596

Observation a472e074-83f3-4c82-bca6-5a29eb6d07bf · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.857178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.960941Z digest=sha256:3d3d329ee7a72f3ebfa14748d1b033b6d67baabbe73fbce424d1de7b00e852b9

Observation a876bf8e-81d5-4f2f-a4b4-c4fa558bb782 · outbound

This paper cites Proceedings of the 2023 conference on empirical methods in natural language processing: system demonstrations , pages=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the 2023 conference on empirical methods in natural language processing: system demonstrations , pages=

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.769058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:37.964397Z digest=sha256:99ba33a584f17c3fa378b12756a89aee1421632b2d07e351cd75a4bf9b58365a

Observation d1136dfb-e2ca-467b-9357-a2659b9d3e15 · outbound

This paper cites Proceedings of the 2024 conference on empirical methods in natural language processing , pages=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the 2024 conference on empirical methods in natural language processing , pages=

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.010486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.010486Z digest=sha256:3e49fd0499cc45fbb856953386834c464a747235b65cf6af28362b590d1773db

Observation 755be9eb-1221-46d8-83ac-6ae17c683be9 · outbound

This paper cites 2024 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2024 , eprint=

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.050499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.050499Z digest=sha256:6c7c16dae96e65f91747bf910e828a0067967f18575445392b451cb799bac892

Observation 4de4fe1b-bc01-438f-a749-415504653898 · outbound

This paper cites 2024 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2024 , eprint=

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.053955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.053955Z digest=sha256:5c8c512273fcb828b39d405b31efb550128e6a3c69756712ff857e88ce8b4ade

Observation c9de6a90-ad9f-4198-a57f-bb2a5ec12917 · outbound

This paper cites International Conference on Learning Representations , volume=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering International Conference on Learning Representations , volume=

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.651943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.057129Z digest=sha256:316eda29a6bcd68d16bc6f7b2f8c3d286ed253d0d7501b4e3d9ecbfb837538f6

Observation 02a3a939-f765-4fe5-8d5d-badca8281be6 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.641702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.061226Z digest=sha256:9c80cfe29cfe9f4bfba622298488dbb6f966abf8c62a23ee0b71fa256f019e9d

Observation a95c83d8-96f4-429a-922a-26b4f8f5e33d · outbound

This paper cites European Conference on Computer Vision , pages=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering European Conference on Computer Vision , pages=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.066004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.066004Z digest=sha256:4a8091269b44225f76f8b14c5889816942b3d5b070eda26f3c752185976ad155

Observation e457d4a3-4834-4943-a18f-918fa659dc98 · outbound

This paper cites 2024 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2024 , eprint=

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.103769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.103769Z digest=sha256:046d80f5a609a0f66b43a71d67e107ee6c0fb48530154dc1db4b2b6ad93d1532

Observation 29b91e4b-5a26-4ecf-a580-ac19761e7bb9 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.619177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.108522Z digest=sha256:96a98f938be526d7506b3b02e8f1ce946c30a454950bf7c13e6faae77e5f2068

Observation 32ad0c3f-bf94-46c8-8c1f-78e37955dde7 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.554916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.112045Z digest=sha256:9ed91a2dcdfd21182e38afa99f2d74415df0c24f9cb9bce1ab23b96dcc6d877b

Observation 8c46febb-1ea0-4954-9982-edb5a21eda24 · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.423055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.116085Z digest=sha256:a6e5ab4776346865ae93ce8578778a3ff931ce1c75f4a048c2c27c1f0e09adaf

Observation a6128d49-3434-457b-a0f0-1eff0c755f34 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.412791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.120013Z digest=sha256:4098c7f3f50c5b1278164af43f4ddab9cdb5a07d5ad958c3b8b9d284764d2190

Observation f622ae6d-5a07-41d3-a504-1e57066c6e82 · outbound

This paper cites Video Summarization with Long Short-Term Memory.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Video Summarization with Long Short-Term Memory

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.402464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.171862Z digest=sha256:7ffd28d6c0de2f1a5d34a4c5e4c14b996bb3edabf1a88ee86d7fc8b0baf2151a

Observation 1b339b60-a989-495d-bdd3-7c3639b04889 · outbound

This paper cites DSNet: A Flexible Detect-to-Summarize Network for Video Summarization , year=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering DSNet: A Flexible Detect-to-Summarize Network for Video Summarization , year=

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.362731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.216920Z digest=sha256:11674b6e102c12b5232de0943c03f9aff8b712ed63b403ebd3ed0cb4ac031121

Observation 4280af97-6077-4899-9eb9-a2c3e2428c90 · outbound

This paper cites IEEE Transactions on Multimedia , year=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering IEEE Transactions on Multimedia , year=

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.236814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.220440Z digest=sha256:f57be03c65b1ef6fac4832f70b7bf7a50c2166bc6f0cde12b4ab8c9eecb43887

Observation a02cea44-4834-4011-b474-6dbaab7c855e · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Advances in Neural Information Processing Systems , volume=

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.224302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.224302Z digest=sha256:74e768e22a6a758516b86ba794e30e64223d4f2400b0f8ed0e99038ae83bddee

Observation e4813424-0ceb-42cd-95ab-a72c4bbdb84a · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.291234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.291234Z digest=sha256:507712b1fae011b7f6358dbf6f061a990f8b38046aaf1698812f3b11dbecf841

Observation 487034f9-8c7f-4ede-bad8-b5b941a54730 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.095211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.320191Z digest=sha256:32ae77554e61f942f8e65a71e477a87c44b59d66a9fed09f86162fe360e593ae

Observation 044124bc-e69f-46e8-b97b-010ee995a53d · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.323888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.323888Z digest=sha256:aa8bc1898e566593d6d5a5bc24e139152bb48541fe495eb4bd6bfd98ed8c549f

Observation baf2a62f-6102-4d3e-b124-538c6fcc3707 · outbound

This paper cites and Han, Rilyn and Fei-Fei, Li and Xie, Saining , title =.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering and Han, Rilyn and Fei-Fei, Li and Xie, Saining , title =

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.327293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.327293Z digest=sha256:0da02ec832de22aa3af0960b88678db22485c5f82c4e2fffecc5897920da29d1

Observation 0875204e-35ff-42b8-89fd-1d1bc0801623 · outbound

This paper cites 2025 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2025 , eprint=

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:40.006036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.331591Z digest=sha256:ce60510e00a5ba3d46a2536c6de974ff07684c6b2478f449c1a4002a02dce199

Observation 628055f0-75db-4fb1-92d2-77cb86298db7 · outbound

This paper cites 2024 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2024 , eprint=

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.370879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.370879Z digest=sha256:210033027a3e1b3f04c5bee876eef30265ee757c6f94216f1ba1b5f44700ad53

Observation 6ae87d8b-7a8c-4712-a36c-41c5d8e70a2e · outbound

This paper cites 2025 , eprint=.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering 2025 , eprint=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.420696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.420696Z digest=sha256:a76d7ecde2d8a9449a5f28e0423dff22ee9ec19c86711acfd4e5c0a69082a8a8

Observation 095ad355-d31e-42e6-8e46-9a1ac0e03a70 · outbound

This paper cites Qwen2.5-VL Technical Report.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Qwen2.5-VL Technical Report

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.425523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.425523Z digest=sha256:24eac44cc95f2d5e7d2e36313114527b11a5b4b06ec24bc59e513b95600c8f90

Observation 4ec6e2ab-553a-45dc-aef0-18518d1d3ca6 · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.430210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.430210Z digest=sha256:dab5ef29011bcc70f39e542b904c73f0d2b56a2ab4a0238aff09d5dd5209a47b

Observation 6b71a19b-365f-48a4-9de3-854f69669728 · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.942750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.432835Z digest=sha256:f584c9042abb7ba4669d2b6b5193bebce788d8431c9daca56ccab785347aa3a7

Observation fb6d0876-1523-4d31-b80d-414336e16598 · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.859460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.498126Z digest=sha256:0818ee66270d55989728a6eed2f84c85a44e385fb23d7946923ca644b4f44ae8

Observation 5eae232a-d8a5-48bc-9ebf-1a15e34654de · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.501680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.501680Z digest=sha256:8472e94dbaf35ed698d01fc1dd8f3cfb4567e25a467cc473d69d2c640530283f

Observation b382ae23-13fc-4e43-b1bb-82ef6d08c681 · outbound

This paper cites Logic-in-Frames: Dynamic Keyframe Search via Visual Semantic-Logical Verification for Long Video Understanding.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Logic-in-Frames: Dynamic Keyframe Search via Visual Semantic-Logical Verification for Long Video Understanding

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.506143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.506143Z digest=sha256:94b899685652a389742520faf95eb6ff47eae1bf6e39e3c940d742852c5fbcf0

Observation d12aa015-26e6-4572-b517-efb079a0cbdb · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.843117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.510086Z digest=sha256:191514fd3ca703dac7f801e5f43965ff5381e040891555bdb7c374f77c8f8025

Observation 381cf954-3567-46fb-8117-99ec00fdf079 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering LLaVA-OneVision: Easy Visual Task Transfer

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.513658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.513658Z digest=sha256:152ab606e1e973eb23d3b48f8669bd92538e4c5fb23c86f85490f33481f5c3b3

Observation e2e61142-1020-484c-b388-4ae9353500c8 · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.517383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.517383Z digest=sha256:79ac72a45c4e14ed1fd4aee6e0fe84549a631d235e4e8e60e8458d6735f2fd42

Observation 21a33fce-14f1-4418-9798-8c8ba66c098f · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.751712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.618037Z digest=sha256:3913217be5a9e0b96d1a19d45281a66a6ee10cd49437785003b98e2aa351eda4

Observation 36d64ae1-cb5a-4418-aca5-3f300da8c1d0 · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.741430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.621710Z digest=sha256:518c60ea2cf07e2e698cbe60da6bbd4ff197325ff7971ea8c859f160852c1cfd

Observation fbb3f12b-3416-48e5-962a-50cfee2e3f3a · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.673024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.625207Z digest=sha256:9dd50d07b992062e59fa352fce8bc4fc85b7d5638efa29c60f5d6147d8bbf77c

Observation d24ab680-ee46-440c-8749-12b49c7398c2 · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.662222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.630154Z digest=sha256:53c175b140b75a58706a604ac71d836c8f8efab5ce9ff598967be0c12f38e744

Observation 0ae16a87-02bb-40df-bd00-4ed4458b3905 · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.652235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.634381Z digest=sha256:d5406c5d495bb441c4b5f02a4ccfa9c54381402d9aff6efb857800cfaa45f931

Observation 4fcb77a8-dc9d-4c0f-9cdc-768e84906422 · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.600497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.744451Z digest=sha256:b3e80c890c62907d46ecf89397f5d41278064ea1f500f3fa09ab2b47869ac8a5

Observation 88791888-b073-4630-b756-7d033096f2b4 · outbound

This paper cites Where to Focus: Query-Modulated Multimodal Keyframe Selection for Long Video Understanding.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Where to Focus: Query-Modulated Multimodal Keyframe Selection for Long Video Understanding

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.813415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.813415Z digest=sha256:a49079526f4a2339e0750615a90fd5f0c018ff8edcbd4342083b5e17b76db02e

Observation 8677b137-e454-49a3-8fc4-0ba484fe82ad · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.590616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.818920Z digest=sha256:76a2628d3f5268e7a7c4795dbdcc3ebf0a5ca38b53e581f3b6a5fdb0cd747ace

Observation a6bff458-a54b-4c55-bb7d-d4cb1917c3c3 · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.553321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.822766Z digest=sha256:d275e75e2264bf6c4b8fb0b528e616f5c50aa771b701ab520a9e75e1cffed29a

Observation c171ed28-6573-459d-8bb4-8da2e75989bb · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.827139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.827139Z digest=sha256:fd62a96bb0af85c162c944d54307a1595b0570d03a8556af036816d8b21018a9

Observation e70afab0-02a9-4ba5-bbf9-a8de4aaf6eae · outbound

This paper cites an unresolved cited work.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:08:39.491721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.906682Z digest=sha256:8ecd9f33994d1e34d99fbf0b276716ac66ddb36cad3ceef0d18f1c3cea90bc58

Observation 90499b80-f861-415c-bcd5-1beff92be612 · outbound

This paper cites MemoryCard: Topic-Aware Multi-Modal Clue Compression for Long-Video Question Answering.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering MemoryCard: Topic-Aware Multi-Modal Clue Compression for Long-Video Question Answering

Reference 73

Resolution
verified exact
local_arxiv, observed 2026-08-15T15:08:39.183218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.973063Z digest=sha256:d4de19e91726df4142b7bb443903589031d6ef5168cfb88d5c19e7db5db3c80c

Observation 29d4acbd-38bb-4148-8f5b-cd22878cb88c · outbound

This paper cites C.; Adeli, E.; Li, F.-F.; Wu, J.; and Li, M.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering C.; Adeli, E.; Li, F.-F.; Wu, J.; and Li, M

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:39.395109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T15:08:38.978117Z digest=sha256:9ed6125a8b58a7fdeca5d7b62299675f3a5744479034a6711f1baec83386a8cd

Observation b4d4bf7c-f593-4ed1-9dac-15e6d61e0d9c · outbound

This paper cites Sigmoid Loss for Language Image Pre-Training.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering Sigmoid Loss for Language Image Pre-Training

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.981773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.981773Z digest=sha256:667aefdcb6abfb59fab256575cce905b840ba37456e0acfac2c47acf83f3c81e

Observation 4c9df4ee-33a4-48cc-8ebb-15aa8ab5d755 · outbound

This paper cites LLaVA-Video: Video Instruction Tuning With Synthetic Data.

Ground, Cover, and Refine: Evidence-Centric Frame Selection for Long-Video Question Answering LLaVA-Video: Video Instruction Tuning With Synthetic Data

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:38.986063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:38.986063Z digest=sha256:a0166fe63c8881b46c775b6e766334cde9ee3db1507d1017b7ff1b62cad4fc46

Pith citing papers

No inbound Pith citation observations are available.