Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:20:36.814781Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 70 of 70 outbound references and 1 inbound Pith citation observation for arXiv:2504.15918.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:20:36.814781Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:55:40.858859Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T19:55:41.494044Z
70 of 70 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 374a4168-67ea-4a1e-869f-e87f082f25bd · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Let the llms talk: Simulating human-to-human conversational qa via zero-shot llm-to-llm interactions
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 22db37f0-4f1e-4017-8f82-570e4c6272eb · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Gpt-3-driven pedagogical agents to train chil- dren’s curious question-asking skills
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7f313568-e6ce-4351-9666-bc3fd9f7be53 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Synthetic dialogue dataset generation using LLM agents
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c53b39f3-0777-4c6c-ae17-307129651210 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Self-RAG: Learning to retrieve, gener- ate, and critique through self-reflection
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b246b9ed-ed5c-4c4f-8c0b-7542e5664981 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Vindlu: A recipe for ef- fective video-and-language pretraining
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f78488ca-a08c-4610-860a-fa13bf3b7da9 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Bert: Pre-training of deep bidirectional trans- formers for language understanding
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d4e2d27b-0db0-4ee8-8de9-4fe33dd54639 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Learning musical representations for music performance question an- swering
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 63b297e1-fe3f-4019-a850-75c978a95f2b · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions From local to global: A graph rag approach to query-focused sum- marization, 2025
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c94dc9e-240e-4906-b76f-5fbf51003fa9 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Hi- erarchical modeling for task recognition and action seg- mentation in weakly-labeled instructional videos
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 748b7648-ecff-480e-84c0-7c4e87d9f227 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions The llama 3 herd of models, 2024
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 5347dc67-045d-4d09-ba8c-e7ddf45deb7e · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions A dataset for medical instructional video classification and question answering
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 91416d4b-a8d6-475e-8187-e9189351de46 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Does prompt formatting have any impact on llm performance?, 2024
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 00828faa-55ba-4925-bc58-a08ecc24d507 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Active retrieval augmented generation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c14ac5e1-4d67-4e30-b8dc-c7d996676847 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Instruction-tuned language models are better knowledge learners
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f9916b4b-48d8-43bd-9574-c4a4533f7681 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Prospector: Improving llm agents with self-asking and trajectory ranking
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 43f47a9b-20e8-41f2-993d-20afe710b60f · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Qube: Question-based belief enhancement for agen- tic llm reasoning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3804e5d9-186a-4d2c-9946-050f6173c8a0 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Dossier at medvidqa 2022: Text-based approaches to medical video answer localization problem
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 50443342-cb3f-4e24-8f78-9a414171e901 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Retrieval-augmented generation for knowledge-intensive nlp tasks
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f6a2fa42-a3d2-4c4b-9428-594b9f8ea07b · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Overview of the nlpcc 2023 shared task: Chinese medical instructional video question answering
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ccc36e41-004d-4ccf-8243-8bafc8025ea6 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Learn- ing to locate visual answer in video corpus using question
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 8db75a8c-6d12-4255-a964-3871bf5c37f8 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Overview of the nlpcc 2024 shared task 7: Multi-lingual medical instructional video question answering
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e8e4e8cb-90cd-4e4d-b011-2a760b0224ea · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions LLaV A-neXT- interleave: Tackling multi-image, video, and 3d in large mul- timodal models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a61b2af5-83d9-4e47-989b-7e3688b0c3c7 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Hello again! llm-powered per- sonalized agent for long-term dialogue
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 62560dd2-b484-42d0-9cef-abe1e29eabfc · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Miller, Sumit Chopra, Marc’Aurelio Ranzato, and Jason Weston
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f0512055-eb1c-454a-bbe1-d8ec78f8413f · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ad59e12c-3519-4dd5-bf81-900b4449e4e0 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Mediq: Question-asking llms and a benchmark for reli- able interactive clinical reasoning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 23dd8833-7f3c-4db8-8d4d-290c4fdf9bde · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Towards visual-prompt temporal answer grounding in instructional video
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0684a7af-088e-493d-8ccd-f8bfc04efeb3 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Fuzzy multimodal graph reasoning for human-centric instructional video grounding
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3d07cffb-231a-4d5c-8da1-90f06ecfad9a · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Video-LLaV A: Learning united visual rep- resentation by alignment before projection
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e55680a5-bdd6-4ea7-864f-f7bc762f4cc9 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions MM-VID: Advancing Video Understanding with GPT-4V(ision)
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d94ee299-635c-4b12-8f37-eedecc69b9f3 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Integrating video re- trieval and moment detection in a unified corpus for video question answering
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e2c386cd-d852-46f2-ba3c-93d63162317e · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Query rewriting in retrieval-augmented large language models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f864ce1e-b94d-4a23-a846-8d0c520141c1 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Query rewriting in retrieval-augmented large language models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d21d769e-e83f-422b-9eed-8170612d5f60 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Query rewriting in retrieval-augmented large language models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7895c9db-c17c-458a-a711-13867f0a5104 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Video-ChatGPT: Towards detailed video un- derstanding via large vision and language models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e7be9eef-d2c8-4355-ae84-f2b2a706849b · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Learn- ing to retrieve videos by asking questions
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ff58acc1-ae80-4ec4-b399-47bb3a5bd7b7 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Evaluating very long-term conversational memory of LLM agents
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 81d1e1cd-7d60-4336-8f28-5f3fd932fc83 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Chatvtg: Video temporal grounding via chat with video dialogue large language models
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation cdc2c8e6-59ca-4b12-a0f8-2902c406c899 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Distilbert, a distilled version of bert: smaller, faster, cheaper and lighter, 2020
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e6724848-406a-4423-93f6-0613a5a69aad · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Hybridrag: Integrat- ing knowledge graphs and vector retrieval augmented gen- eration for efficient information extraction
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ad4036a8-04b3-4c01-adfc-a0350c22b212 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions The art of creative inquiry—from question asking to prompt engi- neering
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 87bb036c-d5d1-4f8b-ad84-4d30a12d02dc · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Rewritelm: An instruction-tuned large language model for text rewriting
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d51c9aa3-e9ab-49ba-b99d-b6a2f21bdeed · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Toward expert-level med- ical question answering with large language models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b33975c-4740-4086-82f7-74fd6544341c · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions R-Bot: An LLM-based Query Rewrite System
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2db706f0-c2c7-41f5-86b4-7cf04f2d15e1 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Rag-adapter: A plug-and-play rag- enhanced framework for long video understanding, 2025
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b05d4636-12f1-4763-b6a2-d002bf7ea955 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Grounded-videollm: Sharpening fine-grained tem- poral grounding in video large language models, 2024
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6b139709-f645-429f-b2cc-a1f145559caf · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Llama3-8b-chinese-chat (revision 6622a23), 2024
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0d46acfc-a995-4a4b-8370-79052ca73c57 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Smarter, better, faster, longer: A modern bidirectional en- coder for fast, memory efficient, and long context finetuning and inference, 2024
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 5ea07396-0bdb-444e-a5d7-b488446568f3 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Visual answer localization with cross-modal mutual knowledge transfer
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 427beb43-e6d2-4524-b9b2-dca724669d09 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions To find where you talk: Temporal sentence localization in video with at- tention based location regression
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ccd2417a-355c-4030-ab57-848da9991451 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Videollama 3: Frontier multi- modal foundation models for image and video understand- ing, 2025
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e5a247c4-ceb5-4692-b423-ddeb8ffcdbd2 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Multi-scale video super-resolution trans- former with polynomial approximation
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f5f76ab3-71f7-4fd0-a3ce-9c84f4562286 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Span-based localizing network for natural language video lo- calization
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b918aff9-c91f-4d66-9545-ab83d94212f2 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Natural language video localization: A revisit in span-based question answering framework
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 55cd55bb-2f3c-4061-a05e-6ef191752d7d · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Fengshenbang 1.0: Being the Foundation of Chinese Cognitive Intelligence
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3777ac06-d384-45f1-b073-021c50ce107e · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions haotian liu, yong jae lee, liangke gui, di fu, jiashi feng, ziwei liu, and chunyuan li
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 63938cf7-aa15-40ac-b172-d429bad4748b · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Llava- next: A strong zero-shot video understanding model, 2024
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1a6632da-4ab1-4f46-9dff-82ef0755be52 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Medrag: Enhancing retrieval-augmented generation with knowledge graph-elicited reasoning for healthcare copilot,
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9f655ec8-399c-4b55-8bd6-ea936e0015c9 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Training-free video temporal grounding using large-scale pre-trained models
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f3b91129-c2ac-4d45-a988-57b19e93abfd · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Cross- task weakly supervised learning from instructional videos
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e686d75a-7c3f-4a5e-b5d8-7ab0fc9b60f1 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Unresolved cited work
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 66befbfc-e568-4043-a2b3-bbe1b9cfd1ac · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Unresolved cited work
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6ee3bf54-6834-4d2a-94be-644cce6e19bc · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Unresolved cited work
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c96aefb2-c9bc-43b3-ad18-8aa7beaad72c · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Unresolved cited work
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 59128a0d-0256-4213-b472-4e348969f07e · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Unresolved cited work
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7635f7af-5976-4c5f-83ac-4a7fd0d8f209 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Unresolved cited work
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 28dae07d-0667-401a-bc5c-87631e57f60a · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Unresolved cited work
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 21ec0141-1324-4470-8476-af9dcadf9ade · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions 是”或“否”回答的问题,比如你可以用“你的 意思是...?
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 05d4d6d5-66cd-4fa1-88d3-90e9dbe248bb · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Unresolved cited work
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6fd4f728-8858-4620-a0b9-82d1a72ecd34 · outbound
Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions Unresolved cited work
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 88150503-529d-4efc-a184-218b1afd31e3 · inbound
M$^3$-Med: A Benchmark for Multi-lingual, Multi-modal, and Multi-hop Reasoning in Medical Instructional Video Understanding Ask2Loc: Learning to Locate Instructional Visual Answers by Asking Questions
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.