Pith. sign in

Paper Citation Record · LEDGER

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos

As of 20 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:2505.01790.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.01790 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:15:24.149486Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact9
  • verified fuzzy11
  • unresolved35
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e3f3df19-9ec7-4e89-ab54-283695590d6f · outbound

This paper cites Computers and Education: Ar tificial Intelligence 7, 100298 (2024).

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos Computers and Education: Ar tificial Intelligence 7, 100298 (2024)

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:23.861392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:23.861392Z digest=sha256:7770893957402e7863897ef3a88ea2fe65bafe33a37e57f2a424b425a5aac56c

Observation 63c82e4a-e06e-45c8-a639-701fbbf3ad91 · outbound

This paper cites OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:23.867723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:23.867723Z digest=sha256:7e88bb20896d4c54da806398db83d3fc724c0a90edc114e53deceabec34b390c

Observation 9b05eb4e-0dbc-401f-86f3-2ab486c758e9 · outbound

This paper cites , Wartena, C.: Question generation capabilities of “small“ large lang uage models.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos , Wartena, C.: Question generation capabilities of “small“ large lang uage models

Reference 3

Resolution
verified exact
doi, observed 2026-08-16T04:15:24.646292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:23.873957Z digest=sha256:ad48651a723520828a2697616c67c18c88d6ad879864cd3215148cee8d980757

Observation 8dc216d1-8823-4945-9468-a6e6fe60b7af · outbound

This paper cites In: Conference on Computing Communication and Networking Technologies, ICCCNT 2023, Delhi, India, July 6-8, 2023.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: Conference on Computing Communication and Networking Technologies, ICCCNT 2023, Delhi, India, July 6-8, 2023

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:23.879180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:23.879180Z digest=sha256:f04681bce50e8414ec7a2944db82655dde6522c85f0d13fecac325d5afe563c0

Observation 868f6dcb-7d43-4aff-82e7-004cd474e11e · outbound

This paper cites In: Conference on User Modeling, Adaptation and Personalization, UMAP 2024, Cagl iari, Italy, July 1-4,.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: Conference on User Modeling, Adaptation and Personalization, UMAP 2024, Cagl iari, Italy, July 1-4,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:15:25.646732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:23.884043Z digest=sha256:eebaae814eabbedb5241f2310519bfba2bfb060b640dabcd6784132cfd0b059f

Observation 1d2e8cd9-e74f-4674-a052-77dd8677fb08 · outbound

This paper cites O’Reilly Media, In c.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos O’Reilly Media, In c

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:15:25.630965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:23.893874Z digest=sha256:a34ddbdf7219e037d050e1dc200b0d630e0718dc87a469c9afb0c28f9df09b73

Observation f74bd273-5329-44d3-8f1d-2fe003c97a8c · outbound

This paper cites an unresolved cited work.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:23.898325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:23.898325Z digest=sha256:a39f7f722007bba08095f23afb23d87250636009669fc3a270b75c481393ffd8

Observation a5a27e75-d2dc-4d93-9981-c503546b56bd · outbound

This paper cites Unleashing the potential of prompt engineering for large language models.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos Unleashing the potential of prompt engineering for large language models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:23.903229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:23.903229Z digest=sha256:7b1764044eaaed615f4f5e781edf00604bdc98847f13efc5f910ff99284ce409

Observation adaad955-8566-48a8-8027-51d66c73c057 · outbound

This paper cites In: International Confe rence on Web and Social Media, ICWSM 2018, Stanford, (California), USA, June 25-28 , 2018.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: International Confe rence on Web and Social Media, ICWSM 2018, Stanford, (California), USA, June 25-28 , 2018

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:23.907975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:23.907975Z digest=sha256:c309cfd8b1287c6ab97c5466b40bb64cc897845f43c22606b3b3b94d7b4683e1

Observation a0b78471-9637-41d3-b532-7eedd68b234b · outbound

This paper cites I n: Language Resources and Evaluation Conference, LREC 2020, Marseille, France, M ay 11-16, 2020.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos I n: Language Resources and Evaluation Conference, LREC 2020, Marseille, France, M ay 11-16, 2020

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:15:25.615927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:23.913114Z digest=sha256:503fcf037c493e7277a35d68d8f2df61fa602985512daab66b5b5556b5684684

Observation b8a9ae60-4457-437e-b00d-ee0a7f586947 · outbound

This paper cites Educ ational Horizons 83(3), 154–159 (2005), http://www.jstor.org/stable/42926529.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos Educ ational Horizons 83(3), 154–159 (2005), http://www.jstor.org/stable/42926529

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:23.918470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:23.918470Z digest=sha256:10a80bbafa0f21e6070399d08e5ea0d66370cb73b8812de9cc2ed404a4f60f22

Observation d2a524b1-69e9-4c89-a686-998db62fcd00 · outbound

This paper cites an unresolved cited work.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos Unresolved cited work

Reference 12

Resolution
verified exact
doi, observed 2026-08-16T04:15:24.593797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:23.923051Z digest=sha256:d87ce12c7e8d3cdef505c04aad430e486bc1be209301f64665940769ede3fb9e

Observation 4533f4a2-e6da-4aa8-97d9-27299db8f6cc · outbound

This paper cites Journal of appl ied psychology 32(3), 221 (1948).

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos Journal of appl ied psychology 32(3), 221 (1948)

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:23.927892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:23.927892Z digest=sha256:d0470ca3ead64e6fd2d6454c485bb280929104f0b0a23f959e062c5ef5bd68d7

Observation c6d320a0-1d6f-4fbc-9aa9-1e5d3e80609d · outbound

This paper cites International Journal o f Advanced Corporate Learning 16(1), 19–27 (2023).

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos International Journal o f Advanced Corporate Learning 16(1), 19–27 (2023)

Reference 14

Resolution
verified exact
doi, observed 2026-08-16T04:15:24.567379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:23.932732Z digest=sha256:16ad11662a5ef4e8ee6eebe86746706f0063605400939dead42dc32f90a45946

Observation bdaf4fe1-a783-48f7-8ef8-5c4628e9b4af · outbound

This paper cites Accurate, Large Minibatch SGD: Training ImageNet in 1 Hour.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos Accurate, Large Minibatch SGD: Training ImageNet in 1 Hour

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:23.937551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:23.937551Z digest=sha256:6abbc3d252ed095f738734cdb89150399ad0d8578dfc582d9657ecca26017f81

Observation a1fd207d-06a9-49eb-965f-3dd9efff3218 · outbound

This paper cites In: ACM/SIGAPP Sym- posium on Applied Computing, SAC 2019, Limassol, Cyprus, Ap ril 8-12, 2019.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: ACM/SIGAPP Sym- posium on Applied Computing, SAC 2019, Limassol, Cyprus, Ap ril 8-12, 2019

Reference 16

Resolution
metadata mismatch
raw_fallback, observed 2026-08-16T04:15:25.174415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:23.942393Z digest=sha256:92211aa7cfcd46e7d94756a4fb79c7550d246d43412134a060f671dc27cf3822

Observation fbc13b5e-9486-4a68-a0cc-0885dbaeb7b9 · outbound

This paper cites Scientific Dat a 10(1), 158 (2023).

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos Scientific Dat a 10(1), 158 (2023)

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:23.947144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:23.947144Z digest=sha256:ab15dd541edcafe731ee771fa147e1c5302271672ac1f61164d055c1a7d6bd4f

Observation 5235225a-8633-43e5-b479-f0b620d055e3 · outbound

This paper cites IEEE Access 11, 20885– 20896 (2023).

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos IEEE Access 11, 20885– 20896 (2023)

Reference 18

Resolution
metadata mismatch
raw_fallback, observed 2026-08-16T04:15:25.092829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:23.951878Z digest=sha256:b7b824623cd44c28b2da6cb7471a1481a0a583b899a3478dddcced978a1b540b

Observation ed1444af-7aa6-40ff-93ce-0b37d9720488 · outbound

This paper cites In: Artificial Intelligence in Education, AIED 2024, Recife , Brazil, July 8-12, 2024.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: Artificial Intelligence in Education, AIED 2024, Recife , Brazil, July 8-12, 2024

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:23.956761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:23.956761Z digest=sha256:3592b4631d20fa720e9dd300d2e94e3e4b3e677907733d3b83a136cfe24a38fc

Observation 7cc26809-51ab-4f14-8f69-c31ed299d267 · outbound

This paper cites Mistral 7B.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos Mistral 7B

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:23.961491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:23.961491Z digest=sha256:6a4a25ddc3b4632218ac970397804980fe17d30502d283ff805d335350e6497a

Observation bb163566-c328-4ee7-8cbd-837ee504e51b · outbound

This paper cites In: IEEE Conference on Computer Vis ion and Pattern Recognition, CVPR 2017, Honolulu, (Hawaii), USA, July 21-2 6, 2017.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: IEEE Conference on Computer Vis ion and Pattern Recognition, CVPR 2017, Honolulu, (Hawaii), USA, July 21-2 6, 2017

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:15:25.599572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:23.966192Z digest=sha256:0e8c3b8f586fbc372f3792276d2330e305ff6ad63becc75f0e0a9c57b199896a

Observation 87c1e9f1-2784-4db4-92cc-8594ea257cdb · outbound

This paper cites In: Worksh op on Statistical Ma- chine Translation co-locared with the ACL Conference, WMT@ ACL 2007, Prague, Czech Republic, June 23, 2007.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: Worksh op on Statistical Ma- chine Translation co-locared with the ACL Conference, WMT@ ACL 2007, Prague, Czech Republic, June 23, 2007

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:15:25.582977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:23.977100Z digest=sha256:99813e7d99b28328757c2dd055f0202aa28e021f836b13e384750b07b3a10680

Observation 6221b65e-18a8-4e91-a259-d34fa6e8a550 · outbound

This paper cites In: Proceedings of the 2018 Conference on Empirical Meth- ods in Natural Language Processing, Brussels, Belgium, Oct ober 31 - Novem- ber 4, 2018.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: Proceedings of the 2018 Conference on Empirical Meth- ods in Natural Language Processing, Brussels, Belgium, Oct ober 31 - Novem- ber 4, 2018

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:23.981662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:23.981662Z digest=sha256:53a5b5dc65a3b44967cca05fb39425cc70c16883eae495597d2301a1654d21e6

Observation 1c9ce369-e86b-4af2-a7dc-730772ed43c4 · outbound

This paper cites In: Association for Computation al Linguistics, ACL 2020, Virtual Event, July 5-10, 2020.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: Association for Computation al Linguistics, ACL 2020, Virtual Event, July 5-10, 2020

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:23.986694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:23.986694Z digest=sha256:9d455ea4842016a9d7c5dabee521a386d1f06df7e8204abd10b4e7ccc084a4b7

Observation 1d3f2296-b83a-4dc4-8676-6527266fba5e · outbound

This paper cites I n: HCI International Conference, HCII 2020, Copenhagen, Denmark, July 19-24, 20 20.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos I n: HCI International Conference, HCII 2020, Copenhagen, Denmark, July 19-24, 20 20

Reference 25

Resolution
verified exact
doi, observed 2026-08-16T04:15:24.485221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:23.992227Z digest=sha256:62091d76e9766c67401f13cb215bd69679591b4bc0088560d669952a04412ff0

Observation 9874c8d7-e323-4541-ab5a-bda65f577efa · outbound

This paper cites In: International Conference on Machine Learning, ICM L 2023, Hon- olulu, (Hawaii), USA, July 23-29 2023.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: International Conference on Machine Learning, ICM L 2023, Hon- olulu, (Hawaii), USA, July 23-29 2023

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:15:25.565861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:23.997122Z digest=sha256:e3a6366fcfe80a3d95669ba3d5e49118c7b6b4f7fa8a95a1dd6d1b9106ef9f69

Observation bc2cfdbd-cac3-4d34-97e3-27ec0138f499 · outbound

This paper cites VideoChat: Chat-Centric Video Understanding.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos VideoChat: Chat-Centric Video Understanding

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:24.002262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:24.002262Z digest=sha256:ff2dfa688636657818a53536554662a4a9da5a775a93db3a2cfc414309407c4a

Observation 8921010e-97e6-45f5-90e7-7a67a8adfd78 · outbound

This paper cites A Novel Approach to Scalable and Automatic Topic-Controlled Question Generation in Education.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos A Novel Approach to Scalable and Automatic Topic-Controlled Question Generation in Education

Reference 28

Resolution
metadata mismatch
local_arxiv, observed 2026-08-16T04:15:24.450054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:24.007527Z digest=sha256:aff9d6747fa1bb19f76ceaa8b747412c37d657e2be2338cbe0a5cef60e000b8b

Observation a263b28a-bdbc-4e71-8158-b0225f184ec1 · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:24.013417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:24.013417Z digest=sha256:2a750a410d59bc75acb2f1f8f1519fb6f1ef561644b3d9d6e6e8598c3dfb8f22

Observation 1b3ef0ee-5713-4ea1-90bb-d59e416ba14d · outbound

This paper cites In: Association for Computational Linguistics, ACL 2004, Barc elona, Spain, July 21-26, 2004.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: Association for Computational Linguistics, ACL 2004, Barc elona, Spain, July 21-26, 2004

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:15:25.549030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:24.018752Z digest=sha256:50dce4fea84767b4a2ba7626bb2714b7e23f65d15aba61679bfd648b1ad32bf1

Observation be3b2869-bd58-4653-b489-8469f30c5bf7 · outbound

This paper cites Improved Baselines with Visual Instruction Tuning.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos Improved Baselines with Visual Instruction Tuning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:24.023195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:24.023195Z digest=sha256:831e76e33ef953915e23e8e0b933bc011f325295ecb85305093bbb4cdeb6fdab

Observation 51a4bb20-cff4-4b49-b311-6584efe9d203 · outbound

This paper cites In: Advances in Neural Information Processing Systems, NeurIPS 2023, New O rleans, (LA), USA, December 10 - 16, 2023.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: Advances in Neural Information Processing Systems, NeurIPS 2023, New O rleans, (LA), USA, December 10 - 16, 2023

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:15:25.532899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:24.028367Z digest=sha256:8d07bc2b4c0ffe53404013958e743d777df3c894dcd6ebb9e9c2a35b3a97e464

Observation e4a58acc-aa56-4f3d-8ace-d31fb1ac8d54 · outbound

This paper cites Valley: Video Assistant with Large Language model Enhanced abilitY.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos Valley: Video Assistant with Large Language model Enhanced abilitY

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:24.032863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:24.032863Z digest=sha256:1f5404164426ac906b7dbd867c222698372236a61c09817f339163a1cb23d03b

Observation 636f3b18-225d-4334-9682-4c1fbe6b1419 · outbound

This paper cites Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:24.037868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:24.037868Z digest=sha256:cf39cead2c4693b5932564db235b3450ce2d28884b43c9933a407f391ba10066

Observation 810cb79a-8a85-4553-a291-2f647be81c23 · outbound

This paper cites https://doi.org/10.1016/j.caeai.2025.100370.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos https://doi.org/10.1016/j.caeai.2025.100370

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:24.043122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:24.043122Z digest=sha256:e672258d196f40f8953aab5dc8ae46ad3f89033fde0a1ce9928b30c7492e6cb3

Observation 49f96d43-eb51-4621-897a-7520541c6d10 · outbound

This paper cites https://doi.org/10.1016/j.compedu.2021.104355.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos https://doi.org/10.1016/j.compedu.2021.104355

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:24.049016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:24.049016Z digest=sha256:be928bafce4358bfcebe89bf5d11916f4073454218b9d842d015481e984ce9a7

Observation 4e54c934-5642-4f1e-8221-ee4cb0517ee6 · outbound

This paper cites , de Souza, A.F., Oliveira- Santos, T.: Automatic multiple-choice question generatio n and evaluation sys- tems based on LLM: A study case with university resolutions.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos , de Souza, A.F., Oliveira- Santos, T.: Automatic multiple-choice question generatio n and evaluation sys- tems based on LLM: A study case with university resolutions

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:15:25.516686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:24.053755Z digest=sha256:0fae37adf963c3e65d59f7ea2d32964803e5dc55b21377fb976bf41babcfe4cc

Observation cd68b808-f4ea-4ee1-b278-09417ff7ae65 · outbound

This paper cites PG-Video-LLaVA: Pixel Grounding Large Video-Language Models.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos PG-Video-LLaVA: Pixel Grounding Large Video-Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:24.058509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:24.058509Z digest=sha256:34fbac55e8ed66edf2e298b4f7375e2b9bc4aeb056f5b22d9d1e2f8f888128e4

Observation eb14dcb7-df7e-47ba-98c4-681709bbe48e · outbound

This paper cites GPT-4 Technical Report.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos GPT-4 Technical Report

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:24.063574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:24.063574Z digest=sha256:733fc929847d03c934a7ca95cbd11315f37dc8bfdc618ba4e4e6a62de81bedf7

Observation 6da932a9-3d70-4870-93e1-99b37f1af481 · outbound

This paper cites Education and Information Technologies 24(1), 805–823 (2019).

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos Education and Information Technologies 24(1), 805–823 (2019)

Reference 40

Resolution
verified exact
doi, observed 2026-08-16T04:15:24.321964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:24.068993Z digest=sha256:fc46e05b5e38c92d9d85e06164f029d8849cfc4fb2cf2370d2c7b828b656e811

Observation 91a2134e-a221-4604-9bbd-c2e7871dd11b · outbound

This paper cites In: Association for Com putational Linguistics, ACL 2002, Philadelphia, (PA), USA, July 6-12, 2002.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: Association for Com putational Linguistics, ACL 2002, Philadelphia, (PA), USA, July 6-12, 2002

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:24.074686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:24.074686Z digest=sha256:ae532706ac0213c95dcdbb32afa695801db1521e3363e3f40c02423e0aa31f2d

Observation 71cce2fe-d5cf-49e6-94b1-14641afa730d · outbound

This paper cites In: Conference on E mpirical Methods in Natural Language Processing, EMNLP 2016, Austin , Texas, USA.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: Conference on E mpirical Methods in Natural Language Processing, EMNLP 2016, Austin , Texas, USA

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:24.080211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:24.080211Z digest=sha256:fbc1905ba89dea2950dbc7e888545c32bf364923007de77aa84e33143bf92302

Observation ff396e65-016f-4d89-a2b5-b6b4af99c348 · outbound

This paper cites In: Conference on Empirical Methods in Natural La nguage Processing, EMNLP 2019, Hong Kong, China, November 3-7, 2019.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: Conference on Empirical Methods in Natural La nguage Processing, EMNLP 2019, Hong Kong, China, November 3-7, 2019

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:24.085459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:24.085459Z digest=sha256:0baee7b9f7494d180b45f1f61c49503fa3395db80c4d66969f80f632491e86d1

Observation 13f0d652-4e9b-41e8-b03c-d9069f3d9bbd · outbound

This paper cites ACM Computing Surveys 55(2), 26:1–26:39 (2023).

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos ACM Computing Surveys 55(2), 26:1–26:39 (2023)

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:24.090213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:24.090213Z digest=sha256:a76469662388cfcabcf4773689070348e2c7ea74b1428630a9d4e0b469290703

Observation 4a5e03e0-2a09-40c5-a8dd-8d764fb66fd5 · outbound

This paper cites In: Artificial Intelligenc e in Education, Using VLMs to Generate Questions for Educational Videos 17 AIED 2024, Recife, Brazil, July 8-12, 2024.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: Artificial Intelligenc e in Education, Using VLMs to Generate Questions for Educational Videos 17 AIED 2024, Recife, Brazil, July 8-12, 2024

Reference 45

Resolution
verified exact
doi, observed 2026-08-16T04:15:24.275682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:24.094848Z digest=sha256:8775147610c6a3c817225326c4c7c9571c08ce706a9c6321fb13f33b2ac9a347

Observation f29f2ad3-3a0a-4d3b-920a-75df25109fc5 · outbound

This paper cites Artificial Intel ligence Review 57(10), 271 (2024).

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos Artificial Intel ligence Review 57(10), 271 (2024)

Reference 46

Resolution
verified exact
doi, observed 2026-08-16T04:15:24.258502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:24.100177Z digest=sha256:3a89bcd22f894d2d04445e98f02b4a15944ca6bbb8143b46278a6be9277ce74d

Observation 8a40fd5f-395e-4a02-bae9-0fb5d694d212 · outbound

This paper cites In: Ar tificial Intelligence in Ed- ucation, AIED 2023, Tokyo, Japan, July 3-7, 2023, Proceedin gs.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: Ar tificial Intelligence in Ed- ucation, AIED 2023, Tokyo, Japan, July 3-7, 2023, Proceedin gs

Reference 47

Resolution
verified exact
doi, observed 2026-08-16T04:15:24.241956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:24.105007Z digest=sha256:400f4d2d6d922bb90c652bf3be8f14dad7c173d2370b6b87d8de23bb7fc2c554

Observation 7c0e8292-8bbe-4899-8e87-c85281351f7c · outbound

This paper cites In: Annual Confere nce on Infor- mation Technology Education, SIGITE 2017, New York, USA, Oc tober 4 - 7, 2017.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: Annual Confere nce on Infor- mation Technology Education, SIGITE 2017, New York, USA, Oc tober 4 - 7, 2017

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:24.109403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:24.109403Z digest=sha256:f59e291ebfe9418915233f44100a9ecb3ccc3193111334d46c02d42048d7aa1d

Observation 20be832d-db44-4b11-ae9c-37f6693ed93d · outbound

This paper cites : Automatic question- naire and interactive session generation from videos.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos : Automatic question- naire and interactive session generation from videos

Reference 49

Resolution
verified exact
doi, observed 2026-08-16T04:15:24.224356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:24.114766Z digest=sha256:7079b647120c29d45a3dc96641184c2572b64a55beee267e50887be335c0db7b

Observation d6bed48f-d654-4c26-9bcb-65564087fa95 · outbound

This paper cites In: ACM on Multimedia Conference, MM 2017, Mountain View, (CA), USA , October 23-27,.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: ACM on Multimedia Conference, MM 2017, Mountain View, (CA), USA , October 23-27,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:15:25.500848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:24.119897Z digest=sha256:945e9c70964654cb1815fb60b07934c7c1f687f78c7598f27ea0888c0231af6a

Observation 3ba8ea4d-63a9-429b-bbc1-441b35474a1f · outbound

This paper cites In: Màrquez, L., Callison-Burch, C., Su, J., Pig hin, D., Marton, Y.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: Màrquez, L., Callison-Burch, C., Su, J., Pig hin, D., Marton, Y

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:24.129526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:24.129526Z digest=sha256:e714a97dfcc4c362bacbb3e85ea58aae8210262d562dfaad30fb749e06d7cf73

Observation b79fb153-4af3-4d0f-8b4e-beb0a357e569 · outbound

This paper cites MCQG-SRefine: Multiple Choice Question Generation and Evaluation with Iterative Self-Critique, Correction, and Comparison Feedback.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos MCQG-SRefine: Multiple Choice Question Generation and Evaluation with Iterative Self-Critique, Correction, and Comparison Feedback

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:24.135060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:24.135060Z digest=sha256:f9208ec1d56ed4c647db015c97e4c8a0a066921297384520a536a4c05c18acc6

Observation 6cd960a3-b2f8-436a-bed0-a3743bfbc9b6 · outbound

This paper cites : Activitynet-qa: A dataset for understanding complex web videos via question answering.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos : Activitynet-qa: A dataset for understanding complex web videos via question answering

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:24.140034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:24.140034Z digest=sha256:247518efbe7b3e13720f1c9e6c8d3a26fad89176689569f52793d4ed4a9b9e1c

Observation 9cc52874-4fc6-46a6-976f-bdcb6a9d1e62 · outbound

This paper cites In: Conference on Empi rical Methods in Nat- ural Language Processing - System Demonstrations, EMNLP 20 23, Singapore, De- cember 6-10, 2023.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: Conference on Empi rical Methods in Nat- ural Language Processing - System Demonstrations, EMNLP 20 23, Singapore, De- cember 6-10, 2023

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:24.144923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:24.144923Z digest=sha256:566dd700d037c67f1f2717e9a43a9f1be085d6bdf7c2ee9be442b8e8693d5edb

Observation 9fc174ac-24db-4425-bd63-a1c608f5e5ab · outbound

This paper cites In: International Conference on Learning Represen- tations, ICLR 2020, Addis Ababa, Ethiopia, April 26-30, 202 0.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos In: International Conference on Learning Represen- tations, ICLR 2020, Addis Ababa, Ethiopia, April 26-30, 202 0

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:15:25.474580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T04:15:24.149486Z digest=sha256:f2b61e501a3c28c2b062a135cc90478af6fa10ed695da205713c398aff0a4ecf

Observation c453569c-ccf4-4c92-a1b2-0766dff87d9a · outbound

This paper cites 1645–1653.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos 1645–1653

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:24.124780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:24.124780Z digest=sha256:b1da2b4a0bb0cba066cc4c2494ea0a90bd0dba03d29211d49d2b70a90baa56c6

Observation 09f58eaa-87be-4097-a607-41f5f388cc7e · outbound

This paper cites https://doi.org/10.1145/3631700.3665233.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos https://doi.org/10.1145/3631700.3665233

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:23.888781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:23.888781Z digest=sha256:27939f85eefd782ef0656e353e87239bf1674709e1e79209f20ca707179f456b

Observation d4946c59-9f71-433b-9bbb-d6e8de624bda · outbound

This paper cites https://doi.org/10.1109/CVPR.2017.571.

Enhancing the Learning Experience: Using Vision-Language Models to Generate Questions for Educational Videos https://doi.org/10.1109/CVPR.2017.571

Reference 5384

Resolution
unresolved
no resolver link, observed 2026-08-16T04:15:23.971259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:15:23.971259Z digest=sha256:73b660457699fc8f8e2a97f54148728903f0a6fc155268ab52e783f3b40b83e3

Pith citing papers

No inbound Pith citation observations are available.