Pith. sign in

Paper Citation Record · LEDGER

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning

As of 7 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 1 inbound Pith citation observation for arXiv:2507.20163.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.20163 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:47:53.840267Z

measured 59 of 59 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T15:55:21.629729Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact1
  • verified fuzzy49
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 615c7d59-3460-4dbd-ba72-8425a78f764c · outbound

This paper cites GPT-4 Technical Report.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:48.338225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:48.338225Z digest=sha256:c5f383a9a20e89c63b074d81fca8fb2b9602f2219e3a10c542242504baf8a86a

Observation c2c43baf-7352-4256-a8b6-28ee0082f827 · outbound

This paper cites Meteor: An automatic metric for mt evaluation with improved correlation with hu- man judgments.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Meteor: An automatic metric for mt evaluation with improved correlation with hu- man judgments

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:03.233657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:48.387302Z digest=sha256:b79e61601be8063ab7984b3e984103ae0461b3e76fd0a0ee256e4e5cfd007941

Observation fff18909-4ae3-4d86-8f4e-47e8ba2e2ad9 · outbound

This paper cites Is space-time attention all you need for video understanding? In ICML, page 4, 2021.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Is space-time attention all you need for video understanding? In ICML, page 4, 2021

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:03.098776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:48.471255Z digest=sha256:9a11704ca5061471c372180dfbb77e002c62618ae6d4b220a8dc9800c60b3994

Observation 411cb742-a6d4-41c2-babd-432823cd6856 · outbound

This paper cites Activitynet: A large-scale video benchmark for human activity understanding.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Activitynet: A large-scale video benchmark for human activity understanding

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:02.909167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:48.557375Z digest=sha256:cdc765670584d05918c537df6783c74ec4cc3e32dd7805c7ec8a47d4f54ba5e2

Observation 0ec81a08-6cf2-43e0-8e5d-3dadb61c85c3 · outbound

This paper cites Quo vadis, action recognition? a new model and the kinetics dataset.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Quo vadis, action recognition? a new model and the kinetics dataset

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:02.760554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:48.651033Z digest=sha256:b983bdf617e8afc3f3a44005edc99374a94963fd8591bde9f1409ae639a9a670

Observation 66722b3d-12f4-41f1-9a81-a015d7c49748 · outbound

This paper cites Collecting highly paral- lel data for paraphrase evaluation.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Collecting highly paral- lel data for paraphrase evaluation

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:02.593360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:48.689591Z digest=sha256:5d106ff0ae77b9460039039a729f41db53a8a7007b5e01c7af41f68cae4094ef

Observation 059320fb-9413-43d9-9774-722ab4638539 · outbound

This paper cites Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:48.721206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:48.721206Z digest=sha256:c6655f53be3ea17e293efd290bfb85b3ac55910d02e65153a1f31fd92ee56cda

Observation e749d5fc-3a1c-493e-90f6-d4cdb1a3f469 · outbound

This paper cites Sportsmot: A large multi-object tracking dataset in multiple sports scenes.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Sportsmot: A large multi-object tracking dataset in multiple sports scenes

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:02.444435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:48.739135Z digest=sha256:a79fad5612d266f076514ebd94dcd49a5557ea44d14c53233ff2c6b9a1a84f9c

Observation 3ddf3f16-8a17-4b3c-9d57-bac93c3d0a91 · outbound

This paper cites A thou- sand frames in just a few words: Lingual description of videos through latent topics and sparse object stitching.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning A thou- sand frames in just a few words: Lingual description of videos through latent topics and sparse object stitching

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:02.255384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:48.785487Z digest=sha256:a59f63abe822ef4c87ee0bc682d391721e902a46f11ab71469a6166225802575

Observation 1ba5df05-4642-4508-a44d-c8cd9c866ffd · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:48.831059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:48.831059Z digest=sha256:1bc3ab54426ace8de26001ffd98cf2e35de0bc69cab25845e02217b891453175

Observation b0c30544-5258-410e-8fa8-503210534ff5 · outbound

This paper cites Soccer captioning: dataset, transformer-based model, and triple-level evaluation.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Soccer captioning: dataset, transformer-based model, and triple-level evaluation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:02.121089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:48.890391Z digest=sha256:386c7fd5d35f00675611c81fbbdfadd2ad4a993ddb2c37144901b32c98584756

Observation 28674be7-f982-4455-aefa-ff76dfeaade9 · outbound

This paper cites an unresolved cited work.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T13:48:01.960339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:48.942982Z digest=sha256:e30df1ac014d3d08a29dfe575a0c9bf61b565b0cf6191d9deea652a76f819b0a

Observation ca1ec4e3-4f73-457f-bbb0-82bcc2a39c01 · outbound

This paper cites Deep residual learning for image recognition.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Deep residual learning for image recognition

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:01.789107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:48.982219Z digest=sha256:929be72f81c833057f83a0c53f9971814784e18f958aa5915c0524a02474bfd3

Observation f9ab540d-86f4-4e49-82d5-b22887b02b66 · outbound

This paper cites Gaussian Error Linear Units (GELUs).

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Gaussian Error Linear Units (GELUs)

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:49.051493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:49.051493Z digest=sha256:18e833e5ec6095005bcbbacf8e7483e438be0c87cd5c94b74c6c4951c790a2fb

Observation 830bdf3f-bf20-47c4-b106-f1922277feca · outbound

This paper cites Overview of temporal action detection based on deep learning.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Overview of temporal action detection based on deep learning

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:01.578230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:49.111293Z digest=sha256:a5868cf63af2e0af0b263b7745680092f68dba4c3749c911c912f0f1b5a4a901

Observation 4d3642e8-9f19-4e66-afe3-f34de4e784d0 · outbound

This paper cites Learn- ing to generate move-by-move commentary for chess games from large-scale social forum data.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Learn- ing to generate move-by-move commentary for chess games from large-scale social forum data

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:01.371450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:49.229047Z digest=sha256:3f3d3fc77abfeb51e1678fa6989fa1d8d6a29b8d1c3e677db9845bffaee05edc

Observation 9b0d2555-6dd6-4586-bb63-a0dd0deded6e · outbound

This paper cites Learning seg- ment similarity and alignment in large-scale content based video retrieval.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Learning seg- ment similarity and alignment in large-scale content based video retrieval

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:01.166288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:49.383887Z digest=sha256:6ad204874820dfc50871d345cafbb6082acef92ffb8afa7966ea033e3c609d65

Observation 9163806d-f645-4003-a9b9-4116062688c1 · outbound

This paper cites Automatic baseball commentary generation using deep learning.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Automatic baseball commentary generation using deep learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:01.005079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:49.465481Z digest=sha256:cf4cf79962b40040f4ba053311a35b49c4ad1da480ebfb67a354f30f3db3a93e

Observation 8d357874-4819-4675-b8b5-8f6966fb7123 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Adam: A Method for Stochastic Optimization

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:49.583990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:49.583990Z digest=sha256:f890da4bf7d8fc2cfe337a6a939418ca7c23d6c21635d1ca60d08e2d61cf1071

Observation 639bf675-0666-4cec-92d7-0ac52890f53c · outbound

This paper cites Video story- telling: Textual summaries for events.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Video story- telling: Textual summaries for events

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:00.869574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:49.829969Z digest=sha256:341cfffcbe55c79f20dbb45d3989899a70b7c574481b97b3b0cd59fa6fb19ac0

Observation 3550829e-2298-4d99-8af8-0e6f3c1021da · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Rouge: A package for automatic evaluation of summaries

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:00.680689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:49.952635Z digest=sha256:da0497fafa198a06708f3035437d22fc88ae8c855406f0c5af7b203de5180ba7

Observation 717f7a0e-aa5a-4888-8221-67085a8a4267 · outbound

This paper cites Swinbert: End- to-end transformers with sparse attention for video caption- ing.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Swinbert: End- to-end transformers with sparse attention for video caption- ing

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:00.501771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:50.083025Z digest=sha256:1e00141cd31cf2d04f79bbd57b4c8749f1f01f07d01f9ec49c27faf1b6e262af

Observation fd23068d-2a81-4dd3-a882-e03f26b28b21 · outbound

This paper cites Video swin transformer.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Video swin transformer

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:00.299165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:50.245337Z digest=sha256:b9283d3b9f7cfeb475d72c8f4e7c3bb93de4a629721ac78f9cd9ae3d3bfb6131

Observation bce3b90b-dbca-4d5a-91c8-665ec89a251a · outbound

This paper cites Llama 3.2 quantized models, 2024.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Llama 3.2 quantized models, 2024

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:00.097823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:50.372060Z digest=sha256:496fde4246770c2408062213633274696d3346088db164711a81b68d9f5ba101

Observation ccd780b4-4273-4abd-a016-7267ffcf0f44 · outbound

This paper cites Soccernet- caption: Dense video captioning for soccer broadcasts com- mentaries.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Soccernet- caption: Dense video captioning for soccer broadcasts com- mentaries

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:59.854049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:50.522585Z digest=sha256:74abec5625d36d6d7d9318557f6d6e4e72e5b344a10344ff711644d403aef813

Observation 0fe48295-4fb5-4b2d-add4-c27f80b8e0b0 · outbound

This paper cites Search-oriented micro-video captioning.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Search-oriented micro-video captioning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:59.618980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:50.650288Z digest=sha256:079a0ab927f0aa2f7becf0908fc752691cea74f096fdc1de63197b05b975160e

Observation b68e4fb3-067f-4dec-a121-f022be7d66ce · outbound

This paper cites Enhancing visual question answering through question-driven image captions as prompts.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Enhancing visual question answering through question-driven image captions as prompts

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:59.373041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:50.783054Z digest=sha256:c0539375e16b98e60f8387580a065c517eaead16e4fff3a8be91b2b8402d5d25

Observation 4a3e2810-37bb-4324-8303-65485346603d · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Bleu: a method for automatic evaluation of machine translation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:59.134945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:51.004697Z digest=sha256:21f6b09bd43b4bfd166981b6eff39f26c3c65f15ffd22eb61fe0474bbad3ed10

Observation b53fca66-fa12-430d-8a6e-7451aac9892a · outbound

This paper cites Identity- aware multi-sentence video description.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Identity- aware multi-sentence video description

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:58.894291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:51.226089Z digest=sha256:e4feb762d5e8b25d5da42ee88308c75567162591db94c4182847fe58d4b8eb9e

Observation 37534dc4-5bcd-4462-9b01-fe8af62f296d · outbound

This paper cites Goal: A challenging knowledge-grounded video captioning benchmark for real- time soccer commentary generation.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Goal: A challenging knowledge-grounded video captioning benchmark for real- time soccer commentary generation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:58.654739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:51.400669Z digest=sha256:f0ae55870141156470cd4aab4df39a66fe151cdfbb8bc05e319b6448f2629ba5

Observation 77af1dcb-5943-4b4e-96b4-73de32742b74 · outbound

This paper cites Sports video captioning via attentive motion representation and group re- lationship modeling.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Sports video captioning via attentive motion representation and group re- lationship modeling

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:58.457349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:51.533250Z digest=sha256:d4f41aed5665cfccc1110bfba7cef0c0d053ca9c172880298ae3f71665b37f2a

Observation f7630cbe-3fa3-4c7e-b011-dcb167702c73 · outbound

This paper cites Language models are unsupervised multitask learners.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Language models are unsupervised multitask learners

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:58.256585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:51.721508Z digest=sha256:5726ef3298c4c3b134e9d2d0fc26ca08058252504db5a61da2469d5e5a2129f6

Observation 5fbd5acf-adb7-4f10-9ac1-637ecc7db084 · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Learn- ing transferable visual models from natural language super- vision

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:58.105707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:51.901836Z digest=sha256:fd6231c11412da2d33cd290166fdbfe3ade54f6a55d8da96a749e0f6cdfedaed

Observation 17fcd46c-ce79-4d66-9dc7-ee5dda44c2ee · outbound

This paper cites MatchTime: Towards Automatic Soccer Game Commentary Generation.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning MatchTime: Towards Automatic Soccer Game Commentary Generation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:52.050904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:52.050904Z digest=sha256:e589227553b588bd7f9fbd671a2d6ebe938091b0edbc2cb778532ca03c734cc7

Observation 9ed940e0-1380-4904-adcc-1aaf416794ad · outbound

This paper cites Towards universal soccer video under- standing.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Towards universal soccer video under- standing

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:57.935512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:52.195571Z digest=sha256:1973946bee6b1580040ffc31ba770a03802d1cb290f4ddc80c619f0df07a7e06

Observation c026b22e-20c4-4f7e-a4c7-94e13b78a84a · outbound

This paper cites Grounding action descriptions in videos.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Grounding action descriptions in videos

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:57.685758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:52.288758Z digest=sha256:92cf060c87765cdda5be0429aaf04bfde424aaeaaa5ebbd880f8c9e2bffb07d8

Observation dd8cd761-f1a1-428d-ac77-a9f617db746e · outbound

This paper cites Timechat: A time-sensitive multimodal large language model for long video understanding.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Timechat: A time-sensitive multimodal large language model for long video understanding

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:57.497245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:52.384320Z digest=sha256:8f2b9e22b03d3d7f24b4931cc3f2e5d14d13a3fe1a626d82c7e18b5480a50f81

Observation cccc58cf-1f89-4681-9e80-55edcc540cee · outbound

This paper cites A dataset for movie description.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning A dataset for movie description

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:57.261849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:52.476419Z digest=sha256:14dc4de980f59b8d418cc14edfad3fac84d807b806488ce922be911a752e3875

Observation b1a26b85-86e0-4762-87ba-0b6a957a1e6c · outbound

This paper cites Accurate and fast compressed video captioning.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Accurate and fast compressed video captioning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:57.090318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:52.570908Z digest=sha256:9872f71029290dd3bee885e36d64f139eb0ead56a53d0db86ad21e23f52ae82b

Observation 358d8369-dea1-4109-938e-9adea259b196 · outbound

This paper cites An overview of the tesseract ocr engine.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning An overview of the tesseract ocr engine

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:56.929106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:52.629207Z digest=sha256:9fff0db14ee3b7eeaef644eadd7b70367ed6218fb56544943bffc204a2d6a36a

Observation f79916d6-9ca0-4e4f-abeb-d43dac9275c1 · outbound

This paper cites Clip4caption: Clip for video caption.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Clip4caption: Clip for video caption

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:56.762986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:52.681290Z digest=sha256:44bb9dcbeee0808652c199daf2ee0183eb898fcef2e7dd510c6c263a977789f4

Observation 37b257fc-62a8-4905-8fe4-d1b9b93c9c43 · outbound

This paper cites Qwen2.5: A party of foundation models, 2024.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Qwen2.5: A party of foundation models, 2024

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:56.634122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:52.753771Z digest=sha256:9ab02bd907bd3bd1ccac849a5a28ca0aece303dfdb625f94c7d2221aec316325

Observation 6ce10f36-56a2-4295-816e-cb5b157b3675 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:52.792329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:52.792329Z digest=sha256:ab479a8f0bd2fed63434efef608f96fcc6dc0a5003cb8cf4b2d92f9fdec90b83

Observation f50975b2-6527-41d1-b5bb-58210b190ba9 · outbound

This paper cites Atten- tion is all you need.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Atten- tion is all you need

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:56.470095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:52.859030Z digest=sha256:62ca0613e1c24618e0df1601fd003b01fd99e42602842302239d3e5e7a1ed623

Observation 1c967b64-34d4-4209-935d-c7c83bc22406 · outbound

This paper cites Player tracking and identification in ice hockey.Expert systems with applications, 213:119250, 2023.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Player tracking and identification in ice hockey.Expert systems with applications, 213:119250, 2023

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:56.298252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:52.919724Z digest=sha256:ffa5ea7633927106a3328116b22c6a4ee32cf8ab83fac45cbd6864741ce58b9a

Observation 98ad022a-2a58-4206-80ae-5b22fdb2ef3a · outbound

This paper cites Cider: Consensus-based image description evalua- tion.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Cider: Consensus-based image description evalua- tion

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:56.092821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:52.975748Z digest=sha256:b5ab0f15f9d2da865919aeb777832bbd142e42251a2dba52074672542d000f42

Observation 265d17ab-2956-4a69-85f1-c0a1c0511237 · outbound

This paper cites Omnivid: A generative framework for universal video understanding.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Omnivid: A generative framework for universal video understanding

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:55.936975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:53.023439Z digest=sha256:17237e2d4bfbfca3a14a6dd7e76ad14b23331d5f17b736e8621639fd50db4057

Observation 371e9520-823e-41f1-aa81-b12c4ac8fd27 · outbound

This paper cites Sports video anal- ysis on large-scale data.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Sports video anal- ysis on large-scale data

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:55.808364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:53.059771Z digest=sha256:7a3226486f17c4d3f92c1302310f0b06153b789cde2d687f32b4170ad720eede

Observation 4e6ee851-9abe-4eab-8126-257f9d148816 · outbound

This paper cites Learning label semantics for weakly supervised group activity recognition.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Learning label semantics for weakly supervised group activity recognition

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:55.644024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:53.106462Z digest=sha256:a52cdb73a7ed5e26572cac5ccedc87202d8a9f48719e9e88faf58fd65ebe4c0d

Observation d8c8a3ac-9f04-42ef-b5ff-18212b579d07 · outbound

This paper cites A simple yet effective knowledge guided method for entity-aware video captioning on a basketball benchmark.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning A simple yet effective knowledge guided method for entity-aware video captioning on a basketball benchmark

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:55.465434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:53.202516Z digest=sha256:2ca626de18a436be4661db2263e0642a43f2f769f5f952bf424f56fd1d2b3441

Observation c9368bf4-c517-4bea-9e43-b0b69e46abe0 · outbound

This paper cites Eika: Explicit & im- plicit knowledge-augmented network for entity-aware sports video captioning.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Eika: Explicit & im- plicit knowledge-augmented network for entity-aware sports video captioning

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:55.304502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:53.279524Z digest=sha256:20cc5b800ca2568a5035a3d01bfa701286048e60020fdcd74d17bc3fbddc44d6

Observation 93235d0d-b85a-4a05-b083-928108e89e78 · outbound

This paper cites Msr-vtt: A large video description dataset for bridging video and language.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Msr-vtt: A large video description dataset for bridging video and language

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:55.163118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:53.363030Z digest=sha256:683f3fd142b52199c3af485fcff7636ee8ad751fdadfefaa25be7af94ff93ffc

Observation 66f835c6-e6b0-4a1e-9771-1efe8ff04925 · outbound

This paper cites Hierarchical modular network for video captioning.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Hierarchical modular network for video captioning

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:54.964288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:53.424019Z digest=sha256:21d5c712738ca54329b70a34ec0330f965241c162663c29739dc01cd578f704b

Observation 4357c369-9a65-48c0-9f0f-736c24d4eb78 · outbound

This paper cites Fine-grained video captioning for sports narrative.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Fine-grained video captioning for sports narrative

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:54.829046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:53.530146Z digest=sha256:cbb00f752392c7250d694e7439d221da29c1bc682428b54b27a89b76f518ff81

Observation 0fe3ff17-e83f-4ba0-9c7d-568a37f848fb · outbound

This paper cites Movie101: A New Movie Understanding Benchmark.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Movie101: A New Movie Understanding Benchmark

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-08-06T13:47:54.028783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:53.617274Z digest=sha256:2749835d1233fb4e08c5e6d548eaa9ec37a1be7efcb332989b5b80e666019d20

Observation c7834a93-67fc-4937-83c0-de9ac951366d · outbound

This paper cites Harnessing large language models for training-free video anomaly detection.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Harnessing large language models for training-free video anomaly detection

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:54.711821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:53.702448Z digest=sha256:ca327303c13c911fd33fcd96e148fa3b754629c9fde753d198903501513346a3

Observation ac9f3192-e1d3-4827-9cf4-099253389e48 · outbound

This paper cites A descriptive basketball highlight dataset for automatic commentary gen- eration.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning A descriptive basketball highlight dataset for automatic commentary gen- eration

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:54.589144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:53.781975Z digest=sha256:8d0294b5866a30a0b6de0fa05ef137bd0d3abdcb3a978780b51f18603f361b3e

Observation 3d044e4e-301e-400c-90f2-2c0b96895b92 · outbound

This paper cites jump ball.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning jump ball

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:54.449519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T13:47:53.840267Z digest=sha256:1c4e874ae18be9a99b0f5c045eb5be3a1f19a6cafc3e301a83b98c982688ef44

Pith citing papers

Observation 950a8896-23ab-4e88-9383-3528f1d53a28 · inbound

Explainable Action Form Assessment by Exploiting Multimodal Chain-of-Thoughts Reasoning cites this paper.

Explainable Action Form Assessment by Exploiting Multimodal Chain-of-Thoughts Reasoning Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-03T15:55:21.629729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:55:21.629729Z digest=sha256:680147c0a9a8eac1eef27b476b95a267e323dc5016b81e1ee6d98ffe71fe74b8