Pith. sign in

Paper Citation Record · LEDGER

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning

As of 7 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 1 inbound Pith citation observation for arXiv:2507.20163.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.20163 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:47:53.840267Z

measured 59 of 59 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T15:55:21.629729Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact1
  • verified fuzzy49
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 615c7d59-3460-4dbd-ba72-8425a78f764c · outbound

This paper cites GPT-4 Technical Report.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:48.338225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:48.338225Z digest=sha256:506e7ee2a9fdedd3331fdea572e0a78d133fa204388e8f2479275c210f832efa

Observation c2c43baf-7352-4256-a8b6-28ee0082f827 · outbound

This paper cites Meteor: An automatic metric for mt evaluation with improved correlation with hu- man judgments.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Meteor: An automatic metric for mt evaluation with improved correlation with hu- man judgments

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:03.233657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:48.387302Z digest=sha256:e9356912a597485aa3e5f5417ef89b4de62b0e2368ba09610ff50e41bfa9d3d3

Observation fff18909-4ae3-4d86-8f4e-47e8ba2e2ad9 · outbound

This paper cites Is space-time attention all you need for video understanding? In ICML, page 4, 2021.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Is space-time attention all you need for video understanding? In ICML, page 4, 2021

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:03.098776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:48.471255Z digest=sha256:0f9964fd8b79b27b90f6ad3ae0d5e91fa512458836d8ab6fd4ab9daca123d204

Observation 411cb742-a6d4-41c2-babd-432823cd6856 · outbound

This paper cites Activitynet: A large-scale video benchmark for human activity understanding.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Activitynet: A large-scale video benchmark for human activity understanding

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:02.909167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:48.557375Z digest=sha256:edf04c307d8c6875570412690cd90027765e848f4e212e5bb82417d7ca44b6f8

Observation 0ec81a08-6cf2-43e0-8e5d-3dadb61c85c3 · outbound

This paper cites Quo vadis, action recognition? a new model and the kinetics dataset.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Quo vadis, action recognition? a new model and the kinetics dataset

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:02.760554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:48.651033Z digest=sha256:f0e26f516909b7b0adfe85c8a6752d512d05bf1317e6d4f2e312dd07276522df

Observation 66722b3d-12f4-41f1-9a81-a015d7c49748 · outbound

This paper cites Collecting highly paral- lel data for paraphrase evaluation.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Collecting highly paral- lel data for paraphrase evaluation

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:02.593360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:48.689591Z digest=sha256:41093a1710a7ba30ed78f13a889c41a92eaa0df9ed06a4c6ff901e70fc571a9f

Observation 059320fb-9413-43d9-9774-722ab4638539 · outbound

This paper cites Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:48.721206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:48.721206Z digest=sha256:c6655f53be3ea17e293efd290bfb85b3ac55910d02e65153a1f31fd92ee56cda

Observation e749d5fc-3a1c-493e-90f6-d4cdb1a3f469 · outbound

This paper cites Sportsmot: A large multi-object tracking dataset in multiple sports scenes.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Sportsmot: A large multi-object tracking dataset in multiple sports scenes

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:02.444435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:48.739135Z digest=sha256:09188f464fb67d3ab14c9cc92f3d93861d5d8cc3964c9836ca3b7eb68b4fbdfa

Observation 3ddf3f16-8a17-4b3c-9d57-bac93c3d0a91 · outbound

This paper cites A thou- sand frames in just a few words: Lingual description of videos through latent topics and sparse object stitching.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning A thou- sand frames in just a few words: Lingual description of videos through latent topics and sparse object stitching

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:02.255384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:48.785487Z digest=sha256:b1e95acd5b15df7f9992edb5d3dafbafb1ac9d1b2abf740419537c6aa258bab8

Observation 1ba5df05-4642-4508-a44d-c8cd9c866ffd · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:48.831059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:48.831059Z digest=sha256:1bc3ab54426ace8de26001ffd98cf2e35de0bc69cab25845e02217b891453175

Observation b0c30544-5258-410e-8fa8-503210534ff5 · outbound

This paper cites Soccer captioning: dataset, transformer-based model, and triple-level evaluation.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Soccer captioning: dataset, transformer-based model, and triple-level evaluation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:02.121089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:48.890391Z digest=sha256:b70c5d53151790eafa2e4d69bd76dbc11c139639c4756c76d84ac743573ed5ec

Observation 28674be7-f982-4455-aefa-ff76dfeaade9 · outbound

This paper cites an unresolved cited work.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T13:48:01.960339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:48.942982Z digest=sha256:57ef7d1f8f8287bda2cd52f5a5f53705f8b29d7ea9d94b5ed067dd356546bfa9

Observation ca1ec4e3-4f73-457f-bbb0-82bcc2a39c01 · outbound

This paper cites Deep residual learning for image recognition.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Deep residual learning for image recognition

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:01.789107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:48.982219Z digest=sha256:52f5e36888fa57512aeb11b32305d11ddbab4bf63d85209fe1b8a38b0a09e5c2

Observation f9ab540d-86f4-4e49-82d5-b22887b02b66 · outbound

This paper cites Gaussian Error Linear Units (GELUs).

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Gaussian Error Linear Units (GELUs)

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:49.051493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:49.051493Z digest=sha256:18e833e5ec6095005bcbbacf8e7483e438be0c87cd5c94b74c6c4951c790a2fb

Observation 830bdf3f-bf20-47c4-b106-f1922277feca · outbound

This paper cites Overview of temporal action detection based on deep learning.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Overview of temporal action detection based on deep learning

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:01.578230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:49.111293Z digest=sha256:c0ab4f32045e831fe770e1c24b8baaab8313bb47973a0bc9ca8eef90d37d9538

Observation 4d3642e8-9f19-4e66-afe3-f34de4e784d0 · outbound

This paper cites Learn- ing to generate move-by-move commentary for chess games from large-scale social forum data.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Learn- ing to generate move-by-move commentary for chess games from large-scale social forum data

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:01.371450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:49.229047Z digest=sha256:70db6b7c0416ebd3d703a22b56f309686d96ff239a8918eca7ffa70163b2110b

Observation 9b0d2555-6dd6-4586-bb63-a0dd0deded6e · outbound

This paper cites Learning seg- ment similarity and alignment in large-scale content based video retrieval.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Learning seg- ment similarity and alignment in large-scale content based video retrieval

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:01.166288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:49.383887Z digest=sha256:c30e393935f8cdec8d65bb0bcc184262089e1c73861b09ebb4baa992bfb508e9

Observation 9163806d-f645-4003-a9b9-4116062688c1 · outbound

This paper cites Automatic baseball commentary generation using deep learning.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Automatic baseball commentary generation using deep learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:01.005079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:49.465481Z digest=sha256:d316c44ce9d8187cdc8b2eb0de8c8b62be3a1ece6d7c3dca3ee33303d91aabfd

Observation 8d357874-4819-4675-b8b5-8f6966fb7123 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Adam: A Method for Stochastic Optimization

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:49.583990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:49.583990Z digest=sha256:f890da4bf7d8fc2cfe337a6a939418ca7c23d6c21635d1ca60d08e2d61cf1071

Observation 639bf675-0666-4cec-92d7-0ac52890f53c · outbound

This paper cites Video story- telling: Textual summaries for events.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Video story- telling: Textual summaries for events

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:00.869574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:49.829969Z digest=sha256:62ef13e6e2f4ef9820d55b0a4f143eba2b6009eb1c37e08b8d02b857dfd73a19

Observation 3550829e-2298-4d99-8af8-0e6f3c1021da · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Rouge: A package for automatic evaluation of summaries

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:00.680689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:49.952635Z digest=sha256:3d5171a8e11a4e73426662b5d5ebed3f7fa00b2d1bca3010fc530247a3c2e4d7

Observation 717f7a0e-aa5a-4888-8221-67085a8a4267 · outbound

This paper cites Swinbert: End- to-end transformers with sparse attention for video caption- ing.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Swinbert: End- to-end transformers with sparse attention for video caption- ing

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:00.501771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:50.083025Z digest=sha256:b03907ae5993730a35cec586f6a52297cdd3410351ea11597e963a456a13963a

Observation fd23068d-2a81-4dd3-a882-e03f26b28b21 · outbound

This paper cites Video swin transformer.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Video swin transformer

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:00.299165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:50.245337Z digest=sha256:ca2edca2396951693df545afd279f20365139b7b8fa51d66d166779a6a4e60c8

Observation bce3b90b-dbca-4d5a-91c8-665ec89a251a · outbound

This paper cites Llama 3.2 quantized models, 2024.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Llama 3.2 quantized models, 2024

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:48:00.097823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:50.372060Z digest=sha256:40506dac3c967adb4971a0472c9b01d299d8f521d96c520f9962fc2db0dd0a2f

Observation ccd780b4-4273-4abd-a016-7267ffcf0f44 · outbound

This paper cites Soccernet- caption: Dense video captioning for soccer broadcasts com- mentaries.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Soccernet- caption: Dense video captioning for soccer broadcasts com- mentaries

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:59.854049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:50.522585Z digest=sha256:843de5368bf6006254cb10f918199a4c90735e8d38cd96bafd1fd821501c5bfd

Observation 0fe48295-4fb5-4b2d-add4-c27f80b8e0b0 · outbound

This paper cites Search-oriented micro-video captioning.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Search-oriented micro-video captioning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:59.618980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:50.650288Z digest=sha256:eca4a32a1abb5c5b99d94707cc4e6126c99ddfb0c0e2747c51eb0190fd4d82ca

Observation b68e4fb3-067f-4dec-a121-f022be7d66ce · outbound

This paper cites Enhancing visual question answering through question-driven image captions as prompts.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Enhancing visual question answering through question-driven image captions as prompts

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:59.373041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:50.783054Z digest=sha256:47235330a3b082a0645b5af646beb22b4267f0a8bdc84d386c3b5a8ba3babef0

Observation 4a3e2810-37bb-4324-8303-65485346603d · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Bleu: a method for automatic evaluation of machine translation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:59.134945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:51.004697Z digest=sha256:521de1444c7c221b0e0fb6f17e71c6f886bbf5b46fefec511430d83d5bd00b5d

Observation b53fca66-fa12-430d-8a6e-7451aac9892a · outbound

This paper cites Identity- aware multi-sentence video description.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Identity- aware multi-sentence video description

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:58.894291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:51.226089Z digest=sha256:97801f43aa80c9d273718b782f005bde0f271a431bc9e583287370a20ce35079

Observation 37534dc4-5bcd-4462-9b01-fe8af62f296d · outbound

This paper cites Goal: A challenging knowledge-grounded video captioning benchmark for real- time soccer commentary generation.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Goal: A challenging knowledge-grounded video captioning benchmark for real- time soccer commentary generation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:58.654739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:51.400669Z digest=sha256:ad3e4678eccd972a822a8bee1e7e0cd87aa77789e3e5fa1ca0026c49a03ceb7d

Observation 77af1dcb-5943-4b4e-96b4-73de32742b74 · outbound

This paper cites Sports video captioning via attentive motion representation and group re- lationship modeling.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Sports video captioning via attentive motion representation and group re- lationship modeling

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:58.457349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:51.533250Z digest=sha256:cd5d6d7d3f3676a7d903006e257ecba5eddea0aaf2e99ef877c2ce9c5630bdc4

Observation f7630cbe-3fa3-4c7e-b011-dcb167702c73 · outbound

This paper cites Language models are unsupervised multitask learners.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Language models are unsupervised multitask learners

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:58.256585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:51.721508Z digest=sha256:13bba08aabcb588f68cb2626a4461fc0a38a20904bc73bf362373ad942bfe9e5

Observation 5fbd5acf-adb7-4f10-9ac1-637ecc7db084 · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Learn- ing transferable visual models from natural language super- vision

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:58.105707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:51.901836Z digest=sha256:708c880a6fe8d0f80034f83b0dcbcd5347f1e23138c40aa3db7ed40f55e22c36

Observation 17fcd46c-ce79-4d66-9dc7-ee5dda44c2ee · outbound

This paper cites MatchTime: Towards Automatic Soccer Game Commentary Generation.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning MatchTime: Towards Automatic Soccer Game Commentary Generation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:52.050904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:52.050904Z digest=sha256:e589227553b588bd7f9fbd671a2d6ebe938091b0edbc2cb778532ca03c734cc7

Observation 9ed940e0-1380-4904-adcc-1aaf416794ad · outbound

This paper cites Towards universal soccer video under- standing.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Towards universal soccer video under- standing

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:57.935512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:52.195571Z digest=sha256:4bf87d865ab1a895090ca8920cd11f033c0666941a896352833bcbaab5451f77

Observation c026b22e-20c4-4f7e-a4c7-94e13b78a84a · outbound

This paper cites Grounding action descriptions in videos.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Grounding action descriptions in videos

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:57.685758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:52.288758Z digest=sha256:e1fed44b562081bb8369305ea100e74388ad16cf16d146bfe8138a24cec57f1c

Observation dd8cd761-f1a1-428d-ac77-a9f617db746e · outbound

This paper cites Timechat: A time-sensitive multimodal large language model for long video understanding.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Timechat: A time-sensitive multimodal large language model for long video understanding

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:57.497245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:52.384320Z digest=sha256:55a144b29e1e1d4709b11a4935352b510b36dc6b27fc040e4d08f7aae7fa957e

Observation cccc58cf-1f89-4681-9e80-55edcc540cee · outbound

This paper cites A dataset for movie description.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning A dataset for movie description

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:57.261849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:52.476419Z digest=sha256:781264580201795d99eb7d7ad713d2d79e85d3a42d07639a2258dac486d3dfd0

Observation b1a26b85-86e0-4762-87ba-0b6a957a1e6c · outbound

This paper cites Accurate and fast compressed video captioning.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Accurate and fast compressed video captioning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:57.090318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:52.570908Z digest=sha256:858681e3a72dbddf6fe44ac71d241751fa2f0a6db8663ce0fbf3118f3036f590

Observation 358d8369-dea1-4109-938e-9adea259b196 · outbound

This paper cites An overview of the tesseract ocr engine.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning An overview of the tesseract ocr engine

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:56.929106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:52.629207Z digest=sha256:3237d472e3927efde85b05e4c03f31ff196fc12649d05d50fba16a149245a4c7

Observation f79916d6-9ca0-4e4f-abeb-d43dac9275c1 · outbound

This paper cites Clip4caption: Clip for video caption.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Clip4caption: Clip for video caption

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:56.762986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:52.681290Z digest=sha256:47925430a3f1088cc2f2eb7c29ae5e83831d54ebf86dcde2ab4931fca531a3fc

Observation 37b257fc-62a8-4905-8fe4-d1b9b93c9c43 · outbound

This paper cites Qwen2.5: A party of foundation models, 2024.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Qwen2.5: A party of foundation models, 2024

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:56.634122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:52.753771Z digest=sha256:17f0712d4ec1170b59dfbeb5ad51915bb9b3ab5ab82a09b5b43b51800e56623f

Observation 6ce10f36-56a2-4295-816e-cb5b157b3675 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T13:47:52.792329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:47:52.792329Z digest=sha256:ab479a8f0bd2fed63434efef608f96fcc6dc0a5003cb8cf4b2d92f9fdec90b83

Observation f50975b2-6527-41d1-b5bb-58210b190ba9 · outbound

This paper cites Atten- tion is all you need.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Atten- tion is all you need

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:56.470095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:52.859030Z digest=sha256:ea74f422f98409535850f020ce0e8129289c3227ea9d8ee29db12dcf8b61d48d

Observation 1c967b64-34d4-4209-935d-c7c83bc22406 · outbound

This paper cites Player tracking and identification in ice hockey.Expert systems with applications, 213:119250, 2023.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Player tracking and identification in ice hockey.Expert systems with applications, 213:119250, 2023

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:56.298252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:52.919724Z digest=sha256:fababfd1679eb2c11911107d61cb09aab2708c21e00f081c0d15a444a714f311

Observation 98ad022a-2a58-4206-80ae-5b22fdb2ef3a · outbound

This paper cites Cider: Consensus-based image description evalua- tion.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Cider: Consensus-based image description evalua- tion

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:56.092821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:52.975748Z digest=sha256:028b49165dbc3d0c061a69a6d1feb353f97d39366cbc1f778d4a23eb352b026e

Observation 265d17ab-2956-4a69-85f1-c0a1c0511237 · outbound

This paper cites Omnivid: A generative framework for universal video understanding.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Omnivid: A generative framework for universal video understanding

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:55.936975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:53.023439Z digest=sha256:f9de025ecfd8c575653046fb94fdef0fd72012cbf1afc940455361f33533d89b

Observation 371e9520-823e-41f1-aa81-b12c4ac8fd27 · outbound

This paper cites Sports video anal- ysis on large-scale data.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Sports video anal- ysis on large-scale data

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:55.808364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:53.059771Z digest=sha256:65b2608ccaa25d649e5a6e0f6703f45a5a0a204c63637064dedfe80cde27e9dc

Observation 4e6ee851-9abe-4eab-8126-257f9d148816 · outbound

This paper cites Learning label semantics for weakly supervised group activity recognition.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Learning label semantics for weakly supervised group activity recognition

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:55.644024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:53.106462Z digest=sha256:fe12d0e0d64641a997eca4f9a0ca4a5bd4b403d14544f081dec0b9e4ad90ce68

Observation d8c8a3ac-9f04-42ef-b5ff-18212b579d07 · outbound

This paper cites A simple yet effective knowledge guided method for entity-aware video captioning on a basketball benchmark.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning A simple yet effective knowledge guided method for entity-aware video captioning on a basketball benchmark

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:55.465434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:53.202516Z digest=sha256:de9fd82e9dee07c7e88cb738339f5340a6725f515dbe084d26e273f01854722f

Observation c9368bf4-c517-4bea-9e43-b0b69e46abe0 · outbound

This paper cites Eika: Explicit & im- plicit knowledge-augmented network for entity-aware sports video captioning.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Eika: Explicit & im- plicit knowledge-augmented network for entity-aware sports video captioning

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:55.304502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:53.279524Z digest=sha256:039a9a6c1d9c1d7de5f5cbef3c41b5be253f155cd8290be7f1a25a52b0865ae2

Observation 93235d0d-b85a-4a05-b083-928108e89e78 · outbound

This paper cites Msr-vtt: A large video description dataset for bridging video and language.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Msr-vtt: A large video description dataset for bridging video and language

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:55.163118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:53.363030Z digest=sha256:f0c6381da7642d0a3bbf7dfdc0ab7a7769b0be31c6d9e4dcb0df7addcb256828

Observation 66f835c6-e6b0-4a1e-9771-1efe8ff04925 · outbound

This paper cites Hierarchical modular network for video captioning.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Hierarchical modular network for video captioning

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:54.964288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:53.424019Z digest=sha256:297f18b6b300a0221aca782a7be8936447b1ba885c1464933de1988db45197bf

Observation 4357c369-9a65-48c0-9f0f-736c24d4eb78 · outbound

This paper cites Fine-grained video captioning for sports narrative.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Fine-grained video captioning for sports narrative

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:54.829046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:53.530146Z digest=sha256:601695feb350785b9b23856802f3e5c5aba93906c45c92713cb9688e0a2945ac

Observation 0fe3ff17-e83f-4ba0-9c7d-568a37f848fb · outbound

This paper cites Movie101: A New Movie Understanding Benchmark.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Movie101: A New Movie Understanding Benchmark

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-08-06T13:47:54.028783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:53.617274Z digest=sha256:53ac6137671467342d7c3616fbf0e35d24ec6a842e4d3f4eabcda0fb19d2cad6

Observation c7834a93-67fc-4937-83c0-de9ac951366d · outbound

This paper cites Harnessing large language models for training-free video anomaly detection.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning Harnessing large language models for training-free video anomaly detection

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:54.711821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:53.702448Z digest=sha256:2ac00bf3c18f15f8b88ee7cc8424dfd831b4c11dfdc147acd4b5068e26fafa3a

Observation ac9f3192-e1d3-4827-9cf4-099253389e48 · outbound

This paper cites A descriptive basketball highlight dataset for automatic commentary gen- eration.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning A descriptive basketball highlight dataset for automatic commentary gen- eration

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:54.589144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:53.781975Z digest=sha256:23d626f2cfa67dda89a7d92ce71b255064f220de1cb554ac2c253f431e11a03c

Observation 3d044e4e-301e-400c-90f2-2c0b96895b92 · outbound

This paper cites jump ball.

Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning jump ball

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:47:54.449519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:47:53.840267Z digest=sha256:573b66cb37bc42e54f8081b154f5e2b90ab5ac00fc0d3588f8d21db55f658cfd

Pith citing papers

Observation 950a8896-23ab-4e88-9383-3528f1d53a28 · inbound

Explainable Action Form Assessment by Exploiting Multimodal Chain-of-Thoughts Reasoning cites this paper.

Explainable Action Form Assessment by Exploiting Multimodal Chain-of-Thoughts Reasoning Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-03T15:55:21.629729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:55:21.629729Z digest=sha256:680147c0a9a8eac1eef27b476b95a267e323dc5016b81e1ee6d98ffe71fe74b8