Pith. sign in

Paper Citation Record · LEDGER

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning

As of 12 August 2026, this Paper Citation Record lists 70 of 70 outbound references and 0 inbound Pith citation observations for arXiv:2412.13543.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.13543 v1

Coverage vector

measured 70 of 70 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T13:05:30.318725Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

70 of 70 outbound references displayed

  • verified exact3
  • verified fuzzy11
  • unresolved56
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 621797f1-6c3a-4f8f-860d-dfef50c566f8 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.981161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.098451Z digest=sha256:97e86e9dec1e5d0e60b8f09d9f14dc16536489bfadab50bacd3deaab97cf8b22

Observation 46a85395-038c-4ba6-843c-8ffbbf226cb8 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.972187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.102460Z digest=sha256:d715e4d66e213ca126e33179e63ccd8466eb354947828fa9ffa26474e10c78fe

Observation 570dc33f-feb1-477e-982b-9a4280959c1a · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.961175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.105913Z digest=sha256:e8de7b959a7eeb4388fdf4cf2b4228d62906909099edce490c3a88884dac1bdd

Observation dbadf8c3-7abf-4b43-bd6d-97455a081214 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.951753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.109727Z digest=sha256:605d0977c61c1972e32ebfab7aef31b4bd55278c211680d6f73dd9c508c826be

Observation 842cb8d9-10b4-44b6-af97-dbc750b8e10e · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.940841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.113082Z digest=sha256:56199b4694914c1ea1d75e1eae85c964575d26d7360f0fc1ebeda840333df81e

Observation 8432f73c-73ba-48d2-955a-73451cec0a78 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.931518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.116379Z digest=sha256:6b355e7ac343bf312d907ad105720b3d95f115384cdb656d9c9757962345b65d

Observation 0f906072-7afb-45c4-910c-c4ae634651cc · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.921799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.119822Z digest=sha256:a3fb80e80c5e367326e5f6f90343ccadfc4584a8a7e04b37e0e3c571de2fb8c5

Observation 7522df0c-121d-40c1-ae6b-3521aa06d7b3 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.912057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.123061Z digest=sha256:65f12ec93fee774ad8ec70aaece5e0fce928f0a4f46ce2d17f7174e97a716be3

Observation 7aadd44e-31d9-4b42-86c3-662695446d4e · outbound

This paper cites F.; Ellis, D.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning F.; Ellis, D

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:05:30.902967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.126024Z digest=sha256:bc9a2087f2c5b70865139103aa727020fc7e0b5be3971f332e43173faa32cc61

Observation 2d030e45-694b-4690-ad3a-e0a2ff197b74 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.893635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.128915Z digest=sha256:2265d95ac8a0ebe0c19f2ed485765965685d223439e17279141602639b4e7ad6

Observation 432e0ef8-a3cf-4b05-bc43-518b6d5f804b · outbound

This paper cites The Kinetics Human Action Video Dataset.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning The Kinetics Human Action Video Dataset

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T13:05:30.131913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:05:30.131913Z digest=sha256:7b69359caff04b08c268739604cdacf346025adfef1618a0786def5901137f54

Observation 321967bc-bbd3-40e2-bf3b-7d2e6688494a · outbound

This paper cites B.; Moon, J.; Choi, J.; and Kim, S.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning B.; Moon, J.; Choi, J.; and Kim, S

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:05:30.884192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.135722Z digest=sha256:d7dc01183a6cdd676bb49f5f7eba841fd28c4096708b7035cd1953caf9a957b1

Observation cdcb2e3a-058d-4721-9efa-37a45bfd6ebe · outbound

This paper cites C.; Lo, W.-Y.; et al.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning C.; Lo, W.-Y.; et al

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:05:30.875021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.138866Z digest=sha256:dc18a850064dc85b6254d20951f563c7720e021c2eaff3dac39573618509556f

Observation e5529ead-eb42-4415-af04-a0e14d8da3da · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T13:05:30.141862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:05:30.141862Z digest=sha256:21e3c98a28a9f908baca523a9e6863ca121122fc75ffc944eae1e074eb7c57f1

Observation 7396238c-efd2-43e4-95ce-82a9d11af8c0 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.859941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.144761Z digest=sha256:dc57b57e4bc335dbcc920f6a88b6b861ac429fac8418629c69322d4fa6d67638

Observation 1d0b9755-324d-422f-a8c3-2fb6bd92ce0b · outbound

This paper cites L.; and Bansal, M.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning L.; and Bansal, M

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:05:30.849962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.147673Z digest=sha256:46764df10104789f31ec3ddc46e4cd4615edc7fe32ca6dec372bdd8fa59a2b4b

Observation 7c866395-42a5-4301-9fdb-b64256c49b98 · outbound

This paper cites L.; and Bansal, M.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning L.; and Bansal, M

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:05:30.839166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.150332Z digest=sha256:c8f6c4fd2626dab18cb0c7d476837bc5ca008dfa300553102d8c6d84179a7392

Observation ac6fb3c8-db07-4f54-b4a1-61dad96d1671 · outbound

This paper cites Progressive Feature Mining and External Knowledge-Assisted Text-Pedestrian Image Retrieval.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Progressive Feature Mining and External Knowledge-Assisted Text-Pedestrian Image Retrieval

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-11T13:05:30.389863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.153175Z digest=sha256:c56ad7c836242ba00aa771ad29db5db9980c8e84b6ea8ca3e2106ce3912116d1

Observation 48415601-61ea-4ca6-a44d-cbbf71e306de · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.829660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.156503Z digest=sha256:9e4015ea39c88f2b040134e1289d5f2066dedaf25b393accc08465e152b9ba96

Observation 8bc981ab-3cdb-48d6-968e-5e77b8612307 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.819345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.159731Z digest=sha256:d7bc9c3fa6b4ed979f13e93017c829919414b41b9332ecd5e59815b834edc6c9

Observation d9c78f73-157a-4adb-b8a5-f6413d93406e · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.810017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.163043Z digest=sha256:b8c28d1cbbb06f2d2adb7dde82ddf5669f1fb755f8611b84d681f760280cc2b8

Observation a12a47b7-7efd-42d3-a236-159d792ae7bd · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T13:05:30.166205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:05:30.166205Z digest=sha256:4d15e52d77d6007cad93f47aace43996f3ab28d70b47dd5189d2ec84f55c5147

Observation 7224c271-c10f-4bc5-b11d-2028699f89d8 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.795369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.169272Z digest=sha256:dccb512da5f167b752f32733222e3ea9d69aab1de1fe542ddb5a46d9b9599e0c

Observation 179b9a92-a563-4919-a0b8-2e9eedb94028 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.786073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.172330Z digest=sha256:bd55dc7edf7d318f4be0e6dc6acb25e1e9603e7a1a7789a755acac93f7e5ea1d

Observation 507c45e5-ba7b-4bd0-a8b8-802d1a0f632e · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.776162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.175565Z digest=sha256:a9cc98826e52232236b50ef0e8a2e5b4ada712a801700d3b632d8da0198caa57

Observation d0c42b9a-eab6-4f20-8e35-2980c6a0711d · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.766160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.178523Z digest=sha256:0970794da0dffb331139c74076f91b8585d2ba0184efe453bc681aed8a9dbbd6

Observation 4db48883-a127-4233-a809-1c978a722af3 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.756559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.181447Z digest=sha256:f8c7042d658502ee20b95dded8df802c1ab8de30573377c9260a031563925692

Observation de6b28d9-7cfb-43d6-85f5-77d6c2694d36 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.746349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.184540Z digest=sha256:23eb2d6a955e0ad85c5f956366d354b5cbef3ccff8513fead70504687ce415c8

Observation ba993c6d-6178-4537-9aed-8b06b7d0de6d · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Representation Learning with Contrastive Predictive Coding

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T13:05:30.187748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:05:30.187748Z digest=sha256:245ccfecf31ab00b932acaa6a2ca97e851e04aab349376415dade4ceeff7f6d5

Observation d9243104-2090-4650-840c-1f2d8d66d667 · outbound

This paper cites a ckstr \.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning a ckstr \

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:05:30.736571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.191315Z digest=sha256:366d732a2000c2a447065811855fbf9ea419bbb7535f5e5e76d5f91cc44f8883

Observation b09caa4e-eb33-4e69-a43c-2dfae1c041c3 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T13:05:30.194852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:05:30.194852Z digest=sha256:db3f501f395ec93193f48f14ae00f7fe58308ef7aa535b6300a6cfc3b3591e47

Observation 85c48127-17b2-498f-b98d-1847bc1448a1 · outbound

This paper cites E.; Neumann, M.; Iyyer, M.; Gardner, M.; Clark, C.; Lee, K.; and Zettlemoyer, L.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning E.; Neumann, M.; Iyyer, M.; Gardner, M.; Clark, C.; Lee, K.; and Zettlemoyer, L

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:05:30.721073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.197899Z digest=sha256:6dcacef3f924d1a9a7302f34d82a2dd97486f342517b4cded91499d7081376c9

Observation 11d9d22a-125e-44fc-b601-c017077bc0ba · outbound

This paper cites W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; et al.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; et al

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T13:05:30.200832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:05:30.200832Z digest=sha256:48e7867f37056e113642591c7883be262ee45259a948ca2461a63195d09b61eb

Observation 1f43b72f-16b2-4d4f-8d1c-aae02b5491b7 · outbound

This paper cites W.; Xu, T.; Brockman, G.; McLeavey, C.; and Sutskever, I.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning W.; Xu, T.; Brockman, G.; McLeavey, C.; and Sutskever, I

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:05:30.705846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.203934Z digest=sha256:753d91715e46ad473310ac67f4001332d6249ea91facb9b80acf13ea2c1568fd

Observation 317c6264-0d7f-474e-8de5-b3009be886c2 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.697089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.207314Z digest=sha256:946b5518213f430bde79cce64f298ba914f35a5c9e1df39f20e0b7a8531d3103

Observation 2c4eed74-6460-4a35-9eaf-25db5747b1d6 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.687783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.210609Z digest=sha256:0f69f5b44760444cfc0ef1831dde3ff95ffda8573f54dad3e7a4b4614866018a

Observation 24b8bfac-d9ce-42ce-9150-e8ca00d4d1ca · outbound

This paper cites S.; and Gong, B.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning S.; and Gong, B

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:05:30.679070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.213667Z digest=sha256:535564bbd4fa6a36bbb74f074b7a1514041cd1ed85b1210fe8161da56ff4457c

Observation efc55409-d1b6-4841-a8bb-34fc11f1be4e · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.669610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.216740Z digest=sha256:b119dcad9f8f358ed726c90e9ce2d4b07a8a41e08330c6b8e6d48f7d599831b4

Observation d0d13d00-66c7-4f14-afb4-98112d5af44d · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.660839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.219711Z digest=sha256:26a89697f1c460e4c0e6a2937771c20bad23deea7b6128d4b21b4884d2d91a34

Observation 1a8bfeb7-dd86-4962-820a-1ff754fa62d4 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.651585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.222672Z digest=sha256:69954de9bfdfdc227864a334c8086a8425a5e9f1ad40052bf6c27c7a5f9cd7f7

Observation 70af3304-33c7-482d-8ae5-b6321f01c5e5 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.642684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.225646Z digest=sha256:df88345a1272808f6c778dde89c313af5f3c0f12a7f0e29a1af22a34cfe1be02

Observation 13f327e9-56a7-4fd4-9ffb-19d013766b01 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.633664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.228749Z digest=sha256:44bccf983df3242e61ef1f0e6fe6b11ade307595e1e2cd19e9cd2b4d9de9b372

Observation cc51fef9-1693-4960-b48d-8860aa2019e8 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.623925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.231795Z digest=sha256:943967d04e416f06ff0d48da22bec92f5a09953b30e6e594f85cb981434e34c1

Observation 453c791e-3257-4725-a86e-cd0983447a08 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.614968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.235073Z digest=sha256:b49b379b468131f288e250a6989141b52a7330f93c282592575289d3cf87026a

Observation 84639bf0-59dc-4504-a8b8-84534fab5558 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.605826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.238018Z digest=sha256:b77ac71ccd564e36a4ffea7aa19dfbd598c4490a55708362c3cbff0028cc20c5

Observation da6a7dcf-ef1f-416f-b9b5-1273ee82172b · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.596246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.241098Z digest=sha256:636a99790b7dd5cbf45bdf2c19e9e87157de7e78a2c6c520376aa08b60c1956d

Observation afc69608-36e0-4da0-9d60-2ab2dcf2e100 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.586905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.244117Z digest=sha256:23765568d1fbe03f660efc743e32add715a82f3fd2faf3497c2fe613e4c74ecc

Observation c9121db9-9e08-42d6-8692-6505a0ed1b7e · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.576808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.247158Z digest=sha256:5e3f213b381ae38319d432fe4c616cc2657d8e6ddc8dafc5e3a7ca75cb6de4ff

Observation f8dbdd3c-08b0-4938-b86f-e5b045e0244b · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.568166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.250144Z digest=sha256:1cc6eb2d0b079660143dd3ebcf83654b5ad8de53bb1150f33c30a70ac59c1ee4

Observation 558ab4ca-667c-49f9-85ac-bc9dbc030ca3 · outbound

This paper cites N.; Kaiser, .; and Polosukhin, I.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning N.; Kaiser, .; and Polosukhin, I

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:05:30.559025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.253170Z digest=sha256:41185643d2e7a2b52c2d53bd160185d882ccca872b3ced746a33228ef35e7fa7

Observation 93e2ec46-2b47-41e0-8cd0-d13736eef4e9 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T13:05:30.256242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:05:30.256242Z digest=sha256:cab51a47004631e007b8aec5e5353b733d79dc169f1cbab86bf775f6e69b844e

Observation 17ccf0fc-06f1-4aa7-a669-eecfa0cbb0ea · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.543656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.259446Z digest=sha256:1b3f43efa7f420d3c856a0a3e3e13828748cd9b317cc5d95db0c51b802292e7d

Observation 80e057c9-42d4-4b7d-80a1-68ecc389297f · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.535610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.262374Z digest=sha256:1a5f9183846fd0de66e01f9e6e9f9992d13d19b39abd4877f273639cb149d10b

Observation d1da524e-6e8d-46e2-9bc4-167b63b1c201 · outbound

This paper cites Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-08-11T13:05:30.367482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.265355Z digest=sha256:e9b9de0a07c3f41f1a54033a56d5e86bcd9f0062e772ac81ecbc42e2838a9f31

Observation a8962afa-2aa9-440b-93cd-72f757818b48 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.527032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.268775Z digest=sha256:206e130ae65222fbb0163efe6b52fd8024461249875b3debbfeaedbe44040d15

Observation 203aae18-a585-4471-a73e-4280fa3e4e43 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.518557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.271964Z digest=sha256:46cedc1d5f6418cbd6ae1c26945f9090bf1e85c50caf9efe22dc4ff707447fee

Observation da782162-49c1-4ed9-8170-6f69ceac2984 · outbound

This paper cites Phrase Decoupling Cross-Modal Hierarchical Matching and Progressive Position Correction for Visual Grounding.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Phrase Decoupling Cross-Modal Hierarchical Matching and Progressive Position Correction for Visual Grounding

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-08-11T13:05:30.354127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.274961Z digest=sha256:7908eb7bf0312889f60c97f337241bcdbcb27eedb4d628f5cf8c440ec277f819

Observation 6f1f3dbd-f444-4f05-8397-c8d2ce1cb1dd · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.509619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.278350Z digest=sha256:dc2cb13a84f9f6798f63d5383f3f7b857a56a8d3111560401033724ee389c344

Observation 1746452d-af49-46d6-b04d-ba1990c53b7a · outbound

This paper cites H.; Miech, A.; Pont-Tuset, J.; Laptev, I.; Sivic, J.; and Schmid, C.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning H.; Miech, A.; Pont-Tuset, J.; Laptev, I.; Sivic, J.; and Schmid, C

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T13:05:30.501011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.281379Z digest=sha256:b8ff59f4a068a57531976f24bbf0d1ecd1c302ecb10933be0528ad77499ed008

Observation 09eebf43-4360-44b5-87a8-c0c41fd14d47 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.492088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.285621Z digest=sha256:462dcdaf23cc0bae184ad8e1409c8836599bff4068857719e3def5a8833aaef1

Observation 8173df16-b73d-4ce2-888b-374a9112814f · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.482993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.289227Z digest=sha256:37ee5c5279da368b89fed32cd1a8a02575f8e58daa9c0d7f23642613fd2827d0

Observation 2af6b1c6-f29f-4868-99f2-d20a0ccbea05 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.473921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.292482Z digest=sha256:ead3c942ad5b4aab4754676f101ecf514f898e649863b96f56492f778ea08487

Observation c00ab8c4-7ec8-4a0f-83b9-3e2119c2b5e6 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.463972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.295720Z digest=sha256:10ec69af397004644597e9f8f3849f1f90db0c7d4a5b316856da37edbc873570

Observation f556a9cb-25c8-456a-b01e-92b96160dd08 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.454283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.299041Z digest=sha256:c22bf841ee82e42dce26b93c2874e66d07502e0aa2fffb3d4f2667f8853f55a2

Observation c5b4cb8f-840d-4c6f-a5b4-35794465efc8 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.445543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.302218Z digest=sha256:b20feea33d62b246b4cda91bb0ffdb87c64b46092d6fb70fa603a43aa85691e5

Observation 399f3877-72aa-4a48-ad97-e56dabdde0ee · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.436792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.305257Z digest=sha256:388d2f5009fa3b268f8d85c3afcecea3e471212837ee20305634950eae61799d

Observation a0b036bc-c213-4e29-878a-b06cc2297a2b · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.427753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.308405Z digest=sha256:e3886b1c2a68f7de01c7623693fe684bb5214302779c2b53e4fbdead91883c46

Observation 7aeaa1e0-1375-4a13-80f6-4b957f735140 · outbound

This paper cites an unresolved cited work.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Unresolved cited work

Reference 68

Resolution
unresolved
raw_fallback, observed 2026-08-11T13:05:30.418347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.311665Z digest=sha256:079f7dfe0bc3c923491650ab43bb225d705322305684990d827acdb1b5ccea32

Observation 27983f8e-47bc-4615-b94a-e6f29b3719f9 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning , " * write output.state after.block = add.period write newline

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-11T13:05:30.315148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:05:30.315148Z digest=sha256:dd0956c160bd77443fcd334e2509fa4b578cdd7f2531b7a2905b5d0152b55b30

Observation 7f2dcf3e-49ac-4e2c-85d1-699328a7c696 · outbound

This paper cites write newline.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning write newline

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-11T13:05:30.318725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:05:30.318725Z digest=sha256:b9fcc799bc89d83d4390e938703c0e12761f02b22796aacf3c8e8f637cf78221

Pith citing papers

No inbound Pith citation observations are available.