Pith. sign in

Paper Citation Record · LEDGER

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding

As of 12 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 1 inbound Pith citation observation for arXiv:2411.17481.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.17481 v1

Coverage vector

measured 68 of 68 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:10:55.840477Z

measured 69 of 69 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T13:05:30.265355Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-11T13:05:30.363755Z

Reference resolution

68 of 68 outbound references displayed

  • verified exact0
  • verified fuzzy59
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation eaef100c-87e0-457b-aabc-b97721150d6c · outbound

This paper cites Real-world anomaly detection in surveillance videos,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Real-world anomaly detection in surveillance videos,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:58.007363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.221759Z digest=sha256:f8423575e28bdae268094e140ed09dfed6ec00372f5b2b563bab6d715b5aebe7

Observation 31b5088c-94aa-4529-b8f3-823ee63a1137 · outbound

This paper cites Toward video anomaly retrieval from video anomaly detection: New benchmarks and model,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Toward video anomaly retrieval from video anomaly detection: New benchmarks and model,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:55.228967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:55.228967Z digest=sha256:11945670cff42e17deda3300184b97aff25fa9af4dc39664e827685ca840da98

Observation 9a11d131-dc00-4c02-96d6-908cedde8615 · outbound

This paper cites Robust multi-drone multi-target tracking to resolve target occlusion: A benchmark,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Robust multi-drone multi-target tracking to resolve target occlusion: A benchmark,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.966703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.235740Z digest=sha256:85c1bd4544cb7f56c607720c410df9228b0aae7f08f542f9d81e237042bc2467

Observation a628c639-8cd9-4491-afed-f84628e65182 · outbound

This paper cites Yolov3-mt: A yolov3 using multi-target tracking for vehicle visual detection,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Yolov3-mt: A yolov3 using multi-target tracking for vehicle visual detection,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.943705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.243540Z digest=sha256:7dc59c118885a7692150176db1d295d2335cf21a70bcbc49e5cd8b1594a6cd0d

Observation 57d5fea4-af98-469e-9c5c-d4d857be7c1b · outbound

This paper cites Deepmtt: A deep learning maneuvering target-tracking algorithm based on bidirectional lstm network,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Deepmtt: A deep learning maneuvering target-tracking algorithm based on bidirectional lstm network,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.914941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.249889Z digest=sha256:ad0badfc397400e2c0204da6e781d56f0ac46e8645293e1c23debede9ed3f08f

Observation 97b607a8-b33d-4e23-9c8b-83c9aadb03a0 · outbound

This paper cites Robust obstacle detection and recognition for driver assistance systems,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Robust obstacle detection and recognition for driver assistance systems,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.892101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.256748Z digest=sha256:c8e291495d997bb100f350c5224e2b70e6c6b88705c366548021a96566cad4ac

Observation 281b165e-5ea9-44c5-a970-973b0f57012b · outbound

This paper cites End-to-end object detection with transformers,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding End-to-end object detection with transformers,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:55.271170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:55.271170Z digest=sha256:ea754d28a15679e0e441d64bee371bb0b9f8921381d43344b9374c871e1514e8

Observation 07177e39-30b7-4ea5-97de-cc8b83974c3b · outbound

This paper cites Pareto refocusing for drone-view object detection,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Pareto refocusing for drone-view object detection,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.851391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.280109Z digest=sha256:238eaffb039e9ad70a0cc65be3c16d6146c240ef140b8174be64bb1894bb987c

Observation 5519f4f9-17fe-4909-a7ab-3c5b52b19743 · outbound

This paper cites Sparse r-cnn: End-to-end object detection with learnable proposals,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Sparse r-cnn: End-to-end object detection with learnable proposals,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.828732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.287581Z digest=sha256:ba33ffe783767e1819e2148ddd64e43621aa1311e1600972ca97359029dca789

Observation 6ab04fce-432d-425c-ac7b-bd8720d7d5c0 · outbound

This paper cites Recent advances for aerial object detection: A survey,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Recent advances for aerial object detection: A survey,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.801878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.296963Z digest=sha256:84addda0a6a3270a7ab1d62ed94cc5973ac3ce40f58b3ac73a8ec2198d2af22c

Observation cd412ab2-9ec9-43ff-9eff-a59525b374f7 · outbound

This paper cites Cascade r-cnn: Delving into high quality object detection,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Cascade r-cnn: Delving into high quality object detection,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.777387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.303386Z digest=sha256:b815eeecbf67f1b9d910a0be77525bc2a2af7bb0fc30d94ec2bd71d498c9f100

Observation af5b13b6-11b2-4950-a99b-9b5046a65f1c · outbound

This paper cites Crnet: Context-guided reasoning network for detecting hard objects,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Crnet: Context-guided reasoning network for detecting hard objects,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:55.309511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:55.309511Z digest=sha256:dea3aaeb15874cc95fff5ccd1a113a2df9e72269518946504014b8142de02225

Observation bd28b2bf-9a4d-4d8e-ae3c-0d2a4ac5240c · outbound

This paper cites Triple adversarial learning and multi-view imaginative reasoning for unsupervised domain adaptation person re-identification,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Triple adversarial learning and multi-view imaginative reasoning for unsupervised domain adaptation person re-identification,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.710758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.314739Z digest=sha256:a2b09e8e354a2b06325f7c643bbbedfd68ed747fd49efc75411ada33408e103a

Observation dce6daa2-8e7f-48ad-9612-fc98dd500735 · outbound

This paper cites Logical relation inference and multiview information interaction for domain adaptation person re-identification,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Logical relation inference and multiview information interaction for domain adaptation person re-identification,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.670988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.323145Z digest=sha256:340ad23ad59fb88bf6000541151397450291314f709b3677e7e0602ffddfb5b8

Observation 779b5200-cfa5-4edd-8ae3-7813fb51ac7f · outbound

This paper cites Attribute-aligned domain- invariant feature learning for unsupervised domain adaptation person re-identification,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Attribute-aligned domain- invariant feature learning for unsupervised domain adaptation person re-identification,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.631604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.331492Z digest=sha256:a557e7805fdd96ea62549dfd368fb9bf2bd8fbdb077d10af963e4cdb80a32508

Observation b04977b5-f773-42a5-8338-e1f790888dc4 · outbound

This paper cites Intermediary-guided bidi- rectional spatial–temporal aggregation network for video-based visible- infrared person re-identification,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Intermediary-guided bidi- rectional spatial–temporal aggregation network for video-based visible- infrared person re-identification,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.605290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.338102Z digest=sha256:8b7e2ee92fb459c76d74d97ff6cad15568bfea28674da07c211c6e15795806a7

Observation 0c2c96ac-253f-426f-ae2f-1b87d47fa8e0 · outbound

This paper cites Video moment retrieval from text queries via single frame annotation,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Video moment retrieval from text queries via single frame annotation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.579605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.346074Z digest=sha256:ae421629533def951275957716dbd651d9a0e9466b1e25f2ff92dd94af1c9fde

Observation a4be4d36-af3a-481d-806f-989a0ead684f · outbound

This paper cites Text-based local- ization of moments in a video corpus,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Text-based local- ization of moments in a video corpus,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.544199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.352979Z digest=sha256:986f89681c200abec1ab44eeb87f9e6167c1b7aa0d7e7bef6b52f6282b5e9d92

Observation 2cc6412e-6b34-4c3f-a1d0-ae662818f792 · outbound

This paper cites Language-guided multi-granularity con- text aggregation for temporal sentence grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Language-guided multi-granularity con- text aggregation for temporal sentence grounding,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.506772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.358800Z digest=sha256:38d811dc2072405b8c14e0b636bbb37a7b15cf1f2dab2b3c53c5448b925d2eaf

Observation b3eddd23-d222-4e7b-aee1-d7939a42857c · outbound

This paper cites Conditional video diffusion network for fine-grained temporal sentence grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Conditional video diffusion network for fine-grained temporal sentence grounding,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:55.366297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:55.366297Z digest=sha256:44d362e14378cfd4e7003e569e6a83449a7fbff1d3dc9a3bc98297ed21092a74

Observation fc8aa203-30da-4a29-bf3b-506ebaec2eed · outbound

This paper cites Relational net- work via cascade crf for video language grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Relational net- work via cascade crf for video language grounding,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.437863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.372296Z digest=sha256:4a4866a6aed87c6ab673a2c11367c7a217ad8067f6dbac7bb41d859270d59f1a

Observation bdd5dff6-2807-4bb7-a096-54df137a0ef1 · outbound

This paper cites Self-supervised learn- ing for semi-supervised temporal language grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Self-supervised learn- ing for semi-supervised temporal language grounding,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.398725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.378760Z digest=sha256:02d5392a7603393a7b3692dcff1b38497d452f6d2b47be5b2bd15077a69bddf6

Observation 5c001614-7cea-4750-8334-0e068401a662 · outbound

This paper cites Zero-shot video moment retrieval with angular reconstructive text embeddings,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Zero-shot video moment retrieval with angular reconstructive text embeddings,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.364436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.385953Z digest=sha256:01cca72a5b680dba8b9324b5537b665a0a715e54a792b5cf30f761aef693cfaf

Observation d09b598e-52ec-4bff-b2ef-afd9a5b3fc8e · outbound

This paper cites Point-supervised video temporal grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Point-supervised video temporal grounding,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.323678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.400824Z digest=sha256:8f3246e50cce46848ccd96861efe05e2b40273b1e48ccfec1c6644b433944dcb

Observation 7064981e-30c1-49d0-9e7d-c41c51d9a195 · outbound

This paper cites Siamese learning with joint alignment and regression for weakly-supervised video paragraph grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Siamese learning with joint alignment and regression for weakly-supervised video paragraph grounding,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.292698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.411581Z digest=sha256:7221766db1bcd68d613a9e7be32e42c8464fb94121b08691949646850590c322

Observation 7247a033-831e-4e9b-b14b-396f4413be51 · outbound

This paper cites Dense events grounding in video,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Dense events grounding in video,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.250727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.419522Z digest=sha256:d46494f3070baaca9179019035ed8d6fa57a7a059d29e36ae8c60d2559892bc5

Observation c41c9aff-f822-4165-9c5d-937bef32f82e · outbound

This paper cites Semi- supervised video paragraph grounding with contrastive encoder,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Semi- supervised video paragraph grounding with contrastive encoder,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.206893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.431154Z digest=sha256:09a5c4a2d26a6bc3b15f7d0025e44d500bb9d249139e6ea50bd0ebc547c87d8f

Observation cb1b510d-c071-465f-b95b-6ca56db34df8 · outbound

This paper cites Gtlr: Graph- based transformer with language reconstruction for video paragraph grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Gtlr: Graph- based transformer with language reconstruction for video paragraph grounding,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.175640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.438415Z digest=sha256:5998f9b30eba29f44aaa013ca5a50953486e4a311acf9a7ebbb1ef3704b40f78

Observation f6fdc42e-b1d2-4a79-a7c7-dab6e355f824 · outbound

This paper cites Hierarchical semantic correspondence networks for video paragraph grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Hierarchical semantic correspondence networks for video paragraph grounding,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.142075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.457277Z digest=sha256:4dec9526e3509671db248ec9e318fa343879e6b73998dcbcdbe13818cd5a6264

Observation 9ac7e611-5dee-4e25-9195-bbfebbbca399 · outbound

This paper cites End-to-end dense video grounding via parallel regression,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding End-to-end dense video grounding via parallel regression,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.106191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.464960Z digest=sha256:091e507603f6eec81d198d8702fec7aab2c3dcb2b7b97e5c1064ee8b145728ba

Observation 89768308-5145-4519-96ce-3ebf6ee12be3 · outbound

This paper cites Joint searching and grounding: Multi-granularity video content retrieval,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Joint searching and grounding: Multi-granularity video content retrieval,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.078811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.476850Z digest=sha256:5e781fe8794a1076cfc8685ccac72e7cb89c66758bb593a99c5a30433e9d9bab

Observation bc09617b-9423-4a3f-b53d-6fb5c988e262 · outbound

This paper cites Learning 2d temporal adjacent networks for moment localization with natural language,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Learning 2d temporal adjacent networks for moment localization with natural language,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.048620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.484261Z digest=sha256:a76df1ebbdcb3bde59469f9fa34111a900a1e4ac5bcc3036d6c4ee7463d5e0aa

Observation abb226eb-ea76-4a74-ba12-1e7d274201af · outbound

This paper cites Multi- stage aggregated transformer network for temporal language localization in videos,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Multi- stage aggregated transformer network for temporal language localization in videos,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.022172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.491258Z digest=sha256:1930617ab3f9f56b03e653ecdbfa35afc55eb2db6e3c0c074dbad9d061204129

Observation 0a93133b-306f-4e67-9be2-1fff525ae7b8 · outbound

This paper cites Structured multi- level interaction network for video moment localization via language query,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Structured multi- level interaction network for video moment localization via language query,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.984458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.498945Z digest=sha256:189e076d020e805eef3d6eec17c59bc9b5562998803ae27adfe3222544db15e4

Observation 904aa02a-57d4-4d36-bf93-60cbc5b7f9a5 · outbound

This paper cites Progressive localization networks for language-based moment localization,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Progressive localization networks for language-based moment localization,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.951168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.507588Z digest=sha256:c9788278feaa149c050210cf845b462694bf9a1674b02ad8f879b4742efb9bd1

Observation bebd8400-6143-45a6-8978-3a0b5e6f5605 · outbound

This paper cites Fast video moment retrieval,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Fast video moment retrieval,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.921860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.517023Z digest=sha256:91db27379b316c738465779f5c76acb0e4fda7c9fcdce4e702c8294724caec93

Observation 745964d0-3fc6-4b3b-840b-42e42513c073 · outbound

This paper cites Exploring optical-flow-guided motion and detection-based appearance for temporal sentence ground- ing,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Exploring optical-flow-guided motion and detection-based appearance for temporal sentence ground- ing,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.891511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.525638Z digest=sha256:708718ae4d2066c5ebf2a0eb870f86cb55128ab3ab94973791b821896c0073f3

Observation 356a9870-e136-4849-a88b-362eb7e7c3ab · outbound

This paper cites Temporally language grounding with multi-modal multi-prompt tuning,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Temporally language grounding with multi-modal multi-prompt tuning,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.865854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.536530Z digest=sha256:668bcc2bcd522b8c91e7878d0cb8791103d7e29b6b95aebd4054390a8cf07076

Observation 16d2b28d-e2b6-454b-b3dc-7e87085f5141 · outbound

This paper cites Dynamic pathway for query- aware feature learning in language-driven action localization,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Dynamic pathway for query- aware feature learning in language-driven action localization,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.838800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.545395Z digest=sha256:3ab9cee0cfdfd3b0b366e9fe30ac0680c50013c3d737f6d9ff2913a0722741ef

Observation bb0952c2-5809-4ae2-8a6f-6a416a57f8c8 · outbound

This paper cites Hierarchical local-global transformer for temporal sentence grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Hierarchical local-global transformer for temporal sentence grounding,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.812687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.555239Z digest=sha256:40072a3c5bfc3b9adefdffbd491b6c13234fa0ca7ccdf59c37cf7d4f99757609

Observation a165414e-3311-494c-836c-e90376b8027d · outbound

This paper cites Local-global video-text interactions for temporal grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Local-global video-text interactions for temporal grounding,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.778771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.566276Z digest=sha256:afdd4eff5ad2ab85d44e8e45a0525fb1a7fac56ddfe7e2b03d87ed6cb5a6601a

Observation 123cc7e8-24fb-45a8-b838-28eac729dba5 · outbound

This paper cites Proposal-free video grounding with contextual pyramid network,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Proposal-free video grounding with contextual pyramid network,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:55.574992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:55.574992Z digest=sha256:9909de052723bf7348fe7066ca1d15501fe875689134d8c01386f31c01378533

Observation 476b4db2-3661-40a4-8445-ebb3ebdfdc58 · outbound

This paper cites Hisa: Hierarchically semantic associating for video temporal grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Hisa: Hierarchically semantic associating for video temporal grounding,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.716128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.585509Z digest=sha256:f2765b3e511d22429ad7b44a0219a58a467e4b9d0327abcfe83b0d455db84cbd

Observation 706cef73-37c1-4842-ad64-6123afa8ecb2 · outbound

This paper cites Siamese alignment network for weakly supervised video moment retrieval,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Siamese alignment network for weakly supervised video moment retrieval,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.685834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.595640Z digest=sha256:2ae75545eb2feeefb91e7b24bfb8c4c2da1ae5adff53271fa811d4d7c02fac9f

Observation ea85aae1-eafe-44a7-a239-8a9909f657b6 · outbound

This paper cites Weakly supervised temporal adjacent network for language grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Weakly supervised temporal adjacent network for language grounding,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.655937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.605247Z digest=sha256:83193d57c1691e074e3beb69025edbbc558952c34b196775aac5b66d4cc2e9b3

Observation d1b290f1-fb02-4584-a06f-def87b6b991b · outbound

This paper cites Asynce: Disentangling false-positives for weakly-supervised video grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Asynce: Disentangling false-positives for weakly-supervised video grounding,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.608089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.614296Z digest=sha256:2e20d9f4a8ce7dfca9ce48506d58ba50ef8c59015b076a8f0d548f43ca0b3ae8

Observation 9be2053a-4c16-4911-a060-554cbd107b99 · outbound

This paper cites Weakly supervised video moment retrieval from text queries,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Weakly supervised video moment retrieval from text queries,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.582232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.622490Z digest=sha256:93deb94b233fbb2889495f63dce46128fb0c2e1a5539da499335798b36b1087c

Observation 6b0e0ca3-3ec3-4576-beba-0d8cbbbba40d · outbound

This paper cites Dual masked modeling for weakly-supervised temporal boundary discovery,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Dual masked modeling for weakly-supervised temporal boundary discovery,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.551350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.631708Z digest=sha256:cffd062a7bd587807e8add11c4019a8934f92eb25be1a7e65928650458043ca8

Observation bc9b5daa-58a0-4065-ac69-ee5e379c77ba · outbound

This paper cites Weakly-supervised video moment retrieval via semantic completion network,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Weakly-supervised video moment retrieval via semantic completion network,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.518228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.640237Z digest=sha256:c70e86c716f99345ddd7e960492c229eab615d9277341c5f1add5cee00f01c75

Observation f60be869-bce3-4bad-9699-25d68017782f · outbound

This paper cites Counterfactual cross-modality reasoning for weakly supervised video moment localization,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Counterfactual cross-modality reasoning for weakly supervised video moment localization,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.486172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.647441Z digest=sha256:eadd4507cddf2032e325f310a8f1d5a49bab04591aecf825435d8da4cefebac0

Observation 3718b914-5303-4331-8f6c-e4a85f011a21 · outbound

This paper cites Weakly supervised video moment localization with contrastive negative sample mining,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Weakly supervised video moment localization with contrastive negative sample mining,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.457021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.658452Z digest=sha256:fbe1dcb01bcf79205be2f1798b547e370d8e2305e8573f0e68ac874ed7ae858f

Observation 591fc544-e965-4f76-b758-056159e74d75 · outbound

This paper cites Weakly supervised temporal sentence grounding with gaussian-based contrastive proposal learning,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Weakly supervised temporal sentence grounding with gaussian-based contrastive proposal learning,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.421029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.666437Z digest=sha256:6f90c495f435a3821fc4dbda09c581a62a1a5184f6742d3befc51e4c3976cb60

Observation fc880122-f95f-43bc-98bc-cbd2ab15241c · outbound

This paper cites Long short-term memory,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Long short-term memory,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:55.683885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:55.683885Z digest=sha256:7f47396ef20b363523191a240bd41f39f55f5e7157d8518bfe2146190a8f3be3

Observation dbbae8f7-b2b0-4e8c-8038-50c52347d6a0 · outbound

This paper cites Distributed representations of words and phrases and their composi- tionality,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Distributed representations of words and phrases and their composi- tionality,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:55.693948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:55.693948Z digest=sha256:0ea402118dadffd5b3fe51529c4b53adc5332d49e2481d8d35a13b9e46070c7d

Observation e41c29e5-d910-4ac9-9293-1e0aade0d9a2 · outbound

This paper cites Learning spatiotemporal features with 3d convolutional networks,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Learning spatiotemporal features with 3d convolutional networks,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.344172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.703166Z digest=sha256:5d2af0a32816fb03f1358d6d9955a1e97a50d8bc1dfc6b8aafff8584d9599950

Observation ace8eb38-e77b-41f8-9fed-81bc565faba9 · outbound

This paper cites Attention is all you need,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Attention is all you need,

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:55.713164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:55.713164Z digest=sha256:d429ec71b74b10d297fc102e8a9e73f2a55651dab175c6459091015fc01b1443

Observation 692669ed-90f7-448b-99b1-334e94fbe8b6 · outbound

This paper cites Momentum contrast for unsupervised visual representation learning,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Momentum contrast for unsupervised visual representation learning,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.274265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.721313Z digest=sha256:56b0b215b4307675beba16e2f9faa6a012b4ea2ba7ce2a206f8037af39b9235b

Observation 73ed2ad6-51cb-4c4a-8ed1-fc9933cb884a · outbound

This paper cites Facenet: A unified embed- ding for face recognition and clustering,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Facenet: A unified embed- ding for face recognition and clustering,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.241021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.735438Z digest=sha256:aff414b34e73ab48c54a2e0f67b82e66450e496177d9a3421a1c04c634c2f422

Observation 86b12f83-b44e-4315-b23d-630af41182fd · outbound

This paper cites Localizing moments in video with natural language,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Localizing moments in video with natural language,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.202818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.743758Z digest=sha256:12d9efb1057b0ec4b65dbf3f7b7dafce3bd55db09a2262da52789d314b3fd462

Observation c592310c-193f-4412-89ab-23c5c461d6fb · outbound

This paper cites Finding Moments in Video Collections Using Natural Language.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Finding Moments in Video Collections Using Natural Language

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:55.756529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:55.756529Z digest=sha256:4797171c6b160a6e89ddab2c0617a53cd345ab8d8dfbfcde93824b4491ce0751

Observation 996b2390-bc7f-47ba-aaee-88c99d94f32c · outbound

This paper cites Tvr: A large-scale dataset for video-subtitle moment retrieval,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Tvr: A large-scale dataset for video-subtitle moment retrieval,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.174848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.766406Z digest=sha256:b370f4e3d17ce7eb51a5eb5dd29b4a3b5fa7cca46ccf18b74f7a92142eae9b71

Observation 6488d2b5-6c46-4b77-a672-f6012f4cb817 · outbound

This paper cites Video corpus moment retrieval with contrastive learning,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Video corpus moment retrieval with contrastive learning,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.138402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.773562Z digest=sha256:47ad0e1c5fce6768f63a5be991918a7141416a15c3a2ad16633c79f802cabd8e

Observation 74012127-5934-42c2-b6d8-75edba98cb61 · outbound

This paper cites Partially relevant video retrieval,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Partially relevant video retrieval,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.110868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.786660Z digest=sha256:9444e356fb34813ed6a4529bc88035e619409bb012b6101943ebbf7a4d5e6617

Observation fb91f54b-b6a5-46d9-ac4a-ae7b47a43678 · outbound

This paper cites Activitynet: A large-scale video benchmark for human activity un- derstanding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Activitynet: A large-scale video benchmark for human activity un- derstanding,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.076158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.798712Z digest=sha256:e29443c2786a0c5a4e20afb1bb89e316bcb1375b841d227819e74844c65e72f0

Observation 94c8de35-35df-4911-8e86-750f68708180 · outbound

This paper cites Script data for attribute-based recognition of composite activities,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Script data for attribute-based recognition of composite activities,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.042480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.806435Z digest=sha256:979d1b6f0215f857cdc624cc52cc8f762ae6e42978d0cff9fa9994482e31c61d

Observation 43826049-c9dc-4d0b-9352-9585362fb38a · outbound

This paper cites Grounding action descriptions in videos,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Grounding action descriptions in videos,

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.013294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.814847Z digest=sha256:e7695ae400ecfd2c9f05a89639f3f7fa4e7627b0b74efdc8e15eee894def52de

Observation a12d2a91-9063-4d06-94b2-c424a82c6d4c · outbound

This paper cites Large-scale video classification with convolutional neural networks,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Large-scale video classification with convolutional neural networks,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:55.987026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.825767Z digest=sha256:59fe164fa9057e4fb44ff3456d263e2aa34121bb6147cbab9da88844a757c9c6

Observation a860967b-f719-4690-8a49-2f9b64fbb613 · outbound

This paper cites Bert: Pre-training of deep bidirectional transformers for language understanding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Bert: Pre-training of deep bidirectional transformers for language understanding,

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:55.959660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.840477Z digest=sha256:3d81547920a75ee371b026c71de1dac239ef2e7ce550222d0e0e50c34bd59669

Pith citing papers

Observation d1da524e-6e8d-46e2-9bc4-167b63b1c201 · inbound

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning cites this paper.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-08-11T13:05:30.367482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.265355Z digest=sha256:e9b9de0a07c3f41f1a54033a56d5e86bcd9f0062e772ac81ecbc42e2838a9f31