Pith. sign in

Paper Citation Record · LEDGER

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding

As of 13 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 1 inbound Pith citation observation for arXiv:2411.17481.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.17481 v1

Coverage vector

measured 68 of 68 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:10:55.840477Z

measured 69 of 69 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T13:05:30.265355Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-11T13:05:30.363755Z

Reference resolution

68 of 68 outbound references displayed

  • verified exact0
  • verified fuzzy59
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation eaef100c-87e0-457b-aabc-b97721150d6c · outbound

This paper cites Real-world anomaly detection in surveillance videos,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Real-world anomaly detection in surveillance videos,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:58.007363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.221759Z digest=sha256:25c9595da10b462f9a5b7a052c8afa27780ddb34006a12c17dc8df4059d36b8d

Observation 31b5088c-94aa-4529-b8f3-823ee63a1137 · outbound

This paper cites Toward video anomaly retrieval from video anomaly detection: New benchmarks and model,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Toward video anomaly retrieval from video anomaly detection: New benchmarks and model,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:55.228967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:55.228967Z digest=sha256:52d87a1c1908c501bb32370711a875d7abc3998f653b19a7d474e620f6665c5e

Observation 9a11d131-dc00-4c02-96d6-908cedde8615 · outbound

This paper cites Robust multi-drone multi-target tracking to resolve target occlusion: A benchmark,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Robust multi-drone multi-target tracking to resolve target occlusion: A benchmark,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.966703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.235740Z digest=sha256:99600465067f584b400e688b9abf4c3e37c102ec2675868470f04c9ccffa13ef

Observation a628c639-8cd9-4491-afed-f84628e65182 · outbound

This paper cites Yolov3-mt: A yolov3 using multi-target tracking for vehicle visual detection,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Yolov3-mt: A yolov3 using multi-target tracking for vehicle visual detection,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.943705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.243540Z digest=sha256:96384159b15d37a9c31e702b98e24a70956693a3b152bafb400e10c7c53c8f65

Observation 57d5fea4-af98-469e-9c5c-d4d857be7c1b · outbound

This paper cites Deepmtt: A deep learning maneuvering target-tracking algorithm based on bidirectional lstm network,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Deepmtt: A deep learning maneuvering target-tracking algorithm based on bidirectional lstm network,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.914941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.249889Z digest=sha256:126e1140b6b31dbfea967eccda0e613e1451fb83092372d9c86aba7315f4f333

Observation 97b607a8-b33d-4e23-9c8b-83c9aadb03a0 · outbound

This paper cites Robust obstacle detection and recognition for driver assistance systems,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Robust obstacle detection and recognition for driver assistance systems,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.892101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.256748Z digest=sha256:f167b87ad8952659a402000932295e509bdffa2556ce52b400e0d8f6e567d27f

Observation 281b165e-5ea9-44c5-a970-973b0f57012b · outbound

This paper cites End-to-end object detection with transformers,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding End-to-end object detection with transformers,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:55.271170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:55.271170Z digest=sha256:e32323283211085181dcc9a6d3175530fb774ae467343fcb33ace7114e28781b

Observation 07177e39-30b7-4ea5-97de-cc8b83974c3b · outbound

This paper cites Pareto refocusing for drone-view object detection,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Pareto refocusing for drone-view object detection,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.851391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.280109Z digest=sha256:30956b1b80ca6724bfeb8ab5672fc44bf6ad1c7aeaddd01ecd88b431730cf0e9

Observation 5519f4f9-17fe-4909-a7ab-3c5b52b19743 · outbound

This paper cites Sparse r-cnn: End-to-end object detection with learnable proposals,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Sparse r-cnn: End-to-end object detection with learnable proposals,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.828732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.287581Z digest=sha256:59ec379a4f36f3d2e646af46452fcbc8c763f6036157be8e39ba4164356dcf7f

Observation 6ab04fce-432d-425c-ac7b-bd8720d7d5c0 · outbound

This paper cites Recent advances for aerial object detection: A survey,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Recent advances for aerial object detection: A survey,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.801878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.296963Z digest=sha256:5d920443b4971dffff5e7be138a9e97865cb0b98dc7275021deec0924200dd81

Observation cd412ab2-9ec9-43ff-9eff-a59525b374f7 · outbound

This paper cites Cascade r-cnn: Delving into high quality object detection,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Cascade r-cnn: Delving into high quality object detection,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.777387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.303386Z digest=sha256:3df727c1b5d06e16707e3b68ab129e9581fb6b2c878fe8935f96926ecd16ac95

Observation af5b13b6-11b2-4950-a99b-9b5046a65f1c · outbound

This paper cites Crnet: Context-guided reasoning network for detecting hard objects,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Crnet: Context-guided reasoning network for detecting hard objects,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:55.309511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:55.309511Z digest=sha256:fca6f6552917088aecc4710c03d7ebb5a95d4810ebfd9f3ffac9982707e4420a

Observation bd28b2bf-9a4d-4d8e-ae3c-0d2a4ac5240c · outbound

This paper cites Triple adversarial learning and multi-view imaginative reasoning for unsupervised domain adaptation person re-identification,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Triple adversarial learning and multi-view imaginative reasoning for unsupervised domain adaptation person re-identification,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.710758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.314739Z digest=sha256:ca452d88b89b0007e28db1bb313516746b2ee668f4a300506850184648dd509f

Observation dce6daa2-8e7f-48ad-9612-fc98dd500735 · outbound

This paper cites Logical relation inference and multiview information interaction for domain adaptation person re-identification,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Logical relation inference and multiview information interaction for domain adaptation person re-identification,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.670988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.323145Z digest=sha256:442575e1bd6c26be7e2ac9b2a8f380170552d58ac2960ee21205895e87a6619a

Observation 779b5200-cfa5-4edd-8ae3-7813fb51ac7f · outbound

This paper cites Attribute-aligned domain- invariant feature learning for unsupervised domain adaptation person re-identification,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Attribute-aligned domain- invariant feature learning for unsupervised domain adaptation person re-identification,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.631604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.331492Z digest=sha256:d7458cb7c495ae02b58bfafde9c339af055d908c3da1f81fb15b473a375f9e30

Observation b04977b5-f773-42a5-8338-e1f790888dc4 · outbound

This paper cites Intermediary-guided bidi- rectional spatial–temporal aggregation network for video-based visible- infrared person re-identification,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Intermediary-guided bidi- rectional spatial–temporal aggregation network for video-based visible- infrared person re-identification,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.605290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.338102Z digest=sha256:8b40a07c5d1b9da221afe1926c9a5bcfa75edbcc512a8e65b306a54488c0fce0

Observation 0c2c96ac-253f-426f-ae2f-1b87d47fa8e0 · outbound

This paper cites Video moment retrieval from text queries via single frame annotation,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Video moment retrieval from text queries via single frame annotation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.579605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.346074Z digest=sha256:a34806eee723833fe15eb8b8d4694457304c756ee8a2643fd4ab7024754efd96

Observation a4be4d36-af3a-481d-806f-989a0ead684f · outbound

This paper cites Text-based local- ization of moments in a video corpus,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Text-based local- ization of moments in a video corpus,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.544199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.352979Z digest=sha256:171c82141595c6813959223219cb840e30db82f8c5017a4dde6405132049d273

Observation 2cc6412e-6b34-4c3f-a1d0-ae662818f792 · outbound

This paper cites Language-guided multi-granularity con- text aggregation for temporal sentence grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Language-guided multi-granularity con- text aggregation for temporal sentence grounding,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.506772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.358800Z digest=sha256:ff8e8b994d34b8058bb883787ec4ad01b05b07dbcf97de73cb12b95bed15d204

Observation b3eddd23-d222-4e7b-aee1-d7939a42857c · outbound

This paper cites Conditional video diffusion network for fine-grained temporal sentence grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Conditional video diffusion network for fine-grained temporal sentence grounding,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:55.366297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:55.366297Z digest=sha256:b579ad5b08c6e49f2f11116514592c62d6d8fed7838613e3faa9bd8cee5d919b

Observation fc8aa203-30da-4a29-bf3b-506ebaec2eed · outbound

This paper cites Relational net- work via cascade crf for video language grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Relational net- work via cascade crf for video language grounding,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.437863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.372296Z digest=sha256:6cd347cd2fa3811971f0c41c13606267e6e51f4d2e042d7811bb1b960bffe390

Observation bdd5dff6-2807-4bb7-a096-54df137a0ef1 · outbound

This paper cites Self-supervised learn- ing for semi-supervised temporal language grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Self-supervised learn- ing for semi-supervised temporal language grounding,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.398725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.378760Z digest=sha256:acbbb828ca3f2e34710ad28b1a9d9e77d6fa5c97e44a292cff0c4a26c6dd5f7c

Observation 5c001614-7cea-4750-8334-0e068401a662 · outbound

This paper cites Zero-shot video moment retrieval with angular reconstructive text embeddings,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Zero-shot video moment retrieval with angular reconstructive text embeddings,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.364436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.385953Z digest=sha256:117ecd8ae2f67f328a83314d4dace3dd0a41e7e19a3c756b8ab25e3d9dbecb96

Observation d09b598e-52ec-4bff-b2ef-afd9a5b3fc8e · outbound

This paper cites Point-supervised video temporal grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Point-supervised video temporal grounding,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.323678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.400824Z digest=sha256:5f783b43cffcd92e3b17a4fc3c3a4aa0d0c80ed0f0f978379b9dceaae1eae62a

Observation 7064981e-30c1-49d0-9e7d-c41c51d9a195 · outbound

This paper cites Siamese learning with joint alignment and regression for weakly-supervised video paragraph grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Siamese learning with joint alignment and regression for weakly-supervised video paragraph grounding,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.292698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.411581Z digest=sha256:1db10e75bf60beb85d15df0a9151016db98f082e597bb2458620287f4b12f82d

Observation 7247a033-831e-4e9b-b14b-396f4413be51 · outbound

This paper cites Dense events grounding in video,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Dense events grounding in video,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.250727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.419522Z digest=sha256:d8aecfc02a69bd1028feffae22deadca047b5061ee27ae20e7c33a8645988117

Observation c41c9aff-f822-4165-9c5d-937bef32f82e · outbound

This paper cites Semi- supervised video paragraph grounding with contrastive encoder,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Semi- supervised video paragraph grounding with contrastive encoder,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.206893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.431154Z digest=sha256:c7f4102b19840589e9af09c75350f31c31d18011009c5111fef0fbf973f4864f

Observation cb1b510d-c071-465f-b95b-6ca56db34df8 · outbound

This paper cites Gtlr: Graph- based transformer with language reconstruction for video paragraph grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Gtlr: Graph- based transformer with language reconstruction for video paragraph grounding,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.175640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.438415Z digest=sha256:252d6adea2507cf537b24bb774cf97e15eeeb70ddd2471d0ba87141c6ffcd934

Observation f6fdc42e-b1d2-4a79-a7c7-dab6e355f824 · outbound

This paper cites Hierarchical semantic correspondence networks for video paragraph grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Hierarchical semantic correspondence networks for video paragraph grounding,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.142075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.457277Z digest=sha256:933d769b4039203b291c114a8b2cd508e3152dab7b685ae3949eef8f76cc8319

Observation 9ac7e611-5dee-4e25-9195-bbfebbbca399 · outbound

This paper cites End-to-end dense video grounding via parallel regression,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding End-to-end dense video grounding via parallel regression,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.106191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.464960Z digest=sha256:6fce7cc4dd919c174250ccf735fff7b818684aa570551a9eeaadb7a4d85aed8a

Observation 89768308-5145-4519-96ce-3ebf6ee12be3 · outbound

This paper cites Joint searching and grounding: Multi-granularity video content retrieval,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Joint searching and grounding: Multi-granularity video content retrieval,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.078811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.476850Z digest=sha256:9710a8888327e1dffecc76b044664e3b08140e1b2ff4bebfda541912350831b1

Observation bc09617b-9423-4a3f-b53d-6fb5c988e262 · outbound

This paper cites Learning 2d temporal adjacent networks for moment localization with natural language,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Learning 2d temporal adjacent networks for moment localization with natural language,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.048620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.484261Z digest=sha256:6ca250371846bc9eec808700b2679725e816ff9b950b6b9b917b0db8e3c2d5d2

Observation abb226eb-ea76-4a74-ba12-1e7d274201af · outbound

This paper cites Multi- stage aggregated transformer network for temporal language localization in videos,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Multi- stage aggregated transformer network for temporal language localization in videos,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:57.022172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.491258Z digest=sha256:49e9281da808475d446c2767f468a8adb38cceb2118bc8946f32c6b55ec5554f

Observation 0a93133b-306f-4e67-9be2-1fff525ae7b8 · outbound

This paper cites Structured multi- level interaction network for video moment localization via language query,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Structured multi- level interaction network for video moment localization via language query,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.984458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.498945Z digest=sha256:f9a0414bd410e563797f404b10645ed7f2767d86f734b7f37b0a027baa79911e

Observation 904aa02a-57d4-4d36-bf93-60cbc5b7f9a5 · outbound

This paper cites Progressive localization networks for language-based moment localization,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Progressive localization networks for language-based moment localization,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.951168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.507588Z digest=sha256:ae1276557a8b4668a235f2b32f838ab1616b9495b9c76e1c1fe3db1b7c4ac4ff

Observation bebd8400-6143-45a6-8978-3a0b5e6f5605 · outbound

This paper cites Fast video moment retrieval,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Fast video moment retrieval,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.921860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.517023Z digest=sha256:828c962dabb98f201a43f93edfaf156dd5c2b9933d8605b9076fe6d0bd50ef7d

Observation 745964d0-3fc6-4b3b-840b-42e42513c073 · outbound

This paper cites Exploring optical-flow-guided motion and detection-based appearance for temporal sentence ground- ing,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Exploring optical-flow-guided motion and detection-based appearance for temporal sentence ground- ing,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.891511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.525638Z digest=sha256:c757d4538517f511449555ff726d79e33269197309a8c48a83be3db47c8a7141

Observation 356a9870-e136-4849-a88b-362eb7e7c3ab · outbound

This paper cites Temporally language grounding with multi-modal multi-prompt tuning,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Temporally language grounding with multi-modal multi-prompt tuning,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.865854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.536530Z digest=sha256:7048698c85e2805e9d957650c8b772fd224fbe388c29306d7c2746db3fec1c57

Observation 16d2b28d-e2b6-454b-b3dc-7e87085f5141 · outbound

This paper cites Dynamic pathway for query- aware feature learning in language-driven action localization,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Dynamic pathway for query- aware feature learning in language-driven action localization,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.838800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.545395Z digest=sha256:564e5436265d6924caf919095ef9c254c24312bf506f3226748f5730804a16bf

Observation bb0952c2-5809-4ae2-8a6f-6a416a57f8c8 · outbound

This paper cites Hierarchical local-global transformer for temporal sentence grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Hierarchical local-global transformer for temporal sentence grounding,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.812687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.555239Z digest=sha256:2beb4a0afbadfade8309df0fffa962844e0c672d86629b6335e305d4b38a5bac

Observation a165414e-3311-494c-836c-e90376b8027d · outbound

This paper cites Local-global video-text interactions for temporal grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Local-global video-text interactions for temporal grounding,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.778771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.566276Z digest=sha256:dbc31f6c730399939f6d97abbc264b12ced68d856266ae596b8f26baec50b082

Observation 123cc7e8-24fb-45a8-b838-28eac729dba5 · outbound

This paper cites Proposal-free video grounding with contextual pyramid network,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Proposal-free video grounding with contextual pyramid network,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:55.574992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:55.574992Z digest=sha256:fc4d4ee8af2a7e8adcd0d9e7427ba6d65ebc8d357d6b53ce79b68fb8f708afbe

Observation 476b4db2-3661-40a4-8445-ebb3ebdfdc58 · outbound

This paper cites Hisa: Hierarchically semantic associating for video temporal grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Hisa: Hierarchically semantic associating for video temporal grounding,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.716128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.585509Z digest=sha256:5fc35babb7e0e2f4fad9106d4709211942684dc5647d2cab8334f332367d4b18

Observation 706cef73-37c1-4842-ad64-6123afa8ecb2 · outbound

This paper cites Siamese alignment network for weakly supervised video moment retrieval,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Siamese alignment network for weakly supervised video moment retrieval,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.685834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.595640Z digest=sha256:44946f638bb939d89cea6b67e3716a66c40d1adee0af16dbcbddbd88781f0ac1

Observation ea85aae1-eafe-44a7-a239-8a9909f657b6 · outbound

This paper cites Weakly supervised temporal adjacent network for language grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Weakly supervised temporal adjacent network for language grounding,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.655937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.605247Z digest=sha256:70c00d9ab8b270daf8c9eb2f9541c3166c8518d7f657fd183e277a6bd49e358b

Observation d1b290f1-fb02-4584-a06f-def87b6b991b · outbound

This paper cites Asynce: Disentangling false-positives for weakly-supervised video grounding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Asynce: Disentangling false-positives for weakly-supervised video grounding,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.608089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.614296Z digest=sha256:505f99c45478f1c69075af3b24b1536a757c9a35686a5303e4820fe130dafc0c

Observation 9be2053a-4c16-4911-a060-554cbd107b99 · outbound

This paper cites Weakly supervised video moment retrieval from text queries,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Weakly supervised video moment retrieval from text queries,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.582232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.622490Z digest=sha256:131cb099e40d3422202698f3a3e1d140fcc2e655ed3e424f3289ba25363ec1f0

Observation 6b0e0ca3-3ec3-4576-beba-0d8cbbbba40d · outbound

This paper cites Dual masked modeling for weakly-supervised temporal boundary discovery,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Dual masked modeling for weakly-supervised temporal boundary discovery,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.551350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.631708Z digest=sha256:87a6b4429cfbc5c4d63595302b8b725ead0ac67cb0ea32e5885b1a75a0748aa9

Observation bc9b5daa-58a0-4065-ac69-ee5e379c77ba · outbound

This paper cites Weakly-supervised video moment retrieval via semantic completion network,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Weakly-supervised video moment retrieval via semantic completion network,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.518228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.640237Z digest=sha256:672155d937eb382513c33279fcec433f385d0a4975cb6b4475c0e8c0aa966c4c

Observation f60be869-bce3-4bad-9699-25d68017782f · outbound

This paper cites Counterfactual cross-modality reasoning for weakly supervised video moment localization,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Counterfactual cross-modality reasoning for weakly supervised video moment localization,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.486172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.647441Z digest=sha256:c679635f5b0a5a3c64da4499d5b29a05eda4e3ecc6f751122077e6c9aa519786

Observation 3718b914-5303-4331-8f6c-e4a85f011a21 · outbound

This paper cites Weakly supervised video moment localization with contrastive negative sample mining,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Weakly supervised video moment localization with contrastive negative sample mining,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.457021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.658452Z digest=sha256:481fc54667578d43007ca9fb26de4046cd37b0778fc174a7c558c49158986b75

Observation 591fc544-e965-4f76-b758-056159e74d75 · outbound

This paper cites Weakly supervised temporal sentence grounding with gaussian-based contrastive proposal learning,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Weakly supervised temporal sentence grounding with gaussian-based contrastive proposal learning,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.421029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.666437Z digest=sha256:df1dc8f611398d534ffc8c54c607a61664820bd9939211e8b440b21066475724

Observation fc880122-f95f-43bc-98bc-cbd2ab15241c · outbound

This paper cites Long short-term memory,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Long short-term memory,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:55.683885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:55.683885Z digest=sha256:a72c3dc7a6709747df7b0a3918f504a090b3ac649f93b27a5bb6d69f90e8b656

Observation dbbae8f7-b2b0-4e8c-8038-50c52347d6a0 · outbound

This paper cites Distributed representations of words and phrases and their composi- tionality,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Distributed representations of words and phrases and their composi- tionality,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:55.693948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:55.693948Z digest=sha256:5bd1cdcc7a844aa6515f982c14def1b0b7bc1749bd8292b4f3db0e561d85b725

Observation e41c29e5-d910-4ac9-9293-1e0aade0d9a2 · outbound

This paper cites Learning spatiotemporal features with 3d convolutional networks,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Learning spatiotemporal features with 3d convolutional networks,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.344172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.703166Z digest=sha256:bc7ccf12903a3be517913762a35d34b769a89fda9048af3983f5e4c4aa883c2c

Observation ace8eb38-e77b-41f8-9fed-81bc565faba9 · outbound

This paper cites Attention is all you need,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Attention is all you need,

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:55.713164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:55.713164Z digest=sha256:7be2dce721a970a8935f4f6d156d94048d62b96458e2f8094a7ee39381630837

Observation 692669ed-90f7-448b-99b1-334e94fbe8b6 · outbound

This paper cites Momentum contrast for unsupervised visual representation learning,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Momentum contrast for unsupervised visual representation learning,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.274265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.721313Z digest=sha256:46afac99f7f6041d6bf3c89bec133fb12afcaee11a42891ca60a138f9b4b7127

Observation 73ed2ad6-51cb-4c4a-8ed1-fc9933cb884a · outbound

This paper cites Facenet: A unified embed- ding for face recognition and clustering,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Facenet: A unified embed- ding for face recognition and clustering,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.241021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.735438Z digest=sha256:77453825ce83a271cad315754ecde62f558e73a9b72a18a1b35219a52ce9a7ed

Observation 86b12f83-b44e-4315-b23d-630af41182fd · outbound

This paper cites Localizing moments in video with natural language,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Localizing moments in video with natural language,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.202818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.743758Z digest=sha256:b74a90f079fc9d9971db151526e509bc2607db03e6f084828190e2f2e260a2e5

Observation c592310c-193f-4412-89ab-23c5c461d6fb · outbound

This paper cites Finding Moments in Video Collections Using Natural Language.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Finding Moments in Video Collections Using Natural Language

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:55.756529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:55.756529Z digest=sha256:8876bcf44e915565bbfc6412ba183d4b7ae4f02b49125353423c3ec60a97459c

Observation 996b2390-bc7f-47ba-aaee-88c99d94f32c · outbound

This paper cites Tvr: A large-scale dataset for video-subtitle moment retrieval,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Tvr: A large-scale dataset for video-subtitle moment retrieval,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.174848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.766406Z digest=sha256:1e388fd7f31a102a429a23e27d65f0023e88c3c5e553084e1b48ede8de61dfe4

Observation 6488d2b5-6c46-4b77-a672-f6012f4cb817 · outbound

This paper cites Video corpus moment retrieval with contrastive learning,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Video corpus moment retrieval with contrastive learning,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.138402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.773562Z digest=sha256:7dcab1b3aca6c20b873d28ed7ae7a20ea63e0e1568944f46cd1c88c3e21c4cf9

Observation 74012127-5934-42c2-b6d8-75edba98cb61 · outbound

This paper cites Partially relevant video retrieval,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Partially relevant video retrieval,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.110868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.786660Z digest=sha256:f069bd11348214808b82c7ea05203a6f2e6c11e1b96022d0a596268cf917ad8d

Observation fb91f54b-b6a5-46d9-ac4a-ae7b47a43678 · outbound

This paper cites Activitynet: A large-scale video benchmark for human activity un- derstanding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Activitynet: A large-scale video benchmark for human activity un- derstanding,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.076158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.798712Z digest=sha256:5b10b5091e6e91ca612d172344a2d182d1aa711f23d8af18e316ddb9530b59aa

Observation 94c8de35-35df-4911-8e86-750f68708180 · outbound

This paper cites Script data for attribute-based recognition of composite activities,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Script data for attribute-based recognition of composite activities,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.042480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.806435Z digest=sha256:c03b6371984bd4e96c5048af68de77e8eafc4826b7a97d0fea5f46c1d8c1662a

Observation 43826049-c9dc-4d0b-9352-9585362fb38a · outbound

This paper cites Grounding action descriptions in videos,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Grounding action descriptions in videos,

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:56.013294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.814847Z digest=sha256:4787bb373a016152453b2478e7e2584d90482c42ebe1950962a8a77433c9f253

Observation a12d2a91-9063-4d06-94b2-c424a82c6d4c · outbound

This paper cites Large-scale video classification with convolutional neural networks,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Large-scale video classification with convolutional neural networks,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:55.987026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.825767Z digest=sha256:d53ceb8c4b91c8ad55af28605192aff135ae7573fcdaf442972eca873a41b08a

Observation a860967b-f719-4690-8a49-2f9b64fbb613 · outbound

This paper cites Bert: Pre-training of deep bidirectional transformers for language understanding,.

Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding Bert: Pre-training of deep bidirectional transformers for language understanding,

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:10:55.959660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T12:10:55.840477Z digest=sha256:2d32ad433eea64b59f2bbfafb8b98625e33565c3566f2ba10309b5cd130af1de

Pith citing papers

Observation d1da524e-6e8d-46e2-9bc4-167b63b1c201 · inbound

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning cites this paper.

Query-centric Audio-Visual Cognition Network for Moment Retrieval, Segmentation and Step-Captioning Dual-task Mutual Reinforcing Embedded Joint Video Paragraph Retrieval and Grounding

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-08-11T13:05:30.367482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T13:05:30.265355Z digest=sha256:99682b5709f8811c57030b618bfa9a323ccdf2ae773890361e088076815bedf3