Pith. sign in

Paper Citation Record · LEDGER

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance

As of 8 August 2026, this Paper Citation Record lists 76 of 76 outbound references and 0 inbound Pith citation observations for arXiv:2506.03589.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.03589 v3

Coverage vector

measured 76 of 76 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:04:06.870962Z

measured 76 of 76 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

76 of 76 outbound references displayed

  • verified exact0
  • verified fuzzy67
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 37d895d2-34b3-445c-af50-209203f129fe · outbound

This paper cites Frozen in time: A joint video and image encoder for end-to-end retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Frozen in time: A joint video and image encoder for end-to-end retrieval

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.225088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.503340Z digest=sha256:e8a3504947df52bbb61fe3741cf3d3aee2da3b0630156d1d4556a4a28af15729

Observation 170f9ec9-47b0-4dda-98db-b60cb63ef068 · outbound

This paper cites Man is to computer programmer as woman is to homemaker? debiasing word embeddings.Annual Conference on Neural Information Processing Systems (NeurIPS), pages 4349–4357,.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Man is to computer programmer as woman is to homemaker? debiasing word embeddings.Annual Conference on Neural Information Processing Systems (NeurIPS), pages 4349–4357,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.208120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.508471Z digest=sha256:67b80d6ca242e97bf8322b17ee77b191101c00c53f39c55c8e76cea60f59df1c

Observation 8d7412a1-a379-4daa-ac8b-fcc9eddebdec · outbound

This paper cites Chen and William B.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Chen and William B

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.192631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.513265Z digest=sha256:cbfd2ee6e0f89f21615679fc49aae4c777cc1248ad34c82f501b978b05164006

Observation 6a7b287e-9a16-4f60-8f42-5e29fb8e4502 · outbound

This paper cites Fine-grained video-text retrieval with hierarchical graph reasoning.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Fine-grained video-text retrieval with hierarchical graph reasoning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.178099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.518420Z digest=sha256:4fc3c53321feeb5c44cd2637f84ccaf024926e3d1a82899c97c3361da9ef03eb

Observation aeb23d3b-4289-410d-97a3-c2c325d0427e · outbound

This paper cites ”factual” or ”emotional”: Stylized image captioning with adaptive learning and attention.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance ”factual” or ”emotional”: Stylized image captioning with adaptive learning and attention

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.162718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.523490Z digest=sha256:6e6db05dd4f600667ba647ea7dc14580d47f38caa8ace43cf66a36f8f1e2d9ca

Observation 4adf04fb-028b-4c48-bb7b-d9aad32dfb76 · outbound

This paper cites Tagging before alignment: Integrating multi-modal tags for video-text retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Tagging before alignment: Integrating multi-modal tags for video-text retrieval

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.146742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.528321Z digest=sha256:bb9d1c0b9411a01b0d239133f0a6e0a6dfb2f608732af5a6c226d28231d67062

Observation be487dbb-982a-48f8-87e5-d3d0c7f9a859 · outbound

This paper cites UATVR: uncertainty-adaptive text-video retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance UATVR: uncertainty-adaptive text-video retrieval

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.130551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.533323Z digest=sha256:86d782c57f3f0087f7ffc1badfb1c1a884a9c2f2b056a99dd82bda7e765aa67c

Observation 1f309bb7-b7a6-4d55-a47b-31324762a0c5 · outbound

This paper cites Multi-modal transformer for video retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Multi-modal transformer for video retrieval

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.113451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.537925Z digest=sha256:e2c43bb3300d1a33f5cfe0d07cf37277ac42369dfa1eca37696c47244f161383

Observation e756754b-57d5-4726-a28c-601d9aea8296 · outbound

This paper cites Word embeddings quantify 100 years of gender and ethnic stereotypes.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Word embeddings quantify 100 years of gender and ethnic stereotypes

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.096333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.543202Z digest=sha256:2c1199f75733ad7d2151cff980fa124da4d7e7cd2d2ed680259cdd0d3dbc94f3

Observation 229b0881-9513-4611-b548-f08ceb170d29 · outbound

This paper cites X-pool: Cross- modal language-video attention for text-video retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance X-pool: Cross- modal language-video attention for text-video retrieval

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.078958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.547830Z digest=sha256:f5581b20ac76f782943fa7849782446bc4b5e0e353a18974a850782609d8eaf8

Observation 74905427-b284-4898-8ca9-7fd21f3fb456 · outbound

This paper cites Girshick, Pieter Noordhuis, Lukasz Wesolowski, Aapo Kyrola, Andrew Tulloch, Yangqing Jia, and Kaim- ing He.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Girshick, Pieter Noordhuis, Lukasz Wesolowski, Aapo Kyrola, Andrew Tulloch, Yangqing Jia, and Kaim- ing He

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.061079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.552444Z digest=sha256:e44f1e08d2da1861f86560a146344f0e3fba022d9233a0b58c1227c33dc5c234

Observation 49c9091a-c126-4c86-8d82-619b9fe343a3 · outbound

This paper cites Mscap: Multi-style image captioning with unpaired stylized text.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Mscap: Multi-style image captioning with unpaired stylized text

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.043200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.557385Z digest=sha256:2ca7dfe1a3edf3af306a15cc7abfbfd802f6f123d96ead5551a124bdb7e020f6

Observation 168b2492-37d5-403c-a6c0-45e32bb122f2 · outbound

This paper cites an unresolved cited work.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:04:08.024603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.561785Z digest=sha256:803e893fcb5ff7850d7e5f0e0647e08479c3faef2b2658d9c2854ee5b04b0f07

Observation 32fd2c9e-cce8-4c7e-a81f-a354ced30798 · outbound

This paper cites Burgess, Xavier Glorot, Matthew M.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Burgess, Xavier Glorot, Matthew M

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.999496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.566038Z digest=sha256:6623b58083d81ac0803810b3f6474db1bc7f2cd5f496acc4498315d1875cc955

Observation ccc6cc9c-7666-406c-bea0-644e84cd562b · outbound

This paper cites Reducing sentiment bias in language models via counterfactual evalu- ation.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Reducing sentiment bias in language models via counterfactual evalu- ation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.976714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.570466Z digest=sha256:3d14110f6cbebb53dac10d203692bd4052693360835873b29bab92c5dcf38cd1

Observation 66ee8173-5bf4-40b3-b739-968d6ad437ab · outbound

This paper cites Narrating the video: Boosting text-video retrieval via comprehensive utilization of frame- level captions.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Narrating the video: Boosting text-video retrieval via comprehensive utilization of frame- level captions

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.958727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.575010Z digest=sha256:052dd638f421d978ce8fdaa145943b45aecc15d3822182a702e7b70f0e7ff743

Observation 8e5c6f52-c8ce-4c68-9668-5c2fbb6de9aa · outbound

This paper cites Imagenet-x: Un- derstanding model mistakes with factor of variation annotations.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Imagenet-x: Un- derstanding model mistakes with factor of variation annotations

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.941350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.579164Z digest=sha256:5f94d6b8f59835a0b3745303af4bf7ae41e2d46bff890a0f732cdd7667e1da4d

Observation 6c97bd6c-e8f0-4378-84b0-a16ec8d047f9 · outbound

This paper cites Clifton, and Jie Chen.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Clifton, and Jie Chen

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.926775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.583124Z digest=sha256:54933460bd243011edab9d2a920b9de94bf3f826b2feb6bb5dbe6ef94e0a5d7d

Observation 7221a011-802b-46be-81b6-246f04669e30 · outbound

This paper cites Video-text as game players: Hi- erarchical banzhaf interaction for cross-modal representation learning.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Video-text as game players: Hi- erarchical banzhaf interaction for cross-modal representation learning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.909624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.588403Z digest=sha256:223d5945fd6dbbdf04f2dcfb350ca1a9fe82fe2e8d346b47676dce402f99d7e1

Observation 0d4653e5-d9fa-4201-b9f7-51a4e8a467e8 · outbound

This paper cites Text-video retrieval with disentangled conceptualization and set-to-set alignment.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Text-video retrieval with disentangled conceptualization and set-to-set alignment

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.893529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.593093Z digest=sha256:dac17b2eed1ed7262a3d4166569d4ce4a36a01131febecbbc6fa4f15634c6121

Observation c677f278-7e7a-4e80-9263-2a88caf0f5c2 · outbound

This paper cites DiffusionRet: Generative text-video retrieval with diffusion model.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance DiffusionRet: Generative text-video retrieval with diffusion model

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.874776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.598200Z digest=sha256:ec8a3a52239f6f1381e3464bb996d8f2cf710b9cdea8017bdff83e07472a3073

Observation b6b19630-9bdb-457e-b0a3-7ffab27740b5 · outbound

This paper cites Disentangled representation learning for non-parallel text style trans- fer.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Disentangled representation learning for non-parallel text style trans- fer

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.857036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.602709Z digest=sha256:f2521719d4d74b5fcb6c601980aa5a939c4399c87df59f877427b7952610504a

Observation 52ff3296-d78e-4a35-98d3-d0b83c12f39a · outbound

This paper cites Kingma and Jimmy Ba.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Kingma and Jimmy Ba

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.838832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.607438Z digest=sha256:1ce61286b02d53c1f53cd552c53e4d3d90098309186fb1cfe7ddc747817c532e

Observation 4f4a0ee0-b635-4cff-ba4d-0bd3ec01c018 · outbound

This paper cites Kingma and Max Welling.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Kingma and Max Welling

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.819429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.613057Z digest=sha256:d88c8025825ce923604b59fbf9597d31f37d44b386943b408df906084e22a943

Observation ce3a2415-943a-4c11-a8e8-ac04ece91984 · outbound

This paper cites Dense-captioning events in videos.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Dense-captioning events in videos

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.803203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.618291Z digest=sha256:35fa097da372eae9fc40318b37786048f1ec9ce8724bd80b6aaa9ea0a8b28126

Observation 73c7429e-10e0-428c-baa8-8799598f98fd · outbound

This paper cites 3db: A framework for debugging com- puter vision models.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance 3db: A framework for debugging com- puter vision models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.787748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.622852Z digest=sha256:bed6bc0dc2f65e6c866b83b2bd6e1b8fda3d2b9f20855cde15fa35813075da31

Observation 0b3362f0-d73c-49bb-9e59-1e991ee45f74 · outbound

This paper cites Berg, Mohit Bansal, and Jingjing Liu.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Berg, Mohit Bansal, and Jingjing Liu

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.770760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.627631Z digest=sha256:29156f347ed7250d779b482db1be152105fe1e475750cfaa41b076a714bc9c50

Observation cc97236b-55e6-4a2f-a837-c24c04300d49 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance LLaVA-OneVision: Easy Visual Task Transfer

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:04:06.632515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:04:06.632515Z digest=sha256:21cced1ea542bd91d7d3eccbbfbe4cf7b86fa48fc1ccea723c8e9dba2d8797ad

Observation 88c616aa-1db2-4987-ad7a-c102da654df9 · outbound

This paper cites Prototype-based aleatoric uncertainty quantification for cross-modal retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Prototype-based aleatoric uncertainty quantification for cross-modal retrieval

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.755279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.637552Z digest=sha256:7d8ab9cca868c7576047c323bfc55a74f1ffafbd778dd2754789fbf040148695

Observation 70563c63-f49a-40d0-ab8f-7faabb02534b · outbound

This paper cites Blip: Boot- strapping language-image pre-training for unified vision-language understanding and generation.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Blip: Boot- strapping language-image pre-training for unified vision-language understanding and generation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.739289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.642098Z digest=sha256:ed149b93b75dd6afb8048f46c64f2afcfc52ce86771bb7dd96dd3fda21976ee0

Observation c4f7c7bc-ac61-4039-8e16-d3bba474639d · outbound

This paper cites Clip-event: Connecting text and images with event structures.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Clip-event: Connecting text and images with event structures

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.723396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.646797Z digest=sha256:d9a20e2169e5dcf4fb26ddc32181d96eb7788aa1703a575cfc2c1c4c8dfc2cb8

Observation ca7d9689-99f7-4d91-8e2a-b97f72c530ac · outbound

This paper cites Towards debiasing sentence representations.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Towards debiasing sentence representations

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.707315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.652173Z digest=sha256:bdea7a1f5512cdf7d37ef6e7bb8388a97ed34a9812eb8a55ee515abca11ca50a

Observation ea17532b-9e5d-43ae-83ad-73170d3913d5 · outbound

This paper cites Text-adaptive multiple visual pro- totype matching for video-text retrieval.Annual Conference on Neu- ral Information Processing Systems (NeurIPS), pages 38655–38666,.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Text-adaptive multiple visual pro- totype matching for video-text retrieval.Annual Conference on Neu- ral Information Processing Systems (NeurIPS), pages 38655–38666,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.691947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.657003Z digest=sha256:848cb91b62efa072aa2b57b1531a4ccb185a0dbbe60dacb3655071946df284fd

Observation a1189552-8ce7-44d3-9b46-8836e4345609 · outbound

This paper cites Disentangled multimodal representation learning for recommendation.IEEE Transactions on Multimedia, 25:7149– 7159, 2022.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Disentangled multimodal representation learning for recommendation.IEEE Transactions on Multimedia, 25:7149– 7159, 2022

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.676034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.661918Z digest=sha256:2d79da29d3e152dbeac1da4f418553d391285f037a646b81c1017126b3111d52

Observation 66de7fee-7f47-4b6a-b8f4-057cc5505204 · outbound

This paper cites Ts2-net: Token shift and selection transformer for text-video retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Ts2-net: Token shift and selection transformer for text-video retrieval

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.659399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.666267Z digest=sha256:95afb8b8b7d8640bb515a6c1312e8d4bff6dfe20a6d1df0c12a5acb523094a1e

Observation 257b646d-e30a-47c2-a928-caf29e8e0780 · outbound

This paper cites A decade’s battle on dataset bias: Are we there yet? InICLR, 2025.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance A decade’s battle on dataset bias: Are we there yet? InICLR, 2025

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.636759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.671726Z digest=sha256:4670a20885c5b129cb5e5d081e9eefaaf214d0c1d935dc1a7dd90bd8d6ea9444

Observation 79ffc9f8-52f8-4cd9-9009-c00c1bf75211 · outbound

This paper cites Clip4clip: An empirical study of clip for end to end video clip retrieval and captioning.Neurocomputing, 508:293–304,.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Clip4clip: An empirical study of clip for end to end video clip retrieval and captioning.Neurocomputing, 508:293–304,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.605157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.676370Z digest=sha256:2429fba113e92c4e7467e010b757c9ee01c1c91f1a94c131a0a1746046e9d0a7

Observation dd2b4a7e-ae1f-48f7-913c-cc1dcfc95902 · outbound

This paper cites Hauptmann, Jo ˜ao F.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Hauptmann, Jo ˜ao F

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.587494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.681627Z digest=sha256:e2dc57c2379c76e0e250d0d195c956e65ca563565f5de5b55498b0a714aac68c

Observation cf747373-4554-4b49-9b0b-f648dfcfa9c0 · outbound

This paper cites Find- ing and fixing spurious patterns with explanations.Transactions on Machine Learning Research, 2022.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Find- ing and fixing spurious patterns with explanations.Transactions on Machine Learning Research, 2022

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.571356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.686895Z digest=sha256:65a59a170abc3aa31b9e6732088289d60273b72605a9e8978a7bd28370f4bfa4

Observation e2456427-48a1-4fd4-9bd1-f3d3ad196fc8 · outbound

This paper cites Learning transferable visual mod- els from natural language supervision.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Learning transferable visual mod- els from natural language supervision

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.538376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.691437Z digest=sha256:ee5de1c500603964388e1b537abcca308c742893f6bb69b63efad9961ab7019f

Observation 01338411-91d9-4d93-9979-a8fe3ff983bc · outbound

This paper cites Movie description.International Journal of Computer Vision, 123: 94–120, 2017.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Movie description.International Journal of Computer Vision, 123: 94–120, 2017

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.515973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.697424Z digest=sha256:ff0141b1c0c07a9b6b02b00c9106dc594a17f7dfac1407db3c7f10a3db2a2dc5

Observation 3af9e6db-9d12-47ed-88b4-a3d6e929a511 · outbound

This paper cites Grad-cam: Visual explanations from deep networks via gradient-based localization.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Grad-cam: Visual explanations from deep networks via gradient-based localization

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.496267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.702516Z digest=sha256:7dae3f8fa02279238f6e5234d2bcf36e6de5b0e99866f21e5ba694b7d74b14d6

Observation 15e59e41-8262-4cc9-91df-69dc30cfa90a · outbound

This paper cites Tempme: Video temporal token merging for efficient text-video retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Tempme: Video temporal token merging for efficient text-video retrieval

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.476355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.707267Z digest=sha256:960f62cc5a232b0d810b557bcae97adbe2df1eb73c46723b957e66a82da79697

Observation 950ccd69-3f02-490d-98bd-d03f335ad74c · outbound

This paper cites In-style: Bridging text and uncurated videos with style transfer for text-video retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance In-style: Bridging text and uncurated videos with style transfer for text-video retrieval

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.459359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.712127Z digest=sha256:b77875367ce63b754d4836748d15d264add0e403b271690354c0dfa1897af98e

Observation 3e400561-23ac-4728-a00b-2b1d578c19c1 · outbound

This paper cites Unbiasing through textual descriptions: Mitigat- ing representation bias in video benchmarks.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Unbiasing through textual descriptions: Mitigat- ing representation bias in video benchmarks

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.442309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.716822Z digest=sha256:0cffdd65860dec898cba211293d67792030b66d8e72b19e0ad0c5be96b13d9eb

Observation 6e37c67d-662e-4f4e-a665-1980d9ed3a21 · outbound

This paper cites Belding, Kai-Wei Chang, and William Yang Wang.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Belding, Kai-Wei Chang, and William Yang Wang

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.423449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.721597Z digest=sha256:d7a13efd8c4ba8cdbedc9d297c1f8e659b8be052dd91a48bc93e4450e2619013

Observation a682861e-2c7c-42a1-9ec7-ffc00682187d · outbound

This paper cites Detach and attach: Stylized image captioning without paired stylized dataset.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Detach and attach: Stylized image captioning without paired stylized dataset

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.408518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.727033Z digest=sha256:c524aded48d247cc7ff92bf49bd3870f02e513a4ff0eeb398627ffa10fe6a705

Observation 9ad7de97-29d6-4cf9-be6b-3cc55a206ddd · outbound

This paper cites Qwen2.5: A party of foundation models, 2024.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Qwen2.5: A party of foundation models, 2024

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.392585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.732765Z digest=sha256:fa9197f2de24d421234812743b4af2a7cfd05d91b33336ca6055896ffe1b9dab

Observation 7ce73968-b1bf-45b8-b9f4-82eb5ddfff49 · outbound

This paper cites Holistic features are almost sufficient for text-to-video retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Holistic features are almost sufficient for text-to-video retrieval

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.377348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.737531Z digest=sha256:66914dfd2a12fece9ea1a218edb2ab7796d6c3c1b229cc938ab99eff3cbab9ce

Observation f57ee609-eb93-4931-8698-1834a4f30bcd · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Representation Learning with Contrastive Predictive Coding

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T11:04:06.743359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:04:06.743359Z digest=sha256:cffcaaa49df0555d8ab21db900e0c0fd9a2b3a02ab5d641428ddd83cf3614f1a

Observation 8d70c2fb-0c01-4c3d-9e5b-2c5c6a4fd046 · outbound

This paper cites Visualizing data using t-SNE.Journal of Machine Learning Research, 9:2579–2605, 2008.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Visualizing data using t-SNE.Journal of Machine Learning Research, 9:2579–2605, 2008

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.361184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.749768Z digest=sha256:9fda18515fd7f33de316723cfe41de698216f64a048c94c2c5f52d556c787dc2

Observation 71d674a7-de73-416f-abeb-1ec3d5191398 · outbound

This paper cites Gomez, Lukasz Kaiser, and Illia Polosukhin.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Gomez, Lukasz Kaiser, and Illia Polosukhin

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.341857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.754364Z digest=sha256:1fbac912deabba8f759c0718c96ceed36f490a8155474b42e9f3db9fda3fb2a5

Observation 1823b0d6-669b-42fb-8ece-118cb8b9eaf3 · outbound

This paper cites Aoe-net: Entities interactions modeling with adaptive attention mechanism for temporal action proposals gener- ation.International Journal of Computer Vision, 131(1):302–323,.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Aoe-net: Entities interactions modeling with adaptive attention mechanism for temporal action proposals gener- ation.International Journal of Computer Vision, 131(1):302–323,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.321829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.759232Z digest=sha256:03313d026b9650d64bc086250827ace888154fd8cb7f6e0df4a0cd6d172c8601

Observation b0069af1-fd92-41d4-9fb4-437b744df6de · outbound

This paper cites Dianat, Majid Rabbani, Raghuveer Rao, and Zhiqiang Tao.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Dianat, Majid Rabbani, Raghuveer Rao, and Zhiqiang Tao

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.303786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.763574Z digest=sha256:c49e6c37620704c5a3747d26d8a9e6f774f9e1c741b83987cee33106e681224e

Observation d9a96611-65c5-4859-850b-dd2da7be35c6 · outbound

This paper cites Towards fairness in visual recognition: Effective strategies for bias mitigation.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Towards fairness in visual recognition: Effective strategies for bias mitigation

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.287611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.768315Z digest=sha256:c4af10ad261023156eb0ec6a3013131aae1a82bf911b28bb7351b69061d4191b

Observation f49c1913-67e1-4d94-9652-25b87cf2e450 · outbound

This paper cites Text proxy: De- composing retrieval from a 1-to-n relationship into n 1-to-1 relation- ships for text-video retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Text proxy: De- composing retrieval from a 1-to-n relationship into n 1-to-1 relation- ships for text-video retrieval

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.272177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.772716Z digest=sha256:e7ab028dc7df1ede6d034ded5e9b549543b198ff4163c848f3d1731da9bfcea5

Observation ea58361d-2d60-4baa-b104-05d3fb5a81f2 · outbound

This paper cites Msr-vtt: A large video description dataset for bridging video and language.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Msr-vtt: A large video description dataset for bridging video and language

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.256994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.777207Z digest=sha256:ada368368bc9b67e3f676d4fe9b34e5705530dfd7ba54bfdf486cd3a1e3a097a

Observation 0417153f-9687-4fe4-9138-a1fec28e17ee · outbound

This paper cites Vltint: visual-linguistic transformer-in-transformer for coherent video paragraph captioning.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Vltint: visual-linguistic transformer-in-transformer for coherent video paragraph captioning

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.241270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.781831Z digest=sha256:d98f835a38cb63bf05c9cff14814b92c0ad0ebe38b6cdee5ed2ee91670814d66

Observation 28fc1343-f53e-4a5a-a9cf-692204ed8500 · outbound

This paper cites Video-text pre-training with learned regions for retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Video-text pre-training with learned regions for retrieval

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.224964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.787253Z digest=sha256:4e3f8c675b7b954bad4cb5e955c8233e5e3247c853156cd22fdc05fd09cf15ce

Observation 64f44cac-f789-4f9a-aef3-9f00e4aede8a · outbound

This paper cites Qwen2 Technical Report.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Qwen2 Technical Report

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T11:04:06.792244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:04:06.792244Z digest=sha256:baf4b50cd762cb5a73f126a45e6aaad06e8d710f89a36989813bbeaaf0f729c4

Observation 0e9bc1af-0aa0-43d7-8d3f-f962a93f0a9f · outbound

This paper cites Dgl: Dynamic global-local prompt tuning for text-video retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Dgl: Dynamic global-local prompt tuning for text-video retrieval

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.209116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.797139Z digest=sha256:ef2150e7c2bea839d87877a819638217929efd3a93e3200428a73a46ebcd4d77

Observation 293d2f09-21d5-489c-9455-d8b1749481f6 · outbound

This paper cites Coca: Contrastive captioners are image- text foundation models.Transactions on Machine Learning Research,.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Coca: Contrastive captioners are image- text foundation models.Transactions on Machine Learning Research,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.190155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.801595Z digest=sha256:1fbddf8b4b5422c654268805f9f88fce977ff6868073c1e9623d7ebe0d604569

Observation 31021580-f50b-4974-aef1-5c64ce9a573a · outbound

This paper cites Gender bias in contextualized word embeddings.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Gender bias in contextualized word embeddings

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.174155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.806141Z digest=sha256:df280a61b1a5f172c319b15e850a546393a369a47d6718aa55cd660c910ced3b

Observation c54c937e-d13e-4ee5-98e0-13998c35818d · outbound

This paper cites slicing mango.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance slicing mango

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.157837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.810912Z digest=sha256:afa05c8902c090f455e68e58469a4af229b62a469c9340d8a61b0c05422bcd08

Observation dd226f14-f365-48b4-be4f-37fa22c6a966 · outbound

This paper cites an unresolved cited work.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:04:07.141223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.815682Z digest=sha256:a0a39c44b235a85db69967107253c16a4b35568094dc04eebdae5c80c116c52e

Observation 7eb032d8-0359-4101-817c-2ed070de52c3 · outbound

This paper cites an unresolved cited work.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:04:07.121550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.820157Z digest=sha256:44bb719ab3cb575008166c63d883195a40ea9792968d9a03fc4de57b7dca19cf

Observation 0e1613af-8c0b-4fca-8d9d-831b4050d33e · outbound

This paper cites an unresolved cited work.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:04:07.103647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.824734Z digest=sha256:831d21d5f31cfb2cdd9b18ed71e4ea24623e1f23938db9f3d244ccb0daac5011

Observation 96e0427a-6615-4e60-baa4-280e41aa8250 · outbound

This paper cites an unresolved cited work.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Unresolved cited work

Reference 68

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:04:07.088783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.829961Z digest=sha256:59f9f2ae6a3c84a0c37ebacb3bec34ff181a62cb16d33161670cd8d99d0f9b63

Observation aab7983f-afef-4b2c-861d-d5e372c60570 · outbound

This paper cites LSMDC 2) He slaps SOMEONE again.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance LSMDC 2) He slaps SOMEONE again

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.073952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.834764Z digest=sha256:024fcc6870c5f36bd68811c9bbb6cbe623604f9ac6628d44b08a25545d8de3da

Observation f2893150-87b1-4c29-afa3-4f4cd3c14f8d · outbound

This paper cites The cup spins across the floor.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance The cup spins across the floor

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.059069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.839675Z digest=sha256:40a1f4afaec4ea4249d7fe4861e516c5022b0f2f70c95469571679ecdaacf55d

Observation dd1016d9-7a51-4d9b-aa3a-cafd0127b0bc · outbound

This paper cites ActivityNet 2) A woman is seen kneeling down next to a man.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance ActivityNet 2) A woman is seen kneeling down next to a man

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.042523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.845267Z digest=sha256:2e421b786b99eb4928529c4d800f73e54c4e761ec5cb1fdeb644f507c4567b09

Observation 34a9de46-5151-42a6-848b-91753b3c4ad1 · outbound

This paper cites an unresolved cited work.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:04:07.022704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.850727Z digest=sha256:87902e9a76a845139545dd7038f3ed7a09f95e5afb39df435cacde65bec15bc7

Observation 82655f37-f877-4303-a711-ed50835fef87 · outbound

This paper cites They first start crouching and hitting the ground with the sticks in a text).

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance They first start crouching and hitting the ground with the sticks in a text)

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.006124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.855384Z digest=sha256:213741707e966fb122b45f143f0c2feb61a601d8f43509eb0cd16128e3173951

Observation e140e28b-5224-4e8d-8f3f-e3e65258d263 · outbound

This paper cites The guitarist is looking straight up.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance The guitarist is looking straight up

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:06.989949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.860495Z digest=sha256:da54974db7130133014503b8c4ae11b4fdcafb02fc412084fb120a1761366ab3

Observation 25e07902-8cda-4a4c-8cc9-94784d975fe5 · outbound

This paper cites White square exits frame left the camera pans back the way it came.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance White square exits frame left the camera pans back the way it came

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:06.974937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.865613Z digest=sha256:2d2e067ca1716f498d128004a99f891ac370f41200cb737090a93abf9c0af18b

Observation 72abfa9a-704b-4ddb-83c7-bdc57d53d2ff · outbound

This paper cites a woman spins around several times very fast.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance a woman spins around several times very fast

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:06.958337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:04:06.870962Z digest=sha256:77f617cea95e6bce8c21860d6008ecfcd6200932427cc0fd973bfc9f498498e2

Pith citing papers

No inbound Pith citation observations are available.