Pith. sign in

Paper Citation Record · LEDGER

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance

As of 22 August 2026, this Paper Citation Record lists 76 of 76 outbound references and 0 inbound Pith citation observations for arXiv:2506.03589.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.03589 v3

Coverage vector

measured 76 of 76 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:04:06.870962Z

measured 76 of 76 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

76 of 76 outbound references displayed

  • verified exact0
  • verified fuzzy67
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 37d895d2-34b3-445c-af50-209203f129fe · outbound

This paper cites Frozen in time: A joint video and image encoder for end-to-end retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Frozen in time: A joint video and image encoder for end-to-end retrieval

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.225088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.503340Z digest=sha256:247a78a89281fec5e1df90df7546edebdfe32bf1b55f357848450b57bd7d11bf

Observation 170f9ec9-47b0-4dda-98db-b60cb63ef068 · outbound

This paper cites Man is to computer programmer as woman is to homemaker? debiasing word embeddings.Annual Conference on Neural Information Processing Systems (NeurIPS), pages 4349–4357,.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Man is to computer programmer as woman is to homemaker? debiasing word embeddings.Annual Conference on Neural Information Processing Systems (NeurIPS), pages 4349–4357,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.208120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.508471Z digest=sha256:764c2dc0b49213e4e56a76a514a298daa3990d95931f92d3df304ad827aca8c5

Observation 8d7412a1-a379-4daa-ac8b-fcc9eddebdec · outbound

This paper cites Chen and William B.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Chen and William B

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.192631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.513265Z digest=sha256:3b24c4d8c785081dac85de3dc95af4e917fe7eac90400061179dc143c99e7319

Observation 6a7b287e-9a16-4f60-8f42-5e29fb8e4502 · outbound

This paper cites Fine-grained video-text retrieval with hierarchical graph reasoning.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Fine-grained video-text retrieval with hierarchical graph reasoning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.178099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.518420Z digest=sha256:b94826b9c5d7993bda751f0984ab385f907d511847db9abca7bab82e5b6b0cb9

Observation aeb23d3b-4289-410d-97a3-c2c325d0427e · outbound

This paper cites ”factual” or ”emotional”: Stylized image captioning with adaptive learning and attention.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance ”factual” or ”emotional”: Stylized image captioning with adaptive learning and attention

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.162718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.523490Z digest=sha256:cb5fe2c7072877a8cd6eee440ce32d4e4c66f0f5c89cc515e7058d0ac3955644

Observation 4adf04fb-028b-4c48-bb7b-d9aad32dfb76 · outbound

This paper cites Tagging before alignment: Integrating multi-modal tags for video-text retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Tagging before alignment: Integrating multi-modal tags for video-text retrieval

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.146742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.528321Z digest=sha256:e6ba68e33a40603ee769285bccd768422526eebee8fc9acff3313d89ebada261

Observation be487dbb-982a-48f8-87e5-d3d0c7f9a859 · outbound

This paper cites UATVR: uncertainty-adaptive text-video retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance UATVR: uncertainty-adaptive text-video retrieval

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.130551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.533323Z digest=sha256:9b1fd930846260f7bfaa0101e348b5ef3fc6667782a829a05345c6020b6250dc

Observation 1f309bb7-b7a6-4d55-a47b-31324762a0c5 · outbound

This paper cites Multi-modal transformer for video retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Multi-modal transformer for video retrieval

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.113451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.537925Z digest=sha256:e1775315afa877d8d7c55eed2c35bfbe40bdd888c807873a22128346bc032e06

Observation e756754b-57d5-4726-a28c-601d9aea8296 · outbound

This paper cites Word embeddings quantify 100 years of gender and ethnic stereotypes.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Word embeddings quantify 100 years of gender and ethnic stereotypes

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.096333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.543202Z digest=sha256:79fb958f1b308de0fa189659121f2e366da9cc2ab85863679da225a4aa08bec0

Observation 229b0881-9513-4611-b548-f08ceb170d29 · outbound

This paper cites X-pool: Cross- modal language-video attention for text-video retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance X-pool: Cross- modal language-video attention for text-video retrieval

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.078958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.547830Z digest=sha256:fff4503146f3ecdabd9a275024e0413c052919b8a99923e800877a35cafefcfb

Observation 74905427-b284-4898-8ca9-7fd21f3fb456 · outbound

This paper cites Girshick, Pieter Noordhuis, Lukasz Wesolowski, Aapo Kyrola, Andrew Tulloch, Yangqing Jia, and Kaim- ing He.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Girshick, Pieter Noordhuis, Lukasz Wesolowski, Aapo Kyrola, Andrew Tulloch, Yangqing Jia, and Kaim- ing He

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.061079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.552444Z digest=sha256:70f0db5a811c2faf4d9c219346e8fcc0f60c3efd4a1f50d371c4b1b17652be98

Observation 49c9091a-c126-4c86-8d82-619b9fe343a3 · outbound

This paper cites Mscap: Multi-style image captioning with unpaired stylized text.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Mscap: Multi-style image captioning with unpaired stylized text

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:08.043200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.557385Z digest=sha256:3f6e1ac011dbe7ce7332c94e1862faff63dd615fd048eb9883afbf49052643e1

Observation 168b2492-37d5-403c-a6c0-45e32bb122f2 · outbound

This paper cites an unresolved cited work.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:04:08.024603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.561785Z digest=sha256:da24ee4e13ac1b49cb5a02c34a2eed4c227891068fb06cb682963576b142c6ce

Observation 32fd2c9e-cce8-4c7e-a81f-a354ced30798 · outbound

This paper cites Burgess, Xavier Glorot, Matthew M.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Burgess, Xavier Glorot, Matthew M

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.999496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.566038Z digest=sha256:d8b9b069672e2535cd7a872d841187b42021cd7498ebdb96e5c546cbfa95ba85

Observation ccc6cc9c-7666-406c-bea0-644e84cd562b · outbound

This paper cites Reducing sentiment bias in language models via counterfactual evalu- ation.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Reducing sentiment bias in language models via counterfactual evalu- ation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.976714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.570466Z digest=sha256:e60e5b50264df4bd7ad8e1d43c89960e3f518cd90de6cd30d8f2b79ddf690fbf

Observation 66ee8173-5bf4-40b3-b739-968d6ad437ab · outbound

This paper cites Narrating the video: Boosting text-video retrieval via comprehensive utilization of frame- level captions.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Narrating the video: Boosting text-video retrieval via comprehensive utilization of frame- level captions

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.958727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.575010Z digest=sha256:5fb64fcd73eeaab6eec6bc4f5ac0a53fd3dc92e84279b89aaafaf1bae2d83e49

Observation 8e5c6f52-c8ce-4c68-9668-5c2fbb6de9aa · outbound

This paper cites Imagenet-x: Un- derstanding model mistakes with factor of variation annotations.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Imagenet-x: Un- derstanding model mistakes with factor of variation annotations

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.941350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.579164Z digest=sha256:adac86f5cf730af4a66fb290ac5b0f38ea0b90e4510cbdb2c420ff9dff304284

Observation 6c97bd6c-e8f0-4378-84b0-a16ec8d047f9 · outbound

This paper cites Clifton, and Jie Chen.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Clifton, and Jie Chen

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.926775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.583124Z digest=sha256:04949b7c9d43466cf51b332c0808ead0fc67e7b6d0499618d0c6e226b00ace1b

Observation 7221a011-802b-46be-81b6-246f04669e30 · outbound

This paper cites Video-text as game players: Hi- erarchical banzhaf interaction for cross-modal representation learning.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Video-text as game players: Hi- erarchical banzhaf interaction for cross-modal representation learning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.909624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.588403Z digest=sha256:49865fab7a2e68f73c64a8496270c03c24627d87452bba97ce30ed4a106be4ac

Observation 0d4653e5-d9fa-4201-b9f7-51a4e8a467e8 · outbound

This paper cites Text-video retrieval with disentangled conceptualization and set-to-set alignment.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Text-video retrieval with disentangled conceptualization and set-to-set alignment

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.893529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.593093Z digest=sha256:670c1d6a17b5d0d70505e6eefd6d401437a75547139f440e98e1cbee129a89a1

Observation c677f278-7e7a-4e80-9263-2a88caf0f5c2 · outbound

This paper cites DiffusionRet: Generative text-video retrieval with diffusion model.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance DiffusionRet: Generative text-video retrieval with diffusion model

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.874776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.598200Z digest=sha256:4a5312552e28a5f8428d7a19c6ad1de7b6788c0df3de9dac8f08bf61d197c72d

Observation b6b19630-9bdb-457e-b0a3-7ffab27740b5 · outbound

This paper cites Disentangled representation learning for non-parallel text style trans- fer.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Disentangled representation learning for non-parallel text style trans- fer

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.857036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.602709Z digest=sha256:f4294b50f4c9f50b31199303937d7eeab8fb406a026e932d44f6b0aac9c5ae71

Observation 52ff3296-d78e-4a35-98d3-d0b83c12f39a · outbound

This paper cites Kingma and Jimmy Ba.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Kingma and Jimmy Ba

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.838832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.607438Z digest=sha256:2af9265c5df869359351148cb87fdef65d46c250b70a89d261fcb8779b5137dc

Observation 4f4a0ee0-b635-4cff-ba4d-0bd3ec01c018 · outbound

This paper cites Kingma and Max Welling.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Kingma and Max Welling

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.819429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.613057Z digest=sha256:03ba00a671c37e8b7c37d90000ac9590b19d04cad680c6c138b556f27da5c47e

Observation ce3a2415-943a-4c11-a8e8-ac04ece91984 · outbound

This paper cites Dense-captioning events in videos.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Dense-captioning events in videos

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.803203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.618291Z digest=sha256:c345c43218a2ecd6f745bca0a81ebac89a9f79e44d4651c401fd8077e09b1c70

Observation 73c7429e-10e0-428c-baa8-8799598f98fd · outbound

This paper cites 3db: A framework for debugging com- puter vision models.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance 3db: A framework for debugging com- puter vision models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.787748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.622852Z digest=sha256:6beaf63a1c81566c940c91a05db148a201eef34ce664424b419e8b924accaba8

Observation 0b3362f0-d73c-49bb-9e59-1e991ee45f74 · outbound

This paper cites Berg, Mohit Bansal, and Jingjing Liu.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Berg, Mohit Bansal, and Jingjing Liu

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.770760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.627631Z digest=sha256:3cab6fe1e0bb4f8bf8baee30a154678702a8401f867bdaa52583f0f49bba8628

Observation cc97236b-55e6-4a2f-a837-c24c04300d49 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance LLaVA-OneVision: Easy Visual Task Transfer

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:04:06.632515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:04:06.632515Z digest=sha256:511512a33c3b66fc7724254e8799457d5ab67a0b60eb13a481bb3caee07ca4d6

Observation 88c616aa-1db2-4987-ad7a-c102da654df9 · outbound

This paper cites Prototype-based aleatoric uncertainty quantification for cross-modal retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Prototype-based aleatoric uncertainty quantification for cross-modal retrieval

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.755279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.637552Z digest=sha256:091e6c11553725cc5ad382a4416e31d4e0678ee50b107daba75389a70d7b7cb1

Observation 70563c63-f49a-40d0-ab8f-7faabb02534b · outbound

This paper cites Blip: Boot- strapping language-image pre-training for unified vision-language understanding and generation.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Blip: Boot- strapping language-image pre-training for unified vision-language understanding and generation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.739289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.642098Z digest=sha256:7507dc27b5d9cb663f8740105db1e5be5e7092b73f1a709a1801cb519beecba3

Observation c4f7c7bc-ac61-4039-8e16-d3bba474639d · outbound

This paper cites Clip-event: Connecting text and images with event structures.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Clip-event: Connecting text and images with event structures

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.723396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.646797Z digest=sha256:6dc0f4740f38fe17e46b58b06f05763f6695f924156df35a1348ba22194398b0

Observation ca7d9689-99f7-4d91-8e2a-b97f72c530ac · outbound

This paper cites Towards debiasing sentence representations.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Towards debiasing sentence representations

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.707315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.652173Z digest=sha256:5f419a424b6d6404571cbd73d75b9a842fa19793ebd7696c2e0df6c5964024b9

Observation ea17532b-9e5d-43ae-83ad-73170d3913d5 · outbound

This paper cites Text-adaptive multiple visual pro- totype matching for video-text retrieval.Annual Conference on Neu- ral Information Processing Systems (NeurIPS), pages 38655–38666,.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Text-adaptive multiple visual pro- totype matching for video-text retrieval.Annual Conference on Neu- ral Information Processing Systems (NeurIPS), pages 38655–38666,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.691947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.657003Z digest=sha256:d63a98eb12b862b693f92b51395ad4a5d401a4d3443c0b7dd21153201c392831

Observation a1189552-8ce7-44d3-9b46-8836e4345609 · outbound

This paper cites Disentangled multimodal representation learning for recommendation.IEEE Transactions on Multimedia, 25:7149– 7159, 2022.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Disentangled multimodal representation learning for recommendation.IEEE Transactions on Multimedia, 25:7149– 7159, 2022

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.676034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.661918Z digest=sha256:00290cfb16a01af4a7b019d6137164d4ca5a444e74fb2a190852e6b656a8890c

Observation 66de7fee-7f47-4b6a-b8f4-057cc5505204 · outbound

This paper cites Ts2-net: Token shift and selection transformer for text-video retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Ts2-net: Token shift and selection transformer for text-video retrieval

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.659399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.666267Z digest=sha256:02bd5254564e5b04aeda8ba38a58fe1dba9951407493e214a9f9d95453a780bd

Observation 257b646d-e30a-47c2-a928-caf29e8e0780 · outbound

This paper cites A decade’s battle on dataset bias: Are we there yet? InICLR, 2025.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance A decade’s battle on dataset bias: Are we there yet? InICLR, 2025

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.636759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.671726Z digest=sha256:2a2b50c59ec0ab7368f06018f9fc435ff82ad93a118635f4485f7d51281ba00f

Observation 79ffc9f8-52f8-4cd9-9009-c00c1bf75211 · outbound

This paper cites Clip4clip: An empirical study of clip for end to end video clip retrieval and captioning.Neurocomputing, 508:293–304,.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Clip4clip: An empirical study of clip for end to end video clip retrieval and captioning.Neurocomputing, 508:293–304,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.605157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.676370Z digest=sha256:137af887cc6d1bc4bb506c681000a585d837faee8f99e5d1ce4cd098d9f3e8f4

Observation dd2b4a7e-ae1f-48f7-913c-cc1dcfc95902 · outbound

This paper cites Hauptmann, Jo ˜ao F.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Hauptmann, Jo ˜ao F

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.587494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.681627Z digest=sha256:a3bea3043f9c048ffc47132126e0e293bd19cabb829cf817e7bc625c56fdd04a

Observation cf747373-4554-4b49-9b0b-f648dfcfa9c0 · outbound

This paper cites Find- ing and fixing spurious patterns with explanations.Transactions on Machine Learning Research, 2022.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Find- ing and fixing spurious patterns with explanations.Transactions on Machine Learning Research, 2022

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.571356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.686895Z digest=sha256:8f9e50501ed87d77911b76ac0b469ca7fe8b61429828578f417f96f940df7c68

Observation e2456427-48a1-4fd4-9bd1-f3d3ad196fc8 · outbound

This paper cites Learning transferable visual mod- els from natural language supervision.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Learning transferable visual mod- els from natural language supervision

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.538376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.691437Z digest=sha256:0450bab1393ed3aadabf0720a3f43999911750c43d62d8a7e6ec23ccab208d61

Observation 01338411-91d9-4d93-9979-a8fe3ff983bc · outbound

This paper cites Movie description.International Journal of Computer Vision, 123: 94–120, 2017.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Movie description.International Journal of Computer Vision, 123: 94–120, 2017

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.515973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.697424Z digest=sha256:452f42d7ca5c39777ccca47fe8b144b2b3f9ee60d3c4677c31f10d9453cc430b

Observation 3af9e6db-9d12-47ed-88b4-a3d6e929a511 · outbound

This paper cites Grad-cam: Visual explanations from deep networks via gradient-based localization.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Grad-cam: Visual explanations from deep networks via gradient-based localization

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.496267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.702516Z digest=sha256:f5a1574e4f259dc00e1aafc35202b9d5e3969e78a3c733b1e3a320ff350c601d

Observation 15e59e41-8262-4cc9-91df-69dc30cfa90a · outbound

This paper cites Tempme: Video temporal token merging for efficient text-video retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Tempme: Video temporal token merging for efficient text-video retrieval

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.476355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.707267Z digest=sha256:4c5c6dafa93c06117ad3f0cc14dc389367a950b23724f62c0ae75106106e481a

Observation 950ccd69-3f02-490d-98bd-d03f335ad74c · outbound

This paper cites In-style: Bridging text and uncurated videos with style transfer for text-video retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance In-style: Bridging text and uncurated videos with style transfer for text-video retrieval

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.459359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.712127Z digest=sha256:2a23e156c0d53fceb98d74581dd6cfd12f0620172d2346155669ec213f9b8e31

Observation 3e400561-23ac-4728-a00b-2b1d578c19c1 · outbound

This paper cites Unbiasing through textual descriptions: Mitigat- ing representation bias in video benchmarks.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Unbiasing through textual descriptions: Mitigat- ing representation bias in video benchmarks

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.442309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.716822Z digest=sha256:2d34cb369944495934036496ad3468e3b4d6721e10eb1f18089ff65825597606

Observation 6e37c67d-662e-4f4e-a665-1980d9ed3a21 · outbound

This paper cites Belding, Kai-Wei Chang, and William Yang Wang.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Belding, Kai-Wei Chang, and William Yang Wang

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.423449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.721597Z digest=sha256:07df8fd899347857ed998bce634619f9eb770d218a2829d15328450ab4bc3c3f

Observation a682861e-2c7c-42a1-9ec7-ffc00682187d · outbound

This paper cites Detach and attach: Stylized image captioning without paired stylized dataset.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Detach and attach: Stylized image captioning without paired stylized dataset

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.408518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.727033Z digest=sha256:876b2a8119858e450509dfcaaf2cad33c23eed8c105bab70aa65533ba31caa67

Observation 9ad7de97-29d6-4cf9-be6b-3cc55a206ddd · outbound

This paper cites Qwen2.5: A party of foundation models, 2024.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Qwen2.5: A party of foundation models, 2024

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.392585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.732765Z digest=sha256:56e2db13ee9f749ef7ac0175cbcfcbf4b45a99044505c3f7f872321651892c87

Observation 7ce73968-b1bf-45b8-b9f4-82eb5ddfff49 · outbound

This paper cites Holistic features are almost sufficient for text-to-video retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Holistic features are almost sufficient for text-to-video retrieval

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.377348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.737531Z digest=sha256:19359cccc2dcfaea20aa9df45372fbad668519a9f07e935dda33242f28faecfb

Observation f57ee609-eb93-4931-8698-1834a4f30bcd · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Representation Learning with Contrastive Predictive Coding

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T11:04:06.743359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:04:06.743359Z digest=sha256:0c7606d46ebd9d25e39f84d4b4125030a13b13067426e8b6fb3d2980eca80bbd

Observation 8d70c2fb-0c01-4c3d-9e5b-2c5c6a4fd046 · outbound

This paper cites Visualizing data using t-SNE.Journal of Machine Learning Research, 9:2579–2605, 2008.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Visualizing data using t-SNE.Journal of Machine Learning Research, 9:2579–2605, 2008

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.361184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.749768Z digest=sha256:1088ed72c237eeaa3fe84e9956250e85bbd94b442cb19e18117cc4e3345f63de

Observation 71d674a7-de73-416f-abeb-1ec3d5191398 · outbound

This paper cites Gomez, Lukasz Kaiser, and Illia Polosukhin.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Gomez, Lukasz Kaiser, and Illia Polosukhin

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.341857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.754364Z digest=sha256:f61f2bf936bdd1743fe5a76e4d2bed806a3bf9f2665b6e2fc20f8ab229340d5b

Observation 1823b0d6-669b-42fb-8ece-118cb8b9eaf3 · outbound

This paper cites Aoe-net: Entities interactions modeling with adaptive attention mechanism for temporal action proposals gener- ation.International Journal of Computer Vision, 131(1):302–323,.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Aoe-net: Entities interactions modeling with adaptive attention mechanism for temporal action proposals gener- ation.International Journal of Computer Vision, 131(1):302–323,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.321829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.759232Z digest=sha256:f166c4de63cdf4b78c983740434d6548e97dac6146f665142c516da472b7045e

Observation b0069af1-fd92-41d4-9fb4-437b744df6de · outbound

This paper cites Dianat, Majid Rabbani, Raghuveer Rao, and Zhiqiang Tao.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Dianat, Majid Rabbani, Raghuveer Rao, and Zhiqiang Tao

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.303786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.763574Z digest=sha256:291ee0794a13322abc99aab67b1e694960eea6846f82c88403e010fee90b9996

Observation d9a96611-65c5-4859-850b-dd2da7be35c6 · outbound

This paper cites Towards fairness in visual recognition: Effective strategies for bias mitigation.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Towards fairness in visual recognition: Effective strategies for bias mitigation

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.287611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.768315Z digest=sha256:f191b260de79efac1e91fbd7fbead72d39cda81e670a7cabb8cd5bc602d2ca50

Observation f49c1913-67e1-4d94-9652-25b87cf2e450 · outbound

This paper cites Text proxy: De- composing retrieval from a 1-to-n relationship into n 1-to-1 relation- ships for text-video retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Text proxy: De- composing retrieval from a 1-to-n relationship into n 1-to-1 relation- ships for text-video retrieval

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.272177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.772716Z digest=sha256:11ea431edc1077ce9a42ee405531fba0e4140ac500d3f043d4df7f0e39c826ff

Observation ea58361d-2d60-4baa-b104-05d3fb5a81f2 · outbound

This paper cites Msr-vtt: A large video description dataset for bridging video and language.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Msr-vtt: A large video description dataset for bridging video and language

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.256994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.777207Z digest=sha256:e662cce5509880b94606c9a87258f51d2d88cfb1a71a5cf6dec02329f46b9fef

Observation 0417153f-9687-4fe4-9138-a1fec28e17ee · outbound

This paper cites Vltint: visual-linguistic transformer-in-transformer for coherent video paragraph captioning.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Vltint: visual-linguistic transformer-in-transformer for coherent video paragraph captioning

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.241270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.781831Z digest=sha256:36eaf546c0ca16c88f73bbaf09627621c2ec15cbd9b36dbf478ebfe745333039

Observation 28fc1343-f53e-4a5a-a9cf-692204ed8500 · outbound

This paper cites Video-text pre-training with learned regions for retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Video-text pre-training with learned regions for retrieval

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.224964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.787253Z digest=sha256:81c4ad71440f129ecd5272a391c78dd91a3e351b9e00d5f8db1660ef20f17f68

Observation 64f44cac-f789-4f9a-aef3-9f00e4aede8a · outbound

This paper cites Qwen2 Technical Report.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Qwen2 Technical Report

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T11:04:06.792244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:04:06.792244Z digest=sha256:c38d0f15c536229dd7ff396263262cee9885e431bce90d1fc64ad5c56392212a

Observation 0e9bc1af-0aa0-43d7-8d3f-f962a93f0a9f · outbound

This paper cites Dgl: Dynamic global-local prompt tuning for text-video retrieval.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Dgl: Dynamic global-local prompt tuning for text-video retrieval

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.209116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.797139Z digest=sha256:14d322da3b189cc807cc91e1a5641d69da8a9849f4a71b979dec4be11dbc5609

Observation 293d2f09-21d5-489c-9455-d8b1749481f6 · outbound

This paper cites Coca: Contrastive captioners are image- text foundation models.Transactions on Machine Learning Research,.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Coca: Contrastive captioners are image- text foundation models.Transactions on Machine Learning Research,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.190155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.801595Z digest=sha256:9cc120cadf88a193882a3dbec36fe4dc4d2c9964d2e8da606699d666ff7096b0

Observation 31021580-f50b-4974-aef1-5c64ce9a573a · outbound

This paper cites Gender bias in contextualized word embeddings.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Gender bias in contextualized word embeddings

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.174155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.806141Z digest=sha256:8a95376cb9a6e075710506134e0748a8a4fe9bf46f565662c16dfedc959fa035

Observation c54c937e-d13e-4ee5-98e0-13998c35818d · outbound

This paper cites slicing mango.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance slicing mango

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.157837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.810912Z digest=sha256:b0225065aa0d9d79bbbb9e5791ed3c877389976011893525d8a9fb56dbb42161

Observation dd226f14-f365-48b4-be4f-37fa22c6a966 · outbound

This paper cites an unresolved cited work.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:04:07.141223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.815682Z digest=sha256:93567c5b7071ddb4088cff6fa1bcbaa4084b4d18f32c759f5ac8e806a9d9a8f8

Observation 7eb032d8-0359-4101-817c-2ed070de52c3 · outbound

This paper cites an unresolved cited work.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:04:07.121550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.820157Z digest=sha256:857e9729af03dfe84d5c4058bbc50f8c1e71baa561f0d334030495b0a9b2a7dc

Observation 0e1613af-8c0b-4fca-8d9d-831b4050d33e · outbound

This paper cites an unresolved cited work.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:04:07.103647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.824734Z digest=sha256:44fd2f73f3dd2e4d853a21e5268b13c51ab39698574a8430d0579838d4c74d8f

Observation 96e0427a-6615-4e60-baa4-280e41aa8250 · outbound

This paper cites an unresolved cited work.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Unresolved cited work

Reference 68

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:04:07.088783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.829961Z digest=sha256:a9f4fe578f017bc4da70c66d5b8263f517fecb8c97ccc9be16b2c3a338771196

Observation aab7983f-afef-4b2c-861d-d5e372c60570 · outbound

This paper cites LSMDC 2) He slaps SOMEONE again.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance LSMDC 2) He slaps SOMEONE again

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.073952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.834764Z digest=sha256:d429c45a379095c135bb4366b3d79cb7c0c121fe1e006a2bb397b84d79ca0f3d

Observation f2893150-87b1-4c29-afa3-4f4cd3c14f8d · outbound

This paper cites The cup spins across the floor.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance The cup spins across the floor

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.059069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.839675Z digest=sha256:0e1b6eeb2dcb0fbf0c2327af17eeb4445c5d1e731f93151f26eaf77a03b0bf2d

Observation dd1016d9-7a51-4d9b-aa3a-cafd0127b0bc · outbound

This paper cites ActivityNet 2) A woman is seen kneeling down next to a man.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance ActivityNet 2) A woman is seen kneeling down next to a man

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.042523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.845267Z digest=sha256:bd729ba74bf3e1abf035f1ba1a9ad77b2624568bfc08a17e52856f5acaaab7f5

Observation 34a9de46-5151-42a6-848b-91753b3c4ad1 · outbound

This paper cites an unresolved cited work.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:04:07.022704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.850727Z digest=sha256:d244e08381fbac625154f9425bba1fef28f1f291e641c6084de18121d1078f83

Observation 82655f37-f877-4303-a711-ed50835fef87 · outbound

This paper cites They first start crouching and hitting the ground with the sticks in a text).

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance They first start crouching and hitting the ground with the sticks in a text)

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:07.006124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.855384Z digest=sha256:517784bb3f0f09f085f609f60ed920e32b2844e314915a62014bf8bf58061da6

Observation e140e28b-5224-4e8d-8f3f-e3e65258d263 · outbound

This paper cites The guitarist is looking straight up.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance The guitarist is looking straight up

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:06.989949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.860495Z digest=sha256:d2b58b3334897b466109745c3f6509e2115c1d1c95aee4d9872722d8a2334445

Observation 25e07902-8cda-4a4c-8cc9-94784d975fe5 · outbound

This paper cites White square exits frame left the camera pans back the way it came.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance White square exits frame left the camera pans back the way it came

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:06.974937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.865613Z digest=sha256:d5d9c3165afb666451e5f003b54f21f9a65b5c021f8b5e5c642954f040647c23

Observation 72abfa9a-704b-4ddb-83c7-bdc57d53d2ff · outbound

This paper cites a woman spins around several times very fast.

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance a woman spins around several times very fast

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:04:06.958337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:04:06.870962Z digest=sha256:348bd2f3486d8d967c72d07f2b0efe2cf9c9527c8bdc04f00c1f2e2801999e20

Pith citing papers

No inbound Pith citation observations are available.