Pith. sign in

Paper Citation Record · LEDGER

T-MARS: Improving Visual Representations by Circumventing Text Feature Learning

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2307.03132.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.03132 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T12:12:30.745548Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:49:44.890982Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7ab61e17-d0ee-4519-8ef5-cf6420d81bda · inbound

Scaling Pre-training to One Hundred Billion Data for Vision Language Models cites this paper.

Scaling Pre-training to One Hundred Billion Data for Vision Language Models T-MARS: Improving Visual Representations by Circumventing Text Feature Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T12:12:30.745548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:12:30.745548Z digest=sha256:0f7ebb90942fb8a68d5e6bd7633bd1654b4bbe99aa0b91d1929bb5c8fc1e25e2

Observation 77f46a19-d285-4d68-a19c-1866b54a68c0 · inbound

Quality over Quantity: Boosting Data Efficiency Through Ensembled Multimodal Data Curation cites this paper.

Quality over Quantity: Boosting Data Efficiency Through Ensembled Multimodal Data Curation T-MARS: Improving Visual Representations by Circumventing Text Feature Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T06:06:17.557918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T06:06:17.557918Z digest=sha256:64a4010ab2079e5271018c0976510b609e1316af69a88919a58e8a1440ea1d22

Observation d9da09a4-3fe4-4be0-b30a-ed005f0366b4 · inbound

What Does the Caption Really Say? Counterfactual Phrase Intervention for Compositional Data Selection in Vision-Language Pretraining cites this paper.

What Does the Caption Really Say? Counterfactual Phrase Intervention for Compositional Data Selection in Vision-Language Pretraining T-MARS: Improving Visual Representations by Circumventing Text Feature Learning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:21:10.270477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T06:19:54.322211Z digest=sha256:caad04628dd951ab090da8e130e674c8e6166ae3f2e9b7ee4c48c709fc49a398

Observation f069d1dc-4508-474e-b932-b754e49d125b · inbound

Data Selection Through Iterative Self-Filtering for Vision-Language Settings cites this paper.

Data Selection Through Iterative Self-Filtering for Vision-Language Settings T-MARS: Improving Visual Representations by Circumventing Text Feature Learning

Reference 77

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:49:44.892311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T09:22:47.537137Z digest=sha256:3ff7a97af7879805f92be14fbfea54d8dba22c34f7acc2a47845d5555a56d17f

Observation 82be6d52-1da5-4d4c-ba4e-2ed518577dd6 · inbound

DataComp-VLM: Improved Open Datasets for Vision-Language Models cites this paper.

DataComp-VLM: Improved Open Datasets for Vision-Language Models T-MARS: Improving Visual Representations by Circumventing Text Feature Learning

Reference 200

Resolution
verified exact
arxiv_id, observed 2026-07-01T15:45:47.689741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T01:16:16.834861Z digest=sha256:1d76c823cb2aa6f9ec14735ab3ead1c95e3b99d7afc94eba40829b8cbe486598

Observation f75d8715-046b-4d2c-9685-fdcde13bd79b · inbound

DataComp-VLM: Improved Open Datasets for Vision-Language Models cites this paper.

DataComp-VLM: Improved Open Datasets for Vision-Language Models T-MARS: Improving Visual Representations by Circumventing Text Feature Learning

Reference 200

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:17:23.898703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-02T21:10:10.548489Z digest=sha256:bff1d24f76baa25482f957871df73317a844244680f79fa5ce43a9d9a37af888