REVIEW 1 cited by
Nomic Embed Vision: Expanding the Latent Space
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
This technical report describes the training of nomic-embed-vision, a highly performant, open-code, open-weights image embedding model that shares the same latent space as nomic-embed-text. Together, nomic-embed-vision and nomic-embed-text form the first unified latent space to achieve high performance across vision, language, and multimodal tasks.
Forward citations
Cited by 1 Pith paper
-
VaRS-Doc: Interpretation-Aware Variant Representations via Latent Self-Probing for Visual Document Retrieval
VaRS-Doc improves visual document retrieval by encoding each page into multiple interpretation-specific variants using latent probing tokens, then letting each query select the best-matching variant.
Discussion (0). Continue with ORCID to comment.