Pith. sign in

Paper Citation Record · LEDGER

LMM-Regularized CLIP Embeddings for Image Classification

As of 11 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 0 inbound Pith citation observations for arXiv:2412.11663.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.11663 v1

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T14:45:44.879975Z

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

18 of 18 outbound references displayed

  • verified exact0
  • verified fuzzy8
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0ccaf34c-13f4-4856-83cb-9026ace0d370 · outbound

This paper cites Learning transferable visual models from natural language supervision.

LMM-Regularized CLIP Embeddings for Image Classification Learning transferable visual models from natural language supervision

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T14:45:44.817992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:45:44.817992Z digest=sha256:456e8d37d2224595234365deccebe4f6afdac180ccc927d7f28b92ab4bc07a40

Observation bff2d027-2e6c-4b3b-8206-82be550f3672 · outbound

This paper cites GestureDiffuCLIP: Gesture Diffusion Model with CLIP Latents.

LMM-Regularized CLIP Embeddings for Image Classification GestureDiffuCLIP: Gesture Diffusion Model with CLIP Latents

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T14:45:44.822200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:45:44.822200Z digest=sha256:e1514dc4cf40f7af8edca1b78ba1284a7762dc629d3fbfe42c1129a0dcab678a

Observation e3017430-f5b6-449f-9971-2adb852e1ad2 · outbound

This paper cites Delving into clip latent space for video anomaly recognition.

LMM-Regularized CLIP Embeddings for Image Classification Delving into clip latent space for video anomaly recognition

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:45:45.069767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T14:45:44.826196Z digest=sha256:56820fb5bae2c389bb1d7113c0ebdf7b13146769e503432ecb5dc8d29af1835a

Observation ad9fdcca-b4c0-493d-8815-ad472a137480 · outbound

This paper cites A Survey of Large Language Models.

LMM-Regularized CLIP Embeddings for Image Classification A Survey of Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T14:45:44.829990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:45:44.829990Z digest=sha256:0c033cf12693f01c893ff91edd1c930075d37b8db54e36f9bbdb2de406f7fed0

Observation 80e931cc-c0db-493e-9f9d-2e8744e85ee5 · outbound

This paper cites A Survey on Multimodal Large Language Models.

LMM-Regularized CLIP Embeddings for Image Classification A Survey on Multimodal Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T14:45:44.834033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:45:44.834033Z digest=sha256:514deca7b2fd59e48ad2a4d3c9db82cd311d35aa9feedf38b14eb18fad840345

Observation 51e91cb0-b2c8-41fb-85bd-152915dd2aba · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

LMM-Regularized CLIP Embeddings for Image Classification MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T14:45:44.838160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:45:44.838160Z digest=sha256:a25668eb1d6a2134f16917fda147cc292fb22ba5741590db07528c691383c953

Observation 7bf1e90a-37f0-4fc5-aaa3-3e90e3a7a762 · outbound

This paper cites Graph embedded convolutional neural networks in human crowd detection for drone flight safety.

LMM-Regularized CLIP Embeddings for Image Classification Graph embedded convolutional neural networks in human crowd detection for drone flight safety

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:45:45.059429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T14:45:44.842426Z digest=sha256:409e386ad8fed877989a7928d7402191ea22a6ca61f01d3cbf84522fb4313d7f

Observation c6f337d1-6a24-4637-85cc-1d5291fac1d6 · outbound

This paper cites Learning to prompt for vision-language models.

LMM-Regularized CLIP Embeddings for Image Classification Learning to prompt for vision-language models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T14:45:44.845777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:45:44.845777Z digest=sha256:871ce393a818385b6acc3757067bf106ad9ad00a8534e7531054cbf356be17a3

Observation f02c98b0-69ba-48dc-9003-aed864977b57 · outbound

This paper cites Conditional prompt learning for vision-language models.

LMM-Regularized CLIP Embeddings for Image Classification Conditional prompt learning for vision-language models

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:45:45.042173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T14:45:44.849154Z digest=sha256:ce03947fffcd1bcedcdf5ffef51ce4ea075b43136cf21c9b99f859cd599816e2

Observation e8c474e3-522b-422d-87c2-a096e4333eca · outbound

This paper cites Language in a bottle: Language model guided concept bottlenecks for interpretable image classification.

LMM-Regularized CLIP Embeddings for Image Classification Language in a bottle: Language model guided concept bottlenecks for interpretable image classification

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:45:45.030547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T14:45:44.852508Z digest=sha256:c6c008f5f299e7b665380e9c3d7e7cd1df4d71826c242cd4144a65fb405977c1

Observation 0140eeeb-c6d1-4c3a-a6d3-7ad68c9eaf54 · outbound

This paper cites Language models are few-shot learners.

LMM-Regularized CLIP Embeddings for Image Classification Language models are few-shot learners

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T14:45:44.855619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:45:44.855619Z digest=sha256:66be8ca2e158d4efa304e10117c627656fa4612234dcb31b9f6d785566cf9a27

Observation 98ff77d4-3c2f-48f5-a56e-80b242af9b1e · outbound

This paper cites En- hancing clip with gpt-4: Harnessing visual descriptions as prompts.

LMM-Regularized CLIP Embeddings for Image Classification En- hancing clip with gpt-4: Harnessing visual descriptions as prompts

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:45:45.012438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T14:45:44.859052Z digest=sha256:e7c42db719a28d7364d74aa2aee77831d250121edd1c34e1fc8198ad5a13076b

Observation 4f2ecd95-2697-4bb9-9e48-2285ca6f157c · outbound

This paper cites GPT-4 Technical Report.

LMM-Regularized CLIP Embeddings for Image Classification GPT-4 Technical Report

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T14:45:44.862809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:45:44.862809Z digest=sha256:9bf3a9a0aa8b895047a816ee886b52c1b5b48326b2d8dd4e361a1e54e22ef38b

Observation b685c802-8fae-4abc-8f85-0037d437b2a5 · outbound

This paper cites Exploiting lmm-based knowledge for image classification tasks.

LMM-Regularized CLIP Embeddings for Image Classification Exploiting lmm-based knowledge for image classification tasks

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:45:45.001520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T14:45:44.866432Z digest=sha256:2354ceaa560ed7a0459cdb3466f06510d499ada7b3dba36a4ea272f07b0abd24

Observation 684881ee-5e3d-472e-af5c-cd6cb9315d5e · outbound

This paper cites Disturbing image detection using lmm-elicited emotion embeddings.

LMM-Regularized CLIP Embeddings for Image Classification Disturbing image detection using lmm-elicited emotion embeddings

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:45:44.990963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T14:45:44.869730Z digest=sha256:333d2d88499cb72a9586121b52fd116971789c88f10181e40cd5cdf0925e5496

Observation 07e99017-8748-4185-9ee3-97db8e30c678 · outbound

This paper cites UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild.

LMM-Regularized CLIP Embeddings for Image Classification UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T14:45:44.872958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:45:44.872958Z digest=sha256:8c1e015f14a5e7e00ab6b54d3a0941090453076aeafa7a46089b65f70e763754

Observation 16b33324-d2fa-4497-9fe9-38f7cc0f2802 · outbound

This paper cites an unresolved cited work.

LMM-Regularized CLIP Embeddings for Image Classification Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-11T14:45:44.980271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T14:45:44.876510Z digest=sha256:d0b5287988cfa9c30ac62da4d19696dc4841b906b920b31e1e065b4951bf9e68

Observation fc70647e-bf9e-4efa-9771-47db5b465892 · outbound

This paper cites Learning from failure: De-biasing classifier from biased classifier.

LMM-Regularized CLIP Embeddings for Image Classification Learning from failure: De-biasing classifier from biased classifier

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:45:44.968822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T14:45:44.879975Z digest=sha256:de4e1d6f679207caca7d6a4b4093571ac4d395a182091787302f1daf2dbc186f

Pith citing papers

No inbound Pith citation observations are available.