Pith. sign in

Llama 3.2: Revolutionizing edge ai and vision with open, customizable models, 2024

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.CV 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

Multi-Modal Language Models as Text-to-Image Model Evaluators

cs.CV · 2025-05-01 · conditional · novelty 6.0

MT2IE uses a single open-source multimodal LLM to generate 20 progressively harder prompts and score image-text consistency, reproducing the 1,600-prompt GenAIBench ranking of 8 text-to-image models.

citing papers explorer

Showing 1 of 1 citing paper.

  • Multi-Modal Language Models as Text-to-Image Model Evaluators cs.CV · 2025-05-01 · conditional · none · ref 2

    MT2IE uses a single open-source multimodal LLM to generate 20 progressively harder prompts and score image-text consistency, reproducing the 1,600-prompt GenAIBench ranking of 8 text-to-image models.