Pith. sign in

Retrieval-based interleaved visual chain-of-thought in real-world driving scenarios

5 Pith papers cite this work. Polarity classification is still indexing.

5 Pith papers citing it

years

2026 3 2025 2

representative citing papers

CaVe-VLM-CoT: An Interpretable Vision-Language Model Framework

cs.AI · 2026-06-16 · unverdicted · novelty 5.0

CaVe-VLM-CoT is a closed-loop agentic-RAG framework with Extractor, Retriever, Solver, Citation Injector and Verifier stages plus 23 metrics anchored by CaVeScore that reports 87.1% accuracy on ScienceQA and 55.2% on MMMU without model changes.

citing papers explorer

Showing 5 of 5 citing papers.