Fine-tuning a LLaVA-style vision-language model on 163k synthetic image-CadQuery pairs yields a model that compiles every test script and matches CAD solids better than general VLMs.
Towards implicit text-guided 3d shape generation
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
CAD-Coder: An Open-Source Vision-Language Model for Computer-Aided Design Code Generation
Fine-tuning a LLaVA-style vision-language model on 163k synthetic image-CadQuery pairs yields a model that compiles every test script and matches CAD solids better than general VLMs.