← back to paper
arxiv: 2608.03357 · 2 revisions
Can Text-to-Image Models Draw from the Right Frame of Reference?