A config- urable library for generating and manipulating maze datasets.arXiv preprint arXiv:2309.10498, 2023

Michael Igorevich Ivanitskiy, Rusheb Shah, Alex F Spies, Tilman Räuker, Dan Valentine, Can Rager, Lucia Quirke, Chris Mathwin, Guillaume Corlouer, Cecilia Diniz Behn, et al · 2023 · arXiv 2309.10498

3 Pith papers cite this work. Polarity classification is still indexing.

3 Pith papers citing it

read on arXiv browse 3 citing papers

representative citing papers

SVoT: State-aware Visualization-of-Thought for Spatial Reasoning via Reinforcement Learning

cs.AI · 2026-06-10 · unverdicted · novelty 6.0

SVoT uses RL with GRPO to train MLLMs on interleaved textual and visual reasoning chains for multi-hop spatial tasks, achieving up to 65% accuracy gains on new domains with quantitative state verification.

Spectral-Progressive Thought Flow for Lightweight Multimodal Reasoning

cs.LG · 2026-06-01 · unverdicted · novelty 6.0

SpecFlow represents intermediate visual thoughts in fixed-size DCT space and uses classifier-free guidance to steer updates from textual thoughts, achieving up to 2.1x lower computation and KV cache costs.

The Confidence Shortcut: A Reasoning Failure Mode of Masked Diffusion Models

cs.AI · 2026-05-27 · unverdicted · novelty 6.0

Confidence-based decoding and training in masked diffusion models shortcut long-range dependencies in reasoning, producing errors on complex inputs that random masking avoids.

citing papers explorer

Showing 3 of 3 citing papers after filters.

SVoT: State-aware Visualization-of-Thought for Spatial Reasoning via Reinforcement Learning cs.AI · 2026-06-10 · unverdicted · none · ref 14
SVoT uses RL with GRPO to train MLLMs on interleaved textual and visual reasoning chains for multi-hop spatial tasks, achieving up to 65% accuracy gains on new domains with quantitative state verification.
Spectral-Progressive Thought Flow for Lightweight Multimodal Reasoning cs.LG · 2026-06-01 · unverdicted · none · ref 11
SpecFlow represents intermediate visual thoughts in fixed-size DCT space and uses classifier-free guidance to steer updates from textual thoughts, achieving up to 2.1x lower computation and KV cache costs.
The Confidence Shortcut: A Reasoning Failure Mode of Masked Diffusion Models cs.AI · 2026-05-27 · unverdicted · none · ref 7
Confidence-based decoding and training in masked diffusion models shortcut long-range dependencies in reasoning, producing errors on complex inputs that random masking avoids.

A config- urable library for generating and manipulating maze datasets.arXiv preprint arXiv:2309.10498, 2023

fields

years

verdicts

representative citing papers

citing papers explorer