Pith. sign in

By analyzing which concepts are invoked, we assess the knowledge dimension activated during problem-solving and whether the LLM navigates these domains coherently

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.CL 1

years

2025 1

verdicts

REJECT 1

representative citing papers

THiNK: Can Large Language Models Think-aloud?

cs.CL · 2025-05-26 · reject · novelty 4.0

THiNK uses a multi-agent, feedback-driven loop of problem revision and GPT-4O-based Bloom's Taxonomy scoring to measure and improve higher-order thinking in LLMs on math word problems.

citing papers explorer

Showing 1 of 1 citing paper.

  • THiNK: Can Large Language Models Think-aloud? cs.CL · 2025-05-26 · reject · none · ref 1

    THiNK uses a multi-agent, feedback-driven loop of problem revision and GPT-4O-based Bloom's Taxonomy scoring to measure and improve higher-order thinking in LLMs on math word problems.