Flow2Code, a new 15-language, 16,866-flowchart benchmark, shows current multimodal LLMs generate code from plain code flowcharts only about half the time on average and substantially worse from pseudocode flowcharts.
Iterate through the list `numbers`
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.SE 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Flow2Code: Evaluating Large Language Models for Flowchart-based Code Generation Capability
Flow2Code, a new 15-language, 16,866-flowchart benchmark, shows current multimodal LLMs generate code from plain code flowcharts only about half the time on average and substantially worse from pseudocode flowcharts.