An LLM agent that writes code and executes it inside the app's runtime can complete tasks that GUI-clicking agents struggle with, with a small case study reporting up to 80% task completion.
(1961) shoebox - ibm archives (78-013)
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.SE 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Every Software as an Agent: Blueprint and Case Study
An LLM agent that writes code and executes it inside the app's runtime can complete tasks that GUI-clicking agents struggle with, with a small case study reporting up to 80% task completion.