Using Natural Language Prompts With AI Models for Low-Cost Assistive Software Design: Exploratory Comparative Evaluation.
Francesc Antoni Bañuls-Lapuerta, Vicent Marti-Miralles, Rómulo Jacobo Gónzalez-García and 1 others
PMID 41875212WHAT IT FOUND
Paid AI models generated code for a gamepad-based computer mouse that worked for 11 to 14 of 16 requested functions.
Free versions managed none to four. The authors say non-technical therapists should not rely on these tools to build assistive software without IT supervision.
Key findings
01Gemini Pro implemented 14 of 16 functions and ChatGPT Pro 11 of 16, while free versions achieved between none and four.
02Free models that produced some working code, such as DeepSeek and GPT Free, lost track of the programming process and generated unstable code a non-technical user would have to debug.
03The authors state professional supervision by a qualified software engineer remains preferable because AI-generated code is generally less efficient than human solutions.
STILL TO COME
How it was doneWhat they found
Read the rest of this summary
You get three full summaries a month, free, and we do not ask for a card. Search, the TL;DRs and your library stay unlimited either way.
What it does not show
Each model was tested only once, and the authors state these systems are nondeterministic, so repeating the same task could give different results. No real patients or therapists used the software, so the study does not show whether a clinician could actually follow the prompts or fix the errors the models produced. The task was one specific gamepad-to-mouse program with 16 fixed functions. Results may not transfer to other assistive software. The free versions were constrained by daily prompt limits, and the authors say they cannot tell whether more time would have produced results as good as the paid versions. The two functions no model achieved, sending an email and building an installer, were blocked by the AI systems for security reasons, not by poor coding ability.
Declared interests
Funded by the Spanish Ministry of Science and Innovation through a research project, in collaboration with the authors' university. The authors declared no conflicts of interest.
The easy way to misread this
Do not read this as evidence that a therapist can hand a prompt to a chatbot and get working assistive software. The authors tested only one program, ran each model once, and found that even the best free tools produced unstable code that would require programming skill to repair. They recommend qualified software engineer supervision.
Summarised by AI from the full paper, without a clinician reviewing it. Check it against the source before it changes what you do. Read it on PubMed →