A team at the University of Utah created a framework to assess how conversational artificial intelligence, including large language models, could automate parts of therapy. The framework describes four categories: scripted systems, AI that evaluates therapists, AI that assists therapists, and AI that provides therapy directly. Each category shows a different level of automation.
The researchers evaluated usefulness and risk and noted that users and health systems may not always know which level they use. The team works with a statewide crisis text line to build tools that review sessions and give feedback. They advise starting with lower-risk tools while studying possible benefits and harms.
Difficult words
- framework — A basic structure to organize ideas or work
- assess — To judge how good or useful something is
- automate — To make a task happen by machine or software
- evaluate — To examine something and decide its qualityevaluates, evaluated
- therapy — Treatment to help a person's mental or physical health
- risk — Possibility that something bad could happen
- feedback — Information given to improve work or behavior
Tip: hover, focus or tap highlighted words in the article to see quick definitions while you read or listen.
Discussion questions
- Do you think AI should give therapy directly? Why or why not?
- Would you feel comfortable if AI gave feedback to a therapist? Why?
- Do you agree with starting with lower-risk tools first? Why or why not?
Related articles
AI audio summaries of research can help — and err
Researchers tested Google’s NotebookLM, which turns research papers into podcast-style audio. The summaries were engaging and clearer for teaching, but every audio overview contained mistakes, so the authors advise reading the original papers to check claims.
Brain predictions use phrases, not just next words
New research shows the human brain anticipates upcoming language by grouping words into grammatical phrases rather than predicting only the next single word. Scientists used brain recordings, behavioral tests and LLM measures across languages.