RNOtherJournal of nursing scholarship : an official publication of Sigma Theta Tau International Honor Society of Nursing2025

From Conversation to Standardized Terminology: An LLM-RAG Approach for Automated Health Problem Identification in Home Healthcare.

Zhihong Zhang, Pallavi Gupta, Jiyoun Song and 2 others

PMID 40785044

WHAT IT FOUND

An automated system classified standard home-care health problems from nurse-patient conversations and matched expert labels better when it was given examples and step-by-step prompts.

It was not tested for patient outcomes or real workflow use.

Key findings

01The automated system classified Omaha health problems from 5118 utterances in 22 encounters, and GPT-4o-mini had the best performance after example prompts and step-by-step reasoning.

02Adding example utterances and step-by-step reasoning improved classification of both problem categories and signs or symptoms, especially for GPT-3.5-mini and GPT-4o-mini.

03The best settings were chosen on one conversation with 114 utterances, then applied to all 5118 utterances.

STILL TO COME

How it was doneWhat they foundWhat it means for RNs

Read the rest of this summary

You get three full summaries a month, free, and we do not ask for a card. Search, the TL;DRs and your library stay unlimited either way.

Already have one?

What it does not show

The evaluation used only 22 selected recordings from 44, so it may not represent all home-care visits. It covered only 28 unique problems and 84 signs/symptoms, not the full Omaha System. The best settings were chosen on one conversation with 114 utterances, so tuning may not generalize. Transcription accuracy was assumed; real speech with impairments, accents, or unclear articulation could degrade performance. The paper says real-world implementation still needs evaluation. The paper says clinician review and training are needed to use outputs.

Declared interests

The authors declared no conflicts of interest.

The easy way to misread this

Do not conclude this home-care documentation tool is ready for clinical use or will improve care. The paper says real-world implementation still needs evaluation, and outputs would require clinician review.

Read it on PubMed →