What we builtTranscription that survives real rooms
Two people, overlapping speech, background noise, and a vocabulary where "hypertension" and "hypotension" differ by one phoneme and mean opposite things. General transcription is excellent at conversational English and noticeably worse here, which is where the tuning went.
What we builtStructuring into SOAP rather than prose
The output has to land in Subjective, Objective, Assessment and Plan as the record system expects them, not as a paragraph somebody then has to cut up. The model is constrained to that structure rather than asked politely for it.
What we builtA verification pass against the transcript
Every clinical claim in the summary is checked back against the words actually spoken. Anything the transcript does not support is dropped rather than softened, and the clinician sees what was dropped.
What we builtHandling built in from the first commit
HIPAA and GDPR requirements shaped where audio lives, how long it lives, who can reach it and what is logged. Retro-fitting that after a working prototype is how projects in this sector die.