Three new transcription models on medical audio
Azure, OpenAI and Google joined the same sealed 1,513-clip clinical board.
See what changed →Add medical speech to the tools you are building—from a weekend prototype to a production healthcare AI product. Use the hosted API or deploy our open model yourself.
For clinicians who code · Healthtech startups · Healthcare AI companies
“Start amoxicillin 500 mg and repeat HbA1c at review.”
Use one medical voice layer for the moments your users already speak through.
Capture speakers, timing and the medical language needed for notes and downstream agents.
Give clinicians a faster input method inside EHR, dental, therapy and specialty workflows.
Feed confirmed medical text into triage, follow-up and care-navigation logic.
Run the open model in your own application when audio cannot leave the environment.
Every plan gets the complete medical API. Payment protects continuity and raises capacity.
For clinicians who code, independent builders and small healthcare AI projects.
For startups and production products that need continuity without a contract.
For healthcare AI companies that need contracted capacity, guarantees or private deployment.
omi-medical-edge-1 has open weights under CC-BY-4.0 and runs on Mac, NVIDIA CUDA or CPU. Inspect the model, test it on your own clinical audio, and keep every byte inside your environment.
Overall transcription, medical terminology and dosage accuracy are scored independently on the same clinical audio.
Practical guides, reproducible comparisons and open model work for people putting medical speech into products. Follow the RSS feed →
Azure, OpenAI and Google joined the same sealed 1,513-clip clinical board.
See what changed →Evaluate clinical accuracy, dosage safety, BAA access, speakers, languages and real price.
Read the guide →What a healthcare builder should verify before sending protected health information.
Use the checklist →The hosted API keeps the operational choices visible, while the open model gives you a path to keep audio entirely inside your environment.
Hosted API audio and transcripts are processed in the European Union.
Sign the BAA and DPA in the console on Builder, before sending PHI.
Choose how long stored job results remain available for retrieval.
Customer audio and transcripts are not used to train Omi models.
Use the hosted API or download the open model.