Medical speech-to-text

Get the medicine right.

Clinical transcription with the top medical-term accuracy on our sealed benchmark and the strongest dosage score of 30 systems.

EU processingSelf-serve BAACustomer data is never used for training
Consultation audio
“Start amoxicillin 500 mg and repeat HbA1c at review.”
✓ Drug and dosage preserved
#1 / 29medical-term accuracy
0.00%drug-name error
97.7%dosage F1 · 86/89 events
Hosted API plans

Build for free. Add a card for production.

Every plan gets the complete medical API. Payment protects continuity and raises capacity.

Builder
$025 pooled audio-hours every month

For clinician hackers, prototypes and small healthcare AI projects.

  • Complete batch and live API
  • 2 live sessions · 1 speaker room
  • No card required · no rollover
  • Self-serve BAA and DPA
Create a Builder key
Enterprise
CustomContract pricing and capacity

For products that need reserved capacity, operational guarantees or private deployment.

  • Custom maximum concurrency
  • Reserved live and speaker capacity
  • Private cloud and supported SDK
  • SLA and integration support
Talk to us
Benchmark results

Every layer of the medical record, measured.

Overall transcription, medical terminology and dosage accuracy are scored independently on the same clinical audio.

Overall word error · lower is better

Azure5.97
Omi5.99
AWS6.12
ElevenLabs6.17

Medical-term error · lower is better

Omi0.94
ElevenLabs0.97
Google1.11
AssemblyAI1.43

Dosage F1 · higher is better

Omi97.7
Deepgram86.8
ElevenLabs85.4
Azure83.3
Trust center

Built for clinical data.

The hosted API keeps the operational choices visible, while the open model gives you a path to keep audio entirely inside your environment.

01

EU processing

Hosted API audio and transcripts are processed in the European Union.

02

Self-serve agreements

Sign the BAA and DPA in the console on Builder, before sending PHI.

03

1–72 hour retention

Choose how long stored job results remain available for retrieval.

04

Never used for training

Customer audio and transcripts are not used to train Omi models.

Start with your own audio.

Use the hosted API or download the open model.