Medical voice infrastructure

Voice AI built for healthcare.

Add medical speech-to-text to a prototype, a startup product or a healthcare AI platform. Start with the hosted API, or inspect and deploy the open model yourself.

For clinicians who code · Healthtech startups · Healthcare AI companies

25 free audio-hours monthlySelf-serve BAANever used for model training
Consultation audio
“Start amoxicillin 500 mg and repeat HbA1c at review.”
✓ Drug and dosage preserved
#1 / 30medical-term accuracy
0.00%drug-name error
97.7%dosage F1 · 86/89 events
Built for the workflow

Turn clinical speech into product data.

Use one medical voice layer for the moments your users already speak through.

Consultations

Ambient documentation

Capture speakers, timing and the medical language needed for notes and downstream agents.

Dictation

Clinical text entry

Give clinicians a faster input method inside EHR, dental, therapy and specialty workflows.

Patient calls

Voice workflows

Feed confirmed medical text into triage, follow-up and care-navigation logic.

Local AI

Private on-device speech

Run the open model in your own application when audio cannot leave the environment.

Choose how it runs

Hosted API or open model.

Hosted API

Start with 25 free hours.

Use batch and live transcription with medical vocabulary, speakers and timestamps.

  • No card required
  • Self-serve BAA and DPA
  • Pay only when usage grows
Open model

Run it in your environment.

Deploy omi-medical-edge-1 on Apple Silicon, NVIDIA CUDA or CPU.

  • CC-BY-4.0 weights
  • MIT runtime
  • Published results for each platform
Benchmark results

Every layer of the medical record, measured.

Overall transcription, medical terminology and dosage accuracy are scored independently on the same clinical audio.

Overall word error · lower is better

Azure5.97
Omi5.99
AWS6.12
ElevenLabs6.17

Medical-term error · lower is better

Omi0.94
ElevenLabs0.97
Google1.11
AssemblyAI1.43

Dosage F1 · higher is better

Omi97.7
Deepgram86.8
ElevenLabs85.4
Azure83.3
Research and evidence

See the work behind the claims.

Open benchmarks, published model results and reproducible evaluation.

Benchmark

Medical speech-to-text benchmark

Thirty systems on the same sealed clinical audio and scorer.

Explore the benchmark →

Start with your own audio.

Use the hosted API or download the open model.