Medical speech-to-text benchmark
Thirty systems on the same sealed clinical audio and scorer.
Explore the benchmark →Add medical speech-to-text to a prototype, a startup product or a healthcare AI platform. Start with the hosted API, or inspect and deploy the open model yourself.
For clinicians who code · Healthtech startups · Healthcare AI companies
“Start amoxicillin 500 mg and repeat HbA1c at review.”
Use one medical voice layer for the moments your users already speak through.
Capture speakers, timing and the medical language needed for notes and downstream agents.
Give clinicians a faster input method inside EHR, dental, therapy and specialty workflows.
Feed confirmed medical text into triage, follow-up and care-navigation logic.
Run the open model in your own application when audio cannot leave the environment.
Use batch and live transcription with medical vocabulary, speakers and timestamps.
Deploy omi-medical-edge-1 on Apple Silicon, NVIDIA CUDA or CPU.
Overall transcription, medical terminology and dosage accuracy are scored independently on the same clinical audio.
Open benchmarks, published model results and reproducible evaluation.
Thirty systems on the same sealed clinical audio and scorer.
Explore the benchmark →Medical-term, drug-name and dosage results for the production API.
Read the model report →Current CUDA, Apple Silicon and CPU accuracy and throughput.
See the runtime results →Use the hosted API or download the open model.