The direct comparison
The benchmark row is Google Chirp 3, not Google's separate legacy medical-conversation or medical-dictation models. Checked 29 August 2026.
Know which Google model you are comparing
Google Cloud lists standard models such as Chirp separately from V1 medical dictation and medical conversation. The tested benchmark row is Chirp 3 because it represents Google's newer general speech stack. If your procurement depends on a named medical SKU, benchmark that exact model and region.
Pricing mechanics
Google's public US list currently shows standard V2 recognition at $0.016 per minute, dynamic batch at $0.003 per minute and the older V1 medical models at $0.078 per minute after the first 60 minutes. Storage, multiple channels and adjacent cloud services can change the total.
Omi publishes two feature-inclusive rates after the recurring allowance: $0.29 per batch audio-hour and $0.45 per live audio-hour. That is easier to model for a product whose primary job is medical transcription.
When Google is the better fit
Google can be the pragmatic choice when your data, identity, monitoring and procurement already live in Google Cloud, or when your application needs its wider language and cloud-service portfolio. Omi is the focused choice when clinical term and dosage accuracy, transparent packaging and a self-deployed open model carry more weight.
Sources
Review the full Omi benchmark and Google Cloud Speech-to-Text pricing. Verify model availability, region and contract before production.