Getting Started
Follow these steps to configure Azure Cognitive Services as your STT provider:1
Add Azure credentials to your vault
Navigate to Integration → Vault in the Rapida dashboard. Add your Azure Subscription Key and the endpoint URL for your Speech resource. The IAM role must have the
Cognitive Services User permission.2
Select Azure as your STT provider
When configuring your assistant, open Audio Settings and choose Azure Cognitive Services as your Speech-to-Text provider.
3
Choose a language
Select the BCP-47 language code for your primary language (e.g.
en-US, es-ES, fr-FR). Azure supports 100+ languages and locales.Supported Models
Key Features
- Real-time streaming: Low-latency partial and final transcripts for live voice applications
- Speaker diarization: Identify and label individual speakers in a conversation
- Custom Speech models: Train on your own audio data to improve accuracy for domain-specific terms
- Profanity filtering: Mask or remove profane words from transcripts
- Phrase lists: Boost recognition accuracy for specific words or phrases
- Content redaction: Automatically redact PII from transcripts
Supported Languages
Azure supports 100+ languages and locales including English (US, UK, AU, IN), Spanish, French, German, Italian, Portuguese, Japanese, Chinese (Simplified, Traditional), Korean, Hindi, and Arabic. See the Azure documentation for the full list.Configuration Options
Notes
- For telephony use cases, set sample rate to 8000 Hz to match PSTN audio.
- Custom Speech models require training data upload via the Azure portal.
- Pricing is per audio hour. See Azure Speech pricing.