Endpoints
Transcribe Audio
This endpoint allows you to transcribe audio files to text using Azure Speech-to-Text services. Only available for public agents.
POST
This endpoint converts audio files to text using Azure’s Speech-to-Text API. It is designed to be used from public-facing chatbot widgets and does not support API key authentication.
Use Cases:
- Enable voice input in chatbot widgets
- Transcribe user audio messages for AI agent processing
- Convert voice notes to text for conversation logs
- Build voice-enabled customer support interfaces
Headers
string
default:"multipart/form-data"
required
Must be set to
multipart/form-data for file uploads.Body (multipart/form-data)
file
required
The audio file to transcribe. Supported formats:
audio/wavaudio/webm
string
required
The ID of a public agent. This is required for all requests. The agent must have its visibility set to “public”.
Response
string
The transcribed text from the audio file.

