Get Started
STT Playground
Record live or upload audio to transcribe with Inworld STT
The STT Playground makes it easy to try out Inworld's Speech-to-Text capabilities through an interactive playground. Use it to test language hints and custom vocabulary, and watch Voice Profile analyze the speaker in real time — all before writing a line of code.

Get Started
Go to Inworld Portal
In Portal, select Speech-to-Text from the left-hand side panel.
Choose a model
Click the model dropdown at the top of the playground to select inworld/inworld-stt-1. See Supported models for its endpoints and language coverage.
Advanced Features
For greater control over transcription, try the following:
- Language hint — The Language dropdown defaults to Auto-detect, which recognizes the spoken language automatically. Select a specific language to give the model a hint when you know the audio's language in advance — this generally improves accuracy, especially for short utterances. See Language Support for how hints behave.
- Custom vocabulary — Click Vocabulary to add key terms (names, jargon, acronyms) that bias recognition toward them. Terms are added as chips, and the button shows how many are active. This is a soft bias rather than a hard keyword lock — test it on the terms the baseline actually misses. The equivalent API field is
prompts. - Voice Profile — The panel on the right analyzes the speaker alongside transcription, showing Age, Accent, Voice Pitch, Vocal Style, and Emotion, each with a confidence score. It populates after transcription begins. See Voice Profiles for the full category reference.
From Playground to Code
Click Integrate (top right) — or Get integration code in the Voice Profile panel — to generate an API request matching your current playground configuration: model, language hint, vocabulary, and Voice Profile settings.
Next Steps
Ready to build? Here's where to go next.