1
Create Account and Access the Dashboard
Begin by visiting the official product page at the provided URL. Click the 'Get Started' or 'Sign Up' button to create your developer or creator account. Universal Dictation on Stream typically requires a brief registration process involving email verification. Once logged in, you will land on the main dashboard. This interface serves as the control center for your streaming sessions. Familiarize yourself with the layout: on the left, you'll find your active sessions and history; on the right, configuration options for language models and output formats. Ensure you navigate to the 'New Stream' or 'Start Session' button to initiate the integration process. The dashboard provides an overview of your usage limits and API keys if you plan to integrate this tool programmatically into a custom broadcasting stack.
Pro Tip
Verify your email address immediately after signing up to unlock full API access and higher transcription limits.
2
Configure Audio Input Settings
Before starting your stream, you must configure the audio source. Universal Dictation on Stream can accept audio via a virtual audio cable (like VB-Cable or BlackHole on macOS) or directly through an API endpoint if you are using a custom encoder. For most users, the 'Virtual Audio Cable' method is recommended. Go to the 'Audio Input' section in the settings. Select your preferred virtual audio device. If you are using OBS Studio, add a 'Audio Input Capture' source and select the virtual cable that Universal Dictation on Stream provides. This allows the AI to intercept your microphone or system audio without altering your physical output. Ensure the sample rate is set to 44.1kHz or 48kHz, as these are standard for high-quality transcription. Incorrect sample rates can lead to synchronization issues between the spoken word and the displayed captions.
Pro Tip
Test your audio input levels before going live. The AI performs best with clear, noise-free audio signals.
3
Set Language and Caption Style
Customize the output to match your brand and audience needs. In the 'Output Configuration' tab, select the primary language for transcription. Universal Dictation on Stream supports multiple languages, but for the highest accuracy, choose the specific dialect (e.g., 'English - US' vs. 'English - UK'). Next, configure the caption style. You can choose between 'Open Captions' (burned into the video) or 'Closed Captions' (sent as a separate data stream like WebVTT). For live streaming on platforms like Twitch or YouTube, Closed Captions are often preferred as they allow viewers to toggle them on or off. Adjust the font size, color, and background opacity to ensure readability against various video backgrounds. You can also enable 'Speaker Diarization' if you have multiple hosts, which will label captions with speaker names (e.g., [Host], [Guest]).
Pro Tip
Use high-contrast colors for captions to ensure they are visible on both light and dark video backgrounds.
4
Integrate with OBS Studio or Streaming Software
The most common use case is integrating with OBS Studio. Universal Dictation on Stream provides a browser source URL or a dedicated plugin. If using the browser source method, copy the provided 'Live Stream Key' from the Universal Dictation dashboard. In OBS, add a new 'Browser' source. Paste the URL, ensuring you include your API key or session token in the query parameters as specified in the documentation. Set the width and height to match your canvas resolution (e.g., 1920x1080). Position the browser source at the bottom of your scene where captions typically appear. If you are using a custom encoder via API, you will need to send your audio stream to the provided WebSocket endpoint. The tool will return the transcription data in JSON format, which you can then render using a simple HTML/CSS overlay in your broadcasting software.
Pro Tip
Keep the OBS browser source updated. If the session expires, the captions will stop appearing.
5
Start the Live Session and Monitor Latency
Once your audio and video sources are configured, click 'Start Live Session' in the Universal Dictation dashboard. Observe the real-time feed in the tool's preview window to ensure accuracy. Then, start your stream in OBS or your preferred platform. Watch for the captions to appear on your live feed. The tool is designed for low-latency processing, typically adding less than 2 seconds of delay. Monitor the 'Latency' metrics in the dashboard. If you notice significant lag, check your internet upload speed and consider reducing the audio bitrate sent to the AI. During the first few minutes, speak clearly and at a moderate pace to allow the AI to calibrate to your voice patterns. Universal Dictation on Stream uses adaptive learning to improve accuracy over the course of the session, so the first few sentences might be slightly less accurate than later ones.
Pro Tip
Avoid speaking over music or loud sound effects, as this can confuse the transcription engine.
6
Post-Stream Analysis and Export
After ending your live stream, terminate the session in the Universal Dictation dashboard. The tool automatically saves the full transcript and the generated caption files. Navigate to the 'History' section to view your past streams. Here, you can download the transcript in various formats, including TXT, SRT, and WebVTT. This is invaluable for creating blog posts, video descriptions, or searchable content from your live broadcast. Review the transcription for any errors. The platform may offer a 'Review & Edit' feature where you can correct misheard words. These corrections help improve the AI's model for future sessions, especially for unique jargon or proper nouns. Exporting these captions allows you to repurpose your live content for on-demand viewing, significantly extending the lifespan and reach of your content.
Pro Tip
Save your corrected transcripts to a glossary to help the AI recognize specific industry terms in future streams.