Enables end-to-end testing of voice-enabled web applications by simulating a synthetic user with a virtual microphone, virtual speakers, a real browser, and a viewable display, allowing audio injection, speech capture/transcription, and browser automation.
Enables speech-to-text transcription, text-to-speech synthesis, and audio analysis using Deepgram's AI models. Supports features like speaker diarization, sentiment analysis, language detection, and various audio processing capabilities.