Tag: speech-to-text

Blog
>
Tag: speech-to-text

ASR speech recognition speech-to-text

Enhanced speech recognition model is now available

62% Word Error Rate (WER) improvement for US English

ASR speech-to-text

Hot Summer Speech-to-Text Updates

Following Google’s release of new Speech API, we are happy to announce improved quality of call records transcription.

TTS streaming gemini elevenlabs voice agent

Introducing Gemini 2.0 Flash Live API Client and ElevenLabs Streaming TTS integration

New integrations for Voice AI have arrived: Google's Gemini 2.0 Flash model, featuring seamless voice-to-voice conversation capabilities and ElevenLabs low-latency streaming speech synthesis are now available for Voximplant developers

elevenlabs voice agent voice ai conversational ai

Introducing integration with ElevenLabs Conversational AI

Connect any Voximplant call to ElevenLabs Conversational AI agents

Voximplant Kit updates. January 2025

New Features in Voximplant Kit: Update overview We are constantly working to improve our product to make it easier to use and more effective for you. In this update, we have added several useful features. Here’s what’s new:

TTS ASR Integration voice ai

OpenAI Client update: gpt-realtime GA alignment

OpenAI has recently announced GA version of their Realtime API that Voximplant now fully supports

events

LEAP 2025: AI, Startups, & the Future of Tech – Voximplant Is There

Discover the future of tech at LEAP 2025 in Riyadh! Join Voximplant as we dive into the latest AI innovations, startup ecosystems, and groundbreaking technologies shaping tomorrow. Don’t miss this chance to network, learn, and transform your business.

Grok Voice Agent API now available in Voximplant

Voximplant now includes a native Grok module that connects any Voximplant call to xAI’s Grok Voice Agent API for real-time, speech-to-speech conversations. With a single VoxEngine scenario, you can interact via audio with Grok over phone numbers, SIP trunks and infrastructure, WhatsApp Business, or WebRTC into Grok — all without building custom media gateways or WebSocket streaming infrastructure.

Ultravox adds SIP to its Voice AI Services using Voximplant

Today Ultravox announced they are directly integrating Voximplant into their platform to provide SIP capabilities. The integration builds on Voximplant’s deep telephony and Voice AI tooling

Deepgram Voice Agent now available in Voximplant

Voximplant now includes a native Deepgram module that connects any Voximplant call to Deepgram’s Voice Agent API for real-time, speech‑to‑speech conversations. You can stream audio from phone numbers, SIP trunks, WhatsApp, or WebRTC into Deepgram’s unified agent environment—combining STT, LLM reasoning, and TTS—and play responses via Voximplant’s serverless runtime with minimal latency.

voximplant kit podcast voximplant-kit-cc-news product management voximplant-kit-automation-news web sdk webrtc video kit-updates call center ios sdk sip voximplant pstn api

Tag: speech-to-text

Enhanced speech recognition model is now available

Hot Summer Speech-to-Text Updates

Sign Up for a free Voximplant developer account or talk to our experts

Introducing Gemini 2.0 Flash Live API Client and ElevenLabs Streaming TTS integration

Introducing integration with ElevenLabs Conversational AI

Voximplant Kit updates. January 2025

OpenAI Client update: gpt-realtime GA alignment

LEAP 2025: AI, Startups, & the Future of Tech – Voximplant Is There

Grok Voice Agent API now available in Voximplant

Ultravox adds SIP to its Voice AI Services using Voximplant

Deepgram Voice Agent now available in Voximplant

Sign Up for a free Voximplant developer account or talk to our experts

Tag: speech-to-text

Sign Up for a free Voximplant developer account or talk to our experts

Sign Up for a free Voximplant developer account or talk to our experts

Contact Us