
High quality Speech Recognition is now available
We are happy to announce the high quality speech recognition for both audio call records transcription and real-time recognition scenarios.

We are happy to announce the high quality speech recognition for both audio call records transcription and real-time recognition scenarios.

The new version of Web SDK will help us to accelerate the development process and includes a lot of new features and improvements.

Now developers can use Promise in their VoxEngine scenarios and we also added Net.httpRequestAsync and Net.sendMailAsync functions.

Your mp3 or ogg files played on VoxEngine scenario level with call.startPlayback or using Player will be played on the Web or Mobile SDK side in HD quality (48KHz), or on SIP side if it does support wideband audio codecs (Speex or Opus).

In HD mode audio is being mixed at 48KHz, all audio sources with lower sample rate will be resampled to 48KHz.

We chose 48 KHz as the base sample rate for HD audio recorder, since WebRTC/Opus can offer this quality, audio from endpoints with lower sample rate will be re-sampled.

Full Featured Instant Messaging

If a call is made in non-P2P mode then its media stream goes via our media servers and we can record it if required.

We've started with audio, then we've added video calls and now it's time to let our developers use instant messaging and presence - two very important features of UC stack.

The new version of our mobile SDK uses WebRTC engine for audio/video processing and supports all features available for WebSDK.

Victor Pascual from Quobis invited us to participate in WebRTC meetup that took place on March 4th in Barcelona, we accepted the invitation and I'm really happy that we did.

Now there is a way to restrict access to VoxImplant HTTP API and only allow it for certain IP addresses or networks when api_key is being used.

Chili Piper is popular, but is it the best for you? This article compares it to competitors like Dashly, Calendly, and others, examining features, pricing, and ideal use cases. Discover the right scheduling tool for your team's needs.

Boost your food tech app in 2024! Learn 12 in-app content tricks from a study of 5000+ stories. Personalize, gamify, and use cross-channel messaging for user retention.

New Features in Voximplant Kit: Update overview. We are constantly working to improve our product to make it easier to use and more effective for you. In this update, we have added several useful features. Here’s what’s new:

New integrations for Voice AI have arrived: Google's Gemini 2.0 Flash model, featuring seamless voice-to-voice conversation capabilities and ElevenLabs low-latency streaming speech synthesis are now available for Voximplant developers

Discover the future of tech at LEAP 2025 in Riyadh! Join Voximplant as we dive into the latest AI innovations, startup ecosystems, and groundbreaking technologies shaping tomorrow. Don’t miss this chance to network, learn, and transform your business.

Voximplant has new realtime speech generation for voice AI from Inworld, our latest Voice AI text-to-speech (TTS) partner. Together, we combine state-of-the-art TTS with carrier-grade connectivity so you can build voice agents that sound like your brand, not a generic robot.

Voximplant now includes a native Cartesia module for streaming, low-latency text-to-speech (TTS). You can use a single VoxEngine API to synthesize speech in real time, connect it to any call (PSTN, SIP, WebRTC, WhatsApp) and control playback from a Large Language Model (LLM) or other source, all inside VoxEngine.

Check out the latest useful Voximplant Kit updates — we developed chat analytics, improved call history, added new tools for supervisors, expanded scenario capabilities, and updated the softphone. Below is a brief overview of the essential enhancements.