For years, DeepL has been the quiet favourite of anyone needing highly accurate, context-aware text translation that does not feel like it was hastily spat out by a machine. However, the language artificial intelligence company is now stepping away from the keyboard and directly into the boardroom. DeepL has officially launched its new Voice-to-Voice product suite, an ambitious real-time spoken translation platform designed to eliminate the friction of multilingual communication in global enterprises.
Bridging Virtual and Physical Conversations
The core premise of the new suite is incredibly straightforward but technically massive: you speak naturally in your native language, and the person on the other end hears you in theirs. Rather than forcing companies to abandon their existing communication platforms, DeepL is integrating its artificial intelligence directly into the environments where work already happens.

For virtual collaboration, the company introduced Voice for Meetings. This specific module hooks directly into corporate staples like Microsoft Teams and Zoom. It allows participants to speak comfortably in their preferred language while others hear the simultaneous translation on their end. For physical, in-person interactions—such as a factory floor safety briefing or an international corporate workshop—DeepL launched Group Conversations. This feature allows participants to instantly join a secure, multilingual session simply by scanning a QR code with their mobile device, ensuring everyone on site receives the simultaneous voice translation at the exact same pace.
Integrating with Internal Systems
Beyond standard meetings, DeepL is also opening up its underlying architecture to developers through the Voice-to-Voice API. This allows massive enterprises, particularly international contact centres and business process outsourcing firms, to embed the real-time translation engine directly into their own internal software. This fundamentally shifts how global teams operate. A customer service department can now hire strictly for technical expertise rather than specific language fluency, as the application programming interface seamlessly bridges the communication gap between the support agent and the caller in real time.
Tackling Corporate Jargon
The true challenge with live voice translation is not just raw speed; it is dealing with heavy industry jargon, complex acronyms, and fast talkers. DeepL is addressing this with a new quality optimisation feature called Spoken Terms. By natively integrating custom corporate glossaries into the voice engine, organisations can ensure that highly specific product names or technical terminology are accurately captured, transcribed, and translated without the artificial intelligence confidently guessing the wrong word.
Accuracy and Enterprise-Grade Privacy
Under the hood, the system currently utilises a highly optimised three-step pipeline: converting speech to text, running it through DeepL’s established translation models, and synthesising it back into voice. According to the company, this reliance on their industry-leading text engine pays off massively. In independent blind evaluations conducted by the language research firm Slator, professional linguists preferred DeepL Voice over the native translation solutions built into Microsoft Teams, Google Meet, and Zoom a staggering 96% of the time. The reviewers specifically cited vastly superior fluency, better contextual accuracy, and a dramatically lower error rate compared to incumbent platforms.
Because the platform is built strictly for enterprise deployment rather than everyday consumers, DeepL has heavily prioritised data privacy. The company explicitly guarantees that it never uses customer voice or transcription data to train its underlying models. All meeting data is processed temporarily in memory and permanently deleted the moment a call ends, ensuring that highly sensitive corporate discussions remain entirely confidential.
Availability and Deployment
If your team is struggling with cross-border communication, the rollout of these features is currently staggered. The Voice for Conversations module and the new Group Conversations feature are generally available right now. The Spoken Terms customisation capability is slated for a broader release on May 7. Meanwhile, the highly anticipated Voice for Meetings module is scheduled to enter an early access program in June. To ensure the technology is accessible to businesses of all sizes, DeepL has also introduced a self-serve model, allowing smaller teams to directly purchase the suite online, start a free trial, and instantly test the deployment without getting locked into massive enterprise contracts.
