An API powered by Google's AI technology allows you to accurately convert speech into text. You can accurately caption your content, provide a better user experience with products using voice commands, and gain insight from customer interactions to improve your service. Google's deep learning neural network algorithms are the most advanced in automatic speech recognition (ASR). Speech-to-Text allows for experimentation, creation, management, and customization of custom resources. You can deploy speech recognition wherever you need it, whether it's in the cloud using the API or on-premises using Speech-to-Text O-Prem. You can customize speech recognition to translate domain-specific terms or rare words. Automated conversion of spoken numbers into addresses, years and currencies. Our user interface makes it easy to experiment with your speech audio.
Learn more

Get quality translations for your app, website, game, supporting documentation, and on. Invite your own translation team or work with professional translation agencies within Crowdin.
Features that ensure quality translations and speed up the process
• Glossary – create a list of terms to get consistent translations
• Translation Memory (TM) – no need to translate identical strings
• Screenshots – tag source strings to get context-relevant translations
• Integrations – set up integration with GitHub, Google Play, API, CLI, Android Studio, and on
• QA checks – make sure that all the translations have the same meaning and functions as the source strings
• In-Context – proofreading within the actual web application
• Machine Translations (MT) – pre-translate via translation engine
• Reports – get insights, plan and manage the project
Crowdin supports more than 30 file formats for mobile, software, documents, subtitles, graphics and assets:
.xml, .strings, .json, .html, .xliff, .csv, .php, .resx, .yaml, .xml, .strings and on.
Learn more
Dubly.AI
Dubly.AI is a browser-based dubbing service. A video goes in, a target language is selected, and the output has the same speaker delivering the same content in that language. The voice is not a stock narrator: the system builds a clone of the original speaker and drives it with the translated script. Mouth movement is regenerated to match, frame by frame.
The hard part of automated dubbing is not the audio. It is keeping the picture credible once the words change — and that fails first on anything other than a centred, unobstructed face. The in-house Lip Sync 2.0 model targets exactly those awkward cases: profile shots, faces partly hidden, rapid close-ups. Output is supported up to 4K.
Accuracy is controlled through a brand glossary, which pins product names and technical terms to a defined spelling so they survive translation intact.
The service reads from more than 100 source languages and delivers into more than 40. Infrastructure sits in Germany, customer material is excluded from model training, and a DPA is issued on request; support is available in German. Users include BMW, RATIONAL, Axel Springer, HAVAS and Liebscher & Bracht. Plans begin at €69 per month on an annual plan, and a free trial covers one short video.
Learn more
Speechmatics
Best-in-Market Speech-to-Text & Voice AI for Enterprises.
Speechmatics delivers industry-leading Speech-to-Text and Voice AI for enterprises needing unrivaled accuracy, security, and flexibility. Our enterprise-grade APIs provide real-time and batch transcription with exceptional precision—across the widest range of languages, dialects, and accents.
Powered by Foundational Speech Technology, Speechmatics supports mission-critical voice applications in media, contact centers, finance, healthcare, and more. With on-prem, cloud, and hybrid deployment, businesses maintain full control over data security while unlocking voice insights.
Trusted by global leaders, Speechmatics is the top choice for best-in-class transcription and voice intelligence.
🔹 Unmatched Accuracy – Superior transcription across languages & accents
🔹 Flexible Deployment – Cloud, on-prem, and hybrid
🔹 Enterprise-Grade Security – Full data control
🔹 Real-Time & Batch Processing – Scalable transcription
🚀 Power your Speech-to-Text and Voice AI with Speechmatics today!
Learn more