Average Ratings 1 Rating

Total
ease
features
design
support

Average Ratings 0 Ratings

Total
ease
features
design
support

No User Reviews. Be the first to provide a review:

Write a Review

Description

Cleanvoice recognizes and modifies stutters to ensure a natural auditory experience. It pinpoints instances of stuttering and refines them for a smoother conversation flow. Eliminating filler words can pose challenges, as their mere removal may render the audio less authentic. Our AI comprehensively analyzes the audio's context, incorporating ambient sounds to enhance the podcast's overall fluidity. Additionally, Cleanvoice can edit filler words across multiple tracks while maintaining perfect synchronization. If you have speakers recorded on separate tracks, we also address mouth noises in every track, ensuring that all your files remain perfectly aligned. This way, you can achieve a polished and professional sound in your recordings.

Description

MAI-Transcribe-2 represents the pinnacle of Microsoft AI's transcription capabilities, engineered to provide rapid and precise speech recognition across various real-world audio scenarios. This model includes features like speaker diarization, enabling it to differentiate between speakers and correctly attribute dialogue, as well as offering word-level timestamps for enhanced alignment, searching, navigation, and editing purposes. Additionally, it utilizes keyword biasing to improve the recognition of specialized terms, abbreviations, and names that may otherwise be challenging to identify from their contextual usage. Developers are afforded the flexibility to select from different transcription styles: a verbatim option that retains filler words and false starts for thorough analysis and compliance, or a clean option that eliminates such elements for clearer captions and more polished published transcripts. Furthermore, the model is adept at handling code-switching, seamlessly transitioning between languages during conversations, even accommodating mixed language combinations like Hinglish and Spanglish, while automatically identifying the language being spoken. This makes MAI-Transcribe-2 an invaluable tool for diverse linguistic environments and applications.

API Access

Has API Yes 

API Access

Has API No 

Screenshots View All

Screenshots View All

Integrations

Adobe Audition Yes 
Adobe Premiere Pro Yes 
Audacity Yes 
DaVinci Resolve Yes 
Microsoft Azure No 
Microsoft Foundry No 
Reaper Yes 

Integrations

Adobe Audition No 
Adobe Premiere Pro No 
Audacity No 
DaVinci Resolve No 
Microsoft Azure Yes 
Microsoft Foundry Yes 
Reaper No 

Pricing Details

€1 per hour
Free Trial No 
Free Version No 

Pricing Details

No price information available.
Free Trial No 
Free Version No 

Deployment

Web-Based Yes 
On-Premises No 
iPhone App No 
iPad App No 
Android App No 
Windows No 
Mac No 
Linux No 
Chromebook No 

Deployment

Web-Based Yes 
On-Premises No 
iPhone App No 
iPad App No 
Android App No 
Windows No 
Mac No 
Linux No 
Chromebook No 

Customer Support

Business Hours No 
Live Rep (24/7) No 
Online Support Yes 

Customer Support

Business Hours No 
Live Rep (24/7) No 
Online Support Yes 

Types of Training

Training Docs Yes 
Webinars No 
Live Training (Online) No 
In Person No 

Types of Training

Training Docs Yes 
Webinars No 
Live Training (Online) No 
In Person No 

Vendor Details

Company Name

Cleanvoice

Country

Romania

Website

cleanvoice.ai/

Vendor Details

Company Name

Microsoft AI

Founded

2024

Country

United States

Website

microsoft.ai/news/mai-transcribe-2-is-the-fastest-most-accurate-and-cheapest-speech-recognition-model-in-the-world/

Product Features

Audio Editing

Audio Effects No 
Batch Processing No 
Export Audio (Multiple File Types) No 
Record Live Audio No 
Record Multiple Simultaneous Tracks No 
Scrub, Search, Bookmark No 
Sound Editing Tools No 
Spectral Analysis / FFT No 
Speech Synthesis (TTS) No 
Swappable Patches No 
Virtual Instruments No 
Virtual Mixing No 
Voice Changer No 

Product Features

Transcription

AI / Machine Learning No 
Annotations No 
Audio/Video File Upload No 
Automatic Transcription No 
Collaboration Tools No 
File Sharing No 
For Manual Transcription No 
Full Text Search No 
Multi-Language Support No 
Natural Language Processing (NLP) No 
Playback Controls No 
Speech Recognition No 
Subtitles No 
Text Editor No 
Timecoding No 

Alternatives

Alternatives