Type a script, click Generate and EditLint AI produces a WAV file using your system's built-in text-to-speech (Windows SAPI or macOS say) and places it on your audio track. Zero cloud, zero API cost.
EditLint AI's voiceover generator uses the text-to-speech engine built into your operating system. On Windows this is Microsoft Speech API (SAPI) with the installed voice pack. On macOS it's the say command using whichever voice you have configured in System Preferences.
The generated WAV file is saved to a temporary folder and immediately imported into your NLE project. The clip is then placed on a dedicated voiceover audio track at the playhead position or at the start of the sequence.
say command with selected voiceSystem TTS voices are not broadcast quality — they lack the warmth, inflection and naturalness of a human VO artist or a premium neural TTS service. But that's precisely the point for several important workflows where speed and zero cost outweigh production quality.
System TTS has clear limitations that make it unsuitable for client-facing deliverables requiring broadcast quality or emotive performance. Understanding when not to use this feature saves time in the long run.
For high-quality neural TTS without sending audio to a cloud service, consider integrating a local neural voice model such as Piper TTS, which EditLint AI supports as an optional local engine upgrade.
Open the Voiceover panel in EditLint AI and type or paste your script. Select a voice from the list of installed system voices, set the speaking speed and choose where the generated clip should be placed — at the current playhead or at the start of the sequence.
Click Generate. EditLint AI passes the script to the system TTS engine (SAPI on Windows, say on macOS) and captures the output as a 48kHz WAV file. For a 60-second script, generation typically takes 2–5 seconds depending on hardware.
The WAV is saved to your project's audio folder, imported into your NLE project bin and placed on a dedicated voiceover track at the configured position. You can then trim, adjust levels and eventually replace it with the real VO when it's recorded.
No cloud. No API key. No waiting. Type your script and it's on the timeline — locally generated using your system's built-in voice engine.