Does the audio leave my computer?
Not when you use the on-device engine. Transcription runs locally; your recordings are not uploaded to transcribe them.
Speech to text
Import a recording or a video, and EdgeSpeak transcribes it on your computer while you review it against playback.
EdgeSpeak is a desktop app that transcribes audio and video locally with Lattice-2 Flash or Pro, which support more than 40 languages. Files stay on your machine.
The transcript follows playback. Click a line to jump to it, fix a word in place, and export when it reads right.
Get word-level timestamps, sentence segmentation and optional speaker labels. Leave the language empty and it is detected from the audio.
Export plain text, SRT, VTT or JSON. The same engine runs from the command line and the OpenAI-compatible local gateway.
edgespeak-cli transcribe meeting.m4a -o meeting.json
Drop in an MP4 or another video file and get the spoken text with timing, ready to become subtitles.
Not when you use the on-device engine. Transcription runs locally; your recordings are not uploaded to transcribe them.
Lattice-2 Flash and Pro support more than 40 languages. The language guide lists the codes and what each capability supports.
Yes. Word-level timestamps are available in the JSON export, the CLI and the local gateway.
Yes. Use edgespeak-cli, or send audio to the OpenAI-compatible local gateway on your computer.