Device-side speech engine

अपने कंप्यूटर में speech AI बड़ा model रखें

EdgeSpeak professional speech models को desktop के लिए compress करता है और meetings, interviews, videos व recordings को locally transcribe करता है. Audio, video और transcripts आपके device पर रहते हैं.

  • Local audio/video transcription available
  • Local gateway और CLI available
  • Future model और speech capability updates
Transcript preview
Transcript previewlocal-transcribe.mov

डेवलपर · CLI & Skills

एक कमांड से CLI इंस्टॉल करें

एक लाइन का जाना-पहचाना इंस्टॉल, बिल्कुल आपके डेवलपर टूल्स की तरह। CLI आत्मनिर्भर है, डेस्कटॉप ऐप की ज़रूरत नहीं, और सीधे टर्मिनल से ट्रांसक्राइब करता है।

macOS या Linux पर Terminal से और Windows x64 पर PowerShell से इंस्टॉल करें। CPU डिफ़ॉल्ट रूप से चलता है; Linux NVIDIA CUDA को अपने आप पहचानता है, जबकि Windows पर CUDA को अलग से चुनना होता है।

macOS · Apple SiliconWindows x64 · CPU / CUDALinux x86_64 · CPU / CUDA
CLI दस्तावेज़ पढ़ें

macOS / Linux · Terminal

curl -fsSL https://edgespeak.com/install.sh | sh

Windows · PowerShell

irm https://edgespeak.com/install.ps1 | iex
edgespeak-cli login
edgespeak-cli transcribe media.mp4 -o media.srt
edgespeak-cli generate "Summarize this transcript" --model Qwen/Qwen3.5-4B

Agent Skills, वही इंजन

EdgeSpeak Skills, Claude Code और Cursor जैसे Skills-समर्थित एजेंटों को CLI के ज़रिए डिवाइस पर ट्रांसक्राइब करना सिखाती हैं। GitHub पर ओपन सोर्स।

npx skills add lattifai/EdgeSpeak

Fast on-device transcription

Audio या video import करें और local high-accuracy transcription पाएं

Meeting, interview या video डालें. EdgeSpeak इसी computer पर speech understanding, transcript generation और text alignment करता है; downstream tools को आगे काम करना हो, तो वे local gateway से transcript ले लेते हैं.

  1. आपके device पर EdgeSpeak

    EdgeSpeak आपके सामने वाले computer पर audio और video को text में बदलता है.

  2. सुनते हुए ठीक करें

    Transcript playback के साथ चलता है, इसलिए review और correction एक ही जगह होते हैं.

  3. Result आगे बढ़ता है

    Text या subtitles export करें, या finished transcript को दूसरे tool में आगे इस्तेमाल करें.

Dedicated on-device model

Lattice-2, desktop के लिए बनाया गया

Lattice-2 को on-device inference के लिए compress और optimize किया गया है, ताकि सामान्य computer meetings, interviews, videos और recordings efficiently process कर सके.

On-device compression और inference adaptation

Speech AI model को desktop execution के अनुकूल बनाया गया है, wait time घटता है, और local transcription external services पर depend नहीं करती.

Flash / Pro model pairing

Flash daily meetings और video transcription के लिए faster response देता है. Pro harder audio, complex accents और higher accuracy ceiling के लिए है, और ज़्यादा local resources इस्तेमाल करता है.

Developer workflows में integrated

Lattice-2 40+ languages, multiple English accents और Chinese dialects support करता है; वही local engine gateway से CLI, agents और automation flows को मिल सकता है.

LOCAL LM + VLM · 0.6B — 27B

उसी EdgeSpeak रनटाइम में स्थानीय भाषा और विज़न।

EdgeSpeak अब 0.6B से 27B तक Qwen और Gemma मॉडल प्रदान करता है। एक समय में एक मॉडल लोड होता है और निष्क्रिय होने पर सो जाता है।

स्थानीय मॉडल से जनरेट करें
Qwen/Qwen3-0.6B0.6B0.4 GB
Qwen/Qwen3.5-0.8B0.8B0.5 GB
Qwen/Qwen3.5-2B2B1.3 GB
google/gemma-4-E2B-it2B3.1 GB
Qwen/Qwen3.5-4Bडिफ़ॉल्ट4B2.7 GB
google/gemma-4-E4B-it4B5.0 GB
Qwen/Qwen3.5-9B9B5.7 GB
google/gemma-4-12B-it12B7.1 GB
Qwen/Qwen3.6-27B27B16.8 GB

टेक्स्ट + इमेज · OpenAI-संगत · लोकल-फर्स्ट

Real desktop app

EdgeSpeak सीधे आपके computer पर चलता है

Audio या video import करें, local model चुनें और result export करें. Automation के लिए local gateway से काम CLI या agents को दें.

Audio, video और transcript एक workspace मेंPlayback, timeline, transcript, export, recent files और active EdgeSpeak model साथ रहते हैं, ताकि tools कम बदलें और context बना रहे.
EdgeSpeak desktop Transcribe screen जिसमें workspace, local EdgeSpeak model status, transcript, playback controls और recent files दिखते हैं.
Task के हिसाब से EdgeSpeak model चुनेंEdgeSpeak 40+ languages, multiple English accents और Chinese dialects support करता है. Flash fast और strong है; Pro ज़्यादा accurate है और ज़्यादा local resources इस्तेमाल करता है.
EdgeSpeak desktop Models screen जिसमें local EdgeSpeak Flash और EdgeSpeak Pro options हैं.
दूसरे tools को भी EdgeSpeak इस्तेमाल करने देंCLI tools, agents और automation audio को EdgeSpeak में भेज सकते हैं और transcript EdgeSpeak से वापस ले सकते हैं.
EdgeSpeak desktop Gateway screen जिसमें same computer पर tools local speech engine इस्तेमाल करते दिखते हैं.

गोपनीयता

Local-first speech processing.

EdgeSpeak is designed so imported media and generated transcripts stay on your device unless you export, upload, or share them.

On-device compression और inference adaptation

Speech AI model को desktop execution के अनुकूल बनाया गया है, wait time घटता है, और local transcription external services पर depend नहीं करती.

Result आगे बढ़ता है

Text या subtitles export करें, या finished transcript को दूसरे tool में आगे इस्तेमाल करें.

Local speech gateway

Agents को EdgeSpeak सीधे call करने दें

EdgeSpeak local, OpenAI-compatible speech API देता है, जिससे CLI, agents और automation उसी computer पर transcribe कर सकें. यह कोई दूसरा cloud नहीं, आपके computer का speech gateway है.

Local OpenAI-compatible endpointCLI / agents / automationEdgeSpeak Flash और Pro
POST /v1/audio/transcriptionslocalhost:1117
curl http://127.0.0.1:1117/v1/audio/transcriptions \
  -H "Authorization: Bearer sk-edgespeak-..." \
  -F file=@meeting.m4a \
  -F model="lattice-2-flash"

Choose by workflow

Local software, cloud APIs, or a model you maintain

The right option depends on where media can go, how much workflow you want ready-made, and who should maintain the speech stack.

Choose by workflowEdgeSpeakCloud transcriptionSelf-managed local model
Source mediaStays on this deviceUploaded to a remote serviceStays on the machine you configure
Review workflowDesktop playback, transcript, and exportDepends on the providerYou build the review surface
Tools and agentsLocal CLI and OpenAI-compatible gatewayHosted APIYou build and maintain the integration
OperationsInstall the app and local modelsManage an account, API keys, and network accessMaintain the runtime, models, and dependencies
Read the full comparison

Early bird

$49 में EdgeSpeak lifetime

Current release में local audio-video transcription, local gateway और CLI शामिल हैं. एक बार खरीदें और future model व speech capability updates पाते रहें.

डाउनलोड

macOS पर EdgeSpeak install करें.

Choose the current macOS build or open the Windows Beta page for the x64 installer and verified setup instructions.

Current macOS build

Use the available desktop build today. Your account keeps purchase, license, and device management in one place.

Windows build

Download the Windows x64 Beta, verify its SHA-256, and follow the installation guide. This Beta uses manual updates.

License and devices

Sign in with the email used for purchase, then manage your license key and activated devices from your account.

FAQ

Frequently asked questions

On-device processing, supported platforms, models, and the lifetime license — answered before you buy.

See all questions
01Does EdgeSpeak upload source media?

Imported media and generated transcripts stay on your device unless you export, upload, or share them yourself.

02Which desktop platforms are available?

EdgeSpeak is available for Apple Silicon Macs with macOS 14.0 or later. A Windows 10/11 x64 Beta is also available for manual installation; Windows in-app updates are not enabled yet.

03What does the lifetime license include?

The early-bird lifetime license is a one-time purchase for permanent use, future model and speech-capability updates, and up to four activated devices.

04What is the difference between Flash and Pro?

Flash is tuned for faster everyday transcription. Pro raises the accuracy ceiling for harder audio and uses more local resources.

05Can CLI tools and AI agents use EdgeSpeak?

Yes. The bundled CLI and local OpenAI-compatible gateway let trusted tools on the same computer call the local speech engine.

Community feedback

Real workflows से EdgeSpeak बेहतर बनाएं

Discord पर real workflows, agent integration needs और product ideas share करें.

  • Real workflows
  • Agent / API
  • Product ideas

आपका feedback सीधे product में जाता है.

Device-side speech engine

अपने कंप्यूटर में speech AI बड़ा model रखें

Transcribe locally, review against an accurate timeline, then export or continue through the local gateway.