EdgeSpeak บนอุปกรณ์ของคุณ
EdgeSpeak เปลี่ยน audio/video เป็นข้อความบนคอมพิวเตอร์ตรงหน้าคุณ
เอนจินเสียงบนอุปกรณ์
EdgeSpeak บีบอัดโมเดลเสียงระดับมืออาชีพสำหรับเดสก์ท็อป และถอดเสียงการประชุม สัมภาษณ์ วิดีโอ และไฟล์บันทึกในเครื่อง ไฟล์เสียง วิดีโอ และข้อความยังอยู่บนอุปกรณ์ของคุณ

นักพัฒนา · CLI & Skills
ติดตั้งด้วยบรรทัดเดียวแบบที่คุ้นเคย เหมือนเครื่องมือนักพัฒนาที่คุณใช้อยู่แล้ว CLI ทำงานได้ในตัวเอง ไม่ต้องใช้แอปเดสก์ท็อป และถอดเสียงได้จากเทอร์มินัลโดยตรง
ติดตั้งผ่าน Terminal บน macOS หรือ Linux และผ่าน PowerShell บน Windows x64 โดย CPU ใช้งานได้ทันที Linux จะตรวจพบ NVIDIA CUDA อัตโนมัติ ส่วน Windows ให้เลือก CUDA เมื่อต้องการ
macOS / Linux · Terminal
curl -fsSL https://edgespeak.com/install.sh | sh
Windows · PowerShell
irm https://edgespeak.com/install.ps1 | iex
EdgeSpeak Skills สอนให้ Claude Code, Cursor และเอเจนต์ที่รองรับ Skills ถอดเสียงบนอุปกรณ์ผ่าน CLI โอเพนซอร์สบน GitHub
npx skills add lattifai/EdgeSpeak
ถอดเสียงเร็วบนอุปกรณ์
ใส่การประชุม สัมภาษณ์ หรือวิดีโอเข้ามา EdgeSpeak จะทำความเข้าใจเสียง สร้างข้อความถอดเสียง และจัดแนวข้อความบนคอมพิวเตอร์นี้ เมื่อเครื่องมือต่อไปต้องทำงานต่อ ก็รับข้อความจาก gateway ในเครื่องได้
EdgeSpeak เปลี่ยน audio/video เป็นข้อความบนคอมพิวเตอร์ตรงหน้าคุณ
ข้อความตามการเล่นเสียง ทำให้ตรวจและแก้ได้ในพื้นที่เดียว
ส่งออกข้อความหรือซับไตเติล หรือให้เครื่องมืออื่นใช้ transcript ที่เสร็จแล้วต่อ
โมเดลเฉพาะบนอุปกรณ์
Lattice-2 ผ่านการบีบอัดและปรับ inference บนอุปกรณ์ ให้คอมพิวเตอร์ทั่วไปประมวลผลการประชุม สัมภาษณ์ วิดีโอ และไฟล์บันทึกได้อย่างมีประสิทธิภาพ
โมเดล AI เสียงถูกปรับให้เหมาะกับการทำงานบนเดสก์ท็อป ลดเวลารอ และการถอดเสียงในเครื่องไม่ต้องพึ่งบริการภายนอก
Flash ตอบสนองเร็วสำหรับการประชุมและวิดีโอทั่วไป ส่วน Pro เหมาะกับเสียงที่ยากกว่า สำเนียงซับซ้อนกว่า และต้องการความแม่นยำสูงกว่า โดยใช้ทรัพยากรในเครื่องมากขึ้น
Lattice-2 รองรับมากกว่า 40 ภาษา หลายสำเนียงอังกฤษ และภาษาถิ่นจีน เอนจินในเครื่องเดียวกันส่งต่อให้ CLI, agent และ automation ผ่าน gateway ได้
LOCAL LM + VLM · 0.6B — 27B
EdgeSpeak มีโมเดล Qwen และ Gemma ตั้งแต่ 0.6B ถึง 27B โหลดครั้งละหนึ่งโมเดลและพักอัตโนมัติเมื่อไม่ได้ใช้งาน
สร้างด้วยโมเดลภายในเครื่องQwen/Qwen3-0.6B0.6B0.4 GBQwen/Qwen3.5-0.8B0.8B0.5 GBQwen/Qwen3.5-2B2B1.3 GBgoogle/gemma-4-E2B-it2B3.1 GBQwen/Qwen3.5-4Bค่าเริ่มต้น4B2.7 GBgoogle/gemma-4-E4B-it4B5.0 GBQwen/Qwen3.5-9B9B5.7 GBgoogle/gemma-4-12B-it12B7.1 GBQwen/Qwen3.6-27B27B16.8 GBข้อความ + รูปภาพ · ใช้ร่วมกับ OpenAI · ภายในเครื่องก่อน
แอปเดสก์ท็อปจริง
นำเข้าเสียงหรือวิดีโอ เลือกโมเดลในเครื่อง แล้วส่งออกผลลัพธ์ เมื่อต้องการ automation ให้ส่งงานต่อไปยัง CLI หรือ agent ผ่าน gateway ในเครื่อง



Workflows
สำรวจไฟล์ส่วนตัว การถอดเสียงและวิดีโอในเครื่อง agent automation และตัวเลือกในเครื่องเทียบกับ cloud
Keep source media on-device and generate the transcript locally.
Review with playback, then export text or subtitles.
Install the EdgeSpeak Skill and let agents call the local CLI.
Choose by media location, processing scale, and integration path.
ความเป็นส่วนตัว
EdgeSpeak is designed so imported media and generated transcripts stay on your device unless you export, upload, or share them.
โมเดล AI เสียงถูกปรับให้เหมาะกับการทำงานบนเดสก์ท็อป ลดเวลารอ และการถอดเสียงในเครื่องไม่ต้องพึ่งบริการภายนอก
ส่งออกข้อความหรือซับไตเติล หรือให้เครื่องมืออื่นใช้ transcript ที่เสร็จแล้วต่อ
Speech gateway ในเครื่อง
EdgeSpeak มี API เสียงในเครื่องที่เข้ากันได้กับ OpenAI ให้ CLI, agent และ automation ถอดเสียงบนคอมพิวเตอร์เครื่องเดียวกัน ไม่ใช่ cloud อีกแห่ง แต่เป็น gateway เสียงของคอมพิวเตอร์คุณ
POST /v1/audio/transcriptionslocalhost:1117curl http://127.0.0.1:1117/v1/audio/transcriptions \ -H "Authorization: Bearer sk-edgespeak-..." \ -F file=@meeting.m4a \ -F model="lattice-2-flash"
Choose by workflow
The right option depends on where media can go, how much workflow you want ready-made, and who should maintain the speech stack.
| Choose by workflow | EdgeSpeak | Cloud transcription | Self-managed local model |
|---|---|---|---|
| Source media | Stays on this device | Uploaded to a remote service | Stays on the machine you configure |
| Review workflow | Desktop playback, transcript, and export | Depends on the provider | You build the review surface |
| Tools and agents | Local CLI and OpenAI-compatible gateway | Hosted API | You build and maintain the integration |
| Operations | Install the app and local models | Manage an account, API keys, and network access | Maintain the runtime, models, and dependencies |
Early bird
เวอร์ชันปัจจุบันมีการถอดเสียงและวิดีโอในเครื่อง gateway ในเครื่อง และ CLI ซื้อครั้งเดียวและรับอัปเดตโมเดลกับความสามารถด้านเสียงในอนาคต
ตอนนี้ $49 ราคาปกติ $99 สิทธิ์ใช้งานตลอดชีพ ซื้อครั้งเดียวในช่วง early bird
For more devices or team purchases: sales@edgespeak.com
ดาวน์โหลด
Choose the current macOS build or open the Windows Beta page for the x64 installer and verified setup instructions.
Use the available desktop build today. Your account keeps purchase, license, and device management in one place.
Download the Windows x64 Beta, verify its SHA-256, and follow the installation guide. This Beta uses manual updates.
Sign in with the email used for purchase, then manage your license key and activated devices from your account.
FAQ
On-device processing, supported platforms, models, and the lifetime license — answered before you buy.
See all questionsImported media and generated transcripts stay on your device unless you export, upload, or share them yourself.
EdgeSpeak is available for Apple Silicon Macs with macOS 14.0 or later. A Windows 10/11 x64 Beta is also available for manual installation; Windows in-app updates are not enabled yet.
The early-bird lifetime license is a one-time purchase for permanent use, future model and speech-capability updates, and up to four activated devices.
Flash is tuned for faster everyday transcription. Pro raises the accuracy ceiling for harder audio and uses more local resources.
Yes. The bundled CLI and local OpenAI-compatible gateway let trusted tools on the same computer call the local speech engine.
เสียงจากชุมชน
แชร์ workflow จริง ความต้องการเชื่อมต่อ agent และไอเดียผลิตภัณฑ์บน Discord
ความคิดเห็นของคุณเข้าสู่การพัฒนาผลิตภัณฑ์โดยตรง
เอนจินเสียงบนอุปกรณ์
Transcribe locally, review against an accurate timeline, then export or continue through the local gateway.