Picovoice/leopard

On-device speech-to-text engine powered by deep learning

What it solves

Leopard 提供了一种直接在设备上进行高准确度语音转文字转录的方法,无需将语音数据发送到云端服务器。这确保了用户隐私并允许离线功能。

How it works

它是一款在设备端处理音频文件或麦克风输入语音转文字引擎。它采用基于模型的做法,用户可以使用默认模型或自定义训练的模型。该引擎需要一个 AccessKey 用于身份验证和授权,该金钥通过 Picovoice license 伺服器进行验证,但实际的语音识别处理过程仍然是 100% 离线。

Who it’s for

开发者正在构建需要跨平台(包括桌面端(Windows, macOS, Linux)、移动端(iOS, Android)、Web 端(Chrome, Safari, Firefox, Edge)以及嵌入式设备(Raspberry Pi))提供私密、高效且准确的语音转文字功能的跨平台应用程序。

Highlights

  • Privacy-First: All voice processing runs locally on the device.
  • Broad Platform Support: Compatible with a wide range of operating systems, browsers, and hardware.
  • Computationally Efficient: Designed to be compact and efficient in its resource usage.
  • Multi-language Support: Supports English, French, German, Italian, Japanese, Korean, Portuguese, and Spanish.