Picovoice/leopard

On-device speech-to-text engine powered by deep learning

What it solves

Leopard は、デバイス上で直接、高精度な音声文字起こしを実行する方法を提供します。クラウドサーバーに音声データを送信する必要がありません。これにより、ユーザーのプライバシーを保護し、オフライン機能を実現します。

How it works

オーディオファイルまたはマイク入力をローカルで処理するデバイス内音声文字起こしエンジンです。モデルベースのアプローチを採用しており、ユーザーはデフォルトモデルまたはカスタムトレーニングされたモデルを使用できます。エンジンは認証と認可のために AccessKey が必要ですが、Picovoice license サーバー経 highlights 経由で検証されます。ただし、実際の音声認識処理は 100% オフラインで行われます。

Who it’s for

デスクトップ (Windows, macOS, Linux)、モバイル (iOS, Android)、Web (Chrome, Safari, Firefox, Edge)、および組み込みデバイス (Raspberry Pi) にわたる、プライバシーを保護し、効率的で正確な音声文字起こし機能が必要なクロスプラットフォーム・アプリケーションを構築する開発者向けです。

Highlights

  • Privacy-First: All voice processing runs locally on the device.

  • Broad Platform Support: Compatible with a wide range of operating systems, browsers, and hardware.

  • Computationally Efficient: Designed to be compact and efficient in its resource usage.

  • Multi-language Support: Supports English, French, German, Italian, Japanese, Korean, Portuguese, and Spanish.