ken107/read-aloud
An awesome browser extension that reads aloud webpage content with one click
What it solves
Read Aloud is a browser extension that converts webpage text into spoken audio. It provides an alternative way to consume web content for users who prefer listening over reading, people with dyslexia or other learning disabilities, and children learning to read.
How it works
The extension integrates into Chrome and Firefox browsers, allowing users to trigger text-to-speech via an extension button or a right-click menu. It utilizes a variety of text-to-speech (TTS) engines, including native browser voices and cloud-based services like Google Wavenet, Amazon Polly, IBM Watson, and Microsoft.
Who it’s for
- Users who want to multitask or give their eyes a rest.
- Individuals with dyslexia or other learning disabilities.
- Children learning to read.
- Anyone who prefers audio consumption of web content.
Highlights
- Support for multiple browsers (Chrome and Firefox).
- Integration with premium cloud TTS providers (Google Wavenet, Amazon Polly, IBM Watson, Microsoft).
- Customizable settings for voice, reading speed, and pitch.
- Text highlighting feature to track reading progress.
- Keyboard shortcuts for playback control (Play/Pause, Stop, Rewind, Forward).
Related
- Project
mengxi-ream/read-frogRead Frog is an open‑source browser extension that uses LLMs and AI TTS to translate, explain, and turn web content into flashcards, offering bilingual view, context‑aware translation, subtitle translation, custom AI actions and support for 20+ AI providers.
- Project
ilyhalight/voice-over-translationA browser extension that adds real-time voice-over translation and AI subtitles to videos across various websites, allowing users to watch foreign content in their native language.
- Project
Migushthe2nd/MsEdgeTTSA simple Azure Speech Service module that uses the Microsoft Edge Read Aloud API to provide text-to-speech synthesis for server-side runtimes.
- Project
rsxdalv/TTS-WebUIA unified web interface and manager for a wide variety of open-source text-to-speech, audio generation, and audio conversion AI models.
- Project
met4citizen/TalkingHeadTalkingHead is an open‑source JavaScript library for browser‑based 3‑D avatars that can lip‑sync, show facial expressions, and be driven by AI services (Google TTS, OpenAI, Whisper, etc.). It works with Mixamo‑rigged GLB models, supports multiple languages, and offers plug‑in modules for TTS, audio‑driven visemes, and advanced gestures. The README provides demos, integration steps, and a large set of configurable options, making it suitable for research, product demos, or hobby projects involving embodied AI.