| audiocpp-android (CaptainArni) |
and |
Android companion to audio.cpp Studio: photograph a book page, OCR it, and listen with follow-along highlighting, background playback and cloned voice |
local-only |
active |
0 |
MIT |
V |
| CopiloTTS |
and/ios |
Kotlin Multiplatform TTS SDK using either native OS TTS or ONNX Runtime. |
hybrid |
? |
29 |
MIT |
M |
| ElevenReader |
and/legacy |
ElevenLabs' consumer reading app: converts articles, PDFs and ePubs to streaming audio in 32+ languages with synced text highlighting. |
cloud-API |
? |
|
Proprietary, f |
L |
| Glyfen (F-Droid) |
and |
"Offline OCR for Android. Search your photos by text, entirely on-device" [vendor]. |
local-only |
active |
|
unknown |
V |
| Imgspeaker |
and |
Android app that performs OCR on Simplified Chinese images and reads the extracted text aloud. |
unknown |
? |
5 |
see repo |
L |
| nabu |
and |
Android on-device test bench for TTS and local LLM chat — Kokoro-82M, Supertonic v2/v3 and Soprano 1.1 through ONNX Runtime, plus a book workflow that |
local-only |
active |
56 |
GPL-3.0 |
M |
| nabu-realtime (richdrummer33 fork of nabu) |
and |
Fork of the Android multi-engine TTS/LLM app nabu, adding chunked playback and background processing. |
local-only |
dormant |
0 |
GPL-3.0 |
M |
| News Reader (livio) (F-Droid) |
and |
"Minimalistic yet powerful news reader with text to speech capability" [vendor]. |
local-only |
active |
|
unknown |
V |
| OCR (Subhamtyagi) (F-Droid) |
and |
"OCR based on Tesseract 5" [vendor]. |
local-only |
active |
|
unknown |
V |
| Offline Translator (F-Droid) |
and |
"On-device translation of text, images and pdf/odt files, with TTS" [vendor]. |
local-only |
active |
|
unknown |
V |
| pdf-reader-aloud |
and |
Android PDF reader (v0.2) that extracts text with PdfBox-Android, chunks it by sentence and speaks it with a bundled Piper en_US-amy-medium voice thro |
local-only |
dormant |
0 |
none declared |
M |
| Poet Assistant (F-Droid) |
and |
"Dictionary and TTS tools for editing poems" [vendor]. |
local-only |
active |
|
unknown |
V |
| Shravan |
and |
Open-source Android app for visually impaired users providing offline real-time text reading (OCR) and object detection with voice output. |
local-only |
? |
6 |
open source (s |
L |
| speakword.koplugin |
and/lin |
KOReader plugin that speaks selected words or sentences on an e-reader via Android system TTS or ElevenLabs. |
hybrid |
? |
0 |
AGPL-3.0 |
L |
| TextLector (nedmah) |
and/ios |
Free offline TTS reader for Android and iOS built with Kotlin Multiplatform, positioned in its README as a Speechify alternative. |
local-only |
active |
7 |
Apache-2.0 |
L |
| tts-tool (felixvonberlin) |
and |
'A simple GUI for android's Text-To-Speech-Engines', published on Google Play as de.favo.ttst. |
os-native |
unknown |
2 |
unknown |
L |
| Udderance AAC (F-Droid) |
and |
"TTS based AAC" [vendor] — augmentative and alternative communication board. |
local-only |
active |
|
unknown |
V |
| Voice Notify (F-Droid) |
and |
"Spoken notifications" [vendor] — speaks incoming Android notifications aloud. |
local-only |
active |
|
unknown |
V |
| Wyoming Android TTS (F-Droid) |
and |
"Use your Android's TTS engines in Home Assistant via the Wyoming protocol" [vendor]. |
local-only |
active |
|
unknown |
V |
| dotool |
X11/Wl/legacy |
uinput-based input simulator that reads commands on stdin and supports keyboard layouts; named alongside ydotool and kwtype as a KDE-Wayland paste met |
local-only |
active |
|
unknown |
L |
| screenshot-osr-linux |
X11 |
Russian-language shell script positioned as an alternative to ABBYY Screenshot Reader: bind a key (e.g. Ctrl+PrtSc), it takes a gnome-screenshot, clea |
local-only |
dormant |
0 |
unknown |
M |
| ttsclip (connbrack) |
X11/Wl |
Bash CLI for Linux that speaks an argument or the clipboard through Amazon Polly, with flags for tempo and for replacing line breaks; the README's own |
cloud-API |
dormant |
1 |
none declared |
M |
| xmag |
X11 |
'Utility to display a magnified snapshot of a portion of an X11 screen' — the X.Org original, hosted only on freedesktop GitLab. |
local-only |
unknown |
0 |
MIT/X11 (X.Org |
V |
| [xterm OSC 52 selection read (printf '\033]52;s;?\007')](https://unix.stackexchange.com/questions/318673/access-highlighted-text-from-script) |
X11 |
Second answer to 'Access highlighted text from script?': normally you cannot, but xterm's Manipulate Selection Data control sequence will reply with t |
local-only |
active |
|
CC-BY-SA (answ |
M |
| ydotool |
X11/Wl/legacy |
'Generic command-line automation tool' that injects input through /dev/uinput, so it works under Wayland where xdotool cannot; requires the ydotoold d |
local-only |
active |
2321 |
AGPL-3.0 |
M |
| a11y-profile-manager |
lin |
GPL-3 tool that 'facilitates the easy enablement of specific settings to improve the accessibility of a desktop environment' — the Ubuntu accessibilit |
local-only |
dormant |
|
GPL-3 (Launchp |
V |
| AbilityMic |
lin |
AAC (augmentative and alternative communication) app with communication boards, OBF/OBZ import-export, TTS synthesis and on-device word prediction. |
hybrid |
commercial-l |
|
Proprietary |
V |
| Accerciser |
lin |
GNOME's interactive AT-SPI tree explorer - the tool you use to see whether an app actually exposes its text and selection to the accessibility bus. |
local-only |
? |
|
BSD-3-Clause |
L |
| access-irc |
lin |
"Accessible IRC client with GTK3 and screen reader support." [vendor, AUR description] |
local-only |
unknown |
|
unknown |
L |
| accessibility-inspector (KDE) |
lin |
KDE tool that inspects an application's accessibility tree - the KDE counterpart to Accerciser / aviewer. |
local-only |
active |
8 |
GPL |
M |
| Accessible Linux distributions (Vinux, Accessible Coconut, Slint |
lin |
Linux distributions preconfigured so Orca, speech-dispatcher, BRLTTY and a working voice are running from first boot, including during installation. |
local-only |
? |
|
varies (GPL-fa |
L |
| ADRIANE (Knoppix Audio Desktop Reference Implementation And Netw |
lin |
Talking Knoppix desktop environment named by a partially-sighted r/linux commenter as their long-time choice ('I have used ADRIENE, as my go-to-distro |
local-only |
active |
|
GPL (Knoppix) |
L |
| Alyssum (icosane) |
lin/win |
'Translate text, speech, books, and documents fully offline with OCR and Whisper.' |
local-only |
unknown |
0 |
unknown |
L |
| AT-SPI a11y-manager + Mutter/KWin accessibility keyboard API |
lin |
The plumbing that restored global screen-reader hotkeys on Wayland: the compositor hands keyboard events only to a privileged accessibility client, ch |
local-only |
? |
|
LGPL-2.1 (AT-S |
M |
| at-spi-dbus |
lin |
'The AT-SPI D-Bus project aims to define D-Bus interfaces used to provide accessibility information to assistive technologies' — the Launchpad-era des |
local-only |
archived |
|
unknown |
V |
| AT-SPI2 (at-spi2-core) |
lin/legacy |
The Linux accessibility bus itself: the D-Bus protocol and registry through which toolkits expose their widget trees and screen readers read them. |
local-only |
? |
30 |
LGPL-2.1 |
M |
| atk (stefan11111) |
lin |
'Minimal implementation of Gnome Accessibility Toolkit'. |
local-only |
unknown |
0 |
unknown |
L |
| atspi (Rust crate, odilia-app) |
lin |
A pure-Rust implementation of the AT-SPI protocol, written for Odilia but usable by any Rust program that wants to read the Linux accessibility tree. |
local-only |
active |
59 |
Apache-2.0 |
M |
| atspi2-rs (wizzwizz4) |
lin |
Rust AT-SPI2 bindings whose own description warns 'Don't use this yet.' |
local-only |
unknown |
0 |
unknown |
M |
| audiobook-converter (Launchpad) |
lin |
'An easy usable application which converts from text formats to mp3 format audio books.' |
unknown |
unknown |
|
unknown |
L |
| Bifrost (Stormux) |
lin |
'A vibe coded, fully screen reader accessible fediverse client.' |
unknown |
unknown |
0 |
unknown |
L |
| Bookstorm |
lin |
'Accessible book reader' from the Stormux Gitea instance. |
unknown |
unknown |
0 |
unknown |
L |
| Cerva |
lin |
Python/GtkBuilder front-end to Festival inspired by the program 'Fala', with Unicode support and save-speech-to-file. |
local-only |
dormant |
|
GPL-3 (Launchp |
V |
| cicero |
lin |
Debian accessibility package at version 0.7.2-4, maintained under the a11y-team group on Salsa; a small self-voicing text tool for blind users. |
local-only |
dormant |
|
unknown |
L |
| Clipboard (Slackadays/Clipboard) |
lin/mac/win/legacy |
A cross-platform clipboard manager operated entirely from the command line, described by its author as a smart clipboard manager [vendor]. |
local-only |
active |
5 |
GPL-3.0 |
L |
| clipboard-watcher (dethos) |
lin |
Tool that reports which applications are reading the clipboard - relevant as a way to observe whether a selection reader is polling. |
local-only |
dormant |
50 |
MIT |
M |
| clipocr (jonnyprouty) |
lin |
'A simple shell script for performing ocr on image displayed on a screen, but not as selectable text.' |
local-only |
unknown |
0 |
unknown |
L |
| crengine-ng / CoolReader-NG |
lin/win/and |
'crengine-ng is cross-platform library designed to implement text viewers and e-book readers' — the maintained continuation of the CoolReader engine, |
local-only |
unknown |
13 |
unknown |
V |
| DAISY Player |
lin |
'Daisy Player for Blinds' — ncurses TUI player for DAISY 2 and 3 talking-book files, on SourceForge and packaged by Debian's a11y-team. |
local-only |
unknown |
|
unknown (no SF |
V |
| Debian a11y-team roster (salsa.debian.org/a11y-team) |
lin |
A single Salsa group holding packaging for ~70 Linux accessibility components, including several with no GitHub presence: cicero, eflite, emacspeak-ss |
local-only |
active |
|
per-package |
M |
Desktop custom-shortcut pipe pitfall - wrap in bash -c '...' |
lin |
Accepted answer: GNOME/KDE custom-shortcut commands are not run through a shell, so a pipeline like xsel --clipboard | festival --tts silently does |
local-only |
active |
|
CC-BY-SA (answ |
M |
| docTR |
lin/win/mac |
Deep-learning document text recognition library (detection + recognition), TensorFlow/PyTorch. |
local-only |
? |
6 |
Apache-2.0 |
M |
| dpScreenOCR |
lin/win |
Cross-platform Tesseract screen OCR with a configurable global hotkey; besides clipboard and history it has a 'Run executable' action that passes the |
local-only |
? |
295 |
Zlib |
M |
| e2spd (mglambda) |
lin |
Low-latency bridge between Emacspeak and the Linux Speech Dispatcher architecture. |
local-only |
? |
1 |
GPL-2.0 |
M |
| EasyOCR |
lin/win/mac |
Ready-to-use OCR for 80+ languages; one of the selectable backends in Translumo and wolfmanstout/screen-ocr. |
local-only |
? |
30 |
Apache-2.0 |
M |
| eBook-speaker |
lin |
Jos Lemmens' tool to 'Read aloud eBooks, text documents or scanned documents using a software speech-synthesizer'; the same author's DAISY Player is p |
local-only |
unknown |
|
LGPL-2.1 / LGP |
V |
| edbrowse |
lin |
Command-line editor/browser designed for blind users (Debian 3.8.12-1), packaged under the Salsa a11y-team. |
local-only |
active |
|
GPL (not re-ve |
V |
| Elevado |
lin |
GNOME 'Accessibility inspector' (10 stars on the feaneron copy, with a Rust variant at qwery/elevado-rust) — a modern successor to Accerciser for brow |
local-only |
unknown |
10 |
unknown |
V |
| emacs-nano-tts-minor-mode |
lin |
Emacs minor mode described as an 'Accessibility tool that reads marked text', from the clipboard-speaker author. |
local-only |
dormant |
0 |
unknown |
V |
| Emacspeak (tvraman/emacspeak) |
lin/mac/legacy |
The Complete Audio Desktop: a full speech-output subsystem for Emacs that speaks everything Emacs displays, with audio formatting and voice-lock. |
local-only |
? |
282 |
none declared |
M |
| Enamy TTS (enamy-tts, enamy-tts-rs, enamy-api, enamy-plaintalk-a |
lin |
C++/Slint on-screen caption overlay (a Rust version is marked DEPRECATED) whose main.cpp posts typed text to a remote EnAmy /api/EnAmy/CreateVoice end |
cloud-API |
dormant |
0 |
LICENSE file p |
M |
| espeaker (taylordotfish) |
lin |
'IRC text-to-speech using espeak'. |
local-only |
unknown |
0 |
unknown |
L |
| espeakup |
lin |
A lightweight connector daemon that lets the in-kernel Speakup console screen reader drive eSpeak/eSpeak NG as its software synthesizer. |
local-only |
? |
1 |
GPL-3.0 |
M |
| espeakup-rs (herman_rimm) |
lin |
Rust reimplementation of espeakup (the Speakup-to-eSpeak connector), with an accompanying PKGBUILD repo. |
local-only |
unknown |
0 |
unknown |
L |
| explique.nvim (mickaelfree) |
lin/mac/win |
Neovim plugin that explains the selected text aloud using an AI model plus local TTS. |
hybrid |
active |
0 |
none declared |
V |
| fenrir (ky1e GitLab copy) |
lin |
GitLab copy of the Fenrir TTY screen reader alongside the chrys87 original — noted only as a forge-location fact about an already-catalogued project. |
local-only |
unknown |
0 |
unknown |
D |
| festcat |
lin |
'Festcat aims to provide high quality Catalan voices to Festival Speech System' — voice-data supply on Launchpad. |
local-only |
unknown |
|
unknown |
V |
| FreeSpeak-AAC (Linux and Android) |
lin/and |
'FreeSpeak-AAC is intended to be a FREE and Open Source alternative to the existing paid AAC software', with separate Linux and Android repos. |
unknown |
unknown |
0 |
unknown |
L |
| Frog |
lin |
GNOME app that extracts text from any image, screen area, video frame or QR code; no speech output found in the README. |
local-only |
? |
897 |
MIT |
M |
| Frog |
lin |
"Extract text from images" [vendor] — screen-region and image OCR that puts the result on the clipboard. |
local-only |
active |
|
MIT |
V |
| gImageReader |
lin |
"A graphical (gtk) frontend to tesseract-ocr" [vendor]. |
local-only |
active |
|
GPL-3.0+ |
V |
| Glate |
lin |
"Translate text and generate speech audio on Linux desktop" [vendor] — translation front-end with a speech-output side. |
cloud-API |
active |
|
proprietary (L |
V |
| gnome-braille (GNOME Archive) |
lin |
Archived GNOME braille library, sibling of gnome-speech. |
local-only |
archived |
0 |
unknown |
L |
| gnome-speech (GNOME Archive) |
lin |
Archived GNOME 2 speech abstraction: 'to provide a simple general API for producing text-to-speech output', version 0.4.25, the layer Gnopernicus and |
local-only |
archived |
0 |
LGPL (GNOME-er |
M |
| gspeech (GNOME Archive) |
lin |
Archived 'gtk module to provide speech output for gtk programs' — you load it into any GTK program and it speaks; ships a simple command server and a |
local-only |
archived |
0 |
unknown |
M |
| gtk-openmary |
lin |
'A front-end for the MARY Java TTS (text-to-speech) engine', modelled on eSpeak's GUI. |
local-only |
dormant |
|
unknown |
V |
| ha-rhvoice |
lin |
Home Assistant integration exposing RHVoice as a local TTS provider. |
local-only |
? |
54 |
MIT |
M |
| HiFi-GAN |
lin |
The reference neural vocoder used as the final stage in a large share of TTS systems. |
local-only |
? |
2 |
MIT |
M |
| hyacinthia (icosane) |
lin/win |
'Simple graphical front-end for F5-TTS'. |
local-only |
unknown |
0 |
unknown |
L |
| hypra-epub-reader |
lin |
Browser-based EPUB reader written for Hachette/Hatier schoolbooks because publisher apps lack Linux support; README states it aims for 'at least basic |
local-only |
active |
0 |
unknown |
M |
| install-spd-piper (alexkuz gist) |
lin |
One-line shell installer that registers Piper as a Speech Dispatcher module and auto-downloads voices, so any Linux screen reader can use it. |
local-only |
? |
|
gist, unstated |
M |
| irssi-autospeak (jticket, Stormux Gitea) |
lin |
'Autospeak For Irssi' — speaks IRC activity from inside the irssi terminal client. |
local-only |
unknown |
0 |
unknown |
L |
| iTTS.py / iTTS (KOLANICH) |
lin |
Jupyter kernel that speaks the text of a cell via speech-dispatcher; a GitLab copy exists at gitlab.com/KOLANICH/iTTS. |
local-only |
dormant |
0 |
unknown (no RE |
V |
| Jormungandr |
lin |
GitLab repository holding a design document for a client/server Linux screen reader - a README enumerating the message verbs each component (Speech, I |
unknown |
dormant |
0 |
A LICENSE file |
L |
| Jovie (KDE Text-to-Speech daemon) |
lin |
KDE's system-tray text-to-speech daemon and subsystem (successor to KTTSD), intended as the standard speech-output layer for KDE applications; effecti |
local-only |
discontinued |
9 |
GPL (KDE) |
M |
| Kate "Speak Text" via Qt Speech |
lin |
Accepted 2025-2026 answer to a writer who wanted chapter-at-a-time readback: paste into Kate and use its Qt Speech text-to-speech, same speechd plugin |
local-only |
active |
|
LGPL/GPL (KDE) |
M |
| KDE Connect |
lin/win/mac/and/legacy |
'Multi-platform app that allows your devices to communicate' [vendor] — includes a clipboard-sync plugin between paired devices on the same network. |
local-only |
active |
3 |
null on the Gi |
L |
| kde-notify-text2speech (boospy) |
lin |
German-language KDE Plasma sound theme that speaks system/program notifications instead of playing tones. |
local-only |
unknown |
0 |
unknown |
L |
| kde-tts-input-method (davidedmundson) |
lin |
KDE repository named kde-tts-input-method with no description and no README — existence confirmed, purpose not. |
unknown |
dormant |
0 |
unknown |
L |
| keyd (rvaiya) |
lin |
Linux key-remapping daemon operating at the evdev layer, so it applies in X11, Wayland and the console alike; can execute commands on a chord. |
local-only |
active |
5827 |
MIT |
M |
| KMag (KDE screen magnifier) |
lin |
KDE's 'Screen magnifier', living in the invent.kde.org accessibility group with many personal forks. |
local-only |
unknown |
4 |
unknown |
V |
| KMouth (KDE) |
lin |
KDE's type-and-say front end for speech synthesizers — a speech-generating utility for people who cannot speak, still actively maintained. |
local-only |
active |
11 |
GPL (KDE) |
M |
| kokoro-debian-tts / piper-speechd / sd-edge-tts / OneNoted speak |
lin |
Speech Dispatcher modules that plug modern neural engines (Piper, Kokoro, Edge, Qwen3-TTS) into the Linux speech stack. |
hybrid |
dormant |
3 |
AGPL-3.0 / GPL |
M |
| lemonade |
lin/mac/win |
'Lemonade is a remote utility tool. (copy, paste and open browser) over TCP.' [vendor] — a daemon on the desktop plus a CLI on the remote box, so a he |
local-only |
dormant |
727 |
MIT |
M |
| libbraille (Savannah) |
lin |
Savannah-hosted 'Braille library' for driving refreshable braille displays. |
local-only |
dormant |
|
unknown |
V |
| libNotify-Speech-Dispatcher |
lin |
'Speaks messages from libNotify, caught by DBus, through speech-dispatcher' — desktop notifications read aloud. |
local-only |
dormant |
|
GPL-3 (Launchp |
V |
| libqaccessibilityclient (KDE) |
lin |
KDE's Qt-flavoured AT-SPI client helper library - the Qt-side counterpart to pyatspi for programs that need to read the Linux accessibility tree. |
local-only |
? |
7 |
LGPL (KDE conv |
M |
| libttstd and speakcid (Savannah) |
lin |
Two Savannah non-GNU projects surfaced by a 'tts' software search: libttstd (a TTS daemon library, by name) and speakcid (a caller-ID speaker, by name |
local-only |
unknown |
|
unknown |
L |
| Lios (Linux Intelligent OCR Solution) |
lin |
OCR suite for Tesseract and Cuneiform aimed at visually impaired users; scans or imports images and can read the recognised text aloud. |
local-only |
dormant |
|
|
L |
| Luwrain |
lin/win |
A Java 'accessible environment' rather than a screen reader: instead of reading a graphical desktop, it replaces the desktop with an entirely non-visu |
local-only |
? |
24 |
GPL-3.0 |
M |
| magnify (amiloradovsky) and gnome-swift-screen-magnifier (Amonit |
lin |
Two GitLab-only screen magnifiers: 'Tiny screen magnifier for X11' and a GNOME Shell magnifier. |
local-only |
unknown |
0 |
unknown |
L |
| Manga_Gaze |
lin/win |
'A Python app to make manga accessible: Turn manga into words' — OCR of manga pages plus translation. |
unknown |
unknown |
0 |
unknown |
L |
| Max TTS |
lin |
GPL-3 'text output application' providing enlarged on-screen text output plus speech output, using MARY TTS as its backend with the stated intent that |
local-only |
dormant |
|
GPL-3.0 (Launc |
V |
| mono-a11y |
lin |
'Enables Winforms and Silverlight applications to be fully accessible on Linux, and allows Assistive Technologies' to reach them — the Mono AT-SPI bri |
local-only |
dormant |
|
unknown |
V |
| Neural-Voices-Speech-Dispatcher |
lin |
Use Microsoft Azure voices with your Linux screen reader. |
cloud-API |
? |
0 |
none declared |
M |
| Newton (Wayland-native accessibility architecture) |
lin |
A proposed replacement for AT-SPI on modern free desktops, by AccessKit's author: push the accessibility tree through Wayland/compositor channels inst |
local-only |
? |
, |
open source (A |
L |
| NoComprendo |
lin |
Qt6 Linux tool by Bruno Anselme that is a voice-command, dictation and text-to-speech utility: Vosk speech recognition, user-recorded commands, automa |
local-only |
dormant |
0 |
unknown |
M |
| NormCap |
lin |
"Capture text from any screen area" [vendor] — OCR screen-capture tool. |
local-only |
active |
2,676 |
GPL-3.0-or-lat |
V |
| nvda2speechd |
lin/win |
Bridge exposing NVDA-style speech output to speech-dispatcher, packaged on Salsa (also surfaced by GitLab.com search). |
local-only |
unknown |
|
unknown |
L |
| OCR Grab |
lin |
Lightweight C/GTK3 + Tesseract X11 utility: interactively select a screen region (with an adjustment mode), OCR it, and copy to the clipboard with an |
local-only |
? |
|
not verified |
V |
| OCRFeeder |
lin |
"The complete OCR suite" [vendor] — document OCR with layout analysis. |
local-only |
active |
|
GPL-3.0+ |
V |
| ocrit2tts |
lin |
'Script to screenshot + OCR webpages that use arrow keys to change pages. Useful for the following process: Screenshot -> OCR -> Text -> Speech'. |
unknown |
dormant |
0 |
unknown |
V |
| ocrizer (Hypra) |
lin |
GPL-3 OCR tool ('ocrize') from Hypra, the French company that packages an accessible MATE-based Debian for blind users; developed on salsa.debian.org, |
local-only |
active |
0 |
GPL-3.0-or-lat |
M |
| OCRmyPDF |
lin/mac/win |
Adds an OCR text layer to scanned PDFs so they become searchable and selectable. |
local-only |
? |
34 |
MPL-2.0 |
M |
| Okular "Speak Text" on a PDF selection (+ the qtspeech-speechd-p |
lin |
Okular reads a right-click text selection aloud through Qt Speech -> Speech Dispatcher, but does nothing until qtspeech5-speechd-plugin (Qt5) or qt6-s |
local-only |
active |
|
GPL-2.0-or-lat |
M |
| Open SAPI (uSpeak) |
lin |
Launchpad project aiming 'to bring the Microsoft Speech Application Programming Interface into use under open source projects', first goal being SAPI |
local-only |
dormant |
|
OSL-3.0 (Launc |
V |
| Orca |
lin |
The GNOME screen reader for Linux; it reads the AT-SPI accessibility tree and speaks through speech-dispatcher, and has no OCR of its own — OCRDesktop |
local-only |
? |
|
LGPL |
M |
| Orca (Codeberg copy, WaylandNewton/orca) |
lin |
A Codeberg-hosted copy of GNOME Orca carrying the upstream description ('Screen reader for graphical applications that use the atspi protocol, via spe |
local-only |
unknown |
0 |
LGPL (upstream |
D |
| orca-controller / Orca-Controller-Client / Remora |
lin |
Go and Python clients that drive the Orca screen reader over D-Bus or named pipes, plus Remora, a companion app for Orca. |
local-only |
? |
1 |
AGPL-3.0 |
M |
| orca-intro-guide |
lin |
New-user guide for the Orca screen reader on Linux. |
local-only |
? |
13 |
none declared |
M |
| orca-remote |
lin |
Bidirectional bridge letting Orca on Linux control, and be controlled by, NVDA Remote-compatible peers over the NVDA Remote v2 protocol — speech, brai |
hybrid |
? |
4 |
LGPL-2.1 |
M |
| orcatutor / orca-teacher |
lin |
Two Launchpad projects teaching Orca use: 'A basic application to guide a user through the user interface using the orca screen reader' and 'An educat |
local-only |
unknown |
|
unknown |
L |
| osc (theimpostor/osc) |
lin/mac/win |
'Access the system clipboard from anywhere using the ANSI OSC52 sequence' [vendor] — a single binary that reads stdin and emits the OSC 52 write seque |
local-only |
active |
145 |
MIT |
M |
| PaddleOCR |
lin/win/mac |
Large OCR toolkit for documents and images, 100+ languages, oriented toward turning PDFs/images into structured data. |
local-only |
? |
87 |
Apache-2.0 |
M |
| ParallelWaveGAN |
lin |
Parallel WaveGAN / MelGAN / HiFi-GAN vocoder implementations with recipes. |
local-only |
? |
1 |
MIT |
M |
| Pied |
lin |
Installs and configures the Piper neural TTS engine to work with Speech Dispatcher, then downloads and manages voices. |
local-only |
active |
|
GPL-3.0+ |
V |
| piper-tts-firefox-reader-mode |
lin |
Guide to adding Piper neural voices to Firefox Reader Mode via Speech Dispatcher on Ubuntu 24.04. |
local-only |
? |
0 |
none declared |
M |
| piper-tts.el (Sshigeru) |
lin/mac |
Emacs package described as 'Run TTS on Emacs regions with piper-tts'. |
local-only |
unknown |
0 |
unknown |
L |
| PipeWire network audio (module-rtp-sink / module-rtp-source, pip |
lin |
PipeWire's own module reference: 'The rtp-sink module creates a PipeWire sink that sends audio RTP packets' with source.ip/destination.ip options and |
local-only |
active |
|
MIT (PipeWire) |
V |
| polyglot-for-orca |
lin |
Orca add-on: automatic language switching, emoji reading and Unicode character pronunciation. |
local-only |
? |
1 |
none declared |
M |
| PulseAudio network audio (module-native-protocol-tcp, module-tun |
lin/legacy |
freedesktop's PulseAudio network documentation: set PULSE_SERVER to a remote host, or load module-tunnel-sink, and audio produced on one machine plays |
local-only |
active |
|
LGPL-2.1+ (Pul |
M |
| pyatspi2 (GNOME) |
lin |
Python client bindings for AT-SPI2; the Text interface is where GetTextSelection / get_text_at_offset live, which is the Linux non-destructive selecti |
local-only |
active |
25 |
LGPL-2.1 |
M |
| pyatspitest (Hypra) |
lin |
'Various tools to access and debug AT-SPI2 (through pyatspi)'. |
local-only |
unknown |
0 |
unknown |
V |
| qml-speechd |
lin |
'QML API bindings for the speech-dispatcher text to speech server' — lets Qt/QML apps (including mobile shells) speak through speechd. |
local-only |
dormant |
|
LGPL-3 (Launch |
V |
| qtatspi (KDE) |
lin |
Qt accessibility bridge plugin that exported Qt widgets onto AT-SPI before the bridge moved into Qt itself. |
local-only |
archived |
7 |
LGPL |
M |
| RapidOCR |
lin/win/mac |
ONNX Runtime / OpenVINO / MNN packaging of PaddleOCR models across many languages and bindings — the usual choice when an app wants offline OCR withou |
local-only |
? |
7 |
Apache-2.0 |
M |
| readaloud (SourceForge) |
lin |
'Reading program for plain text books. Readaloud is an aide for learning reading also. It helps by displaying text in large fonts and speaking a word |
local-only |
unknown |
|
GPL-2.0 (SF ca |
V |
| repy |
lin/mac/win |
Rust terminal EPUB reader inspired by epy, 10 stars: it parses a book file and reads it inside its own TUI, toggling TTS with !, chunking sentence b |
hybrid |
active |
10 |
none declared |
M |
| retro-tts-pack |
lin |
Retro speech synthesizers packaged for Speech Dispatcher and Orca on Linux. |
local-only |
? |
2 |
none declared |
M |
| retro-tts-pack / Apple-Eloquence-ELF / ViaVoice-SPD |
lin |
Projects that bring classic screen-reader voices (Eloquence, ViaVoice, retro synths) to Linux Speech Dispatcher and Orca. |
local-only |
? |
16 |
other / GPL-2. |
M |
| RHVoice-English fixes (Stormux) |
lin |
'English fixes for RHVoice' — a voice-data patch set living only on the Stormux Gitea. |
local-only |
unknown |
0 |
unknown |
L |
| Roc Toolkit |
lin/mac/and |
'Real-time audio streaming over the network' [vendor] — a toolkit and set of tools for sending audio between machines with loss recovery and latency c |
local-only |
active |
1 |
MPL-2.0 |
M |
| SAK (Speecher Assistive Kit) |
lin/win/mac |
Pascal library whose description says 'With sak, your application becomes assistive directly, without changing anything in your code' — bundles eSpeak |
local-only |
dormant |
0 |
split: /sak pe |
M |
| screengrab-ocr (bhvsh) and screenshot-ocr-linux (emonbhuiyan) |
lin |
Two further GitLab-only screenshot-OCR utilities: 'A Python application that captures screenshots and performs optical character recognition (OCR) on |
local-only |
unknown |
0 |
unknown |
L |
| sd-piper |
lin |
Speech Dispatcher driver implementing Piper neural voices for the Orca screen reader. |
local-only |
? |
1 |
GPL-3.0 |
M |
| seed-tts-eval |
lin |
ByteDance's TTS evaluation benchmark and test sets. |
local-only |
? |
1 |
none declared |
M |
| selenium-webdriver-at-spi (KDE) |
lin |
'Selenium/Appium WebDriver implementation based on AT-SPI Accessibility' — drives Linux desktop apps through the same accessibility tree a selection r |
local-only |
unknown |
0 |
unknown |
V |
| selenium-webdriver-at-spi (KDE) |
lin |
A WebDriver implementation backed by AT-SPI, used to drive and assert on KDE apps through the accessibility tree. |
local-only |
active |
7 |
GPL |
M |
| sight-free-talon |
lin/mac/win |
Bridges the Talon voice-control engine to screen readers, TTS and braille, so voice commands can trigger reading actions. |
local-only |
? |
24 |
GPL-3.0 |
M |
| Silero-TTS-Service |
lin/legacy |
Silero TTS backend service for Home Assistant / Rhasspy. |
local-only |
? |
60 |
MIT |
M |
| simple-orca-plugin-system (chrys87) |
lin |
'a simple but powerfull pluginsystem for the orca screenreader', from the Fenrir/Jormungandr author. |
local-only |
unknown |
0 |
unknown |
L |
| Snapcast |
lin/mac/and/legacy |
'Synchronous multiroom audio player' [vendor] — a server/client pair that distributes an audio stream to multiple networked clients in sync. |
local-only |
active |
7 |
GPL-3.0 |
M |
| sonarmacs |
lin |
'Speech-dispatcher enabled Emacs' — an Emacs speech layer wired to speech-dispatcher rather than to Emacspeak's own servers. |
local-only |
unknown |
|
unknown |
V |
| sonic / libsonic |
lin/and |
Library and CLI to speed up or slow down speech without pitch distortion — the playback-rate layer behind many screen readers; also on Launchpad as 's |
local-only |
unknown |
|
Apache-2.0 (up |
V |
| speakerd |
lin |
Python reimplementation of Speech Dispatcher whose README states 'The purpose is not to replace the original program but to use it for testing or POC' |
local-only |
dormant |
0 |
GPL-3.0 (READM |
M |
| Speech Dispatcher (speechd) |
lin/legacy |
The Linux speech-output middleware layer: a common device-independent API that multiplexes many client applications onto pluggable synthesis modules. |
local-only |
? |
|
GPL |
M |
| Speech Dispatcher (speechd) |
lin/legacy |
The Linux speech broker: one stable protocol (SSIP) in front of many synthesis engines, with spd-say as a general text-to-speech CLI. |
local-only |
? |
329 |
GPL-2.0 (daemo |
M |
| Speech Dispatcher upstream on Savannah (speechd) |
lin |
The Savannah group registration for speechd — a forge-location fact about an already-catalogued project: its canonical home has historically been free |
local-only |
active |
|
GPL/LGPL |
M |
| Speech Viewer / speechd-up / Emacspeak bridges |
lin |
Debug and integration utilities around Speech Dispatcher — a window that displays what would be spoken, a Speakup console-reader bridge, and Emacs spe |
local-only |
? |
0 |
BSD-2-Clause / |
L |
| speech-dispatcher-rs / speech-dispatcher-sys (ndarilek and forks |
lin |
Rust bindings to speech-dispatcher, with GitLab copies under ndarilek, TTWNO, mcb2003, Caellian, cpu and MallocVoidstar. |
local-only |
unknown |
0 |
unknown |
V |
| speechd-java |
lin |
The official Java client library for Speech Dispatcher, developed separately from the main speechd tree. |
local-only |
? |
3 |
not reported b |
M |
| speechd-up |
lin |
The other Speakup connector: bridges the Speakup kernel screen reader to Speech Dispatcher, so the console reader can use any speechd output module. |
local-only |
? |
|
GPL-2.0-or-lat |
L |
| Spiel / libspiel |
lin |
A newer GNOME-adjacent speech framework: a Web-Speech-API-shaped client library over D-Bus speech providers, integrated with GStreamer - positioned as |
local-only |
? |
53 |
LGPL-2.1 |
M |
| SRAL (Screen Reader Abstraction Library) |
lin |
Library that abstracts over screen readers and speech engines so an application can speak through whatever the user is running. |
local-only |
archived |
21 |
MIT |
V |
| ssip-client (Rust) — TTWNO and lp-accessibility copies |
lin |
'Speech Dispatcher SSIP client library in rust' — the client library under the Odilia/Rust Linux-a11y line, present on GitLab under both TTWNO (Tait H |
local-only |
unknown |
0 |
unknown |
V |
| Stormux |
lin |
Self-hosted-Gitea Arch-based accessible Linux distribution described as 'The continuation of the F123Light project'; the same forge hosts ~50 blind-ac |
local-only |
active |
0 |
unknown |
V |
| Subtitle Reader (Hackaday project) |
lin |
Hobby project that reads TV subtitles off the screen by OCR and speaks them aloud, for children and visually impaired viewers. |
local-only |
? |
|
not verified |
L |
| Surya |
lin/win/mac |
OCR with layout analysis, reading order and table recognition in 90+ languages. |
local-only |
? |
21 |
Apache-2.0 |
M |
| tacotron (keithito) |
lin |
TensorFlow Tacotron with pretrained models; source of the widely reused text-normalization code. |
local-only |
? |
3 |
MIT |
M |
| Tacotron 2 (NVIDIA) |
lin |
PyTorch Tacotron 2 with faster-than-realtime inference. |
local-only |
? |
5 |
BSD-3-Clause |
M |
| Tacotron-2 (Rayhane) |
lin |
TensorFlow Tacotron-2 implementation. |
local-only |
? |
2 |
MIT |
M |
| Talking DOSBox |
lin |
"A talking dosbox with screen reader." [vendor, AUR description] — DOSBox fork that speaks its emulated screen. |
local-only |
dormant |
|
unknown |
L |
| Tesseract OCR |
lin/win/mac |
A high-adoption open-source OCR engine; hand it an image, it returns text. Used by dpScreenOCR, NormCap, Capture2Text, TextSnatcher, Frog, gImageReade |
local-only |
? |
76 |
Apache-2.0 |
M |
| Text Grabber (GNOME Shell extension) |
lin |
OCR grab extension; listing carries a maintainer-wanted notice. [vendor] |
local-only |
unknown |
|
unknown |
V |
| text-to-speech-ubuntu |
lin |
Minimal Ubuntu selection reader: select text with the mouse, press a key, and xsel piped to espeak reads it — no OCR. |
local-only |
? |
29 |
none declared |
M |
| TextSnatcher |
lin |
"Snatch Text with just a Drag" [vendor] — drag a region, get its text. |
local-only |
active |
|
GPL-3.0-or-lat |
V |
| tolk-spd-proxy |
lin |
A drop-in tolk.dll replacement for Wine prefixes that forwards screen-reader speech calls (mostly from audio games) to speech-dispatcher on the Linux |
local-only |
unknown |
|
unknown |
V |
| tolk2spd |
lin |
Bridge exposing the Windows Tolk screen-reader API on top of Linux Speech Dispatcher. |
local-only |
? |
0 |
0BSD |
M |
| triggerhappy (wertarbyte) |
lin |
Lightweight hotkey daemon that reads evdev directly, so it works with no X server at all. Pre-dates Wayland but is the same evdev-level answer as hkd. |
local-only |
dormant |
296 |
GPL-2.0 |
M |
| vimspeak |
lin/mac/win |
Vim plugin connecting vim to espeak: it overrides s and S so that s{motion} reads the implicated text aloud (e.g. s} reads the paragraph). |
local-only |
unknown |
|
unknown |
V |
| WaveGlow |
lin |
Flow-based vocoder from NVIDIA. |
local-only |
? |
2 |
BSD-3-Clause |
M |
| wavenet_vocoder |
lin |
WaveNet vocoder in PyTorch. |
local-only |
? |
2 |
NOASSERTION |
M |
| waynotify |
lin |
An accessible notification daemon for Wayland that routes desktop notifications through AT-SPI so a screen reader announces them. |
local-only |
? |
3 |
MIT |
M |
| wl-clipboard / wl-clipboard-rs |
lin |
The Wayland copy/paste utilities — wl-paste --primary is the Wayland equivalent of xsel/xclip and is the acquisition primitive every Wayland sel |
local-only |
active |
2397 |
GPL-3.0 |
M |
| wlr-data-control / ext-data-control Wayland protocols |
lin |
The privileged Wayland protocol extensions that let a non-focused client read the clipboard and PRIMARY selection at all; ext-data-control is the upst |
local-only |
active |
|
MIT |
L |
| xctts |
lin |
X-Chat script using eSpeak to give each IRC user a distinct voice by varying espeak parameters, so users can be told apart while doing something else. |
local-only |
dormant |
|
GPL-3 (Launchp |
V |
| xdg-desktop-portal (frontend) |
lin |
The D-Bus service that front-ends every portal interface (Screenshot, ScreenCast, GlobalShortcuts, RemoteDesktop, InputCapture) and routes each call t |
os-native |
active |
819 |
LGPL-2.1 |
M |
| xdg-desktop-portal Screenshot / ScreenCast interfaces |
lin |
The sandboxed D-Bus route by which a Wayland application may capture the screen at all — the OCR-reader equivalent of the clipboard problem, and the r |
local-only |
? |
|
LGPL-2.1 |
L |
| ABBYY FineReader |
win/mac |
Long-established commercial OCR and PDF conversion suite; converts scans and images into editable/searchable text with high accuracy. |
local-only |
? |
|
proprietary, p |
V |
| access-bridge-explorer (google) |
win |
Explorer for the accessibility tree of Java Access Bridge-enabled applications - the separate accessibility stack Java desktop apps use on Windows. |
local-only |
dormant |
131 |
Apache-2.0 |
M |
| Accessibility Insights for Windows |
win |
Microsoft's UIA tree inspector and automated-check tool for Windows apps - the Windows counterpart to Accerciser. |
local-only |
? |
536 |
MIT (repo repo |
M |
| AccessibleRunner / CommandRunner |
win |
Runs console commands and presents their output in a screen-reader-friendly window; a Python version and a C# version by the same author. |
local-only |
? |
0 |
MIT |
M |
| AccessKit |
win/mac/lin/and |
Cross-platform accessibility infrastructure for UI toolkits: one tree model that AccessKit projects onto Windows UIA, macOS AX, and Linux AT-SPI. |
local-only |
active |
1 |
BSD-3-Clause ( |
M |
| accesskit-python |
win/mac/lin |
Python bindings for AccessKit. |
local-only |
active |
6 |
BSD-3-Clause |
M |
| ADELE-TEAM |
win |
French free reading/writing assistant for DYS profiles: vocalised reading with follow-along, rapid navigation, highlighter-style marking, passage extr |
local-only |
unknown |
|
free |
V |
| adispeak / adispeak-2 |
win |
'Solution to make AdiIRC interface with screen readers' — two Codeberg repos from the same author. |
local-only |
unknown |
0 |
unknown |
L |
| Adobe Acrobat Reader — Read Out Loud (built-in) |
win/mac |
Acrobat Reader's View > Read Out Loud reads a clicked paragraph or from the cursor to end of document, using the OS voices. |
os-native |
? |
|
Proprietary (A |
L |
| AI Text Tools (Netropolitan) |
win |
'AI-powered text transformation for Windows. Select text anywhere, press a hotkey, and transform it using OpenAI, Anthropic Claude, Google Gemini, or |
hybrid |
active |
3 |
not read |
V |
| AOK application suite — MyEdit Neo / MyWord7 / MyBook Neo / MyNe |
win |
The application layer 高知システム開発 ships around PC-Talker: a voice text editor (MyEdit Neo), voice word processor (MyWord7), barrier-free reading app (MyB |
local-only |
commercial-l |
|
proprietary (N |
M |
| arboard (1Password) |
win/mac/lin |
Rust clipboard crate maintained by 1Password; the crate many Tauri/Rust readers use to read text after a simulated copy. |
local-only |
active |
959 |
Apache-2.0 / M |
M |
| ARIA-AT (Assistive Technology ARIA Experience Assessment) |
win/mac |
W3C community-group test suite measuring how JAWS, NVDA and VoiceOver actually behave on ARIA patterns - the closest thing to a neutral capability com |
local-only |
? |
181 |
W3C document a |
M |
| AT Driver (W3C aria-at-automation) |
win/mac |
A draft protocol for introspecting and remote-controlling assistive technology over a bidirectional channel - a WebDriver-shaped standard for screen r |
local-only |
? |
|
W3C / communit |
M |
| AT Guys |
win |
US assistive-tech reseller that was the retail front for the Code Factory Eloquence + Vocalizer NVDA bundle. |
unknown |
unknown |
|
n/a (reseller) |
L |
| AutoHotkey non-clipboard selection capture via `ControlGet, Sele |
win |
Stack Overflow answer showing the only Windows route that avoids the clipboard entirely - it works only when the focused control matches Edit\d+, with |
local-only |
active |
|
CC-BY-SA (answ |
M |
| aviewer (ThePacielloGroup) |
win |
Inspector that dumps what MSAA, IAccessible2, UI Automation, ARIA and the HTML DOM each expose for a given Windows control - the tool that answers 'do |
local-only |
dormant |
163 |
none |
M |
| Blio (KNFB Reading Technology / Ray Kurzweil, with Baker & Taylo |
win/and/ios |
Kurzweil's colour, layout-preserving ebook platform launched in 2010 with text-to-speech using downloadable 'Samantha' and 'Tom' voices; its distribut |
hybrid |
discontinued |
|
proprietary, f |
M |
| Brazilian Portuguese GitHub reader cluster: vinimostaco/incipit |
win/mac/lin/web |
Small Portuguese-described projects: a Tauri+Python desktop app that reads PDFs, EPUBs and text aloud (incipit), a text reader with voice selection, a |
local-only |
unknown |
0 |
mixed |
L |
| calibre E-book viewer Read aloud (built-in) |
win/mac/lin/legacy |
calibre's bundled e-book viewer reads the open book aloud on Ctrl+S, highlighting the current word or sentence, using the Piper neural engine locally |
local-only |
? |
22 |
GPL-3.0 (calib |
M |
| CaptiOCR |
win |
Real-time screen text extraction: pick a rectangular region, and it repeatedly screenshots and Tesseract-OCRs it, stitching the text into a continuous |
local-only |
? |
19 |
MIT |
M |
| Chatterb0x |
win |
Small Python tray utility that fronts Chatterbox TTS. |
local-only |
? |
3 |
MIT |
L |
| chrome-extension-speak-selection (belcrod5) / obsidian--speak-se |
win/mac/lin |
Three further single-surface 'speak selection' implementations — a Chrome extension, an Obsidian plugin, and a JS selection-TTS project. |
unknown |
? |
0 |
unspecified |
L |
| Clavier+ (gryder.org) |
win |
French global-hotkey utility: launch a program or type text from a key combination, described as a single EXE under 200 KB with no registry storage an |
local-only |
active |
|
open source |
M |
| clipboard (CrossCopy) |
win/mac/lin |
Clipboard API with text and image read/write/watch for macOS, Windows and Linux. |
local-only |
active |
25 |
MIT |
M |
| Clipboard TTS (clipboardtts.com) |
win |
Commercial clipboard-to-speech desktop product; I reached only the download page, so capability claims beyond 'reads the clipboard' are unverified. |
unknown |
unknown |
|
proprietary |
L |
| clipboard-rs (ChurchTao) |
win/mac/lin |
Cross-platform Rust clipboard API covering text, image, rich text, HTML, files and change monitoring. |
local-only |
active |
176 |
Apache-2.0 |
M |
| clipboard_listener (orderweaver) |
win/mac/lin |
Cross-platform Rust crate for listening to clipboard events. |
local-only |
dormant |
2 |
none |
M |
| clipboard_watcher (leanflutter) |
win/mac/lin |
Flutter plugin that raises an event when the system clipboard changes, on desktop targets. |
local-only |
dormant |
60 |
MIT |
M |
| clipboardEnhancement (NVDA add-on) |
win |
NVDA add-on extending clipboard handling and announcement, listed in the NVDA add-ons directory alongside Clipspeak and Autoclip. |
local-only |
? |
|
unknown |
L |
| ClipboardTTS (Daichi09) |
win |
Small Windows program that reads the current clipboard text through Microsoft SAPI; the README's stated use case is visual-novel text hookers that pus |
os-native |
dormant |
4 |
Unlicense |
M |
| ClipboardWatcher (DissectMalware) |
win |
Windows utility that monitors textual data pasted into the clipboard. |
local-only |
dormant |
29 |
none |
M |
| Clipspeak (NVDA add-on) |
win |
NVDA add-on that announces clipboard operations — cut, copy, paste, undo, redo — rather than reading the clipboard's contents. |
local-only |
? |
7 |
check repo |
M |
| copypasta (alacritty) |
win/mac/lin |
Cross-platform Rust system-clipboard crate maintained alongside Alacritty; exposes X11 PRIMARY as a distinct clipboard. |
local-only |
dormant |
351 |
MIT / Apache-2 |
M |
| Descolada/UIA-v2 |
win |
'UIAutomation library for AHK v2, based on thqby's UIA library' - the route by which an AutoHotkey selection reader can read text from a UIA TextPatte |
local-only |
active |
460 |
MIT |
M |
| Deskflow |
win/mac/lin |
'Share a single keyboard and mouse between multiple computers.' [vendor] — the maintained successor line to Synergy/Barrier; clipboard sharing between |
local-only |
active |
28 |
GPL-2.0 |
L |
| DirectShell |
win |
Posted to HN as "a new software primitive that replaces AI screenshot agents"; repo description is only "Because i could not Ressist". |
unknown |
? |
21 |
NOASSERTION |
L |
| document-reader-nvda-addon |
win |
NVDA add-on for reading multiple document formats accessibly. |
local-only |
? |
3 |
GPL-2.0 |
M |
| Dolphin EasyReader (winget) |
win |
Accessible-library reader from Dolphin Computer Access, packaged for winget. |
unknown |
unknown |
|
proprietary |
L |
| Dolphin GuideConnect |
win |
Windows product for people with sight loss that replaces the desktop with its own spoken, magnified menu environment for email, web, documents, scanni |
local-only |
commercial-l |
|
Commercial; so |
V |
| DOSVOX (Instituto Tércio Pacitti / NCE, UFRJ) |
win |
Brazilian speech-based computing environment for blind users from UFRJ, running on Windows: a shell plus 80+ of its own speech-driven programs (editor |
local-only |
unknown |
|
free distribut |
M |
| DSpeech |
win |
Portable freeware Windows text-to-speech application (v1.74) that can capture and speak clipboard content, mix and juxtapose several SAPI voices, chan |
os-native |
unknown |
|
Freeware (prop |
M |
| Dual Voice for NVDA |
win |
'Dual Voice for NVDA is an open source speech driver for NVDA screen reader. This lets you use two separate voices for reading non-Latin and Latin lan |
local-only |
unknown |
|
unknown |
L |
| Dys-Vocal (Dyslogiciel) |
win |
French compensation software for dyslexia and dyspraxia: text formatting (coloured syllables, dyslexia fonts) plus modules for speech-synthesis readin |
local-only |
commercial-l |
|
proprietary |
M |
| Easy Screen OCR |
win/mac |
Commercial screenshot-OCR utility that grabs a snapshot of the screen and extracts the text for editing. |
cloud-API |
? |
|
proprietary, f |
V |
| EdSharp |
win |
A self-voicing text editor distributed through the blind-user software scene, catalogued on Blind Help Project. |
local-only |
unknown |
|
unknown |
L |
| enigo (enigo-rs) |
win/mac/lin |
Cross-platform input simulation in Rust; the crate behind the simulated Ctrl+C in several Rust selection readers. |
local-only |
active |
1761 |
MIT |
M |
| eSearch |
win/mac/lin |
Cross-platform screenshot tool with offline OCR, screen translation, live-text and search — Windows, macOS and Linux. |
local-only |
? |
6921 |
GPL-3.0 |
M |
| Firefox Reader View — Narrate (built-in) |
win/mac/legacy |
Firefox's Reader View has carried a Narrate (listen) button since 2016 that reads the distilled article using the platform speech synthesis. |
os-native |
? |
|
MPL-2.0 (Firef |
V |
| FlaUI |
win |
.NET wrapper over Windows UIA2/UIA3 for UI automation, including text-pattern access to selections. |
local-only |
? |
3 |
MIT |
M |
| Fluent Search — Screen Search with OCR |
win |
Windows launcher whose screen-search feature OCRs on-screen text so you can act on text that isn't selectable. |
local-only |
? |
|
proprietary (f |
V |
| global-hotkey (tauri-apps) |
win/mac/lin |
Rust crate providing global hotkeys for desktop apps on Windows, macOS and Linux; the crate Tauri-based readers use. Its Linux backend is X11-based, w |
local-only |
active |
259 |
MIT / Apache-2 |
M |
| Hebrew NVDA packaging — תשר אופטיקה געש (TSR Gaash) |
win |
Israeli AT supplier that ships NVDA with its own selected add-ons and bundled Hebrew and English speech engines, stating the add-ons improve NVDA's be |
local-only |
commercial-l |
|
GPL (NVDA) plu |
M |
| huaiyinfeilong/xyOCR · yizilian-iren/HoverDict · zjmxczhy/foo_sp |
win/mac |
Three Chinese-described accessories: an OCR plug-in for NVDA (xyOCR), a macOS hover-to-look-up-and-speak tool (HoverDict), and a foobar2000 plug-in th |
local-only |
unknown |
9 |
unknown |
L |
| IBM Home Page Reader |
win |
Self-voicing web browser grown out of Chieko Asakawa's work at IBM Japan - it began as a Netscape extension and became an Internet Explorer plug-in, r |
local-only |
discontinued |
|
proprietary co |
M |
| ICE Book Reader Professional |
win |
Russian e-book reader whose feature list includes read-aloud through SAPI 4.0 and SAPI 5.1, export of a book to MP3/WAV, splitting a book into several |
local-only |
dormant |
|
freeware (Russ |
M |
| Input Leap |
win/mac/lin |
'Open-source KVM software' [vendor] — the Barrier fork that preceded the Deskflow line; shares keyboard, mouse and clipboard across machines. |
local-only |
archived |
8 |
NOASSERTION on |
M |
| InputBot (obv-mikhail) |
win/legacy |
Rust library for global hotkeys and input simulation. |
local-only |
active |
462 |
MIT |
M |
| Interception (oblitum) |
win |
Windows kernel-level input interception driver and API; the layer AutoHotInterception and some Windows remappers sit on. |
local-only |
dormant |
1957 |
Zlib |
M |
| islambenmebarekdz-collab/Daftari |
win |
Arabic-first, NVDA-native Markdown note-taking app for Windows — an application built so an Arabic screen-reader user can work in it, rather than a re |
local-only |
unknown |
0 |
unknown |
L |
| Israeli AT distributors: AccessMind (בראש נגיש) · Let's Talk טכנ |
win/ios |
Two Israeli assistive-technology vendors whose blindness product pages are built around JAWS for Hebrew-speaking users, one of them advertising a JAWS |
local-only |
commercial-l |
|
proprietary |
M |
| IUIAutomationTextPattern::GetSelection (Windows UI Automation) |
win |
The Windows equivalent of the macOS AX call: 'Retrieves a collection of text ranges that represents the currently selected text in a text-based contro |
os-native |
active |
|
|
M |
| java-access-bridge-wrapper (robocorp) |
win |
Python wrapper around the Java Access Bridge Windows DLL, giving programmatic reads of Java app UI text. |
local-only |
active |
20 |
Apache-2.0 |
M |
| JAWS Tandem |
win |
A Freedom Scientific deployment and licensing page: adding paid 'remote authorization' to an existing JAWS or Fusion licence lets a copy of JAWS runni |
hybrid |
commercial-l |
|
commercial; a |
M |
| jintellitype (melloware) |
win |
Java API for registering Windows global hotkeys. |
local-only |
active |
180 |
Apache-2.0 |
M |
| jkeymaster (tulskiy) |
win/mac/legacy |
Java global-hotkey registration via JNA for X11, Windows and macOS. |
local-only |
active |
241 |
LGPL-3.0 |
M |
| kanata (jtroo) |
win/mac/lin |
Cross-platform keyboard remapper with layers, tap-hold and command execution; runs on Windows, macOS and Linux below the window system. |
local-only |
active |
7722 |
LGPL-3.0 |
M |
| Korean JAWS (한국어 JAWS) — Siloam revision |
win |
The Korean revision of JAWS, produced by the same Siloam institute that developed 드림보이스; the institute's board hosts the installers and a troubleshoot |
local-only |
active |
|
proprietary |
M |
| Kurzweil 3000 / Kurzweil 1000 / OpenBook |
win/mac/legacy |
Long-established scan-and-read products for print-disabled users: OCR a scanned or imported document, then read it aloud with synchronized highlightin |
hybrid |
commercial-l |
|
proprietary, p |
V |
| MAGic |
win |
Screen magnification software with speech output for low-vision users (Freedom Scientific). |
os-native |
? |
|
Proprietary, p |
L |
| Microsoft Copilot Vision |
win/ios/and |
Windows/Edge feature where Copilot looks at the current screen or app window and answers questions about it in voice — an OS-level 'point at the scree |
cloud-API |
? |
|
commercial/bun |
V |
| Microsoft Edge Read Aloud |
win/mac/lin |
Browser-native read-aloud with natural voices and word highlighting for web pages, EPUB and PDFs — but only where a text layer already exists; an imag |
hybrid |
? |
|
proprietary (b |
V |
| Microsoft Edge Read Aloud (browser built-in) |
win/mac/legacy/and/ios |
Read Aloud is built into Edge itself: Ctrl+Shift+U (Cmd+Shift+U on macOS) reads the page, a PDF, or a selection, with a playback toolbar and neural vo |
hybrid |
? |
|
Proprietary (b |
V |
| Microsoft Immersive Reader |
win/mac/web |
Reading-mode surface across Edge, Word, OneNote and Teams that reflows the document and reads it aloud with line focus, syllable splitting and a dysle |
hybrid |
? |
|
proprietary (f |
V |
| Microsoft Office Read Aloud / Speak command |
win/mac/web |
Built into Word, Outlook, OneNote and PowerPoint: Read Aloud reads the document or the selection with word highlighting, and the older Speak command r |
hybrid |
? |
|
proprietary (p |
V |
| Microsoft Speech Platform voice packages (msspeech-tts-* on Choc |
win |
Per-voice Chocolatey packages that install Microsoft Speech Platform runtime voices (Helen, ZiraPro, Hazel, Heather, Hayley, Heera, Hanna, Herena, Hel |
local-only |
active |
|
Microsoft EULA |
M |
| mIRC With Speech (MWS) |
win |
mIRC script that pushes IRC messages, joins/parts and notices out through the user's existing screen reader (JAWS and NVDA supported) rather than mIRC |
local-only |
unknown |
|
unknown (no SF |
V |
| mircwithspeech (fudge333 GitLab) |
win |
'Accessible IRC mIRC script for screen reader users' — a GitLab copy of the SourceForge MWS line. |
local-only |
unknown |
0 |
unknown |
L |
| MyStudyBar |
win |
CALL Scotland page describing MyStudyBar, a portable Windows launcher bar bundling literacy freeware that can be run from a USB pendrive, with two tex |
local-only |
dormant |
|
Free of charge |
M |
| Nattiq (Acapela Arabic voice supply for Dolphin) |
win |
Named on Dolphin's own voice-inventory page as the supplier and seller of the Acapela Arabic voice Leila — an example of regional voice sub-distributi |
local-only |
commercial-l |
|
Voices bundled |
L |
| NaturalReader |
win/mac/legacy |
Commercial TTS reader with free tier, spanning a desktop app, web app and Chrome extension, with OCR for scanned documents. |
hybrid |
? |
|
Proprietary, f |
L |
| NaturalVoiceSAPIAdapter |
win |
A SAPI 5 engine shim that republishes Windows 11 Narrator natural voices, Edge Read Aloud online voices, and (with your key) Azure voices to any SAPI |
hybrid |
? |
889 |
MIT |
M |
| NHotkey (thomaslevesque) |
win |
Managed .NET library for global hotkeys in WinForms and WPF. |
local-only |
active |
342 |
Apache-2.0 |
M |
| NoMachine NX clipboard (EnableClipboard server/client/none) |
win/mac/lin/and/ios/web |
NoMachine's own knowledge base states 'By default users can copy and paste from locale to the session and vice-versa', and documents EnableClipboard=s |
unknown |
commercial-l |
|
proprietary (f |
M |
| NVDA 'tesseractOCR' add-on (Rui Fontes) |
win |
NVDA add-on that OCRs a FILE rather than the screen: Windows+Control+R runs Tesseract over the image or PDF file currently selected, Windows+Control+W |
local-only |
active |
14 |
GPL-2.0 |
M |
| NVDA Brasil (nvda.com.br) |
win |
Brazilian NVDA community portal distributing the official installer alongside curated add-ons; its download page was current for NVDA 2026.1.1 as of 2 |
local-only |
active |
|
GPL (NVDA) |
M |
| NVDA Controller Client |
win |
NVDA's official external-process API: a DLL any program can call to make a running NVDA speak text or SSML, braille a message, cancel speech, or repor |
local-only |
? |
|
Distributed wi |
M |
| NVDA Controller Client (bindings cluster) |
win |
The NVDA-supplied DLL API that lets any third-party program push a string into NVDA's speech queue; found in this sweep as language bindings rather th |
local-only |
dormant |
2 |
|
M |
| NVDA Remote Access (built-in) |
win |
NVDA's now-built-in remote feature: an NVDA on one machine speaks what an NVDA on another machine sees, mediated by a relay server - NVDA-to-NVDA, not |
hybrid |
? |
|
GPL-2.0-or-lat |
M |
| NVDA Remote Access add-on (NVDARemote) |
win |
NVDA add-on (v2.6) that connects two machines each already running NVDA, relaying keyboard and braille input to the controlled machine and its speech |
hybrid |
dormant |
78 |
GPL-2.0 |
M |
| nvda-autohotkey (hi5) |
win |
Library letting AutoHotkey scripts send text to NVDA for speech and braille output. |
local-only |
? |
4 |
LGPL-2.1 |
L |
| nvda-testing-driver |
win |
Build functional tests that drive the NVDA screen reader. |
local-only |
? |
48 |
GPL-3.0 |
M |
| nvda2speechd |
win/lin |
Bridge letting Windows applications route speech through Linux's Speech Dispatcher. |
local-only |
dormant |
12 |
GPL-3.0 |
L |
| Obsidian-pdf-read-aloud |
win/mac/lin |
Obsidian plugin that reads PDFs opened inside the vault using the browser Web Speech API: play/pause/stop, skip by a configurable number of sentences, |
local-only |
active |
1 |
MIT |
M |
| Oculos |
win/mac/lin |
"If it's on the screen, it's an API" — exposes any desktop app's UI automation tree over REST and MCP, in Rust. |
local-only |
? |
126 |
MIT |
M |
| Parsec |
win/mac/lin/and/web |
Low-latency remote-desktop / game-streaming product; its public help centre was reachable but I could not locate a clipboard-behaviour or accessibilit |
hybrid |
commercial-l |
|
proprietary |
L |
| peterrc87/TCA_Portapapeles · rayo-alcantar/KillProcess |
win |
Two Spanish-language NVDA add-ons: a clipboard add-on and one that ends the focused process. [lang: Spanish; English UI: no] |
local-only |
unknown |
2 |
unknown |
L |
| pipe2textbox |
win |
Pipes arbitrary command-line output into a read-only Windows textbox purely so a screen reader can read it - a one-trick bridge from stdout to accessi |
local-only |
? |
13 |
MIT |
M |
| Piper-Tray |
win |
Windows system-tray utility that drives Piper TTS. |
local-only |
? |
33 |
none declared |
L |
| POMXARK/SmartDictor · pegas365i4/voicing_text_in_cmd · ikopylov1 |
win |
Three Russian-described hobby projects: recognition and voicing of text taken from the screen (SmartDictor), a self-described simplified Balabolka ana |
local-only |
unknown |
0 |
unknown |
L |
| PowerTalk |
win |
'PowerTalk automatically speaks Microsoft PowerPoint presentations. For presenters who find speaking difficult, audiences containing people with visua |
os-native |
unknown |
|
GPL-2.0 (SF ca |
V |
| Premier Literacy suite (Universal Reader, Scan and Read Pro, Tal |
win/legacy |
US literacy-software suite for dyslexic and low-vision users bundling a general reader, a scan-and-read tool, a talking word processor and a browser t |
local-only |
unknown |
|
proprietary; c |
L |
| Py TTS |
win |
'Py TTS is a text to speech (TTS) software which uses Microsoft's SAPI to convert text into spoken audio. You can listen as the program reads out the |
os-native |
dormant |
|
GPL-3 (Launchp |
V |
| pyia2 (illinois-dres-aitg) |
win |
Python interface to the MSAA and IAccessible2 interfaces - the accessibility API Firefox and LibreOffice expose on Windows, distinct from UIA. |
local-only |
dormant |
10 |
none |
M |
| PyScreenReader |
win/mac/lin |
"A cross-platform Python library that wraps native accessibility APIs to collect widget tree information on the screen" [vendor] — parses UI propertie |
local-only |
active |
|
MIT |
V |
| pyscrout |
win |
Python library that sends text to be spoken and shown in braille by a screen reader, positioned by its README as a maintained replacement for accessib |
local-only |
active |
0 |
unknown |
M |
| Python-UIAutomation-for-Windows (yinkaisheng) |
win |
Python 3 wrapper of Microsoft UIAutomation covering MFC, WinForms, WPF, Modern UI and more. |
local-only |
active |
3547 |
Apache-2.0 |
M |
| pywinauto |
win |
Python GUI automation for Windows with a UIA backend - a practical way to script 'get the selected text out of the focused control' without writing a |
local-only |
active |
6 |
BSD-3-Clause |
M |
| QHotkey (Skycoder42) |
win/mac/legacy |
Global shortcut library for desktop Qt applications on Windows, macOS and X11. Qt-native; relevant to the KDE stack because Qt Speech-based readers ne |
local-only |
active |
677 |
BSD-3-Clause |
M |
| Qt Speech (QTextToSpeech) |
win/mac/lin/and/ios |
Qt's cross-platform TTS abstraction: one API that maps onto SAPI/WinRT on Windows, AVSpeechSynthesizer/NSSpeechSynthesizer on macOS, speech-dispatcher |
local-only |
active |
39 |
LGPL-3.0 / GPL |
M |
| Quicker (getquicker.net) |
win |
Chinese Windows action-launcher platform; its community forum carries a user request titled 建议增加划词朗读 (add select-and-read-aloud), placing it in the tr |
local-only |
active |
|
proprietary (f |
M |
| Quill (Community-Access) |
win |
Screen-reader-first writing, review and document-intelligence environment for Windows with guided diagnostics and format workflows. |
unknown |
? |
42 |
MIT |
M |
| RDAccess (Remote Desktop Accessibility for NVDA) |
win |
NVDA add-on that carries screen-reader output across a remote-desktop session, so the remote machine's accessibility information is spoken by the loca |
local-only |
? |
15 |
GPL-2.0 |
M |
| rdev (Narsil) |
win/mac/lin |
Rust library to listen for and send keyboard/mouse events on macOS, Windows and Linux. |
local-only |
active |
738 |
MIT |
M |
| RDP clipboard redirection as a selection-acquisition route (meth |
win/mac/lin |
Over RDP the clipboard is shared between guest and host by default, so a copy-then-speak reader running on the LOCAL machine can read text selected in |
local-only |
? |
|
n/a |
D |
| Read Text Extension for LibreOffice / Apache OpenOffice (jimholg |
win/mac/legacy |
Office-suite extension that reads the SELECTION in Writer, Calc, Draw, Impress or Web Writer, or the clipboard contents, by handing the text to an ext |
hybrid |
? |
26 |
none declared |
M |
| ReadAny |
win/mac/lin/ios/and |
Cross-platform e-book application (Tauri desktop plus React Native mobile) that opens EPUB, PDF, MOBI, AZW, AZW3, FB2, FBZ, CBZ, TXT and UMD, and can |
hybrid |
active |
2162 |
NOASSERTION (G |
M |
| Remote Incident Manager (RIM) — Pneuma Solutions |
win/mac |
Cross-platform remote desktop built so blind, low-vision and sighted technicians can drive the same session on equal terms; the accessibility is in th |
hybrid |
? |
|
commercial |
V |
| remoteSpeechControl (NVDA add-on) |
win |
NVDA add-on that manages speech during remote sessions — optionally mutes the controlled machine so only the local box speaks; uses NVDA's bundled _re |
local-only |
? |
3 |
GPL-2.0 |
M |
| renpy-nvda-bridge |
win |
'Enables RenPy to output speech and braille via NVDA' — game-engine to screen-reader bridge. |
local-only |
unknown |
0 |
unknown |
L |
| RustDesk |
win/mac/lin/and/ios/web |
Self-hostable remote-desktop application positioned as an alternative to TeamViewer/AnyDesk/Splashtop [vendor]; clipboard sync between endpoints is pa |
hybrid |
active |
120 |
AGPL-3.0 |
L |
| RVC WebUI |
win/lin |
Train a retrieval-based voice-conversion model from under 10 minutes of voice data. |
local-only |
? |
37245 |
MIT |
M |
| SAPI-POC / SAPI-Bridge (AceCentre) |
win |
Design + proof-of-concept for a Windows SAPI DLL that forwards every speak call to a background pipe service holding TTS engines warm in memory. |
hybrid |
? |
1 |
open source (A |
M |
| scr-ocr |
win |
Windows area-screenshot tool that lays recognized text over the screenshot as real selectable text — a local imitation of macOS Live Text, built on El |
local-only |
? |
0 |
MIT |
M |
| Screenshot OCR / SnapOCR / MiniSnip / HushSnap / SnipText / Long |
win/lin/mac |
A large cluster of small hotkey-to-clipboard OCR utilities found repeatedly during the sweep: press a hotkey, drag a box, get the text on the clipboar |
local-only |
? |
0 |
mixed: some MI |
L |
| ScreenTranslate |
win |
Lightweight Windows screen OCR and translation utility. |
unknown |
? |
2 |
none declared |
V |
| ServantVoice (winget) |
win |
Voice application published under the ServantVoice publisher in winget-pkgs. |
unknown |
unknown |
|
unknown |
L |
| ShareX |
win |
Windows capture suite with region capture and built-in OCR among many post-capture actions; no text-to-speech feature found in its README. |
local-only |
active |
39 |
GPL-3.0 |
M |
| Simple TTS Reader |
win |
SourceForge directory entry: 'a small utility that reads text from your clipboard using Microsoft Speech API. Whenever you copy any text, the app inst |
local-only |
unknown |
|
unknown |
L |
| SnapX |
win/lin/mac |
Cross-platform fork/successor effort of the ShareX capture tool. |
local-only |
active |
998 |
|
L |
| SofTalk |
win |
Long-running Japanese desktop reading app built on AQUEST's AquesTalk; its developer CNCC announced on 2022-07-23 that support for the AquesTalk middl |
local-only |
dormant |
|
freeware |
L |
| Speakonia (CFS-Technologies, Chris Schuster) |
win |
Free Notepad-like SAPI 4/5 front end from 2002, now permanently frozen: the vendor FAQ opens 'PLEASE NOTE: Development of this program has been discon |
local-only |
discontinued |
|
freeware for p |
M |
| SpeechCore (still-standing88) |
win/lin |
C++ cross-platform speech library; the fourth abstraction layer named in Prism's README. |
local-only |
dormant |
1 |
|
L |
| Sprint Plus (Jabbla, distributed in France by Cimis) |
win |
Belgian-origin reading and writing aid for DYS users sold in the French market; positioned as 'Aide à la lecture'. [lang: French, Dutch; English UI: y |
local-only |
commercial-l |
|
proprietary |
L |
| stt2tts-mcp |
win/mac/lin |
Local-first STT/TTS MCP server with hot-swappable engines (faster-whisper, piper, kokoro, coqui, ollama, lmstudio, openai). |
hybrid |
? |
2 |
MIT |
M |
| Talon Voice |
win/mac/legacy |
Cross-platform hands-free computer control by voice, eye tracking and noise — a full alternative input stack. |
local-only |
? |
|
Closed source; |
L |
| Text2Speech (SourceForge, C#) |
win |
'Text2Speech is a small and easy to use Text To Speech (TTS) application written in C#. It uses the Microsoft .NET Framework 2.0 to run.' Distinct fro |
os-native |
unknown |
|
GPL-2.0 (SF ca |
V |
| text2speech.org scraping function for AutoHotkey |
win |
An r/AutoHotkey regular's WinHttpRequest function that queues text on text2speech.org and plays the result, with session-cookie handling and a tooltip |
cloud-API |
unknown |
|
n/a |
M |
| textsnap |
win/mac/lin |
Snap any image, screenshot or webpage into plaintext with one command — CPU-only ONNX/PaddleOCR, no GPU, no cloud. |
local-only |
? |
179 |
MIT |
M |
| textspeech and txtreader (SourceForge minor Windows tier) |
win |
Two minimal SourceForge Windows entries found while probing the older freeware tier: 'Enjoy text to speech application for windows' (needs .NET 3.5) a |
os-native |
unknown |
|
unknown (no SF |
V |
| trinity (zhubby) |
win/mac/lin |
Rust/egui always-on-top desktop assistant: Ctrl+Shift+T translates the current selection with DeepL, Ctrl+Shift+V opens a clipboard history picker, ho |
cloud-API |
active |
1 |
MIT |
M |
| TTS-WebUI Ignition (winget) |
win |
Launcher/installer package for the TTS-WebUI project, in the winget community repo. |
unknown |
unknown |
|
unknown |
L |
| UIAComWrapper (TestStack) |
win |
COM-to-.NET adapter for the Windows Automation API 3.0 COM interfaces. |
local-only |
dormant |
82 |
none |
M |
| Umi-OCR |
win/lin |
Free offline OCR software with screenshot OCR, batch image import, PDF recognition and QR handling; no speech output found in the README. |
local-only |
? |
46 |
MIT |
M |
| UniversalSpeech |
win |
'Make popular screen readers speak in your application' - an older C++ abstraction over screen readers plus direct SAPI/native synthesis, with a Pytho |
local-only |
dormant |
48 |
MIT |
M |
| vik-ma/screenshot-OCR |
win/lin |
Desktop front-end for Tesseract that lets you mark a section of the screen instead of loading an image file. |
local-only |
? |
10 |
GPL-3.0 |
M |
| Vision Assistant Pro (NVDA add-on) |
win |
NVDA add-on wrapping Google Gemini as an in-screen-reader copilot, including instant translation of selected text, dictation and CAPTCHA solving. |
cloud-API |
? |
|
unknown |
L |
| vision-access-nvda-addon |
win |
NVDA add-on that generates descriptive narration of on-screen graphics for blind and low-vision users. |
unknown |
? |
1 |
none declared |
V |
| VisionAssistantPro |
win |
NVDA add-on adding AI vision (image/screen description), translation, dictation and CAPTCHA solving. |
cloud-API |
? |
47 |
GPL-2.0 |
L |
| VoiceBroker (AceCentre) |
win |
Attempt at a Windows SAPI bridge — a Python COM server registering as a SAPI voice so that any SAPI-consuming app can use online/neural voices. |
hybrid |
? |
3 |
open source (A |
M |
| VoiceLink (ManveerAnand/VoiceLink) |
win |
Windows shim that exposes Kokoro and other open-source voice models as system voices so apps like Thorium Reader, Edge and Narrator can use them. |
local-only |
? |
21 |
MIT |
V |
| VoiceWave LocalCore (winget) |
win |
Voice-related package published under the VoiceWave publisher in winget-pkgs. |
unknown |
unknown |
|
unknown |
L |
| vs-read-aloud |
win |
Adds 'Read Aloud Selected Text' support to Visual Studio. |
os-native |
? |
0 |
check repo |
L |
| w32uiautomation (hnakamur) |
win |
Go bindings for Windows UI Automation; README marks it unmaintained. |
local-only |
archived |
48 |
MIT |
M |
| whkd (LGUG2Z) |
win |
Hotkey daemon for Windows, config-file driven, from the komorebi author. |
local-only |
active |
920 |
MIT-ish (NOASS |
M |
| Windows Magnifier reading |
win |
The built-in Windows magnifier can read text on screen aloud from a chosen start point, using the accessibility text layer rather than OCR. |
os-native |
? |
|
proprietary (b |
L |
| Windows Narrator |
win |
The screen reader built into Windows; it reads UI elements and text, and offers AI-generated image descriptions, but no documented user-invoked OCR of |
hybrid |
? |
|
proprietary (b |
L |
| Windows PowerToys — Text Extractor |
win |
Microsoft PowerToys utility: a global shortcut lets you drag a box over any screen region and OCRs the pixels straight to the clipboard. |
os-native |
? |
110 |
MIT (PowerToys |
D |
| Windows Snipping Tool — Text Actions |
win |
The built-in Windows 11 snipping tool recognizes text in a capture and lets you copy it; no speech. |
os-native |
? |
|
proprietary (b |
V |
| Windows UI Automation + SAPI 5 / OneCore (platform stack) |
win |
The Windows halves of the stack every Windows reader stands on: UIA (and legacy MSAA/IAccessible2) for reading the UI tree, SAPI 5 and the OneCore/Nat |
os-native |
? |
|
proprietary (p |
D |
| Windows.Media.Ocr (Windows OCR API) |
win |
The OS-provided OCR engine on Windows 10 and later, free and offline, driven by installed language packs — the recognizer behind PowerToys Text Extrac |
os-native |
? |
|
proprietary (p |
M |
| wolfmanstout/screen-ocr |
win/lin/mac |
Python library to perform OCR on portions of the screen with a choice of backends (Tesseract, WinRT, EasyOCR); a library, not an app. |
local-only |
? |
50 |
Apache-2.0 |
M |
| WordTalk |
win |
Free Microsoft Word add-in that speaks the document, paragraph, sentence or word from the cursor with word-by-word highlighting; the vendor's site sta |
os-native |
discontinued |
170,000 |
free (propriet |
M |
| wsay |
win |
'Windows say' CLI; v1.5.0 added clipboard playback and piping specifically in response to an r/AutoHotkey thread asking for 'say text from clipboard'. |
local-only |
active |
172 |
BSD-3-Clause |
M |
| WYNN / WYNN Wizard / WYNN Reader (Freedom Scientific) |
win |
Freedom Scientific's literacy-side scan-and-read product for learning disabilities, sold alongside JAWS and OpenBook and positioned against Kurzweil 3 |
local-only |
discontinued |
|
proprietary co |
L |
| zhuguohui/PageReader · lixiaowang/PDFReciter · johnsmith2078/pdf |
win/mac/lin/and |
Chinese document-scoped readers: word-by-word read with auto-scroll (PageReader), drag-a-box over PDF text to read it (PDFReciter), a PyQt6 PDF reader |
hybrid |
active |
1111 |
MPL-2.0 (Color |
M |
| Zotero ZoTTS |
win/mac/lin |
Zotero 7 plugin: inside a Reader tab Ctrl/Cmd+S speaks the selected text, the selected annotations' text, or the whole paper when nothing is selected, |
unknown |
dormant |
208 |
AGPL-3.0 |
M |
| Говорилка (Govorilka) |
win |
Russian freeware text-to-speech reader in circulation since 1999: reads text aloud, writes the reading to WAV/MP3 at raised speed with size-based spli |
local-only |
dormant |
|
freeware |
M |
| עלמה רידר / Alma Reader (עלמגו / Almago) |
win/legacy |
Hebrew reading software built on Hebrew and English linguistic analysis; Tel Aviv University's law library describes it as aimed at users with learnin |
local-only |
commercial-l |
|
proprietary |
M |
| 文字朗读神器 (Microsoft Store, publisher site zonboapp.com) |
win |
Microsoft Store app whose Chinese listing describes a copy-then-press-the-button workflow: '在其它任何软件,直接复制想朗读的文字内容,点朗读按钮即可自动开始朗读' (in any other software |
cloud-API |
unknown |
|
proprietary |
V |
| Chrome extension: Talkie (Chrome) |
ext |
Chrome Web Store read-aloud/text-to-speech extension surfaced in the store's own search index. |
hybrid |
unknown |
|
unknown |
L |
| Chrome extension: Text to Speech (TTS) (Chrome) |
ext |
Chrome Web Store read-aloud/text-to-speech extension surfaced in the store's own search index. |
hybrid |
unknown |
|
unknown |
L |
| Chrome extension: Text to Speech AI Voice Generator |
ext |
Chrome Web Store read-aloud/text-to-speech extension surfaced in the store's own search index. |
hybrid |
unknown |
|
unknown |
L |
| Chrome extension: Text to Speech TTS AI Reader |
ext |
Chrome Web Store read-aloud/text-to-speech extension surfaced in the store's own search index. |
hybrid |
unknown |
|
unknown |
L |
| Chrome extension: Voice Out - Read Aloud Text |
ext |
Chrome Web Store read-aloud/text-to-speech extension surfaced in the store's own search index. |
hybrid |
unknown |
|
unknown |
L |
| Firefox add-on: BeeLine Reader |
ext |
"BeeLine's color gradient makes reading faster/easier fo[r some readers]" [vendor]. |
hybrid |
active |
|
unknown |
V |
| Firefox add-on: ChatGPT Reader & Transcriber |
ext |
"Free AI text to speech (tts) and speech to text (stt) w[orkflow]" [vendor]. |
hybrid |
active |
|
unknown |
V |
| Firefox add-on: Circle reader |
ext |
Reader-mode extension that extracts page content for distraction-free reading [vendor]. |
hybrid |
active |
|
unknown |
V |
| Firefox add-on: Gemini Reader |
ext |
"Natural Gemini AI text to speech (TTS) for the web, PDF[s]" [vendor]. |
hybrid |
active |
|
unknown |
V |
| Firefox add-on: Google Reader: Free Natural AI Text to Speech |
ext |
"Natural Gemini AI text to speech (TTS) for the web, PDF[s]" [vendor]. |
hybrid |
active |
|
unknown |
V |
| Firefox add-on: Intelligent Speaker |
ext |
"Intelligent Speaker: smart reader running on high-profile tt[s engines]" [vendor]. |
hybrid |
active |
|
unknown |
V |
| Firefox add-on: Native text to speech (tts) |
ext |
"Website and PDF text to speech reader. Uses installed v[oices]" [vendor]. |
hybrid |
active |
|
unknown |
V |
| Firefox add-on: Picture Reader |
ext |
"Picture Reader is a browser plug-in that extracts all t[ext from images]" [vendor]. |
hybrid |
active |
|
unknown |
V |
| Firefox add-on: Read Aloud ff |
ext |
"Reads text marked by the user aloud." [vendor]. |
hybrid |
active |
|
unknown |
V |
| Firefox add-on: Talkie |
ext |
"Select text on any web page, and have the computer read [it aloud]" [vendor]. |
hybrid |
active |
|
unknown |
V |
| Firefox add-on: Text to Speech (TTS) |
ext |
"Text to Speech is a text to speech engine with natural [voices]" [vendor]. |
hybrid |
active |
|
unknown |
V |
| orator-chrome-extension |
ext |
Catalogued from a single sentence in MyronKoch/orator-macos's README describing a browser-extension sibling that reads web pages with Kokoro and Super |
unknown |
unknown |
|
unknown |
L |
| AceCentre TextAloud (iOS) |
ios |
Swift iOS app that reads text out sentence by sentence, paragraph by paragraph or word by word. |
os-native |
? |
11 |
open source (A |
L |
| Amazon Kindle apps — Assistive Reader / VoiceView |
ios/and/legacy |
Amazon's in-app text-to-speech for Kindle books, with real-time highlighting and speed control on iOS, Android, Mac and Fire tablets. |
os-native |
? |
|
Proprietary |
V |
| Dolphin EasyReader |
ios/and/win/ext |
Free accessible-book reader that connects to about 50 talking-book libraries, and can also read text pasted in from the clipboard — but the free tier |
hybrid |
commercial-l |
|
Free app with |
M |
| Elocance |
ios/and |
French mobile application, free to download on iOS and Android, that converts documents and other content into synthesised speech; catalogued by the F |
hybrid |
unknown |
|
freemium |
V |
| KNFB Reader |
ios/and/win |
National Federation of the Blind's app that photographs printed text and converts it to speech. |
hybrid |
? |
|
Proprietary, p |
L |
| Learning Ally |
ios/and/legacy |
US non-profit audiobook library for print-disabled students, largely human-narrated rather than synthesised, delivered through its own reading apps. |
cloud-API |
commercial-l |
|
Membership/sub |
M |
| Listen2 Reader |
ios |
AppleVis directory entry, vendor copy: 'Listen2 runs neural voice models locally on your device, giving you natural-sounding speech without ongoing co |
local-only |
commercial-l |
|
proprietary |
V |
| Lumyeye |
ios/and |
A subscription app pitched explicitly as the replacement for reading-machine hardware — the commercial shape that is eating this category. |
hybrid |
unknown |
|
Subscription, |
V |
| Readify: AI Natural Read Aloud |
ios |
AppleVis directory entry for an iOS reader supporting PDF, EPUB, TXT, MOBI and AZW. |
unknown |
commercial-l |
|
proprietary |
V |
| Seeing AI / Envision / Lookout / Google Lens |
ios/and |
Camera-first mobile OCR readers that photograph text in the physical world (or an image) and speak it; several also describe scenes. |
hybrid |
? |
|
proprietary (S |
V |
| SpeakCamera |
ios |
Pair of free iPhone Shortcuts: 'SpeakCamera' turns on voice announcements for the camera so text the camera sees is read aloud, and 'StopSpeech' stops |
os-native |
dormant |
2 |
none declared |
V |
| Speech Central |
ios/mac/win/and |
Commercial cross-platform reading app (iOS, macOS, Windows, Android) marketed for visual impairment and dyslexia, covering web pages, documents and eb |
hybrid |
? |
|
Proprietary, f |
L |
| Voice Dream Reader |
ios/mac |
Accessibility-first document reader for Apple platforms with synced word highlighting and broad document import; long-standing dyslexia-community stap |
hybrid |
? |
|
Proprietary, s |
L |
| AccessLint screenreaders (Auto-VO / VoiceOver.js) |
mac |
A Node CLI and library that starts VoiceOver from the command line, drives it with AppleScript, and dumps every announcement as text. |
local-only |
? |
183 |
MIT |
M |
| Aftertone |
mac/win/lin |
On-device daemon that speaks a short summary after a coding agent (Cursor, Claude Code) answers, using Supertonic ONNX. |
local-only |
? |
11 |
MIT |
L |
| agent-desktop |
mac/win/lin |
A native desktop-automation CLI that exposes OS accessibility trees as structured JSON with deterministic element refs, built for AI agents rather tha |
local-only |
active |
1 |
Apache-2.0 |
L |
| Apple Books — Speak Screen / VoiceOver reading (built-in) |
mac/ios |
Apple ships no dedicated in-app 'read this book aloud' button in Books; reading is done through the OS Spoken Content / VoiceOver layer. |
os-native |
? |
|
Proprietary (b |
L |
| Apple Speech Synthesis Provider Audio Unit |
mac/ios/legacy |
AppleVis forum post announcing Apple's API that lets third-party synthesisers register as system voices for accessibility features, shipped with iOS 1 |
os-native |
active |
|
Apple platform |
L |
| Apple Vision framework (VNRecognizeTextRequest) / VisionKit Live |
mac/ios |
Apple's on-device text recognition API, the engine behind Live Text and behind every macOS OCR utility in this list (TRex, Textinator, macOCR, ocrit, |
os-native |
? |
|
proprietary (p |
M |
| aria-at-automation-driver |
mac/win |
W3C WebSocket server letting clients observe what a screen reader enunciates and simulate user input. |
local-only |
? |
10 |
NOASSERTION |
M |
| ax-kit / AXKit (Akazm) |
mac |
Fork-derived Swift wrapper for the macOS accessibility client APIs (originally forked from AXSwift). |
local-only |
dormant |
3 |
none |
M |
| AXSwift |
mac |
A Swift wrapper over the macOS Accessibility client API, giving typed access to attributes like the focused element and its selected text. |
local-only |
dormant |
413 |
MIT |
M |
| BetterPopupTranslateSelection.spoon (rshlin) |
mac |
Hammerspoon Spoon that shows a popup with a context menu when text is selected, for translation. |
local-only |
dormant |
0 |
|
L |
| BetterSwiftAX (beeper) |
mac |
Swift wrapper around the macOS AX APIs, maintained by Beeper. |
local-only |
active |
5 |
none |
M |
| BetterTouchTool |
mac |
macOS automation app that since v5.177 ships a predefined 'OCR / Recognize / Extract Text from Clipboard Contents / Image' action, documented as combi |
os-native |
? |
|
proprietary, p |
M |
| Cheese! OCR |
mac |
macOS hotkey OCR that recognizes a selected screen area entirely on-device via Apple Vision and copies the text. |
os-native |
? |
|
proprietary |
V |
| Clicknow |
mac |
Commercial macOS selection tool: select text in any app and get AI translation, explanation, summary or search - named by MoePeek's README as the robu |
cloud-API |
commercial-l |
|
proprietary (m |
V |
| clipboard-tts (khiet) |
mac |
macOS command-line script that reads the clipboard with pbpaste, synthesizes it locally with Kokoro, and plays it through mpv over mpv's JSON IPC sock |
local-only |
active |
0 |
None declared |
M |
| clipper (wincent) |
mac/lin |
'Clipboard access for local and remote tmux sessions' [vendor] — a listener on the local machine that accepts text over a socket or TCP and places it |
local-only |
active |
687 |
BSD-2-Clause |
M |
| CursorBounds (Aeastn) |
mac |
Swift package that retrieves the on-screen position and bounds of the text caret in macOS apps via the Accessibility API - the primitive a floating 's |
local-only |
active |
120 |
none |
M |
| debrief (rs07-git) |
mac |
macOS floating read-along panel for Claude Code responses, with word highlighting and attention chimes; the text is handed to it by a hook rather than |
local-only |
active |
1 |
MIT |
M |
| Desktop Reader (hlindquist/bookreaderpackages) |
mac/win |
Packaged desktop application described by its author as a screen-region OCR reader with AI text-to-speech, distributed as a macOS arm64 DMG and a Wind |
unknown |
active |
0 |
none declared |
L |
| DFAXUIElement (DevilFinger) |
mac |
Objective-C helper for driving AXUIElement from macOS apps. |
local-only |
dormant |
61 |
none |
M |
| Dolphin EasyReader |
mac |
"Global links to over 50 libraries... Access magazines, newspapers and periodicals or import files and read clipboard text." [vendor] |
os-native |
commercial-l |
|
proprietary |
V |
| Drafts — Speak Selection |
mac/ios |
The Drafts editor has an Editor > Speak Selection menu item that opens a speech interface with pause/resume. |
os-native |
? |
|
proprietary |
V |
| FluidVoice (altic-dev) |
mac |
macOS dictation app with on-device speech-to-text and AI enhancement - the speech-to-text counterpart that Murmur was written as the inverse of. |
local-only |
active |
9451 |
GPL-3.0 |
M |
| get-selected-text (yetone) |
mac/win/lin |
Tiny Rust library that obtains the currently selected text on macOS, Windows and Linux behind one call — the reusable L1 primitive underneath a large |
local-only |
dormant |
209 |
unknown/NOASSE |
M |
| gruut |
mac/win/lin |
Tokenizer, text cleaner and phonemizer for many languages. |
local-only |
? |
331 |
MIT |
M |
| Hammerspoon |
mac |
macOS Lua automation framework with global hotkey binding and a built-in speech module, commonly paired with an external Vision-OCR binary to assemble |
os-native |
? |
16 |
MIT |
M |
| hammerspoon-penguin-click (KirkAlton-Class7) |
mac |
Hammerspoon Spoon adding Linux-style primary selection and middle-click paste to macOS. |
local-only |
active |
0 |
|
L |
| hifi-gan-bwe |
mac/win/lin |
Bandwidth extension for audio (upsampling speech quality). |
local-only |
? |
225 |
MIT |
M |
| Highlight AI |
mac/win |
Desktop AI assistant that reads on-screen context across every app so you can ask about what you are looking at without copying it into a chat box. |
cloud-API |
? |
|
commercial |
V |
| HotKey (soffes) |
mac |
Swift wrapper for Carbon RegisterEventHotKey global shortcuts on macOS. |
local-only |
dormant |
1077 |
MIT |
M |
| isimud |
mac |
"AI-native macOS menu bar text-to-speech and MCP server for agents" [vendor, crates.io description]; repo tagline "Have your agents speak to you". |
hybrid |
active |
0 |
MIT |
V |
| kAXSelectedTextAttribute / AXUIElementCopyAttributeValue (macOS |
mac |
The macOS AX client call every non-destructive macOS selection reader is built on: ask the focused AXUIElement for its kAXSelectedTextAttribute. Requi |
os-native |
active |
|
|
M |
| Keyboard Maestro |
mac |
macOS macro automation with published community macros that screenshot a user-selected area and OCR it; its shell-script action can pipe the result to |
os-native |
? |
|
proprietary, p |
V |
| KeyboardShortcuts (sindresorhus) |
mac |
Swift package adding user-customisable global keyboard shortcuts to a macOS app, with a recorder UI. |
local-only |
active |
2689 |
MIT |
M |
| kokoro-tts-mcp (scottschram) |
mac |
MCP server exposing Kokoro-82M on Apple Silicon via MLX to coding agents, able to speak the command line and the clipboard. |
local-only |
active |
3 |
MIT |
V |
| local-voice-reader |
mac |
Electron reading app for Apple Silicon macOS: paste text or open a .md/.txt file and it reads it aloud in a voice cloned from a 3-15 second recording, |
local-only |
active |
1 |
MIT |
M |
| Lue (paulilaaso/lue) |
mac/lin/legacy |
Terminal e-book reader for EPUB/PDF/DOCX/HTML/RTF/TXT/MD with word-level highlighting synchronized to speech, auto-scroll, and swappable TTS backends. |
hybrid |
? |
796 |
GPL-3.0 |
M |
| macOCR |
mac |
macOS tool/CLI that gets any text on your screen into the clipboard using Apple's Vision OCR. |
os-native |
? |
2 |
none declared |
M |
| macOS Accessibility (AX) API |
mac |
The AXUIElement C API that VoiceOver, Read & Speak, and third-party Mac selection readers all use to ask the focused app what its selected text is. |
os-native |
? |
|
proprietary (p |
D |
| macOS AppleScript clipboard-hack selection capture (System Event |
mac |
The question's own code: back up the clipboard, send Cmd+C via System Events, delay 1, read, delay 1, restore - the asker reports 'This doesn't wo |
local-only |
active |
|
CC-BY-SA (answ |
M |
| macOS Services "Add to Music as Spoken Track" (formerly iTunes) |
mac |
Answer's workaround for the missing pause: select text in a browser, right-click, 'add to iTunes as spoken track', then use the media keys - a native |
os-native |
active |
|
n/a |
M |
| macos-accessibility-client (x3ro) |
mac |
Rust wrapper around the macOS accessibility-client APIs; get-selected-text's README points at it for the permission prompt. |
local-only |
active |
31 |
none |
M |
| macos_accessibility_client (ahkohd) |
mac |
Node.js wrapper around the same macOS accessibility-client APIs, for Electron-shaped apps. |
local-only |
dormant |
7 |
none |
M |
| Magnet (Clipy) |
mac |
Global-hotkey library for macOS from the Clipy clipboard-manager project. |
local-only |
active |
450 |
MIT |
M |
| Marker (Mazide) |
mac |
Swift menu-bar app that 'Watches text selections system-wide via the Accessibility API' and stores them in its own buffer; README states 'Strict separ |
local-only |
active |
4 |
MIT |
M |
| mekedron/ocr |
mac |
Single-binary Go tool for macOS: capture a screen region, recognize with Apple Vision, pipe to pbcopy or wire to a Hammerspoon hotkey. |
os-native |
? |
3 |
MIT |
M |
| mkhd (Miigon) |
mac |
Layer-based hotkey daemon for macOS. |
local-only |
dormant |
17 |
MIT |
M |
| Multi OCR (Alfred workflow) |
mac |
"Run OCR on screenshots, images, and PDFs" [vendor]. |
local-only |
active |
|
unknown |
V |
| Narrly: Read Aloud PDF & Text |
mac |
"Reads PDF, EPUB, Word, RTF, Text and images... Speech is synthesized on your device... Supports TTS for 50+ languages." [vendor] |
os-native |
commercial-l |
|
proprietary |
V |
| node-get-selected-text (yetone) |
mac/win/lin |
The Node binding of the same cross-platform selected-text primitive, for Electron and Tauri-shaped desktop readers. |
local-only |
dormant |
41 |
none declared |
M |
| npm: @guidepup/virtual-screen-reader |
mac/win/lin |
"Virtual Screen Reader driver for unit test automation." [vendor]. |
unknown |
active |
|
unknown |
V |
| npm: speech-rule-engine |
mac/win/lin |
"A standalone speech rule engine for XML structures" [vendor] — turns MathML into spoken descriptions. |
unknown |
active |
|
unknown |
V |
| Obsidian plugin: AI Selection Toolbar |
mac/win/lin |
"AI-powered toolbar for selected text with TTS, translation, explanation, and wor[d lookup]" [vendor] |
hybrid |
active |
|
unknown |
V |
| Obsidian plugin: Aloud |
mac/win/lin |
"Speak text from your notes. Converts text to speech in real-time using lifelike [voices]" [vendor] |
hybrid |
active |
|
unknown |
V |
| Obsidian plugin: Apple TTS |
mac/win/lin |
"Read notes aloud using macOS native text-to-speech." [vendor] |
hybrid |
active |
|
unknown |
V |
| Obsidian plugin: Chatty |
mac/win/lin |
"Allows you to listen to your notes using text-to-speech. Uses the browser's buil[t-in engine]" [vendor] |
hybrid |
active |
|
unknown |
V |
| Obsidian plugin: Edge TTS |
mac/win/lin |
"Read notes aloud using Microsoft Edge Read Aloud API (free, high quality text-to-[speech])" [vendor] |
hybrid |
active |
|
unknown |
V |
| Obsidian plugin: Eleven Labs |
mac/win/lin |
"Turn your notes into text-to-speech audio files with Eleven Labs." [vendor] |
hybrid |
active |
|
unknown |
V |
| Obsidian plugin: Hermes TTS |
mac/win/lin |
"Generate lightweight audio from a markdown note and prepend timestamped metadata" [vendor] |
hybrid |
active |
|
unknown |
V |
| Obsidian plugin: Lens OCR |
mac/win/lin |
"Capture screen regions and digitize handwritten notes via native macOS and Windo[ws OCR]" [vendor] |
hybrid |
active |
|
unknown |
V |
| Obsidian plugin: Local Voiceover - Private TTS |
mac/win/lin |
"Speak selected text with local Inflect Micro v2 synthesis." [vendor] |
hybrid |
active |
|
unknown |
V |
| Obsidian plugin: Murmur |
mac/win/lin |
"Read your notes aloud with karaoke-style highlighting. Multiple TTS providers su[pported]" [vendor] |
hybrid |
active |
|
unknown |
V |
| Obsidian plugin: Note Reader |
mac/win/lin |
"Provides text-to-speech (TTS) for notes or clipped articles by reading them alou[d]" [vendor] |
hybrid |
active |
|
unknown |
V |
| Obsidian plugin: Open Reader |
mac/win/lin |
"Obsidian local TTS plugin for reading selected text and Markdown notes aloud" [vendor] |
hybrid |
active |
|
unknown |
V |
| Obsidian plugin: Read Along |
mac/win/lin |
Sentence-by-sentence read-aloud with the current sentence highlighted; exports notes to an offline-readable page [vendor, Chinese listing]. |
hybrid |
active |
|
unknown |
V |
| Obsidian plugin: Text Extractor |
mac/win/lin |
"A (companion) plugin to facilitate the extraction of text from images (OCR) and [PDFs]" [vendor] |
hybrid |
active |
|
unknown |
V |
| Obsidian plugin: Text to Speech (joethei) |
mac/win/lin |
"Hear your notes." [vendor] |
hybrid |
active |
|
unknown |
V |
| Obsidian plugin: Text2Audio |
mac/win/lin |
"Convert text to speech." [vendor] |
hybrid |
active |
|
unknown |
V |
| Obsidian plugin: Voice (chrisurf) |
mac/win/lin |
"Listen to your notes as natural speech with text-to-speech (TTS). Read notes alo[ud]" [vendor] |
hybrid |
active |
|
unknown |
V |
| OCR (Alfred workflow, alanhe) |
mac |
"Take a screenshot and copy its text to the clipboard" [vendor]. |
local-only |
active |
|
unknown |
V |
| OCR Light (Alfred workflow) |
mac |
"Copy screenshot text to the clipboard" [vendor]. |
local-only |
active |
|
unknown |
V |
| OCR Text Recognition: Textify |
mac |
Mac App Store OCR text-recognition utility (iSolid SPRL). |
unknown |
unknown |
|
proprietary |
L |
| ocrit |
mac |
Command-line utility that performs OCR on image files using Apple's Vision framework, outputting text or files. |
os-native |
? |
187 |
BSD-2-Clause |
M |
| Omnivox |
mac/win/lin |
'Omnivox is an Emacspeak TTS server designed to be cross platform.' |
unknown |
unknown |
|
unknown |
V |
| OneTalker |
mac/win/lin |
'OneTalker is a free, open-source Augmentative and Alternative Communication (AAC) desktop app' — 31 stars, the highest-starred TTS-relevant repo foun |
unknown |
unknown |
31 |
unknown |
V |
| OpenPhonemizer |
mac/win/lin |
Permissively-licensed espeak-compatible IPA phonemizer based on DeepPhonemizer. |
local-only |
? |
112 |
BSD-3-Clause-C |
M |
| OwlOCR |
mac |
macOS OCR app for screenshots, PDFs and local AI processing. |
local-only |
? |
|
proprietary, p |
V |
| owocr |
mac/win/lin |
"Multi-service, multi-platform optical character recognition" [vendor] — OCR daemon that can watch the clipboard or a screen area and feed recognised |
hybrid |
active |
283 |
GPL-3.0 |
V |
| phonemizer |
mac/win/lin |
Text-to-phoneme converter wrapping espeak-ng, festival and segments. |
local-only |
? |
1 |
GPL-3.0 |
M |
| PopClip (relationship anchor, already-known) |
mac |
Named alongside Clicknow in MoePeek's alternatives section as the macOS selection-action baseline; recorded here only as the traversal edge. |
local-only |
commercial-l |
|
proprietary |
L |
| PopClip extension: Say |
mac |
Official PopClip extension that speaks the selected text through the macOS say command; voice and rate can be overridden per extension setting, and |
os-native |
active |
|
see extension |
M |
| portable-translator (Zxx20061021) |
mac |
macOS desktop translator whose selection watcher tries AXSelectedText first and falls back to other macOS-specific APIs. |
hybrid |
? |
|
check repo |
M |
| Prizmo |
mac |
"Scanning application with Optical Character Recognition (OCR)" [vendor] — Creaceed's macOS scanning/OCR app. |
local-only |
commercial-l |
|
proprietary, p |
V |
| PyScreenReader (PyPI) |
mac/win/lin |
Cross-platform Python library wrapping the native accessibility APIs to collect the on-screen widget tree — the Python counterpart to pywinauto/FlaUI/ |
local-only |
? |
|
unknown |
L |
| qtspeech (Homebrew formula) |
mac/lin |
"Enables access to text-to-speech engines" [vendor] — the Qt Speech module. |
local-only |
active |
|
LGPL-3.0-only |
V |
| Raycast / Alfred / Hammerspoon / skhd / Karabiner-Elements (trig |
mac |
General macOS hotkey and launcher layers that supply L2 only — you bind them to a say, pbpaste | say, or CLI-reader invocation to assemble a sele |
os-native |
? |
12 |
mixed — Hammer |
M |
| Raycast extension: Easy OCR |
mac |
Raycast extension with one command: select a screen area, Tesseract extracts the text. [vendor, store manifest] |
local-only |
active |
|
see store list |
V |
| Raycast extension: Read My Screen |
mac |
Raycast extension that captures a screen region, a window, the full screen, a clipboard image, a local image file, or the active browser tab's text an |
cloud-API |
active |
62 |
not declared i |
M |
| Raycast extension: ScreenOCR |
mac |
Raycast extension doing local OCR of a captured screen area or the entire screen, with a language picker and a barcode mode. [vendor, store manifest] |
local-only |
active |
|
see store list |
V |
| raycast-tts |
mac |
Raycast extension for macOS with a local Python WebSocket server running Piper: you TYPE into its text area and it speaks each word as you press space |
local-only |
active |
0 |
none declared |
M |
| read-aloud.el |
mac/lin/win |
Emacs package that does live speech-to-TEXT transcription: it captures microphone audio, detects voice activity, and sends the transcript into an Emac |
hybrid |
active |
2 |
GPL-3.0-or-lat |
M |
| readio (hrhrng/readio) |
mac/lin/win |
Rust terminal e-book reader (EPUB/PDF/Markdown/text) with local-model read-aloud that highlights the spoken sentence and the current word. |
local-only |
? |
7 |
MIT |
M |
| Selection (pot-app) |
mac/win/lin |
The cross-platform selection-acquisition crate extracted from pot-desktop: get the text selected by the cursor, on X11, Wayland, Windows and macOS. |
local-only |
dormant |
70 |
GPL-3.0 |
M |
| selection-translator (huacius) |
mac |
Minimal macOS selection translator for English reading and learning: select a word or phrase, get a translation card with UK/US pronunciation display |
hybrid |
active |
2 |
MIT |
M |
| Shottr |
mac |
Free-to-use macOS screenshot app with a hotkey-driven OCR: press a hotkey, select an area, text is copied to the clipboard. |
os-native |
? |
|
proprietary (f |
V |
| Sinter |
mac/lin/win |
Research 'Accessible Remote Desktop Protocol': a remote-desktop system that transmits UI semantics rather than pixels so assistive technology works ov |
local-only |
? |
3 |
NOASSERTION |
M |
| skhd (asmvik) |
mac |
Hotkey daemon for macOS: a config file of key : command lines that runs any shell command on a global chord. Gives a Speak-Selection pipeline its ke |
local-only |
active |
8063 |
MIT |
M |
| skhd.zig (jackielii) |
mac |
Zig port of skhd for macOS. |
local-only |
active |
607 |
MIT |
M |
| SnapPop (gradinnovate) |
mac |
macOS background utility described as a PopClip alternative: it detects a text selection and shows a floating menu of quick actions such as copy and s |
local-only |
dormant |
1 |
Apache-2.0 |
M |
| Speak Text Pro |
mac |
"Convert web pages text into speech! Get your Mac talking with this app." [vendor] |
os-native |
dormant |
|
proprietary |
V |
| speak-app (eliefrancis5) |
mac |
169-line macOS menu-bar app in Python: click the menu-bar icon, TYPE text into a box, press Speak, and it shells out to say with a chosen voice and |
local-only |
active |
0 |
none declared |
M |
| speak-selected-text-sublime |
mac |
Sublime Text plugin that pipes the editor selection to the Mac say command. |
os-native |
? |
11 |
MIT |
L |
| Speech Central: Text to Speech (Mac App Store listing) |
mac |
Paid Mac App Store text-to-speech reader; surfaced in the macSoftware search for 'screen reader' at $9.99 with 0 US ratings recorded. |
unknown |
commercial-l |
|
proprietary |
L |
| star (Speaking Terminal Access Reader) |
mac/win/lin |
"star is an accessible, GUI-first document reader and Markdown authoring tool with built-in text-to-speech" [vendor, PyPI long description]. |
local-only |
active |
|
GPL-3.0-or-lat |
V |
| star / star-reader (leavesofgrass) |
mac/win/lin |
'Speaking Terminal Access Reader' — a GUI-first accessible document reader that opens PDF, Word, EPUB, PowerPoint, web pages and spreadsheets, reads t |
local-only |
? |
, |
GPL-3.0-or-lat |
M |
| swiftmac |
mac |
'Swiftmac TTS server for emacspeak' — a Swift-written Emacspeak speech server for macOS, hosted only on sourcehut. |
os-native |
unknown |
|
unknown |
V |
| SwiftTranslate (blackTDE) |
mac |
macOS global translation utility explicitly inspired by PopClip and Raycast Translate: select text anywhere, hotkey, translated popup. |
hybrid |
active |
0 |
MIT |
M |
| Text Scraper (johnbean393) |
mac |
macOS menu-bar app that OCRs the text on your displays on Cmd+Ctrl+C and presents it grouped for selection and copying; it has no speech of its own an |
local-only |
dormant |
10 |
none declared |
M |
| Text to Speech PDF Reader |
mac |
"The app will highlight words as it reads and scroll the page automatically... Choose a primary and secondary voice." [vendor] |
os-native |
commercial-l |
|
proprietary |
V |
| Textinator |
mac |
macOS status-bar app that automatically runs Vision OCR on every screenshot you take and puts the text on the clipboard. |
os-native |
? |
203 |
MIT |
M |
| textra |
mac |
macOS CLI converting images, PDFs and audio to text using Apple's APIs. |
os-native |
? |
755 |
MIT |
M |
| TextSniper |
mac |
Proprietary macOS menu-bar OCR: hotkey, drag over any screen region — including video frames and locked UI elements — and the text lands on the clipbo |
os-native |
? |
|
proprietary, p |
V |
| TRex |
mac |
macOS menu-bar OCR: hotkey, drag a region, text on the clipboard, using Apple's on-device Vision framework. |
os-native |
? |
1 |
MIT |
M |
| tts.nvim |
mac/lin |
Neovim plugin that speaks the visual selection, the current line, paragraph or section from the buffer, with a play/stop/queue keymap set and a choice |
hybrid |
active |
11 |
None declared |
M |
| Viz (alienator88) |
mac |
Swift macOS snip-to-text tool (v2.3.3) built on the Apple Vision framework: extract text, QR codes, barcodes and colours from a screen snippet on a cu |
local-only |
dormant |
526 |
NOASSERTION (G |
M |
| Voice Reader . |
mac |
"Reads copied text from emails, apps, notes, etc. Supports reading PDF files. Reads website content. Works even without an internet connection." [vend |
os-native |
commercial-l |
|
proprietary |
V |
| VoxBar (+ voxbar-ob Obsidian plugin) |
mac |
Local macOS TTS app driven from Obsidian: read the selection, the whole note, or fetch and read the article at a selected source URL. |
local-only |
? |
0 |
check repo |
M |
| VS Code extension: Text to Speech Preview |
mac/win/lin |
"Preview selected text or the active document with local syst[em TTS]" [vendor]. |
local-only |
active |
|
unknown |
V |
| WebOutLoud: Web Text to Speech |
mac |
Reads web-page text aloud inside its own browser view. [vendor listing] |
os-native |
commercial-l |
|
proprietary |
V |
| Xpop-lineage note: PopClip alternatives that do NOT speak |
mac |
Bare macOS selection-popup experiment surfaced by an 'alternative to PopClip' README search; no speech capability found. |
local-only |
active |
0 |
none declared |
L |
| Ability (TBosak/ability) |
legacy |
Browser extension bundling accessibility controls, including TTS, for users with varying degrees of ability. |
local-only |
? |
22 |
none declared |
V |
| ABiSee Eye-Pal Solo / Eye-Pal ACE / Zoom-Ex |
legacy |
The previous generation of camera reading machines; ABiSee was absorbed by Freedom Scientific and the Eye-Pal line is marked discontinued by retailers |
local-only |
discontinued |
|
n/a |
M |
| Access Lens — Kane, Frey & Wobbrock, CHI 2013 |
legacy |
Computer-vision gesture tracking that lets a blind user perform accessible touch gestures on paper documents and physical objects that have no screen |
local-only |
dormant |
|
research proto |
M |
| Access Overlays — Kane, Morris, Perkins, Wigdor, Ladner & Wobbro |
legacy |
Three software overlays (edge projection, neighbourhood browsing, touch-and-speak) that give blind users spatial access to large touch screens. |
local-only |
dormant |
|
research proto |
M |
| Adapted Digital Exams (CALL Scotland / SQA) |
legacy |
Scottish exam-board guidance on running SQA digital question papers, including which text readers to use and how to configure Adobe Reader around them |
unknown |
active |
|
Public sector |
M |
| adult-sharks/spatial-screen-reader · leelaeloo/Senior-OCR-Projec |
legacy |
Two Korean student/hobby projects: a Chrome extension using OpenCV as a visual-assistance aid, and 읽어드림, a document-reading app aimed at older users. |
hybrid |
unknown |
2 |
unknown |
L |
| Anki-TTS-Edge (EllisMorrow/Anki-TTS-Edge) |
legacy |
Edge-TTS tool that both generates Anki card audio and doubles as an immersive reader with real-time highlighting, click-to-play and navigation. |
cloud-API |
? |
14 |
NOASSERTION (u |
V |
| Apache Guacamole (clientless HTML5 remote desktop gateway) — cli |
legacy |
Browser-based RDP/VNC/SSH gateway whose manual documents a clipboard text area: 'Text copied/cut within Guacamole will appear here' and 'text that is |
hybrid |
active |
3 |
Apache-2.0 |
M |
| Artificial Analysis Speech Arena Leaderboard |
legacy |
Second crowd-voted TTS leaderboard on Hugging Face Spaces, comparing systems using each provider's own native voices. |
cloud-API |
? |
|
see Space |
L |
| Assistive-Webdriver |
legacy |
Tool for automating end-to-end web application tests driven through a real screen reader in a VM. |
local-only |
? |
28 |
see repo |
L |
| Audiobook Read-Along plugin for KOReader (stradichenko/audiobook |
legacy/and/lin |
KOReader plugin adding offline TTS with synchronized word highlighting, automatic page turns, and Bluetooth control on e-ink devices. |
local-only |
? |
108 |
AGPL-3.0 |
M |
| AuraLens |
legacy |
Flutter prototype assisting blind and low-vision users with real-time scene understanding and text reading (OCR), backed by Gemini models. |
cloud-API |
? |
7 |
see repo |
L |
| Auris (nikhilprasanth/Auris) |
legacy |
Offline audiobook-style reader for EPUB/PDF/TXT with local OmniVoice TTS, per-character voices, and synced text highlighting. |
local-only |
? |
20 |
MIT |
M |
| Bhola (2022) — 'Effect of Text-to-speech Software on Academic Ac |
legacy |
Small pretest-posttest experiment giving one group four months of text-to-speech software and comparing achievement against an untreated control. |
unknown |
active |
|
open access |
M |
| Bonifacci, Colombini, Marzocchi, Tobia & Desideri (2022) — 'Text |
legacy |
Experiment measuring whether text-to-speech changes how often a student's attention drifts off task during reading, alongside comprehension. |
unknown |
active |
|
open access (W |
M |
| Bookshare Reader |
legacy/ios/and |
Bookshare's own free reading tool for its accessible-book library, usable in a web browser, on iOS and Android, and on Alexa-enabled devices. |
cloud-API |
active |
|
Free to Booksh |
M |
| Brèthes, Cavalli, Denis-Noël, Melmi, El Ahmadi, Bianco & Colé (2 |
legacy |
Regression study of which cognitive skills predict text reading fluency versus text reading comprehension in dyslexic and non-dyslexic university stud |
unknown |
active |
|
open access (F |
M |
| calibre-tts-ebook-viewer (christineye) |
legacy |
Pre-built-in-TTS calibre viewer plugin using Windows SAPI 5, with paragraph highlighting, click-a-paragraph-to-start select mode, and customizable hot |
os-native |
? |
31 |
none declared |
M |
| CALL Scotland text-reader catalogue |
legacy |
Practitioner-written catalogue of 14 Windows and cross-platform text readers, each with its own page giving the literal keystroke sequence a user perf |
unknown |
active |
|
Public sector |
M |
| Capti Voice |
legacy/win/mac/ios/and |
Named in an Apple StackExchange answer as a free reader that 'highlights, pauses, etc.' and syncs to phone; the same answer reports it 'can crash a bi |
hybrid |
commercial-l |
|
proprietary (f |
L |
| Cavalli, Colé, Brèthes, Lefèvre, Lascombe & Velay (2019) — 'E-bo |
legacy |
Study of how the reading medium affects long-text comprehension specifically in dyslexic adults. |
unknown |
active |
|
paywalled (Spr |
L |
| ChatGPT_ReadAloud |
legacy |
Chrome extension restoring a Read Aloud button in ChatGPT with auto-play and playback controls. |
unknown |
? |
3 |
Apache-2.0 |
M |
| Chen, Hung & Jian (2026) — 'The effects of text-to-speech on rea |
legacy |
Eye-tracking experiment comparing silent reading with text-to-speech across dyslexia, ADHD-with-reading-difficulty, and typically developing groups. |
unknown |
active |
|
open access (S |
M |
| Citrix ICA clipboard redirection policy (Restrict session clipbo |
legacy |
Citrix documents session-to-client clipboard as an administrator-controlled ICA policy: 'When the Restrict session clipboard write setting is Enabled, |
unknown |
commercial-l |
|
proprietary (C |
V |
| Citrix virtual channel allow list (Virtual Apps and Desktops 210 |
legacy |
From Citrix Virtual Apps and Desktops 2109 onward a virtual-channel allow list restricts third-party virtual channels by default; the NVDA rdAccess do |
unknown |
active |
|
n/a |
M |
| Clicker and DocsPlus (Crick Software) |
legacy |
Literacy-support writing environments for schools with built-in speech feedback; named by Adapted Digital Exams among the commercial tools that add te |
hybrid |
commercial-l |
|
Commercial; sc |
M |
| Clicky (farzaa) / Clicky for Windows |
legacy |
AI companion that sits by the cursor: hold a hotkey, ask about what is on your screen, and it sees the screen and talks back — with a community Window |
hybrid |
? |
7 |
MIT |
M |
| Clinton-Lisell & Litzinger (2026) — 'Decoding digital reading: a |
legacy |
Network meta-analysis ranking paper against computers, tablets, e-readers and smartphones for reading comprehension, with scrolling as the moderator. |
unknown |
active |
|
open access (S |
M |
| Clinton-Lisell (2021) — 'Listening Ears or Reading Eyes: A Meta- |
legacy |
Meta-analysis comparing comprehension when the same material is read versus listened to, across age groups. |
unknown |
active |
|
paywalled (SAG |
M |
| Clinton-Lisell (2023) — 'Reading while listening meta-analysis' |
legacy |
Meta-analysis of audio-assisted reading — seeing the text and hearing it at the same time, which is exactly the Speak-Selection condition. |
unknown |
active |
|
open access (O |
M |
| cliphist (sentriz) |
legacy |
Wayland clipboard-history daemon driven by wl-paste --watch; README states clipboard content is 'preserved byte-for-byte' and it has 'No concept of |
local-only |
active |
1520 |
GPL-3.0 |
M |
| clipmenu (cdown) |
legacy |
X11 clipboard-history manager built on clipnotify + dmenu. |
local-only |
active |
1253 |
Public Domain |
M |
| clipnotify (cdown) |
legacy |
Tiny X11 utility that blocks until the clipboard or PRIMARY selection changes, then exits - the event primitive a shell-script selection reader loops |
local-only |
dormant |
256 |
Public Domain |
M |
| clipvault (Rolv-Apneseth) |
legacy |
Wayland clipboard-history manager, cliphist-inspired. |
local-only |
active |
115 |
MIT |
M |
| Creative TextAssist / Texto'LE (DECtalk on Sound Blaster ASP) |
legacy |
DECtalk-derived TTS bundled with Sound Blaster 16 and AWE32 cards, with a Windows 'TextAssist TextReader' front end - and a hard dependency on the car |
local-only |
discontinued |
|
proprietary, b |
M |
| Do GUI Agents Believe Their Eyes? Diagnosing State-Belief Relian |
legacy |
Diagnostic benchmark that measures, with paired single-channel interventions, whether a multimodal agent's belief about what is on screen comes from t |
unknown |
active |
|
unknown |
M |
| DSpeech (Dimio) |
legacy |
Italian freeware SAPI front end with speech-recognition-driven branching, long distributed as a portable app; the author's own site is unreachable fro |
local-only |
dormant |
|
freeware |
M |
| dxhd (dakyskye) |
legacy |
X11 hotkey daemon whose config is a shell script. |
local-only |
dormant |
100 |
GPL-3.0 |
M |
| EASTIN (European Assistive Technology Information Network) |
legacy |
Federated European assistive-product database aggregating national AT registries under an ISO 9999 classification. |
unknown |
active |
|
EU-funded netw |
L |
| Edge-endpoint proxies on Cloudflare Workers (DIYgod/cloudflare-e |
legacy |
Serverless relays that expose the Edge Read Aloud endpoint (often in an OpenAI-compatible shape) from a Worker, mainly to route around region blocks a |
cloud-API |
? |
199 |
see repos |
? |
| Emmabuntüs accessibility work |
legacy |
A Linux distribution's accessibility programme, integrating screen-reading and speech synthesis choices at distro level. |
local-only |
? |
|
open (distro) |
L |
| Envision Glasses (Envision, on Google Glass Enterprise Edition 2 |
legacy |
Head-worn camera reader built on discontinued Google hardware — Instant Text, Scan Text and Batch Scan, plus an LLM question-answering mode over what |
hybrid |
commercial-l |
|
Hardware purch |
M |
| Explain Code Audio (JetBrains Marketplace plugin) |
legacy |
JetBrains IDE plugin that explains selected code and plays the explanation back as text-to-speech audio. |
unknown |
? |
|
Not verified |
L |
| Faster Text-to-Speeches: Enhancing Blind People's Information Sc |
legacy |
Experiment comparing raising the speech rate of one voice against running two or three voices at once, to find which lets a listener scan faster witho |
local-only |
dormant |
|
n/a (study) |
M |
| Foliate (johnfactotum/foliate) |
legacy |
GTK e-book reader for Linux whose read-aloud is delegated to speech-dispatcher with output modules such as espeak-ng. |
local-only |
? |
8614 |
GPL-3.0 |
M |
| Freedom Scientific RUBY 10 HD Speech (Vispero) |
legacy |
Handheld video magnifier with an OCR speech mode — magnification first, reading aloud second. |
local-only |
commercial-l |
|
Commercial har |
L |
| GNOME Shell issue #5559 - "Select to speak a specific UI element |
legacy |
Open upstream feature request from a visually impaired user asking GNOME for select-to-speak with word highlighting and a Ctrl-to-stop key. |
local-only |
active |
|
n/a |
M |
| GNOME Speaks |
legacy |
GNOME Shell extension adding dictation and 'text-to-speech readback' powered by Azure Speech Services, wired through a D-Bus service to a four-project |
cloud-API |
active |
1 |
GPL-3.0 |
M |
| Google Chrome — Listen to this page / Reading mode (built-in) |
legacy/ext |
Chrome's own read-aloud: 'Listen to this page' on Android and a listen control inside desktop Reading mode, with text highlighting and auto-scroll. |
hybrid |
? |
|
Proprietary (b |
V |
| Google Docs — Accessibility > Speak > Speak selection (built-in) |
legacy |
Google Docs can speak the current selection via the Accessibility menu, once screen-reader support is enabled in the document. |
hybrid |
? |
|
Proprietary (G |
L |
| Grid 3 (Smartbox / thinksmartbox) |
legacy |
Windows AAC (augmentative and alternative communication) application with a Computer Control feature for driving the whole desktop, and a speech layer |
hybrid |
commercial-l |
|
proprietary, c |
L |
| Guidepup |
legacy |
Screen-reader DRIVER for test automation — programmatically drives VoiceOver and NVDA from Node.js and captures what they speak. |
local-only |
? |
500 |
MIT |
L |
| Guidepup Playwright |
legacy |
Playwright integration for the Guidepup screen-reader automation library. |
local-only |
? |
80 |
MIT |
L |
| Harpo Software / Speech2Go (harposoftware.com) |
legacy |
A live retail counter for Nuance and IVONA desktop voices — the channel that still sells voices whose original vendors have shut their own doors. |
local-only |
commercial-l |
|
Per-voice prop |
M |
| hkd (aaronamk) |
legacy |
Display-server-agnostic hotkey daemon; README: 'Works in Xorg, Wayland, and the TTY (using libevdev)'. Requires adding your user to the input group. |
local-only |
active |
25 |
MIT |
M |
| Hui & Godfroid (2025) — 'Listening, Reading, or Both? Rethinking |
legacy |
Registered report testing whether reading while listening improves comprehension; it found the opposite of its own preregistered hypothesis. |
unknown |
active |
|
open access |
M |
| HumanWare Victor Reader Stream 3 |
legacy |
Tactile-keypad DAISY/audio player with on-device TTS for text files — you load content into it; it has no camera and no view of your computer. |
hybrid |
commercial-l |
|
Commercial har |
M |
| hyprland-global-shortcuts-v1 (Hyprland Wayland protocol) |
legacy |
Hyprland's own Wayland protocol for global shortcuts, used underneath its GlobalShortcuts portal backend; a registered shortcut appears in `hyprctl gl |
os-native |
active |
|
|
L |
| Interaction Proxies for Runtime Repair and Enhancement of Mobile |
legacy |
Strategy that inserts a proxy between an application's real interface and the interface a person perceives, so a third party can re-map an interaction |
local-only |
dormant |
|
research proto |
M |
| ISO/IEC 13066 family — 'Information technology — Interoperabilit |
legacy |
The international standard family that defines what an application must expose so an assistive technology can read it — Part 1 the general requirement |
unknown |
active |
|
paid standard |
L |
| iSpeak It (ZappTek) |
legacy |
Mac utility (2004-2006 era) that turned documents, web pages and RSS feeds into MP3/AAC tracks in iTunes using Mac OS X's own text-to-speech, with a r |
local-only |
discontinued |
|
proprietary sh |
M |
| J-Say and J-Dictate (Hartgen Consultancy) |
legacy |
Commercial JAWS scripting products that bridge JAWS with Dragon speech recognition so a blind user can dictate and hear results; listed for DSA fundin |
local-only |
commercial-l |
|
Commercial add |
M |
| Jump Desktop (RDP, VNC, Fluid) |
legacy |
macOS/iOS/Android remote-desktop client supporting RDP, VNC and its own 'Fluid Remote Desktop' protocol; its own support article confirms the client r |
hybrid |
commercial-l |
|
proprietary (p |
M |
| kaccessible (KDE) |
legacy |
Historical KDE QAccessibleBridgePlugin providing focus tracking and a screen reader inside the Qt accessibility bridge; superseded when the bridge mov |
local-only |
discontinued |
4 |
LGPL |
M |
| KDE Discuss t/18444 - "Method to trigger speech dispatcher to re |
legacy |
Jul 2024 KDE thread asking exactly the project question - Meta+Space passing the selection to spd-say - and receiving no reply at all. |
local-only |
dormant |
|
n/a |
M |
| KDE Klipper D-Bus clipboard read (qdbus org.kde.klipper /klipper |
legacy |
A Russian-language Linux forum thread on reading the primary selection under Wayland; the poster reports reading the buffer straight out of Klipper ov |
local-only |
unknown |
|
n/a |
L |
| KDE post-Jovie speech story (QtSpeech + Okular + KMouth + spd-sa |
legacy |
Jeremy Whiting's 2021 statement of what replaced Jovie on KDE, verbatim: notifications go through QtSpeech via each application's notification configu |
local-only |
active |
|
|
M |
| Keelor, Creaghead, Silbert, Breit & Horowitz-Kraus (2023) — 'Imp |
legacy |
Five-condition experiment separating the effect of text-to-speech itself from the effect of one of its presentation features — synchronised word highl |
unknown |
active |
|
paywalled (Spr |
M |
| Kindle device text-to-speech (2009-2013 era) and its VoiceView s |
legacy |
Amazon shipped read-aloud on Kindle e-readers from 2009, removed it from later hardware, then reintroduced audio only through a separately purchased U |
hybrid |
active |
|
proprietary, d |
M |
| kokoro-tts-chrome-extension (paul-rinaldi) |
legacy |
Chrome extension wiring the Kokoro TTS model into the browser (no description or license published). |
unknown |
? |
0 |
none declared |
L |
| KWtype |
legacy |
'Virtual keyboard input tool for KDE Wayland' - the KWin-specific answer where wtype's wlroots protocol is unavailable. |
local-only |
active |
5 |
MIT |
M |
| lectern (Acumane) |
legacy |
Listen to PDFs with natural TTS and read-along text prompts. |
local-only |
? |
11 |
see repo |
L |
| Lecteur PDF accessible (RGAA Checker) |
legacy |
Free in-browser accessible PDF reader with reflowable text, dyslexia-friendly typography profiles, OCR for scanned documents, dual view and annotation |
unknown |
? |
|
free (terms un |
L |
| Lectura (Dolphin service menu, espeak) |
legacy |
Named by a KDE Discuss responder as 'a dolphin service menu called lectura that uses espeak to speak text from a file using the context menu'. |
local-only |
active |
|
n/a |
L |
| lefthk (leftwm) |
legacy |
Hotkey daemon in Rust from the LeftWM project. |
local-only |
active |
29 |
Rust |
M |
| Loquendo TTS Director |
legacy |
The Java, multi-platform, client-server prompt-authoring suite inside the Loquendo TTS SDK - a listen-and-edit tool for tuning recorded-sounding promp |
local-only |
discontinued |
|
proprietary, p |
M |
| Markdown Read Aloud (Robin-Reiche/markdown-read-aloud) |
legacy |
VS Code extension that renders a Markdown file as a reader view and speaks it with Edge neural voices, highlighting the spoken sentence and auto-scrol |
cloud-API |
? |
3 |
MIT |
M |
| Markit and Talkit — Shi, Zhao & Azenkot, UIST 2017 |
legacy |
Toolkit pair for attaching audio annotations to 3D-printed models: a sighted maker marks regions and writes text, then a blind user touches the printe |
local-only |
dormant |
|
research proto |
M |
| Mayari (BoltzmannEntropy) |
legacy |
Native macOS document read-aloud and audiobook workspace for PDF, DOCX and EPUB. |
local-only |
? |
5 |
see repo |
L |
| Microsoft Agent |
legacy |
The animated-character runtime that gave a generation of Windows apps a speaking assistant over SAPI 4; Microsoft's own support article states it 'has |
os-native |
discontinued |
|
proprietary, b |
M |
| Microsoft Reader (.lit) |
legacy |
Microsoft's pre-Kindle ebook reader for PC and Pocket PC, whose PC version carried an optional text-to-speech plug-in; announced for discontinuation i |
local-only |
discontinued |
|
proprietary, f |
M |
| Microsoft Speech API 4 (SAPI 4) as a substrate |
legacy |
The 1998 speech interface that every Windows reader in this historical set was built on, and the reason so many of them are unusable today - modern sc |
os-native |
discontinued |
|
proprietary, p |
M |
| Microsoft Word / Outlook / PowerPoint / OneNote — Speak and Read |
legacy |
Office ships two distinct features: Speak reads only the text you select; Read Aloud reads the whole document from the cursor with a floating player. |
os-native |
? |
|
Proprietary (M |
V |
| MorseWriter / EyeCommander / FaceCommander (AceCentre) |
legacy |
Alternative-input assistive tools from the same charity: eye-movement tracking, facial-gesture control, and Morse-to-text entry. |
local-only |
? |
58 |
open source (A |
L |
| MS-RDPEA — Remote Desktop Protocol: Audio Output Virtual Channel |
legacy |
Microsoft's published RDP audio channel spec, described as the extension 'which transfers audio data from the server to the client' [measured] — the d |
unknown |
active |
|
Microsoft Open |
M |
| MS-RDPECLIP — Remote Desktop Protocol: Clipboard Virtual Channel |
legacy |
Microsoft's published RDP clipboard channel spec; abstract reads 'enables users to seamlessly transfer data via the system clipboard between applicati |
unknown |
active |
|
Microsoft Open |
M |
| MsEdgeTTS (Migushthe2nd) |
legacy |
Node/TypeScript client for the same Edge Read Aloud endpoint, for stacks that cannot shell out to Python. |
cloud-API |
? |
335 |
see repo |
? |
| neoreader |
legacy |
Screen reader for Neovim that reads code with language-aware structure, including infix operators in Haskell/Scala and Python AST analysis. |
local-only |
? |
29 |
see repo |
L |
| neru (y3owk1n) |
legacy |
Navigate your entire screen without touching the mouse. |
local-only |
? |
565 |
see repo |
L |
| NextUp — Cerence voice store (for TextAloud) |
legacy |
Cerence Vocalizer Embedded voices sold per-voice to consumers, with the catch that they are locked to NextUp's TextAloud application. |
local-only |
commercial-l |
|
Proprietary, p |
M |
| Nova (bigduu) / Screenhand / Orbination Desktop Vision |
legacy |
A cluster of MCP servers giving an LLM agent eyes and hands on the desktop — screenshots, Apple Vision or Windows OCR, UI Automation trees, and mouse/ |
hybrid |
? |
15 |
varies |
L |
| obsidian-speechify-reader (TheShiningVampire) |
legacy |
Obsidian plugin that hands the current note to the Speechify browser extension — Alt+A to listen, Alt+S to save to the Speechify library — with no API |
cloud-API |
? |
0 |
MIT |
V |
| obsidian-tts-kokoro (yuengling) |
legacy |
Fork of joethei's Obsidian Text to Speech plugin retargeted at the Kokoro model. |
unknown |
? |
0 |
GPL-3.0 |
D |
| obsidian-tts-reader (10x-oss) |
legacy |
Obsidian plugin that reads Markdown notes aloud starting from the current cursor position. |
unknown |
? |
0 |
MIT |
V |
| OCR - Image Reader (Chrome extension) |
legacy |
Chrome extension using tesseract.js to recognize text in images on a page, injecting the library on demand and removing it afterwards. |
local-only |
? |
|
not verified |
V |
| Odiofy |
legacy |
Named in one r/accessibility comment as 'a decent app which is free and can be downloaded and installed on your local machine'; I could not verify a p |
unknown |
unknown |
|
unknown |
L |
| openai-edge-tts (travisvn) |
legacy |
A local server that speaks the OpenAI /v1/audio/speech protocol but synthesises through the free Edge Read Aloud endpoint - a drop-in shim that makes |
hybrid |
? |
2 |
GPL-3.0 |
M |
| opendataloader-pdf |
legacy |
Open-source PDF parser producing AI-ready structured data, with automated PDF accessibility remediation. |
local-only |
? |
28 |
open source (s |
L |
| OpenGuider |
legacy |
Desktop AI companion that watches the screen, listens to the user's voice, and guides them step by step with spoken actions. |
hybrid |
? |
164 |
unknown |
L |
| OpenReader (richardr1126/openreader) |
legacy |
Self-hostable Next.js read-along document reader for EPUB/PDF/DOCX/TXT/MD with word-by-word highlighting derived from Whisper alignment, plus audioboo |
hybrid |
? |
489 |
MIT |
M |
| OpenWebTTS |
legacy |
Self-described open-source Speechify alternative: read PDFs and EPUBs with local models. |
local-only |
? |
68 |
MIT |
L |
| Optelec ClearReader+ (Basic / standard / Advanced) — Vispero |
legacy |
Portable platen-camera reading machine: press one button, it photographs the page and reads it aloud; the Advanced model drives an external monitor so |
local-only |
commercial-l |
|
Commercial har |
M |
| OrCam Read 3 |
legacy |
Handheld AI reading device: point it at printed or digital text — including a computer screen — press the button, and it reads aloud from that point; |
local-only |
unknown |
|
commercial |
V |
| OrCam Technologies — vision product line (MyEye, MyReader, Read |
legacy |
The vendor behind OrCam Read closed its entire low-vision division in July 2024 and pivoted to hearing; a purchase decision here is a decision about a |
local-only |
discontinued |
|
n/a |
M |
| org.freedesktop.portal.GlobalShortcuts |
legacy |
The D-Bus portal interface (documented at version 2) through which a Wayland application registers shortcuts that fire 'regardless of the focused stat |
os-native |
active |
|
|
M |
| [OSC 52 clipboard escape sequence (ESC ] 52 ; c ; base64)](https://terminfo.dev/extensions/osc-52-clipboard) |
legacy |
A terminal escape sequence that lets a program inside a remote session write to (and optionally read) the LOCAL terminal's system clipboard; the refer |
local-only |
active |
|
n/a — a contro |
M |
| Overlay Translator |
legacy |
No-root Android real-time screen translator: OCRs the screen (on-device or cloud), translates, overlays the result in place, and can speak it via TTS. |
hybrid |
? |
518 |
Apache-2.0 |
L |
| phonemizer.js |
legacy |
eSpeak NG phonemization in JavaScript. |
local-only |
? |
49 |
Apache-2.0 |
M |
| Pied |
legacy |
Flutter desktop GUI that installs and configures Piper as a Speech Dispatcher back-end on Linux, then downloads and manages voices for it. |
local-only |
active |
290 |
GPL-3.0 |
M |
| Piper Reader for Obsidian (dCO2/obsidian-piper-reader) |
legacy |
Obsidian plugin that POSTs the selected text to a local Python bridge which drives Piper in Docker over Wyoming TCP and plays the returned WAV. |
local-only |
? |
0 |
none declared |
M |
| pocket-tts-browser-extension (hiCozyty) |
legacy |
Browser extension front-ending the Pocket TTS engine, claiming ultra-fast free synthesis. |
unknown |
? |
0 |
MIT |
L |
| Prefab layers and prefab annotations — Dixon et al., UIST 2014 ( |
legacy |
Extension of the Prefab pixel-reverse-engineering line adding layered interpretation and annotation of recognised interface structure. |
local-only |
dormant |
|
research proto |
L |
| Prefab: Implementing Advanced Behaviors Using Pixel-Based Revers |
legacy |
System that recognises widgets from the pixels a toolkit painted, so behaviours can be added to applications built with any toolkit on any windowing s |
local-only |
dormant |
|
research proto |
M |
| primary-selection-unstable-v1 (Wayland protocol) |
legacy |
The Wayland protocol that carries the X11-style PRIMARY (highlight) selection. Its compositor coverage is far wider than data-control's, but it is a n |
os-native |
active |
|
MIT (Red Hat) |
M |
| PRISM (ethindp/prism) |
legacy |
Platform-agnostic Reader Interface for Speech and Messages — one API that refracts a string out to whichever screen reader or TTS backend is present, |
local-only |
active |
58 |
MPL-2.0 |
M |
| pulseaudio-network (ferdiu) |
legacy |
'A simple client-server program to easily share TCP pulseaudio sinks over the network' [vendor], packaged as RPM and DEB. |
local-only |
active |
0 |
MIT |
M |
| pwWebSpeak (The Productivity Works, later isSound) |
legacy |
1996 non-visual browser built on 'first order design' - HTML converted straight to structured audio via a rule base (the Tag Language Definition), byp |
local-only |
discontinued |
|
proprietary co |
M |
| python-global-shortcut-portal (marvin1099) |
legacy |
Pure-Python client for org.freedesktop.portal.GlobalShortcuts; description states it 'Lets any application register and receive global keyboard shortc |
local-only |
active |
0 |
AGPL-3.0 |
M |
| qt_wayland_globalshortcut_via_portal (slbtty) |
legacy |
Minimal Qt demonstration of registering a global shortcut on Wayland through the GlobalShortcuts portal. |
local-only |
dormant |
1 |
|
M |
| rclipd (pmkap) |
legacy |
Clipboard-manager daemon targeting compositors that implement ext/wlr-data-control. |
local-only |
active |
0 |
none |
M |
| RD Pipe (rd_pipe-rs) |
legacy |
'Windows Remote Desktop Services Dynamic Virtual Channel implementation using named pipes, written in Rust' [vendor] — the library the NVDA rdAccess a |
local-only |
active |
9 |
AGPL-3.0 (as r |
M |
| react-speech-highlight (albirrkarim) |
legacy |
React / vanilla-JS text-to-speech component that highlights the word and the sentence currently being spoken — the reusable implementation of the kara |
hybrid |
? |
188 |
unknown |
M |
| read-aloud-best-practices (Readium) |
legacy |
Documentation project recording best practices for implementing a read-aloud feature in reading apps. |
unknown |
? |
14 |
see repo |
L |
| read-aloud-local |
legacy |
Repository named read-aloud-local; carries no description or README summary. |
unknown |
? |
6 |
MIT |
L |
| read-aloud.el (gromnitsky) |
legacy |
Emacs package that speaks the word at point, the selected region, or the whole buffer through an external CLI TTS engine such as speech-dispatcher or |
local-only |
? |
35 |
MIT |
M |
| Read-It-Out (Spartan-71) |
legacy |
Reads any webpage or article aloud with natural AI voices. |
hybrid |
? |
2 |
see repo |
L |
| ReadEasy Evolve / Evolve ECO / Evolve MAX (VisionAid; sold by Hu |
legacy |
The other current standalone reading-machine family: fold-out camera arm over a document, one button, reads aloud — ECO is A4/13 MP, MAX is A3/18 MP. |
local-only |
commercial-l |
|
Commercial har |
V |
| Readest (readest/readest) |
legacy |
Cross-platform modern e-book reader with multilingual TTS and read-along narration that highlights text in step with EPUB 3 Media Overlays. |
hybrid |
? |
23223 |
AGPL-3.0 |
M |
| reading-for-listeners |
legacy |
Deep-learning accessibility application that turns PDFs into audio files, with OCR improvement and inflection-aware TTS. |
local-only |
? |
25 |
AGPL-3.0 |
L |
| Readiris Pro (IRIS / Canon) |
legacy |
Commercial OCR package on the DSA approved list, used to turn scanned or photographed print into text that a separate reader then speaks. |
local-only |
commercial-l |
|
Commercial; DS |
M |
| readium/speech |
legacy |
TypeScript library for implementing read-aloud on the Web, from the Readium ebook-standards project. |
local-only |
? |
21 |
see repo |
L |
| ReadSpeaker |
legacy |
Commercial text-to-speech service, largely sold to publishers to add read-aloud to their own websites and products. |
cloud-API |
? |
|
Proprietary, c |
L |
| remiforall/dys-play |
legacy |
French PWA reading aid for dyslexic users combining local OCR, speech synthesis, adapted fonts and a stated zero-data-collection design. [lang: French |
local-only |
active |
0 |
AGPL-3.0 |
M |
| Remmina |
legacy |
GTK remote-desktop client for RDP, VNC, SPICE, X2Go, SSH and plain HTTP; its RDP plugin carries a clipboard structure (rf_clipboard) in the published |
local-only |
active |
|
GPL-2.0-or-lat |
M |
| RFB / RFC 6143 ServerCutText + Extended Clipboard pseudo-encodin |
legacy |
The VNC wire protocol carries selection text, not only pixels: RFC 6143 section 7.6.4 ServerCutText says 'The server has new ISO 8859-1 (Latin-1) text |
unknown |
active |
|
IETF RFC (Info |
M |
| Rhasspy (v2) and Rhasspy 3 |
legacy |
'Offline private voice assistant for many human languages' [vendor] — carried an HTTP TTS API alongside STT and intent handling; the v3 rewrite is a s |
local-only |
archived |
2 |
MIT |
M |
| Robust Annotation of Mobile Application Interfaces in Methods fo |
legacy |
Methods for identifying the same screen and the same element across app versions, so accessibility annotations attached by third parties survive inter |
local-only |
dormant |
|
research proto |
M |
| Saladict / 沙拉查词 · 沙拉翻译 · Read Frog 陪读蛙 · DualRead |
legacy |
Chinese-market select-a-word dictionary and translation extensions covering 中英日韩法德西 with pronunciation playback; the reading is a pronunciation featur |
hybrid |
active |
|
mixed |
L |
| Say It |
legacy |
Free browser page that OCRs an uploaded or drawn image and speaks the recognized text. |
local-only |
? |
|
not verified |
V |
| Scanning for Digital Content: How Blind and Sighted People Perce |
legacy |
Journal extension testing the concurrent-speech scanning result across both blind and sighted listeners. |
local-only |
dormant |
|
n/a (study) |
M |
| Scanning Pens Ltd |
legacy |
The distributor that puts C-Pen hardware into UK/EU/US schools and workplaces — the reason C-Pen appears under several storefront names. |
unknown |
commercial-l |
|
n/a |
L |
| Screen Point-and-Read (Tree-of-Lens agent) |
legacy |
Research system for the 'ScreenPR' task: given a screenshot and a point the user indicates, a multimodal LLM agent reads out the content at that point |
local-only |
? |
31 |
see repo |
M |
| Screen Point-and-Read / Tree-of-Lens (ToL) agent — Fan et al., a |
legacy |
Research system defining the 'ScreenPR' task — given a screenshot plus a point the user indicated, generate a spoken-style description of the content |
hybrid |
dormant |
31 |
see repo |
M |
| Screen Recognition: Creating Accessibility Metadata for Mobile A |
legacy |
On-device model that detects UI elements from an iOS app's rendered pixels and generates accessibility metadata to feed VoiceOver where the app suppli |
local-only |
active |
|
not released ( |
M |
| Screen2AX — Muryn et al., arXiv:2507.16704 (MacPaw) |
legacy |
Framework that builds a tree-structured macOS accessibility hierarchy from a single screenshot using vision-language and object-detection models, to s |
local-only |
active |
32 |
see repo |
M |
| ScreenTrack — Hu & Lee, CHI 2020 |
legacy |
Software that screenshots the computer at regular intervals and turns the capture history into a time-lapse the user scrubs to find and re-open a docu |
local-only |
dormant |
|
research proto |
M |
| SelectON (emvaized/selecton-extension) |
legacy |
Browser extension for Chrome and Firefox that shows a configurable popup with actions whenever text is selected on a page. |
local-only |
active |
128 |
NOASSERTION |
M |
| selsync (Stoica-Mihai) |
legacy |
C daemon that mirrors PRIMARY into CLIPBOARD by speaking ext-data-control-v1, wlr-data-control v2 or X11 XFIXES directly (no wl-copy/xclip subprocess) |
local-only |
active |
0 |
MIT |
M |
| shotkey (phenax) |
legacy |
Small X hotkey daemon with modes and key chords. |
local-only |
active |
47 |
MIT |
M |
| Sikuli — Yeh, Chang & Miller, UIST 2009 |
legacy |
Search and automation of graphical interfaces by screenshot: take a picture of a button or icon, and use that image both as a help-system query and as |
local-only |
dormant |
|
research relea |
M |
| Slide Rule — Kane, Bigham & Wobbrock, ASSETS 2008 |
legacy |
Audio-based multi-touch interaction techniques that made a touch screen usable without sight — the ancestor of the touch-explore-then-confirm pattern |
local-only |
dormant |
|
research proto |
M |
| speakable (tollwerk) |
legacy |
Simple, privacy-friendly on-page screen-reader / TTS player built on the native browser Web Speech API. |
local-only |
? |
24 |
see repo |
L |
| speakers (OneNoted) |
legacy |
Local Linux TTS daemon plus a Speech Dispatcher bridge built on Qwen3-TTS, in Rust; targets Hyprland/niri/Wayland. |
local-only |
? |
3 |
none declared |
M |
| Speech Dispatcher network mode (SPEECHD_ADDRESS=inet_socket:HOST |
legacy |
The shipped spd-say man page documents 'SPEECHD_ADDRESS ... specifies TCP endpoint where speech-dispatcher is listening and to which spd-say should co |
local-only |
active |
|
GPL-2.0-or-lat |
M |
| Speech Kit for Obsidian (brittain9/speech-kit-obsidian-plugin) |
legacy |
Combined dictation, transcription, translation and note-listening plugin for Obsidian with a managed local model catalog (formerly Local Dictation). |
hybrid |
? |
11 |
MIT |
M |
| SpeechCore / UniversalSpeech / Tolk |
legacy |
Three older screen-reader abstraction libraries — the prior art that PRISM sets out to unify, still embedded in many accessible games and apps on Wind |
local-only |
dormant |
104 |
varies per pro |
L |
| speechd-el (brailcom/speechd-el) |
legacy |
Emacs speech and Braille output interface that routes Emacs output to Speech Dispatcher. |
local-only |
? |
16 |
GPL-3.0 |
M |
| Speechify |
legacy/mac |
Commercial cross-platform reading product (web, mobile, browser extension) with an OCR path for scanned/printed text. |
cloud-API |
? |
|
Proprietary, f |
L |
| speechify.nvim (HmZyy) |
legacy |
'speechify powered text-to-speech inside Neovim' [vendor, repo description] — editor-scoped read-aloud for the current buffer/selection. |
hybrid |
active |
1 |
not read |
V |
| SPICE VD_AGENT clipboard (spice-vdagent agent protocol) |
legacy |
SPICE's guest-agent protocol carries clipboard text both ways with a symmetric GRAB / RELEASE / REQUEST / CLIPBOARD message set, VD_AGENT_CLIPBOARD_UT |
unknown |
active |
|
SPICE project |
M |
| splash-damage (insidewhy) |
legacy |
Keyboard remapper daemon for Wayland via evdev/uinput with per-app exclusions. |
local-only |
active |
0 |
none |
M |
| Stormux (Linux for the blind on Raspberry Pi) |
legacy |
Accessible Linux distribution for Raspberry Pi presented at a German accessibility event; a sibling of the accessible-distro class already catalogued |
local-only |
unknown |
|
open source |
L |
| SUGILITE — Li, Azaria & Myers, CHI 2017 |
legacy |
Programming-by-demonstration system that automates arbitrary Android apps through the accessibility API, generalising a script from a single spoken co |
local-only |
dormant |
|
research proto |
M |
| sxhkd (baskerville) |
legacy |
X11 hotkey daemon: binds chords to shell commands via XGrabKey. X11 only. On Wayland it has no equivalent grab, which is the constraint recorded in th |
local-only |
active |
2951 |
BSD-2-Clause |
M |
| Tailo |
legacy |
Study-support tool carried on the DSA approved list in two categories, OCR and Research. |
cloud-API |
commercial-l |
|
Commercial sub |
L |
| Talkify |
legacy |
JavaScript text-to-speech library with text highlighting, for embedding reading into web pages. |
hybrid |
? |
240 |
none declared |
M |
| Task Mode: Dynamic Filtering for Task-Specific Web Navigation us |
legacy |
System that uses an LLM to filter a web page down to the elements relevant to a stated goal, so a screen-reader user does not traverse minutes of irre |
hybrid |
active |
|
unknown |
M |
| tesseract.js |
legacy |
Pure-JavaScript/WASM OCR for 100+ languages; the engine behind browser and Electron OCR tools such as scr-ocr. |
local-only |
? |
39 |
Apache-2.0 |
M |
| Text-to-speeches: Evaluating the Perception of Concurrent Speech |
legacy |
Earlier experiment in the same line, testing whether blind listeners can pick out the relevant channel among two, three or four simultaneous speech st |
local-only |
dormant |
|
n/a (study) |
M |
| Texthelp OrbitNote |
legacy |
Browser-based PDF workspace from Texthelp with read-aloud and annotation, bundled on the DSA list alongside EquatIO. |
cloud-API |
commercial-l |
|
Commercial sub |
M |
| Texthelp ReachDeck (formerly Browsealoud) |
legacy |
Website-owner-deployed accessibility toolbar that speaks page content to visitors; the reading capability is bought by the site publisher, not install |
cloud-API |
commercial-l |
|
Commercial sub |
M |
| TextyMcSpeechy |
legacy |
Toolkit for training custom Piper voice models from your own recordings or RVC voices, offline and even on a Raspberry Pi. |
local-only |
? |
697 |
see repo |
L |
| The accessibility of digital technologies for people with visual |
legacy |
Scoping review of first-hand accessibility experiences of blind and visually impaired users of digital technology, organised against the WCAG guidelin |
unknown |
active |
|
open access (C |
M |
| The Impact of Element Ordering on LM Agent Performance — Zhang e |
legacy |
Study of how much the ORDER in which on-screen elements are presented matters when an agent has only pixels and no hierarchy to inherit an order from. |
unknown |
active |
|
unknown |
M |
| The State of Modern AI Text To Speech Systems for Screen Reader |
legacy |
First-hand January 2026 evaluation by a screen-reader user who built NVDA add-ons for Supertonic and Kitten TTS and documents why modern neural TTS fa |
local-only |
? |
|
CC0 1.0 |
M |
| Thorium Reader (edrlab/thorium-reader) |
legacy/mac/lin |
Cross-platform Readium-based EPUB reading app whose stated accessibility approach is to work WITH NVDA, JAWS or Narrator rather than to embed its own |
local-only |
? |
2818 |
BSD-3-Clause |
M |
| Thunder-RJ (RJ Cooper & Associates) |
legacy |
Rebadged distribution of the free Thunder Windows screen reader offered by a US assistive-technology reseller. |
local-only |
active |
|
Free/low-cost |
M |
| TTS Arena V2 (TTS-AGI) |
legacy |
Crowd-sourced blind A/B benchmark for TTS models with an Elo leaderboard, on Hugging Face Spaces. |
cloud-API |
? |
|
see Space |
M |
| tts-wrapper (willwade) |
legacy |
Unified Python interface over many online and offline TTS engines, with a documented feature matrix covering streaming and word-boundary events per en |
hybrid |
? |
39 |
MIT |
M |
| TTSReader for Obsidian (sundy-li/obsidian-ttsreader) |
legacy |
Obsidian plugin that reads the selection, the current note, or pasted text, with a 'Read the selected text' command in the palette and the editor righ |
hybrid |
? |
0 |
MIT |
M |
| TTSVoicesAvailable (AceCentre) |
legacy |
Small API (plus a Streamlit front-end) that enumerates the TTS voices available across providers, including offline voices. |
hybrid |
? |
1 |
open source (A |
L |
| Understanding Blind and Low Vision Users' Attitudes Towards Spat |
legacy |
Formative study with a custom desktop screen reader that adds spatial input and output to web navigation, and a report on how blind and low-vision use |
local-only |
dormant |
|
research proto |
M |
| US state AT Act lending-library catalogues (TechOWL/myatprogram, |
legacy |
Network of state Assistive Technology Act programmes that lend software and devices for trial before purchase; TechOWL alone states its Pennsylvania l |
unknown |
active |
|
US government- |
L |
| vim-oscyank |
legacy |
'A Vim plugin to copy text through SSH with OSC52' [vendor] — makes an editor yank inside a remote session land on the local clipboard. |
local-only |
active |
740 |
BSD-2-Clause |
M |
| vim-piper |
legacy |
Vim plugin that speaks buffer or selected text through Piper. |
local-only |
? |
29 |
see repo |
L |
| Virtual Screen Reader (guidepup) |
legacy |
A simulated screen reader implemented in JavaScript for unit-testing accessibility, with no real assistive technology involved. |
local-only |
? |
400 |
MIT |
L |
| VizLens — Guo et al., UIST 2016 |
legacy |
Mobile application that lets a blind user photograph a physical interface (microwave, kiosk, thermostat), has crowd workers label it once, then speaks |
hybrid |
dormant |
|
research proto |
M |
| VMware / Omnissa Horizon clipboard redirection (client-side bloc |
legacy |
Horizon's client GPO catalogue exposes a setting described as 'Whether block clipboard redirection to client side when client doesn't support audit' — |
unknown |
commercial-l |
|
proprietary (O |
L |
| Voicing (ainure-git/voicing) |
legacy |
VS Code extension that reads SELECTED TERMINAL OUTPUT aloud at ~2x with pause/stop, using the OS local voice, plus dictation into the terminal. |
os-native |
? |
1 |
MIT |
M |
| Vorleser XL (in-media KG / MEDIAKG) |
legacy |
German commercial Vorleseprogramm marketed for reading PDF, Word, ePub, web pages, e-mails and plain text and for converting them to MP3; a distributo |
local-only |
commercial-l |
|
proprietary |
V |
| VSCode Read Aloud Text (azu/vscode-read-aloud-text) |
legacy |
VS Code extension that speaks the whole document, from the cursor, or just the selection, using OS TTS, and highlights the text currently being read. |
os-native |
? |
15 |
MIT |
M |
| Wayland cross-application selection read - the privileged-operat |
legacy |
On Wayland, reading another application's selection from a background process is a privileged operation, not a normal one. wl-clipboard's man page BUG |
os-native |
active |
|
|
M |
| wayland-clipboard-listener (Decodetalkers) |
legacy |
Rust library and CLI that emits an event on every Wayland clipboard or PRIMARY change via the data-control protocols. |
local-only |
active |
15 |
GPL-3.0 |
M |
| web-speech-recommended-voices (Readium) |
legacy |
Curated data set of recommended voices for the Web Speech API, per language and platform. |
local-only |
? |
65 |
see repo |
L |
| WebAnywhere — Bigham, Prince & Ladner, W4A 2008 |
legacy |
Screen reader that ran inside the browser with nothing installed on the machine, so a blind user could get speech on a computer that was not theirs. |
hybrid |
dormant |
|
open-source re |
M |
| What Frustrates Screen Reader Users on the Web: A Study of 100 B |
legacy |
Survey and time-diary study of what actually goes wrong for screen-reader users, and how much of their working time it costs. |
unknown |
dormant |
|
n/a (study) |
L |
| Windows App (Microsoft) — macOS client, formerly Microsoft Remot |
legacy/win/web |
Microsoft's own cross-platform feature matrix marks 'Clipboard - bidirectional' supported on macOS and describes it as 'Redirect the clipboard on the |
hybrid |
commercial-l |
|
proprietary (f |
M |
| Wisp AI Assistant (SunnyLich) |
legacy |
'Wisp - A hotkey-driven AI overlay for your desktop. Press a key, pick an intent, and Wisp reads the right context, then streams an answer' [vendor, r |
hybrid |
active |
8 |
not read |
V |
| wl-clipboard-rs (YaLTeR) |
legacy |
Rust reimplementation of the wl-clipboard logic as a library plus wl-clip-persist-style binaries; speaks the data-control protocols directly instead o |
local-only |
active |
507 |
Apache-2.0 |
M |
| wlr-data-control-unstable-v1 (Wayland protocol) |
legacy |
The wlroots-originated predecessor of ext-data-control, same purpose (privileged clipboard/PRIMARY read by an unfocused client); still the version man |
os-native |
active |
|
MIT-style (pro |
M |
| Wood, Moxley, Tighe & Wagner (2018) — 'Does Use of Text-to-Speec |
legacy |
Meta-analysis of studies testing whether text-to-speech and read-aloud tools improve reading comprehension for students with reading difficulties. |
unknown |
active |
|
paywalled (SAG |
M |
| wtype |
legacy |
'xdotool type for wayland' - the wlroots-compositor route for synthetic typing, named as the sway/Hyprland option in Handy's paste-methods documentati |
local-only |
active |
549 |
MIT |
M |
| Wyoming protocol (OHF-Voice/wyoming, formerly rhasspy/wyoming) |
legacy |
'Peer-to-peer protocol for voice assistants' [vendor] — a JSONL-over-TCP framing used to expose TTS (and STT/wake-word) services as network endpoints |
local-only |
active |
387 |
MIT |
M |
| wyoming-satellite |
legacy |
'Remote voice satellite using Wyoming protocol' [vendor] — a thin networked audio endpoint that plays synthesised speech on a device separate from the |
local-only |
archived |
1 |
MIT |
M |
| wyoming_openai |
legacy |
'OpenAI-Compatible Proxy Middleware for the Wyoming Protocol' [vendor] — bridges Wyoming clients to any OpenAI-shaped speech endpoint. |
hybrid |
active |
204 |
Apache-2.0 |
M |
| xdg-desktop-portal-cosmic |
legacy |
Per-compositor portal backend for COSMIC. Its src/ directory contains access.rs, app.rs, buffer.rs, documents.rs, file_chooser.rs, screencast.rs and s |
os-native |
active |
102 |
|
M |
| xdg-desktop-portal-gnome |
legacy |
Per-compositor portal backend for GNOME / Mutter. GlobalShortcuts is tracked as work item 47 in the GNOME GitLab project; a GNOME Discourse thread sta |
os-native |
active |
|
|
L |
| xdg-desktop-portal-hyprland |
legacy |
Per-compositor portal backend for Hyprland. Its src/portals directory contains GlobalShortcuts.cpp, InputCapture.cpp, Screencopy.cpp and Screenshot.cp |
os-native |
active |
473 |
|
M |
| xdg-desktop-portal-kde |
legacy |
Per-compositor portal backend for KDE Plasma / KWin. Its src/ directory contains globalshortcuts.cpp, inputcapture.cpp, remotedesktop.cpp, screencast. |
os-native |
active |
83 |
|
M |
| xdg-desktop-portal-lxqt |
legacy |
Per-compositor portal backend for LXQt. Interface coverage not inspected. |
os-native |
active |
31 |
|
L |
| xdg-desktop-portal-wlr |
legacy |
Per-compositor portal backend for wlroots compositors (Sway, river and relatives). Its README states 'Currently it only implements the following porta |
os-native |
active |
725 |
|
M |
| xdotool (jordansissel) |
legacy |
X11 fake keyboard/mouse input and window management; the tool a shell-script selection reader uses to synthesise Ctrl+C. |
local-only |
active |
3837 |
BSD-3-Clause |
M |
| Xpra (persistent remote applications for X11; shadow mode for X1 |
legacy |
Remote-application/screen-sharing system whose usage docs list clipboard and audio as independently switchable forwarded features, with documented inv |
local-only |
active |
2 |
GPL-2.0 |
M |
| XuGaoFeng-Victor/mcp-point-reader |
legacy |
An MCP server that captures text via the clipboard and screenshots and offers translation and text-to-speech over it. [lang: Chinese + English descrip |
hybrid |
unknown |
2 |
unknown |
L |
| Zoomax low-vision devices with OCR read-aloud |
legacy |
Chinese low-vision magnifier maker whose devices include an OCR-to-speech function; the budget end of the reading-hardware market. |
local-only |
unknown |
|
Commercial har |
L |
| Zotero TTS Reader (zcyisiee/zotero-tts-reader) |
legacy |
Zotero 7 plugin that speaks text you select in Zotero's PDF reader through the OpenAI TTS API, with citation stripping and a side-panel player. |
cloud-API |
? |
5 |
none declared |
M |
| Zotero-TTS-Plugin (Echo-Lian) |
legacy |
A second, separate Zotero plugin adding a TTS function to the reference manager. |
unknown |
? |
0 |
none declared |
L |
| Edge read-aloud endpoint reuse cluster (CN): guozhigq/ReadAloud |
web/win/mac/lin |
Chinese-language projects that wrap Microsoft Edge's 大声朗读 speech endpoint: a PWA front end (138 stars), a Cloudflare-Workers/Vercel/Docker HTTP forwar |
cloud-API |
active |
229 |
mixed open sou |
M |
| Speechify / ElevenReader / Matter / VocalVia / Readox (Product H |
web |
Product Hunt's text-to-speech category page as captured in round 1, naming eight commercial products (ElevenLabs, Murf AI, Matter, VocalVia, Cartesia, |
unknown |
unknown |
|
n/a (directory |
L |
| vocodex (C043) |
web |
Work-in-progress self-hostable TTS client, surfaced by a 'like Speechify' README search. |
local-only |
dormant |
22 |
MIT |
L |
| a-haute-voix (LaPelle) |
? |
French accessibility tool described as extracting text from a PDF and re-rendering it for easier reading. |
unknown |
unknown |
0 |
unknown |
L |
| TuxReader (Savannah) |
? |
Savannah non-GNU project registered as 'TuxReader'; name and hosting confirmed by the Savannah software search, description not retrievable in this pa |
unknown |
unknown |
|
unknown |
L |