Credits and licenses

PageVoice runs inference on your device. Only fixed model and runtime files are downloaded. Your book text is never part of those requests. The optional device-speech engine uses your operating system; its privacy and offline availability depend on that platform.

Speech is generated. Voices do not imply the participation or endorsement of a real speaker. Use audio lawfully and respect the restrictions attached to its model and the copyright of the book.

Models and voices

ComponentLicense and source
Kokoro English, US and UK voice stylesApache-2.0. hexgrad, ONNX conversion. License. English phonemization only in this app.
Piper DaveFX, Spain SpanishMIT voice repository; CC0 speech dataset. Model card, Pinned weights and source.
Piper Sharvard, two Spain Spanish speakersMIT voice repository; CC-BY-3.0 corpus. Attribution: Vincent Aubanel, Maria Luisa GarcĂ­a Lecumberri and Martin Cooke, Sharvard corpus. Model card. Speaker gender/age is not independently verified.
Supertonic 2, F1 and M1OpenRAIL-M, including use restrictions. Supertone weights. Upstream archived September 2026. Experimental comparison; not a measured quality improvement.
Multilingual MiniLM meaning searchApache-2.0. Sentence Transformers, Xenova ONNX conversion. Matches are inferred, with source citations.
Multilingual name detectionAFL-3.0. Davlan, Xenova conversion. Trained on news; names and speakers in fiction can be wrong.
English and Spanish OCR dataApache-2.0. Tesseract tessdata, integer LSTM data distributed by @tesseract.js-data/eng and /spa 1.0.0. License.

Runtimes and redistribution

Model licenses and library licenses are separate. PageVoice's original code is MIT; downloaded third-party software retains its own license.

Exact model revisions, file sizes and SHA-256 hashes are checked into web/src/offline/model-assets.json. Runtime versions are pinned in the npm lockfile and hashes in runtime-assets.json. No models download automatically.

Back to PageVoice