PageVoice runs inference on your device. Only fixed model and runtime files are downloaded. Your book text is never part of those requests. The optional device-speech engine uses your operating system; its privacy and offline availability depend on that platform.
Speech is generated. Voices do not imply the participation or endorsement of a real speaker. Use audio lawfully and respect the restrictions attached to its model and the copyright of the book.
| Component | License and source |
|---|---|
| Kokoro English, US and UK voice styles | Apache-2.0. hexgrad, ONNX conversion. License. English phonemization only in this app. |
| Piper DaveFX, Spain Spanish | MIT voice repository; CC0 speech dataset. Model card, Pinned weights and source. |
| Piper Sharvard, two Spain Spanish speakers | MIT voice repository; CC-BY-3.0 corpus. Attribution: Vincent Aubanel, Maria Luisa GarcĂa Lecumberri and Martin Cooke, Sharvard corpus. Model card. Speaker gender/age is not independently verified. |
| Supertonic 2, F1 and M1 | OpenRAIL-M, including use restrictions. Supertone weights. Upstream archived September 2026. Experimental comparison; not a measured quality improvement. |
| Multilingual MiniLM meaning search | Apache-2.0. Sentence Transformers, Xenova ONNX conversion. Matches are inferred, with source citations. |
| Multilingual name detection | AFL-3.0. Davlan, Xenova conversion. Trained on news; names and speakers in fiction can be wrong. |
| English and Spanish OCR data | Apache-2.0. Tesseract tessdata, integer LSTM data distributed by @tesseract.js-data/eng and /spa 1.0.0. License. |
Model licenses and library licenses are separate. PageVoice's original code is MIT; downloaded third-party software retains its own license.
Exact model revisions, file sizes and SHA-256 hashes are checked into web/src/offline/model-assets.json. Runtime versions are pinned in the npm lockfile and hashes in runtime-assets.json. No models download automatically.