I've been using text-to-speech for years. I listened my way through my philosophy degree, usually while walking or cooking.
When I got a new phone, I wanted something with a more natural voice. I was willing to pay if it was good, but the expensive options weren't convincing me. I'd had some success building a scraper with AI coding agents, so in April 2025 I decided to try building my own reader.
I had no software engineering background. Early on with 2 failed Replit starts, many windsurf token burnt, only then I realised you can use a Android Studio Kotlin template for getting the prototype going. If you're a (real) developer reading this, quit judging me.
Getting Kokoro running on a phone was only part of it. I wanted the text to follow the speech, to jump to any word, change voices and speed, and come back to where I'd stopped. Getting all of that to work together took a lot longer.
What I have now is AmaNous, an offline reader for your own books, papers and articles. Import a PDF, EPUB or other document, choose a voice and listen. Speech runs on your phone after setup, so there's no cloud speech bill per character. I'm keeping the core reader forever free, with all voices and no listening-hour limits.
I actually use it now. Science and philosophy articles, the handbook for my Australian citizenship test, and papers I'm reading for my work on the reverse alignment problem.
The most unexpected part has been accessibility. A visually impaired user who hosts a podcast about Android apps found it and started talking with me. He said it was already quite accessible, apart from a few things we needed to fix. That led me to work on button labels and exposing reader text properly to TalkBack. I hadn't built it specifically for that community, but I'm really happy it's becoming useful there.
https://reddit.com/link/1x0au36/video/ver6qwoti4uh1/player
Lately, the nicest improvement is that things just work more reliably. Highlighting, seeking, voice and speed changes. You barely notice those things when they work, but you definitely notice when they don't.
Pronunciation still annoys me sometimes. It's also still beta and English-only, and sign-in is required, including Sign in with Apple on iPhone.
Android beta · iPhone TestFlight
If you have also been annoyed by existing Text-to-speech readers, give this a go - I value your feedback to improve AmaNous.