Solo dev here: I spent 8 months building a PDF-to-audiobook app, and I even run my own GPU server for the voices

Wait 5 sec.

I'm Emin, a solo developer from Turkey. For the past 8 months I've been building Obook, an app that takes any PDF or ebook and turns it into an audiobook you can actually enjoy listening to. Why I built it: I have way more books and PDFs than time to read them. Existing TTS apps either sounded robotic, choked on messy PDFs (headers, page numbers, and footnotes read out loud mid-sentence), or cost a fortune. What it does: Import a PDF and it gets cleaned up and split into proper sections, so headers, page numbers, and junk get skipped 32 narrator voices across 20+ languages Word-by-word highlighting, so you can read along while you listen (shown in the video) [Public domain library / any other feature you want to highlight] The fun/painful part: I built everything myself, including the React Native app, the Node.js backend, and the TTS servers. Instead of paying per character to a TTS API, I run the voices on my own GPU server, which is the only way the economics work for long books. Where it's at: Live on iOS and Android, with about 100 users so far. I'd genuinely love feedback, especially: Does the voice quality hold up for a full chapter? Anything confusing in the first 2 minutes of using it? iOS: IOS_LINK · Android: Andoid_Link Happy to answer anything about the tech stack or running my own GPUs!   submitted by   /u/Puzzleheaded-Boot525 [link]   [comments]