Paper2Audio vs TTSReader
Audio
Voice quality

Fine for study listening, clear enough that I can follow a methods section on a walk. It's not in the same league as ElevenReader or Matter's HD audio though. Prosody is flatter, less “someone is narrating this,” more “competent TTS that doesn't make me angry.” App Store people sometimes call the voices top-notch; I think they're reacting to the parsing win and grading the voice on a curve. If voice candy is why you open the app, look elsewhere.

Free = Web Speech API / installed system voices, unlimited, quality depends on your browser/OS (Chrome gets Google voices). Fine for proofreading and short listens; not Matter/ElevenLabs. Premium neural voices (Azure, Google, OpenAI, xAI/Grok, proprietary) are the ones that sound modern, free accounts only get a tiny neural tease (~5k premium characters total per their pricing docs). You can mix voices/languages/speeds in one file for dialogs.
Voice selection

A handful of named styles (Narrator, Teacher, Orator, etc.). Enough to switch when one gets old. Tiny catalog next to Speechify/ElevenReader. I wouldn't buy Plus for more voice options, there aren't really “more.”

Huge catalog once you're on Premium, including pasting any Azure Voice Gallery name. Free tier still has a lot of system options. Good for multilingual and classroom use cases they market hard.
Pronunciation

Built to handle jargon, decimals, abbreviations, symbols better than generic readers. Physics/legal students specifically praise formulas and dense prose. Still not magic on every symbol-heavy worksheet, their own FAQ admits math-problem sheets and weird scans can struggle.

Offline audio

Whole document gets generated, then mobile apps download for offline. Hiking / subway / airplane without babysitting a stream. This is why people who bounced off live-only apps stick around.

System voices work offline via the browser. Premium MP3/WAV export is the "listen later / publish" path, commercial rights come with Premium per their license. Export has been Windows-focused historically; check current player docs if you're on Mac.
Document processing
PDFs

Research PDFs are the home field. Strips footnotes, headers, reference junk; summarises tables/figures/math instead of reading them cell by cell into nonsense. Multi-column academic layouts are why the product exists. Free cap is 250 pages / 100 MB, plenty for papers; Plus goes to 1000 pages.

Upload a PDF and it extracts text for listening / copy-out. Works for many text-layer docs; scanned/layout-heavy academic PDFs still get the usual TTS mess (headers, columns, footnotes). Reddit comparisons sometimes ding it for not skipping citations/footnotes, expect a paste-box extractor, not a journal-native parser.
Ebooks

EPUB chapter detection + skipping extras. People report huge files (one reviewer: 900-page EPUB) working. Fiction/nonfiction both fine; the academic brain still shows in how it cleans layout.

EPUB import turns Project Gutenberg-style books into listenables. DOCX/TXT/CSV too. Solid "make this file speak" coverage for a browser tool.
Web articles

Web articles get junk stripped and image summaries. Good enough that it's not “papers only” anymore, still not Matter for newsletter inbox / social discovery. Use it when the article is long and messy, not when you want a Pocket replacement.

Paste a URL into the player, or use the Chrome/Edge extension to listen on-page. Extension claims skip irrelevant chrome and only injects when you trigger it. Official V3 extension (WellSource Ltd.) is Premium-oriented with a trial, watch out for copycat "TTS Reader" listings in the Chrome Web Store.
OCR & scans

Scanned PDFs are a real strength in user reports I've read: the same file that felt mangled in Speechify, NaturalReader, or ElevenReader often came through cleaner here. Lower-quality scans still vary (they say so). Not every photo of a whiteboard will be perfect.

Office docs

Word / pasted text / legal docs show up in the supported list. Slides and weird forms are weak spots per their FAQ, don't expect PowerPoint magic.

Import workflow

Upload, link, extension, Drive integration paths on web. Share sheet on phone. Then wait for the pipeline. Once it's done, listen anywhere.

Layout fidelity

Two layers: clean audio of the main text, plus Reader View / original PDF when you need to see a figure. Listening.com-style clean narration without giving up visual context. That combo is why I rank it above pure skip-chrome tools for many paper workflows.

Listening workflow
Bookmarks & notes

Highlights and notes are there. Not a Readwise second-brain, enough to mark the paragraph that mattered on the train.

Persistence of text + last position is the main workflow win. Not a full read-later library with tags/playlists. Premium adds storage/sharing of rendered text-audio pages (teachers share, students listen free is a marketed pattern).
Cross-device sync

Web + iOS + Android library and progress sync on free. No Apple-only trap.

Offline listening

Offline after generation is first-class on mobile. Generate on wifi, listen on the trail. Plus can export M4A out of the app; free keeps playback inside.

UI / UX
iOS app

Feels closer to a podcast/audiobook player than a bloated study suite. Upload or share a doc, wait for generation, then listen with highlights/notes. Sync with web so the commute pickup actually works. PhD-type reviews harp on switching between listen and read on different devices without losing place, that's the daily loop.

Android app

Same library on Play. Google reviewers call out offline downloads and figure/math handling. Not an Android-only special, just… actually available, unlike Matter.

Web app

Web is where I dump papers from the laptop. Library, collections, search, share links. Browser extension can grab pages (including some paywalled flows) without the download dance. Fine, not flashy.

The web player is the product: big text box, voice picker, play/export. Rich-text editor, sentence highlight + auto-scroll while speaking, font/theme tweaks, word/character counter for writers. Remembers text + caret position if you close the tab (Chrome/Safari mobile too). No ads on the main experience, that alone puts it ahead of a lot of "free" TTS pages.
Library & organisation

Collections, TOC jumps, library search. Enough for a semester pile without turning into Readwise.

Reading view

Reader View is the differentiator vs “audio-only academic” apps. Reformats the doc for the phone, keeps figures/tables/math inline, word highlight while it speaks. You can flip back to the original PDF when you need the real layout. People who've tried the same scan elsewhere and got messy extraction specifically call out this reformatting + transcript switch.

Ease of use

Pick Full Text / Long Summary / Short Summary, optional “additional context,” hit go. Fewer knobs than Voice Dream, and to my ear less of the loud marketing push I associate with Speechify. The tradeoff is you wait for processing before you listen, not instant stream-as-you-scroll.

Paste → play is genuinely zero-friction. No account required for free voices. Import PDF/DOCX/EPUB/TXT when you have a file. Proofreading people love editing while it speaks. Not a polished "library + collections" UX, if you want that, you're in the wrong app.
Platforms & ecosystem
Platform coverage

Web, iOS, Android, browser extension. Broader than Matter/Voice Dream, focused enough that it doesn't feel like Speechify's everything-store.

Web is the real product (Chrome/Firefox/Safari, Chromebooks, mobile Safari/Chrome tabs). Official Chrome/Edge extension for on-page listen. The Play listing for Android TTSReader Pro looks largely dormant to me (last meaningful update around 2020 last I checked), so I'd treat phone use as "open the website," not a polished native daily driver. No modern iOS native app that I've found.
Browser extension

Extension for grabbing pages into the pipeline quickly. Handy next to upload-from-disk.

Official V3 extension (WellSource Ltd., ID to match from ttsreader.com/x) is privacy-pitched, no inject until you ask. Built for paying users: sign-in + Premium after trial for the AI-voice experience. Free unlimited listening still lives mainly in the web player with system voices. Ignore similarly named copycats in the Chrome Web Store.
Pricing & limits
Free tier

This is the standout vs almost everyone else on this site. 56 hours/week of generation at 1x (they frame it as ~8 hrs/day average), the same voices free users and Plus users get (no robotic bait tier), no ads, offline on mobile, summaries, sync. Not a 10-minute demo. Personal use is meant to live here.
Limits that matter: 250-page PDFs, 100 MB, web articles capped shorter than Plus. Free content may be used to improve their parsing models (anonymised / not shared as your doc, still a privacy checkbox some people care about). Voices aren't the best in class, the free plan is great because of hours + features, not because it sounds like ElevenLabs.

Unlimited free/system voices for listening in the reader, no ads, no login required for the basic loop. Neural voices are a tiny tease (~5k premium characters total on free, per their pricing page, confirm). Best free utility TTS page I've used for proofreading and one-offs.
Paid value

Plus is the professional tier, about $20/mo or ~$16/mo yearly (~$192/yr on web; App Store IAPs can differ). Triples weekly hours (168), bigger files (1000-page PDFs), M4A export, Markdown transcript export, priority support, and no model-training on your uploads.
Honest take: for students / personal paper piles, free is so good that Plus feels optional, “ok paid plan,” not a must-buy. You upgrade for work compliance, exports, privacy, or monster files. Compared to Listening's $39/yr (no real free) or Speechify's ~$139, the free story is what wins; Plus is fine, not the cheapest paid seat in the category.

Premium $10.99/mo or $99/yr with 1M premium characters/month, plus pay-as-you-go credit packs ($10/200k, ~$32/1M, ~$300/10M, confirm on their pricing docs). Unlocks neural voices, MP3/WAV export, sharing, commercial use, extension. Cheap vs Speechify if you mainly need export + nicer voices. Character metering matters for long books. A "Premium+" tier has appeared in tables as coming-soon, don't plan on it yet.
Limits & upsells

Terms push professional / funded research use onto Plus, trust-based. Generation is metered weekly (not listen time). Processing delay before first play. No slides support. If you need instant browser-read of every tab with zero wait, this isn't that tool.

Study extras
Extra section
Short & long summaries

Short / Long Summary modes cut listen time hard (their Spence demo markets an 87% cut on long summary). Denser than fake “AI podcast” skits, aimed at learning, not entertainment. I use long summary to triage, full text when it matters.

Figure, table & math summaries

Instead of reading a table cell-by-cell into nonsense, it summarises visual elements. Same idea for figures and math. This is the feature people coming from weaker academic TTS setups notice first, in my experience.

Additional context

Optional extra definitions / context while listening, helpful on jargon-heavy papers without leaving the audio. Definition-style extras show up in Play praise too.

Steven's top picks

Best premium text to speech: ElevenLabs voices, word-synced reading view, and a daily-driver app across phone and web.
1st place
Best free text to speech for everyday listening. Handles articles, ebooks, and docs, with a particular edge on complex PDFs and research papers.
2nd place
Audio-first read-it-later with excellent HD voices for articles and newsletters. The Apple listen-later pick when the pile is web reading.
3rd place