About BooksReddit
Ask Reddit for a book recommendation and you'll get the same handful of answers, upvoted for the thousandth time. That repetition is annoying in a thread — but across seven years and millions of comments, it's data. BooksReddit reads the threads so you don't have to: 1,142,341 book mentions across 41 subreddits, ranked by how often each book comes up. Published excerpts, upvotes, and sentiment from scored excerpts provide context; they do not change the ranking.
BooksReddit is a 2026 rebuild of the original WordPress booksreddit.com (active 2014–2022). The original aggregated Amazon book mentions across Reddit, ranked by upvotes and mentions, and organized them by subreddit. This version keeps that posture and adds three things:
- Freshness. The pipeline processes versioned cached batches from the indexed subreddits, and the site rebuilds when a reviewed batch lands. The original's data was static.
- Provenance. Every book page shows actual Reddit comments with permalinks and timestamps — not just a mention count. Read what redditors said before you buy.
- Sentiment. The same tracked mention count can accompany very differently toned published evidence. We score selected displayed excerpts; unscored excerpts remain “Not scored” and do not enter the average.
Why books from non-book subs?
Half the value of the original site was tracking book mentions in specialty subreddits rather than the obvious literature ones. The specialty subs are where practitioners actually cite the books they read: engineers in r/ExperiencedDevs, therapists in r/therapists, investors in r/Bogleheads. We started with programming and now cover eight themes — from philosophy to mental health — with more arriving when the data justifies them.
Editorial integrity
We use Amazon and Bookshop.org affiliate links. We don't accept paid placements; we don't display prices (Amazon's terms forbid caching them); we don't host Amazon's product images. The rankings use mention-count order in the stated scope, then the source sample is synthesized into editorial summaries. The methodology page documents the process and its limits. Published excerpts retain their source links, and aggregate counts are available in the public dataset. That aggregate download excludes excerpts, raw comments, and comment-level matches, while its totals can include partial source units.
What's next
We publish on a weekly cadence and are testing evidence-first discovery: compact recommendation receipts, community comparisons, reproducible data stories, and portable charts other sites can cite. The methodology page has the technical detail.