// the find
rhiever/reddit-analysis
A Python script that parses post titles, self-texts, and comments on reddit and makes word clouds out of the word frequencies.
A small CLI (`word_freqs`) that uses PRAW to pull titles, selftexts, and comments from a subreddit or user's history, counts word frequency, and dumps a CSV you then paste into a word-cloud generator. Useful for someone who wants a quick, disposable look at what a subreddit talks about without building anything themselves.
Does exactly one thing and the usage is a single command once installed. Supports PRAW's multiprocess mode out of the box so you don't get your Reddit API access throttled or banned running it repeatedly. Ships as an actual pip package (`redditanalysis`) with a `setup.py` and a `tests.py`, not just a loose script.
README still carries Python 2.7/3.5 badges and documents the old username-as-user-agent auth scheme, not Reddit's OAuth2 flow that PRAW has required for years — this will not run against the current API without patching. It doesn't actually generate a word cloud; it hands you a CSV and tells you to paste it into wordle.net by hand, which has been dead/redirected for a long time, so half the stated purpose is broken. GPLv3-licensed and last touched in 2023 with no real commit activity since — this is effectively an unmaintained 2016-era script, not something to build a pipeline on.