← All communities
Reddit community

r/reinforcementlearning

Reinforcement Learning. Reinforcement learning is a subfield of AI/statistics focused on exploring/understanding complicated environments and learning how to optimally acquire rewards. Examples are AlphaGo, clinical trials & A/B tests, and Atari game playing. Founded 2012 · 90.2K members

90.2K members. The best post we have captured here reached 7.6K upvotes, which works out to 84 per 1,000 members.

84peak upvotes per 1,000 members
90.2Kmembers
7.6Kbest post
410most comments
397posts we track

Open r/reinforcementlearning on Reddit ↗

What actually lands here

The mix of post types among everything we have captured. Posting a link where almost nothing but images lands is wasted effort, and this is the fastest way to see it.

30% 25% 24% 13% 8%
Text 30% Video 25% Link 24% Image 13% Question 8%

The topics it labels

Flairs are the community's own categories, and the closest thing a subreddit has to telling you what it wants.

Robot 29 DL 18 P 18 D 13 R 13 Multi 10 Bayes 3 Psych 3 DL, MF, R 2 N, P 2

What the ceiling looks like

The biggest posts we have captured here. Links go to the original on Reddit.

Google copied our open-source code, removed our engineers’ names, and gave us zero credit one year after we beat them on their own benchmark

↑ 7.6K 💬 410 image 2026

We beat Google Deepmind but got killed by a chinese lab

↑ 795 💬 33 text 2025

Trained a PPO agent to beat Lace in Hollow Knight: Silksong

↑ 552 💬 66 video 2025

IT'S LEARNING!

↑ 549 💬 23 image 2025

I literally build the jev architecture one year back and made it open-sourced

↑ 519 💬 51 text 2026

PPO Ping Pong

↑ 359 💬 25 video 2025

Why is RL fine-tuning on LLMs so easy and stable, compared to the RL we're all doing?

↑ 355 💬 42 question 2025

Andrew G. Barto and Richard S. Sutton named as recipients of the 2024 ACM A.M. Turing Award

↑ 348 💬 14 link 2025

Can RL redefine AI vision? My experiments with partial observation & Loss as a Reward

↑ 324 💬 60 video 2025

Built a custom robotic arm environment and trained an AI agent to control it

↑ 314 💬 18 video 2025

Communities its size

Within four times the member count either way, so the comparison is fair. Sorted by upside.

r/reinforcementlearning, in numbers

Answered from the posts ViralHunt holds from this community, not from Reddit's own about page.

What is r/reinforcementlearning?
r/reinforcementlearning ("Reinforcement Learning") describes itself this way: Reinforcement learning is a subfield of AI/statistics focused on exploring/understanding complicated environments and learning how to optimally acquire rewards. Examples are AlphaGo, clinical trials & A/B tests, and Atari game playing. It was founded in 2012 and has 90.2K members.
How many members does r/reinforcementlearning have?
r/reinforcementlearning has about 90.2K members. ViralHunt has tracked 397 of its posts; the biggest reached 7.6K upvotes, which is 84 per 1,000 members.
What is the most viral post on r/reinforcementlearning?
"Google copied our open-source code, removed our engineers’ names, and gave us zero credit one year after we beat them on their own benchmark" reached 7.6K upvotes and 410 comments in 2026, the highest score among the posts we hold from r/reinforcementlearning.
What subreddits are similar to r/reinforcementlearning?
Communities of a comparable size we track: r/greenland, r/Leakednews, r/LetsDiscussThis, r/stevehofstetter, r/StrikeAtPsyche, r/InvictaSolaris.
What kind of posts do best on r/reinforcementlearning?
Of the posts we hold, 30% are text posts and 25% are video posts.
When is the best time to post on r/reinforcementlearning?
Across 397 timestamped posts we hold, the slots whose posts averaged the highest score are Monday at 20:00 UTC (201 avg upvotes), Monday at 15:00 UTC (192 avg upvotes), Wednesday at 19:00 UTC (174 avg upvotes). The busiest day is Tuesday. For Reddit as a whole, see the best time to post on Reddit page.
How many upvotes does a post need to reach the top of r/reinforcementlearning?
The 20th-best post we hold from r/reinforcementlearning has 222 upvotes, so that is roughly the bar for the top twenty. We captured 84 posts from it in the last 30 days.

See what is climbing here before you write

ViralHunt tracks Reddit and eight other networks, records how every post grows, and puts what is taking off in front of your team in one list.

Start free →

Every figure here is what ViralHunt has captured from r/reinforcementlearning, not everything the community has ever posted. We deliberately do not publish an average score: our sample is weighted toward a community's best posts, so an average would be flattering and wrong. Peak per 1,000 members is computed the same way for every community listed. See how we measure.