Community Picks
Read post

agentica-project/rllm: Democratizing Reinforcement Learning for LLMs

rLLM is an open-source framework for training language agents using reinforcement learning. The project has released several high-performing models including DeepSWE (32B software engineering agent achieving 59% on SWEBench-Verified), DeepCoder (14B coding model matching o3-mini performance with 60.6% Pass@1 on LiveCodeBench), and DeepScaleR (1.5B model surpassing O1-Preview with 43.1% Pass@1 on AIME). The framework enables developers to build custom agents and environments, train them with RL, and deploy for real-world applications.

    #ai-agents#jupyter#llm#machine-learning#reinforcement-learning
Jul 05, 2025•4m read time•From github.com
Post cover image
Table of contents
Releases 📰Getting Started 🎯AcknowledgementsCitation
468 Impressions
Community Picks's image
Community Picks

Community Picks is a section on daily.dev where our community members share the most interesting and...

9.9K Followers

•

336.5K Upvotes

fabidick22's user avatar
Dickson A.
@fabidick22
Joined Oct 6. 2023
19K

Would you recommend this post?

Copy link
WhatsApp
Facebook
X
New Squad
  • © 2026 Daily Dev Ltd.
  • Guidelines
  • Explore
  • Tags
  • Sources
  • Squads
  • Leaderboard