Asynchronous, decentralized reinforcement-learning framework for post-training.

Are you the maintainer?
Claim this page →Prime Rl
Asynchronous, decentralized reinforcement-learning framework for post-training.
01 / About
What Prime Rl is.
02 / Discussion CREDIBILITY-GATED
Discussion
Reading is open to everyone. Posting and voting need a verified identity or a GitHub grade of B or higher.
- No discussions yet.
03 / Related
More around Prime Rl.
More from the same builder
HarnessPrime Agent
Coding
PrimeIntellect-ai
A self-improving RLM agent for coding workflows and long-running autonomous tasks.
rlmself-improvementreplcli
FrameworkVerifiers
Monitoring
PrimeIntellect-ai
Library of shareable RL environments and rubric-based verifiers for agents.
reinforcement-learningevaluation
Similar tools
FrameworkART (Agent Reinforcement Trainer)
Training
OpenPipe
Open-source GRPO reinforcement-learning trainer for multi-step LLM agents
reinforcement-learninggrpofine-tunelora
ToolRay
InferenceTrainingDeployment
ray-project
Distributed runtime and ML libraries for scaling Python and AI workloads from a laptop to a cluster
distributed-computingreinforcement-learninghyperparameter-tuning
ToolLlamaFactory
Training
hiyouga
Fine-tune and post-train open LLMs across many families through a command line or Gradio web UI, with no training code required
fine-tunelorareinforcement-learninggui
ToolRLinf
Training
RLinf
RLinf: Reinforcement Learning Infrastructure for Embodied and Agentic AI
reinforcement-learningdistributed-trainingroboticsfine-tune
ToolrLLM
TrainingQA
rllm-org
rLLM bolts RL onto agents you already wrote. verl, trlx, OpenRLHF make you rewrite the agent into their pipeline. A deco
reinforcement-learningagent-trainingsandboxdistributed-training
ToolSurogate
TrainingDeployment
invergent-ai
Fine-tuning a small model on a few hundred rows costs a fraction of sending the same task through an API. Soup runs that
fine-tunegrpodpolora
04 / Build
Build with Prime Rl.
Browse the catalogue for frameworks, tools, and harnesses, each scored on real GitHub credibility.
Get Prime Rl →