GitHub
rllm
观察池 · 暂无正式排名
该项目当前未满足“两个有效维度 + 两种数据源”的主榜门槛。
— 未排名
可观测采用度缺失 —
动量当前有效 · 2026-09-12 12
关注度当前有效 · 2026-09-12 24
信号可信度 依据当前有数据的独立评分维度数量计算。
中2/3 · 1 种数据源
项目介绍
Agentic RL on any harness, with any backend, on any benchmark.
rLLM is an open-source framework for training language agents with reinforcement learning. Bring any harness, run it in any sandbox, and switch training backends with one flag — the same agent code drives both eval and training.
rLLM requires Python >= 3.11. You can install it either directly via pip or build from source.
This installs dependencies for running rllm CLI with the tinker backend (single-machine, Tinker API). For other backends:
For building from source or Docker, see the installation guide.
Define a rollout (your agent) and an evaluator (your reward function), then hand them to the trainer:
During training,…
各数据源
5.8k Star
- Star 5.8k
- Fork 615
- 提交 1.9k
- 发布 4