rlvr
Here are 148 public repositories matching this topic...
Awesome List for Agentic RL
-
Updated
Aug 28, 2026 - HTML
[EMNLP'25] s3 - ⚡ Efficient & Effective Search Agent Training via RL for RAG (RLVR for Search with Minimal Data)
-
Updated
Nov 5, 2025 - Python
Evidence-first infrastructure for reproducible, isolated executable evaluation and online rewards.
-
Updated
Sep 2, 2026 - Python
Official repository for "RLVR-World: Training World Models with Reinforcement Learning" (NeurIPS 2025), https://arxiv.org/abs/2505.13934
-
Updated
Oct 28, 2025 - Python
Curated papers, taxonomy, benchmarks, and decision guides for credit assignment in reasoning and agentic LLM reinforcement learning.
-
Updated
Aug 3, 2026 - Python
[ICLR 2026] An official implementation of "CapRL: Stimulating Dense Image Caption Capabilities via Reinforcement Learning"
-
Updated
Jun 23, 2026 - Python
🌱 A little course on Reinforcement Learning Environments for evaluating and training Language Models
-
Updated
May 27, 2026 - Python
Scalable pipeline for synthesizing verifiable RLVR training data for computer-use agents
-
Updated
Aug 13, 2026 - Python
RL study guide — foundations through RLHF, DPO, GRPO, RLVR, agentic RL, and offline RL. Hand-written CS294 notes, 19 lecture drafts, 5 tested exercises, citations that resolve.
-
Updated
Jul 1, 2026 - Python
Reinforcement Learning Short Course
-
Updated
Aug 26, 2026 - Jupyter Notebook
CUA-Gym-Hub: mock web apps as reproducible RL training environments for computer-use agents
-
Updated
Sep 3, 2026 - JavaScript
A curated list of awesome resources about reward construction for AI agents. This repository covers cutting-edge research, and practical guides on defining and collecting rewards to build more intelligent and aligned AI agents.
-
Updated
Sep 1, 2025
Procedural data generators for verifiable reasoning, synthetic pretraining, post-training, evaluation, and RL.
-
Updated
Sep 11, 2026 - Python
PersonaMem-v2: Towards Personalized Intelligence via Learning Implicit User Personas and Agentic Memory
-
Updated
Sep 6, 2026 - Python
🐝 SwarmBench: Benchmarking LLMs' Swarm Intelligence
-
Updated
May 21, 2025 - Python
Add this topic to your repo
To associate your repository with the rlvr topic, visit your repo's landing page and select "manage topics."