RunRL docs · Oct 2025
Official product update announcing LoRA support, automatic MoE detection, dataset wizard beta, REST API parity with the UI, SFT warmups, and MCP tool support.
Public company, workplace, funding, and market signals
Updated Jul 30, 2026
RunRL is a San Francisco-based YC Spring 2025 startup building reinforcement learning as a service for improving LLMs and AI agents with custom reward functions and RL fine-tuning.
Primary product
Reinforcement learning as a service platform for training and improving LLMs and AI agents
Founded
2025
Headquarters
San Francisco, California, United States
Team size
1-10
Industry
Artificial intelligence
Sub-industry
Reinforcement learning platform / AI agent optimization
Business model
0 jobs at RunRL
Check back later for new openings
Stage
Seed
Total raised
$500K
Latest round
Seed · Jun 2025
Latest amount
$500K
Jun 2025 · Y Combinator
Jan 2025 · Y Combinator
Investors
Public materials emphasize a research-driven, hands-on engineering culture focused on making RL 'just work' through automation, self-serve tooling, and rapid experimentation.
Compensation
No public employee salary bands or benefits page were found. Public compensation-like signals are product pricing: $80 per node-hour for self-serve usage, with enterprise/custom pricing.
Pricing
Usage-based pricing at $80 per node-hour, with free slow inference for deployed models and custom enterprise pricing.
Differentiators
Technology
Customers
Competitors
RunRL docs · Oct 2025
Official product update announcing LoRA support, automatic MoE detection, dataset wizard beta, REST API parity with the UI, SFT warmups, and MCP tool support.
Hacker News · Sep 2025
Launch discussion describing RunRL's RL workflow, pricing at $80/node-hour, and use cases such as antiviral design, formal verification, browser agents, and tool use.
Andrew Gritsevskiy LinkedIn · May 2025
A technical experiment post showing how RunRL uses Qwen3-8B with an LLM-as-judge reward setup and iterative RL to optimize outputs.
Y Combinator · May 2025
YC launch post describing RunRL as a platform for reinforcement learning on AI, with custom reward functions and specialist-model use cases.
RunRL LinkedIn · May 2025
RunRL announced its YC X25 launch and said founders Derik and Andrew previously ran a research lab together; the post also referenced early customer gains and the founder backstory.