Your task.
Mastered by your agent.

Arena turns your raw data into agents, specialized at your task. Skip the reliance on someone else's frontier model.

Build, train and ship agents with Arena
Powered by reinforcement learning

The smarter way to build agents

STANDARD APPROACHES

  • Frontier models are generally good, but aren't specialized at what matters to you.

  • Agents stop learning after training and deployment.

  • Hosted fine-tuning and models carry a supply chain and data security risk.

  • Costly and time consuming training runs.

  • Tied to one frontier model provider where the cost scales with usage.

  • Small models fine-tuned on your task become experts.

  • Agents keep learning.
    Live results feed the next run.

  • Training runs in your environment; model weights remain yours.

  • 10x faster automated evolutionary training.

  • Use any open-source model, served on your own infrastructure.

How Arena fits into your workflow

  • CONNECT & VALIDATE

    Bring your data; use Arena to prepare it as LLM datasets and RL environments. Validate everything with feedback before training begins.

  • CONFIGURE ONCE

    Select algorithms, rewards, constraints, and objectives; enable evolutionary tuning to explore promising configurations automatically.

  • SCALE TRAINING

    Distribute across available GPUs; monitor metrics, sample efficiency, and checkpoints in real time.

  • DEPLOY & IMPROVE

    One-click promote to production; track performance, roll back, or keep training on live feedback.

> import agilerl

The framework Arena runs on.
Open-source.

  • Open-source framework with docs, examples, and community support

  • Single and multi-agent support across on/off-policy, offline RL, bandit and LLM training

  • Python-first, compatible with your data and environments

  • Works with your cloud compute and scales to multi-GPU

10X FASTER

Training with open-source v2

Used by leading research
labs and institutions

400,000+

Downloads from the community

Built for teams shipping superhuman agents

Transforming Autonomous Systems

Enabling breakthrough results in training AI agents for complex aerial interception missions with RTDynamics

READ CASE STUDY

Accelerating Financial AI Development

Substantially cutting compute expenses and boosting training speed for RL workflows with Warburg AI

READ CASE STUDY

Optimising Logistics Efficiency

Dramatically increasing utilisation and reducing training time for complex bin-packing with Decision Lab

READ CASE STUDY

Latest updates from AgileRL

Introducing the Arena Client: Reinforcement learning at scale, from your terminal

READ MORE

How we built a robust and scalable async-RL system that beats TRL and ART by 7x

READ MORE

How to pick an RL algorithm and reward system for multi-turn LLM training

READ MORE

Watch your agent master your task, live

Bring a task, dataset, or environment. We'll train, tune, and deploy an agent on it with you in a live session.

Book a demo