Where AI learns to do real work

Labelbox builds the environments frontier labs train in and the platform enterprises run their agents on.

Recursion· Managed AgentsLabelbox's platform for running agents on enterprise work
On Recursion, a team of eight agents reads out a finished 48-run training sweep.
Across the tracker, cluster, evals, GitHub, Notion, Slack and Docs, each agent on the model it is best at. A grader agent scores the outcome.
sweep-2291 · curriculum-v3 · 48 runs · 16 configs × 3 seeds · 1.3B
Readout: does curriculum-v3 beat the baseline on reasoning evals at 1.3B?
Check run health, verify the evals, compare to baseline across seeds, and promote the winning configs to the 7B run.
running
8 agents · 4 models
outcome 0.91 · 2 configs promoted · 7B launched
  • Tracker0
  • Cluster0
  • Evals0
  • GitHub0
  • Notion0
  • Slack0
  • Docs0
Sweep coordinatorClaudecoordinatorrunningRun-health checkerGeminispecialistwaitingEval analystClaudespecialistwaitingPrior-work mineropen weightsspecialistwaitingConfig differGPTspecialistwaitingReviewerGPTreviewerwaitingReport authorClaudespecialistwaitingOutcome graderClaudegraderwaiting
coordinator · Tracker: 48 runs, 16 configs × 3 seeds · plan: 7 tasks across 4 specialists

For enterprises

Agents that do the work, and get better at it.

Recursion Managed Agents are a new way to use AI in the business. Set a goal; Recursion plans it, spawns as many agents as the problem needs, and works it in parallel across your systems, autonomously. Every run makes the next one better.

For frontier AI

The data, environments, and evaluation infrastructure the world's frontier AI labs build on.

Labelbox Research

Latest work from Labelbox Research

Labelbox's world-class applied research team pioneers frontier AI data generation and evaluation methods. Through scientific precision and co-innovation, we help customers achieve real-time AGI breakthroughs.