Managed AI agents for every team

Recursion Managed Agents puts AI agents to work inside the apps your business already uses: Salesforce, Zendesk, NetSuite, Slack, Jira and 100+ more. They start on a schedule, a message or an API call, finish the job, and get better every run.

Recursion Managed Agents· Sessions
Recursion Managed Agents· Sessionsincident INC-4821
INC-4821 · SEV-2 · checkout-api · 5xx at 6× baseline since the 14:05 deploy
Incident: checkout is failing for some customers after the 14:05 deploy
Find the root cause, stop the errors, ship a tested fix, count the customer impact and keep the status page current.
running
8 agents · 4 models
outcome satisfied 5/5 · mitigated in 26 min · fix in review
  • PagerDuty0
  • Datadog0
  • GitHub0
  • LaunchDarkly0
  • Stripe0
  • Zendesk0
  • Statuspage0
  • Slack0
Incident coordinatorClaudecoordinatorrunningLogs and tracesGeminispecialistwaitingDeploy diffopen weightsspecialistwaitingFeature flagsGPTspecialistwaitingReproduceClaudespecialistwaitingCustomer impactGPTspecialistwaitingFix authorClaudespecialistwaitingOutcome graderClaudegraderwaiting
coordinator · PagerDuty INC-4821 acknowledged · checkout-api 5xx at 6× baseline since 14:07 · plan: 5 checks
Checkout errors spike after a deploy.
PagerDuty pages the on-call: checkout is failing at six times the normal rate since the 14:05 deploy. The coordinator reads the alert, the error graphs and the deploy, then plans five checks.

One goal in. A fleet of specialists out.

Recursion Managed Agents runs persistent, long-horizon agents as one fleet inside the apps your business already uses. Hand them a goal, not a script. Sandboxes, credentials, subagents, memory and grading are all handled for you.

Works in your apps

100+ integrations across CRM, support, finance, engineering, data and HR. Agents take real actions, limited to the tools you allow.

Runs on its own, for hours or days

Starts on a schedule, a Slack message, a webhook or an API call. Sessions pause on dependencies and pick up where they left off.

Multi-agent by design

A coordinator plans the work and spawns specialists in parallel. Big jobs get more agents, not longer prompts.

Every outcome graded

A separate grader checks each session against your rubric and sends it back until every criterion passes. Every finished session feeds the agent’s memory.

Know exactly how your agents are performing

A coordinator breaks down each request and sends the right checks to specialist agents with the models and apps they need. The specialists gather the evidence, and the coordinator brings it together into a recommendation.

A separate grader evaluates the result against your rubric, criterion by criterion, using the evidence gathered along the way.

Set the standard, measure every outcome, and manage agent quality at scale.

Recursion Managed Agents· Sessions
Recursion Managed Agents· Sessionscoordinator · 6 specialists · 1 grader
Vendor review coordinatorClaude
JiraPROC-1182
Can we onboard Kestrel Analytics as a data processor by Friday?
planreading PROC-1182 and 14 files
Sanctions
Litigation
Security review
Financials
Insurance and DPA
References
recommendation
Approve with conditions: a new pen test within 90 days.
SanctionsGeminiWeb
not started
LitigationOpen weightsWebGoogle Drive
not started
Security reviewClaudeGoogle DriveGmail
not started
FinancialsGPTNetSuiteGoogle Drive
not started
Insurance and DPAClaudeDocuSignGoogle Drive
not started
ReferencesGPTGmail
not started
Outcome graderClaude
your rubricwaiting for a recommendation
No sanctions or watchlist matches
evidence: OFAC, EU and UK lists
No open litigation
evidence: court records, settlement letter
Security evidence is current
evidence: SOC 2 report; pen test → condition
Financially sound
evidence: audited FY2025 statements
DPA and insurance meet policy
evidence: signed DPA, certificate of insurance
satisfied 5/5approve with conditions

See your entire agent fleet at a glance

See every agent at work in one place — from vendor reviews and support tickets to invoices and security alerts. Know what’s done, what’s in progress, and how well each agent is performing.

Get a live view of your operations and the quality of every run.

Recursion Managed Agents· Sessions
Recursion Managed Agents· Sessions144 sessions this week · 127 completed · 10 running
  • Vendor Risk Review
  • Support Triage
  • Invoice Matching
  • Security Alert Triage
  • brighter = better graded

Chat agents wait for a prompt. Business work doesn’t.

A chat agent8 tasks this week
Every task starts with someone asking
Recursion Managed Agents2,316 sessions this week
SchedulesEventsAPI calls
One week of agent activity. A chat agent works only when someone asks: eight tasks, one at a time, in weekday working hours, and nothing at night or on the weekend. With Recursion Managed Agents, agents start on schedules, events and API calls and work around the clock, many sessions at once: 2,316 in the same week.

Business work repeats

Invoices to match, tickets to triage, vendors to review, a pipeline report every Monday. Most knowledge work comes back on a schedule or arrives in a queue, follows the same steps each time, and there’s more of it than your team can get to.

Chat agents wait for a person

A chat agent is built for someone to type a request and wait for the answer. A person still has to remember to ask, check the result and copy it where it belongs, one task at a time. Nights and weekends, nothing happens.

Recursion Managed Agents runs it on its own

It’s a new way to deploy agents: they start on a schedule, an event or an API call, run hundreds of sessions in parallel when the queue is long, and deliver the results into your apps. Nobody has to press go.

Work that runs in the background

Each agent has one job, the apps it may use and what starts it: a schedule, an alert, an event or an API call. No one has to open a chat and ask. These are a few of the jobs teams run on it; any work that keeps coming back can be one.

  • Ticket moved to ReadyLinearGitHubVercelSlack
    Delivers: A tested pull request with a preview link, ready for review
    Engineering

    Feature Builder

    Builds each ready ticket into a working feature, with tests, in its own sandbox.

  • New issue in SentrySentryGitHubLinear
    Delivers: A pull request with the fix and a test that proves it
    Engineering

    Bug Fixer

    Reproduces each new bug, finds the cause and writes the fix with a regression test.

  • New package version releasedGitHubJenkinsSlack
    Delivers: One ready-to-merge pull request per package, tests green
    Engineering

    Dependency Upgrader

    Upgrades dependencies as new versions ship, fixes what breaks and reruns every test.

  • PagerDuty alertPagerDutyDatadogSentryGitHub
    Delivers: A root-cause draft and the suspect commit before on-call is up
    Engineering

    Incident Investigation

    Correlates alerts, logs and recent deploys the moment an alert fires.

  • CI failure on mainGitHubJenkinsSlack
    Delivers: A pull request with the fix and 50 green runs in a row
    Engineering

    Flaky Test Fixer

    Reproduces each flaky test, finds the race and fixes it.

  • Spec approved in NotionNotionGitHubVercel
    Delivers: A live prototype on a preview link, with the spec beside it
    Product

    Prototype Builder

    Turns each approved spec into a working prototype your team can click through.

  • Request tagged for the roadmapLinearGongNotion
    Delivers: A draft spec with the problem, the customer evidence and open questions
    Product

    Spec Writer

    Writes the first spec for each request that makes the roadmap, from the calls and tickets behind it.

  • Pull request mergedGitHubLinearSlack
    Delivers: A docs pull request that matches what shipped
    Product

    Docs Updater

    Updates your docs whenever a merged pull request changes how the product works.

  • Feature released in LinearLinearSalesforceGmail
    Delivers: A personal note to each requester, queued for the account owner
    Product

    Close the Loop

    When a requested feature ships, writes to every customer who asked for it.

  • Mondays 07:00GongSalesforceNotion
    Delivers: Requests ranked by the customers and revenue behind them
    Product

    Feedback Digest

    Reads the week’s sales calls, reviews and feature requests for what customers keep asking for.

  • Issue labeled sweepLinearGitHubDatabricks
    Delivers: A results table with the winning config, rerun to confirm
    Research

    Experiment Sweep

    Runs dozens of experiment configs at once and compares every run with the baseline.

  • Pull request touching promptsGitHubGoogle BigQuerySlack
    Delivers: A per-task score diff, and a blocked merge if anything regresses
    Research

    Eval Regression Watch

    Reruns your eval suite on every model or prompt change.

  • New dataset versionDatabricksGitHubSlack
    Delivers: A checkpoint, its eval report and a model card
    Research

    Model Trainer

    Launches fine-tuning runs, watches the curves and stops bad runs early.

  • New model releasedDatabricksGoogle BigQuerySlack
    Delivers: A leaderboard on your tasks, with cost and latency for each model
    Research

    Benchmark Runner

    Runs every candidate model against your own benchmarks in parallel, overnight.

  • Mondays 09:00NotionGitHubSlack
    Delivers: A digest of what reproduced, with the code and numbers
    Research

    Research Digest

    Reads new papers and repos on your topics and reruns the key result.

  • Alert webhookOktaGoogle Cloud LoggingSlack
    Delivers: Every alert investigated, with a recommended containment step
    Security

    Threat Investigation

    Investigates every identity, mailbox and endpoint alert as it fires.

  • New pull requestGitHubJira
    Delivers: Real vulnerabilities only, each with proof and the fix
    Security

    Code Security Review

    Reviews every pull request and reproduces what it finds.

  • Critical CVE publishedGitHubJiraSlack
    Delivers: A tested patch for every exposed service, most urgent first
    Security

    Vulnerability Patcher

    Finds every service a new critical vulnerability touches and patches it.

  • Nightly 02:00Google Cloud LoggingGitHubJira
    Delivers: Each real exposure, ranked, with the fix as a pull request
    Security

    Cloud Posture Review

    Checks every cloud account for misconfigurations and exposed data.

  • First day of each quarterOktaGitHubGoogle Drive
    Delivers: Every stale or risky grant, with a revoke list for approval
    Security

    Access Review

    Reviews who can reach what, and flags access no one uses or should have.

  • Mondays 06:00SalesforceSnowflakeGoogle Drive
    Delivers: An updated forecast with what moved since last week, and why
    Finance

    Forecast Refresh

    Rebuilds the revenue forecast from pipeline, usage and bookings.

  • Books closed for the monthSnowflakeLookerGoogle Drive
    Delivers: Commentary for the CFO, with the numbers behind each driver
    Finance

    Variance Analysis

    Explains every material gap between budget and actuals, driver by driver.

  • Discount requested in SalesforceSalesforceSnowflakeSlack
    Delivers: Approve or counter, with the margin impact and precedent deals
    Finance

    Deal Desk Review

    Checks every non-standard deal against pricing policy and margin targets.

  • Contract signed in DocuSignDocuSignSalesforceGoogle Drive
    Delivers: The treatment for each contract, with the clauses behind it
    Finance

    Revenue Recognition Review

    Reads every signed contract for terms that change how its revenue is recognized.

  • Earnings release publishedWebGoogle DriveSlack
    Delivers: A one-page note on each company you cover within an hour of its release
    Finance

    Earnings Review

    Reads the release, the filing and the call transcript, and compares the results with your estimates.

  • Meeting on the calendarSalesforceGongGoogle Calendar
    Delivers: A one-page brief in the rep’s inbox an hour before every call
    Sales

    Account Research

    Builds a brief from the CRM, past calls, filings and news.

  • Deal moves to commitSalesforceGongSlack
    Delivers: The risks, missing stakeholders and next steps for the rep
    Sales

    Deal Review

    Reads every call and email in a late-stage deal and checks it against your sales process.

  • Daily 06:00SalesforceAmplitudeWeb
    Delivers: The accounts ready to expand, with the evidence for each
    Sales

    Expansion Signals

    Watches usage, hiring and news across your accounts for signs they’re ready to grow.

  • RFP added in SalesforceSalesforceGoogle DriveConfluenceSlack
    Delivers: A first draft of every answer, sourced from past wins, with gaps flagged
    Sales

    RFP Response

    Drafts RFP and security questionnaire answers from your approved content.

  • 90 days before renewalSalesforceAmplitudeGong
    Delivers: A renewal brief for each account: usage, open issues, sentiment and a plan
    Sales

    Renewal Risk Review

    Reads usage, account history and call notes before every renewal.

  • Mondays 08:00WebGongNotionSlack
    Delivers: A weekly brief on competitor moves and what they mean for open deals
    Strategy

    Competitive Intelligence

    Monitors competitor sites, pricing pages, job posts and mentions in your sales calls.

  • Data room sharedBoxGoogle DriveNotion
    Delivers: A diligence memo with every red flag sourced to the document and page
    Strategy

    Deal Diligence

    Reads the whole data room: financials, customer contracts, IP, employment and litigation.

  • Daily 07:00WebNotionSlack
    Delivers: New targets worth a look, with the case for each one
    Strategy

    Acquisition Screening

    Screens newly funded companies in your target markets against your criteria.

  • End of each quarterGongSalesforceNotion
    Delivers: A win-loss report with the calls behind every reason
    Strategy

    Win-Loss Analysis

    Reads the calls, notes and competitor mentions from every closed deal.

  • Two weeks before each board meetingSnowflakeSalesforceGoogle Drive
    Delivers: A first draft of the deck, every number tied to its source
    Strategy

    Board Deck Prep

    Pulls the quarter’s metrics, wins and risks into your board deck template.

  • API call from the procurement portalSalesforceDocuSignGoogle Drive
    Delivers: Approve, approve with conditions or reject, with evidence for every check
    Procurement

    Vendor Risk Review

    Runs sanctions, legal, security, financial and reference checks in parallel.

  • Daily 07:00WebGoogle DriveSlack
    Delivers: An alert with the evidence when a supplier becomes a risk
    Procurement

    Supplier Risk Watch

    Monitors key suppliers for financial distress, sanctions and outages.

  • Sourcing event closesGoogle DriveBoxSlack
    Delivers: A scored comparison and a recommended supplier, with the trade-offs
    Procurement

    Bid Comparison

    Compares supplier proposals line by line against your requirements and budget.

  • 60 days before a contract renewsDocuSignSnowflakeWeb
    Delivers: Your leverage, a target price and a walk-away point for every term
    Procurement

    Negotiation Prep

    Builds a negotiation brief from past spend, contract terms and market benchmarks.

  • Month-end closeSnowflakeLookerSlack
    Delivers: Savings ranked by value, with the spend behind each one
    Procurement

    Spend Analysis

    Reviews each month’s spend for duplicate vendors, price creep and missed discounts.

  • A scheduleAn alertAn eventAn API call
    And many more

    Any job that keeps coming back

    These are a few of the jobs teams run. Describe yours, what starts it and what good looks like, and an agent takes it from there.

Works in the apps your business runs on

Connect 100+ apps across CRM, support, finance, HR, engineering and data, then choose exactly which tools each agent may call: read only, changes data, or destructive. Credentials stay in a vault, and every tool call is recorded in the session transcript.

Recursion Managed Agents
Salesforce
Salesforce
Built-in connector · vaulted credential
Tools this agent can call4 of 6 on
  • find_account
    read only
  • get_case
    read only
  • list_opportunities
    read only
  • log_activity
    changes data
  • update_opportunity
    changes data
  • delete_record
    destructive

Credentials stay in the vault. The agent sees only the tools you turn on, and every call lands in the session transcript.

Native integrations

Slack, GitHub, Jira, Confluence, Google Cloud Logging, Loom and LaunchDarkly are built into Recursion Managed Agents. Start an agent from a Slack message and get the result back where your team works.

100+ built-in connectors

Salesforce, Zendesk, NetSuite, Workday, Snowflake and the rest of the catalog. Connect once, then turn tools on per agent.

Your own MCP servers

Point an agent at any MCP server, internal or third party, and set the same per-tool permissions.

Browser tasks, handled on a schedule

Some work only happens in a browser: vendor portals behind a login, sites with no API, forms that have to be filled in by hand. With computer use, the agent gets its own browser to sign in, search, fill in forms and download files. Put it on a schedule and it does the rounds every week on its own. Every click is saved as a screenshot you can replay.

Recursion Managed Agents· Sessions
Recursion Managed Agents· SessionsRFP Sweep · scheduled run · computer use
New tab
New tabSearch or type a URLAgent driving
Agent
Computer · every action saved with a screenshot0 frames
Live
RFP Sweepsession 7318running
every Monday 07:00computer useSalesforceSalesforce
Agent narration
Monday 07:00 run. Checking 6 bid portals for new RFPs that fit our fleet software. None has an API, so I’ll sign in to each one.
Runsevery Monday 07:00
Sep 14
0 new
Sep 21
1 new
Sep 28
running
next · Oct 5
07:00
Every Monday at 07:00, with no one watching, the RFP Sweep agent signs in to six government bid portals with the team's vendor login, searches for fleet telematics RFPs, downloads the bid documents and adds each match to Salesforce with a go/no-go summary. The run finds three new RFPs, skips one already tracked, checks all six portals and schedules the next run for the following Monday.

Your agent fleet, in your pocket

Your agents keep working after you log off. Check on them from your iPhone or iPad: see every session on one board, follow a run as it happens, and open the reports and files they made. Widgets on your Home Screen and Lock Screen show what’s running at a glance.

Recursion Managed Agents on iPhone: a vendor review session, with the main agent and each check on its own track
Recursion Managed Agents on iPhone: the agents list, each agent with its model and team
Recursion Managed Agents on iPhone: the fleet board, every session this week as one grid

Fine-tune specialist models and add them to the fleet

Memory makes each run better. Fine-tuning makes the model itself better. When a job runs at volume, Recursion Managed Agents turns its graded runs into environments and evals, trains a specialist with reinforcement learning, and adds it to your fleet when it beats the model the agent runs today.

  1. 1Graded runs
    12,480 graded runs

    Every session scored against your rubric

  2. 2Environments
    • Refund above policy limit
    • Duplicate charge dispute
    • Order number missing
    • +1,237 more

    Real tickets become repeatable scenarios

  3. 3Evals
    312 held-out tickets
    • Right resolution
    • Policy followed
    • No invented facts
    • Escalates when needed

    Held-out checks on every criterion

  4. 4RL training
    reward0.61 → 0.88

    Rewarded for passing your rubric

  5. 5In the fleet
    Support Triage
    Claude Opus 4.8Support specialist v3
    84% resolved · $32 per 1k tickets

    Promoted when it beats the model it replaces

Support Triage's 12,480 graded runs become environments from real tickets, then evals on 312 held-out tickets, then a reinforcement learning run that lifts reward from 0.61 to 0.88. The resulting model, Support specialist v3, replaces Claude Opus 4.8 in the fleet, and its own graded runs train the next version.

The specialist always beats the generalist

On the job it was trained for, the specialist resolves more tickets than frontier models, invents less and answers faster, at a fraction of the cost.

lower cost per task with fine-tuned models
5–10×
of leading US AI labs trust Labelbox Horizon's RL environments
90%+

Support Triage after fine-tuning: more tickets resolved, fewer hallucinations, faster responses and lower inference cost than frontier models

Resolution rate

Support specialist v3

84%

GPT-5.5

76%

Claude Opus 4.8

73%

Reduction in hallucinations

Support specialist v3

72%

GPT-5.5

46%

Claude Opus 4.8

41%

Cost per 1,000 support tickets

Support specialist v3

$32

GPT-5.5

$158

Claude Opus 4.8

$176

Time to first token

Support specialist v3

0.42s

GPT-5.5

1.10s

Claude Opus 4.8

1.28s

Managed agents, without the lock-in

Agent platforms from the AI labs run only their own models, and they’re built for developers. Recursion Managed Agents runs any model on any job, works in the apps your teams already use, and turns every graded run into a better specialist.

  • Built in
  • Partly, or you build it
  • Not offered
  • Models

    Recursion Managed AgentsBuilt in: Any model, per agentFrontier, open-weight or your own fine-tuned models
    Other managed agentsNot offered: One lab’s modelsLocked to the provider
  • Multi-agent

    Recursion Managed AgentsBuilt in: A team of specialistsEach with its own model, tools and access
    Other managed agentsPartly, or you build it: SubagentsAll on the same lab’s models
  • Integrations

    Recursion Managed AgentsBuilt in: 100+ business apps, built inPer-tool permissions, credentials in a vault
    Other managed agentsPartly, or you build it: Bring your ownYou connect and maintain MCP servers
  • Grading

    Recursion Managed AgentsBuilt in: Every run gradedA separate grader, revising until it passes
    Other managed agentsPartly, or you build it: Not always built inOften yours to build
  • Memory

    Recursion Managed AgentsBuilt in: Learns automaticallyEvery run makes the next one better
    Other managed agentsPartly, or you build it: Varies by providerMemory stores or session history
  • Training

    Recursion Managed AgentsBuilt in: Specialist models from your runsGraded runs become evals and RL training
    Other managed agentsPartly, or you build it: Separate fine-tuning, where offeredYou bring the data and graders
  • Built for

    Recursion Managed AgentsBuilt in: Every teamNo-code console, iPhone app, CLI and SDKs
    Other managed agentsPartly, or you build it: DevelopersAPIs, SDKs and a console
Contact us

Other managed agents: the leading AI labs' agent platforms, based on their public documentation as of September 2026.

Real access to your systems. Never the keys.

Every session runs in an isolated sandbox with only the tools and credentials you grant it. Recursion Managed Agents attaches each credential to the approved call, so the model never sees it, and keeps it to the hosts you allow. Every model turn and tool call lands in an audit-ready transcript, backed by Labelbox's enterprise security and compliance program.

  • SOC 2 Type II
  • GDPR
  • CCPA
  • NIST 800-171
  • PCI DSS
A support triage agent runs in an isolated sandbox with no credentials inside. Each tool call goes through Recursion Managed Agents, which attaches a short-lived token from the vault to approved calls to Salesforce (read only) and Zendesk (changes data), blocks a request to a host that is not on the allowlist, and records every call in the session transcript.

Put your first agent to work

Start with one job your team repeats every week. Connect the apps it touches, describe what a good result looks like, and set a schedule. From then on, it gets done before anyone thinks to ask.