Factory
by Factory
Agent-native software development — Droids that take whole engineering tasks
Factory builds Droids, autonomous software-development agents that take an entire task — a migration, a refactor, a code review, an incident triage — rather than autocompleting the next line. The same agent runs in a desktop app, a CLI, the browser and Factory-managed cloud machines, against whichever frontier model you choose, with enterprise controls for regulated engineering organisations.
Factory sells an agent-native software development platform built around Droids: autonomous coding agents that take on a complete task — a refactor, a framework migration, a code review, an incident triage — rather than autocompleting the next line in an editor. The deliberate design choice is surface- and model-independence. The same Droid runs in the Factory desktop app, in the Droid CLI, in a browser at app.factory.ai, and on Factory-managed cloud machines called Droid Computers, and it can be pointed at different frontier models rather than being welded to a single vendor's. Droids connect to repositories and the systems around them — GitHub, issue trackers and incident-management tooling — so an agent can read a ticket, open a branch, and hand back a reviewable pull request, and the documented automation targets are triage, code review, QA, documentation and incident response. Factory's pitch to enterprises is orchestration, policy and auditability rather than raw model quality: Business and Enterprise tiers add SSO, SAML and SCIM, zero data retention, audit logging, customer-managed encryption keys, data-residency guarantees, dedicated compute with partitioned inference, and on-premise deployment. On benchmarks, Factory's entries took first and third place on Terminal-Bench at 58.8% and 52.5% task success, and its earlier Code Droid posted a state-of-the-art SWE-bench result in 2024. Founded in San Francisco in 2023, the company raised a $50 million Series B in September 2025 from NEA, Sequoia Capital, NVIDIA and J.P. Morgan after a $15 million Sequoia-led Series A, with angels including Frank Slootman, Nikesh Arora and Aaron Levie. Named users span Blackstone, Adyen, Wipro, Comarch, Groq, Chainguard, Podium, MongoDB, EY, Bayer, Zapier and Clari.
A VP Engineering or platform lead at a regulated enterprise who wants autonomous coding agents under SSO, audit logging and data-residency controls, and refuses to standardise on a single model vendor.
Delegate whole engineering tasks — migrations, refactors, reviews, incident triage — to agents that run identically in the CLI, IDE, browser and cloud, with an audit trail an enterprise can defend.
At a Glance
- Category
- Developer Tools
- Pricing
- Subscription, Contact for pricing
- Target Market
- CTOs, VPs of Engineering, Enterprise Developers, Platform and DevEx Teams
- Deployment
- Cloud-first, Hybrid, API-based
- Founded
- 2023
- Headquarters
- San Francisco, CA, United States
Key Features
- ✓Droids (autonomous task agents)
Agents that own an entire unit of work end to end rather than suggesting the next line of code.
- ✓Surface independence
The same Droid behaves identically in the desktop app, CLI, browser and Factory-managed cloud machines, so workflows do not fork.
- ✓Droid Computers
Factory-managed cloud machines that run agents in the background, freeing the developer's local environment during long tasks.
- ✓Model independence
Agents can be pointed at different frontier models, so a model change does not require replacing the tooling around it.
- ✓Repository and workflow integrations
Connects to GitHub, issue trackers and incident-management tools so agents work inside existing engineering process.
- ✓Enterprise identity and data controls
SSO, SAML and SCIM, zero data retention, audit logging, customer-managed encryption keys and data-residency guarantees.
- ✓On-premise and dedicated compute
Enterprise tier offers on-premise deployment and partitioned inference for organisations that cannot use shared capacity.
- ✓Agent-readiness dashboard
Usage statistics and billing tracking that show where agents are actually being used across a team.
Capabilities
Use Cases
- •Large-scale codebase migration
Run agents across repositories to move a framework or language version, with each change surfacing as a reviewable pull request.
- •Automated code review pass
A review Droid inspects incoming pull requests and leaves findings before a human reviewer spends time on them.
- •Incident triage and response
An agent pulls context from logs, tickets and source, drafts a diagnosis, and shortens the path to a fix during an outage.
- •Backlog burn-down of mechanical work
Delegate dependency bumps, test coverage gaps and documentation debt that engineers routinely deprioritise indefinitely.
- •Regulated-industry AI rollout
Deploy agents under zero data retention, audit logging and data residency so security review can actually approve the tool.
Ideal For
Best For
- ✓Large-scale framework or language migrations across many repositories where the work is mechanical but the volume is prohibitive
- ✓Automated first-pass code review and QA that runs before a human reviewer is asked to look
- ✓Incident response and triage, where an agent gathers context across logs, tickets and code before paging an engineer
- ✓Regulated engineering organisations needing zero data retention, customer-managed keys, data residency or on-premise deployment
- ✓Teams that want to switch between frontier models without changing their agent tooling
Not Ideal For
- ✗Individual developers and small teams on a budget — there is no free tier, entry is $20/month, and the useful cloud-machine capacity starts at $100/month
- ✗Organisations already deeply bought into GitHub Copilot or a cloud vendor's bundled agent, where distribution and procurement inertia usually beat a standalone tool
- ✗Buyers who need published, self-serve team pricing: Business and Enterprise are contact-sales only
- ✗Greenfield prototyping where an interactive in-editor assistant is a better fit than long-running autonomous task execution
Integrations
Deployment
Market Analysis
Pros
- ✓Strongest independent evidence on this list is benchmark data, not marketing: first and third on Terminal-Bench at 58.8% and 52.5% task success
- ✓Model and surface independence is a genuine hedge — the buyer is not stranded when the leading model changes
- ✓Enterprise security posture is unusually complete for a company at this stage: zero data retention, customer-managed keys, data residency and on-premise
- ✓Named production users across finance, pharma and consulting (Blackstone, Adyen, EY, Bayer, MongoDB, Wipro) indicate real procurement, not pilots only
Cons
- ✗Distribution is the structural risk: GitHub Copilot, Google and Microsoft bundle agentic features into tools enterprises already deploy, and independent analysis calls that a formidable advantage a startup struggles to overcome
- ✗Benchmark leadership is not production performance — the same analysis warns that real environments bring flaky test suites, mixed stacks and human handoffs that Terminal-Bench does not model
- ✗Headline outcome claims such as 31x faster feature delivery and 96.1% shorter migration times come from Factory's own case studies and are not independently verified
- ✗Practitioner discussion is thin: Hacker News threads about Droids consistently draw single-digit points, so there is little unfiltered production feedback to weigh against the vendor's numbers
- ✗No free tier and no self-serve team plan, so evaluation requires either a personal $20/month subscription or a sales cycle
Pricing
Pro
From $20/mo
- ✓Desktop, CLI and SDK access
- ✓Cloud and local background agents
- ✓Billing tracking and usage statistics
- ✓Agent-readiness dashboard
Plus
From $100/mo
- ✓Everything in Pro
- ✓Approximately 5x Pro usage limits
- ✓Droid Computers (Factory-managed cloud machines)
Max
From $200/mo
- ✓Everything in Plus
- ✓Approximately 10x Pro usage limits
- ✓Early access to new features
Business
Contact for pricing
- ✓Up to 150 seats
- ✓Custom usage limits
- ✓SSO, SAML and SCIM
- ✓Zero data retention and audit logging
- ✓Dedicated onboarding and support
Enterprise
Contact for pricing
- ✓Unlimited seats
- ✓On-premise deployment
- ✓Dedicated compute with partitioned inference
- ✓Customer-managed encryption keys and data residency
- ✓Sub-organisations and full admin controls
- ✓Dedicated account manager and priority SLA
Individual tiers are published and flat — $20, $100 and $200 per month — but they are differentiated by opaque 'usage limits' expressed only as multiples of Pro (roughly 5x at Plus, 10x at Max), so the real cost of a heavy agent workload is not knowable before you run it. There is no free tier and no advertised trial. Team pricing is seat-based and entirely contact-sales: Business caps at 150 seats, Enterprise is unlimited, and everything a security review will demand — on-premise deployment, customer-managed encryption keys, data residency, sub-organisations — sits behind the Enterprise quote.
Security & Compliance
Sources
This page was written from 6 sources, 4 on domains other than factory.ai.
Stay Ahead of the Curve
Weekly enterprise AI insights for technology leaders. No spam, no vendor pitches—unsubscribe anytime.
SubscribeRelated Products
Confident AI
Hosted LLM evaluation, observability and red teaming built on the open-source DeepEval framework
Respan
Observability, evals and an LLM gateway for AI agents in one control plane
Meta Muse Code
Meta's terminal coding agent for large repositories, with persistent background agents and the most aggressive token pricing in the category
Niteshift
The full-stack cloud for coding agents — real environments, verified pull requests
Mentioned In
ABB and NVIDIA Close the Sim-to-Real Gap: 80% Faster Robot Deployment
ABB and NVIDIA Close the Sim-to-Real Gap. For CFOs and finance leaders: cost implications, budget planning, and ROI benchmarks from enterprise AI deployments.
March 14, 2026Enterprise AINTT DATA + NVIDIA AI Factories: Closing the Pilot-Production Gap
Enterprise AI analysis: NTT DATA + NVIDIA AI Factories. Strategic insights, ROI considerations, and implementation guidance for technical and business leader...
March 16, 2026AI InfrastructureNutanix Agentic AI Stack: Scaling Enterprise AI Factories at Lower Cost per Token
Enterprise AI analysis: Nutanix Agentic AI Stack. Strategic insights, ROI considerations, and implementation guidance for technical and business leaders eval...
March 16, 2026NVIDIANVIDIA GTC 2026 Day 1: $1 Trillion Revenue Path, Vera Rubin Platform, and OpenClaw Partnership
NVIDIA GTC 2026 Day 1. For CFOs and finance leaders: cost implications, budget planning, and ROI benchmarks from enterprise AI deployments.
March 16, 2026Enterprise AIThe Enterprise AI ROI Era Arrives: What 4,000 Deployments Tell Us
The Enterprise AI ROI Era Arrives. For CTOs and platform teams: architecture decisions, integration challenges, and deployment strategies for production AI.
March 17, 2026Enterprise AIIBM and NVIDIA Close the Pilot-to-Production Gap: 83% Cost Savings at Nestlé
IBM and NVIDIA Close the Pilot-to-Production Gap. For CFOs and finance leaders: cost implications, budget planning, and ROI benchmarks from enterprise AI dep...
March 17, 2026GeelyGeely Expands NVIDIA Partnership Across Physical, Enterprise, and Industrial AI
Geely Auto Group announced a major expansion of its NVIDIA partnership spanning autonomous vehicles, cloud-based agentic AI, and factory automation—signaling a transformation from car manufacturer to AI organization.
March 22, 2026NVIDIANVIDIA GTC 2026 Final Roundup: $1 Trillion Revenue, 50x Performance Leap, and the Groq Acquisition That Changes Everything
NVIDIA GTC 2026 roundup: $1T revenue forecast, 50x performance leap, Groq acquisition. For enterprise leaders: strategic implications of accelerated computin...
March 22, 2026Enterprise AIVerily's $300M Round: Alphabet's Healthcare AI Exit
Alphabet sold its controlling stake in Verily after 10 years. The $300M precision health AI round with Series X Capital signals enterprise healthcare is standalone.
March 20, 2026ROILenovo Plus NVIDIA Hybrid AI Cuts Costs 8x With ROI in Six Months
Lenovo Plus NVIDIA Hybrid AI Cuts Costs 8x With ROI in Six Months. For enterprise decision-makers: strategic analysis, cost implications, and implementation ...
March 22, 2026Enterprise AIByteDance's $2.5B Bet on AI Infrastructure — What It Means for Enterprise Buyers
ByteDance's $2.5B Bet on AI Infrastructure — What It Means for Enterprise Buyers. For CFOs and finance leaders: cost implications, budget planning, and ROI b...
March 13, 2026Agentic AIAgentic AI Market Explodes: $139B by 2034
Agentic AI Market Explodes. For CFOs and finance leaders: cost implications, budget planning, and ROI benchmarks from enterprise AI deployments.
March 15, 2026AI agentsAI Agent Adoption in 2026: What NVIDIA's Research Reveals Ab
AI Agent Adoption in 2026: What NVIDIA's Research Reveals Ab Key insights for enterprise AI leaders on what this means and what to do next.
March 16, 2026AI AgentsWhy Enterprise AI Agents Fail: 2026 Gartner and IDC Data
Gartner and IDC 2026 data: 89% of enterprise AI agent pilots stall. The 3 failure modes every CIO must address before scaling AI agents to production.
March 16, 2026Vendor SelectionGPT-5.4 Mini and Nano Launch: How OpenAI Just Undercut Anthropic 5x (And Why Google Still Wins on Price)
GPT-5.4 Mini and Nano Launch. For enterprise decision-makers: strategic analysis, cost implications, and implementation guidance for AI investments and vendo...
March 22, 2026GoogleGoogle Stitch Made Figma Drop 8%: AI Design Just Got Real
Google Stitch's March 2026 update with AI-powered voice design and design agents caused Figma stock to drop 8%. Enterprise teams must now decide when complexity moats become liabilities in conversational AI workflows.
March 22, 2026DellDell AI Factory: 2.6x ROI Across 4,000 Deployments
Dell AI Factory deployment data from 4,000 customers shows 2.6x first-year ROI, 12x faster data indexing, and 19x faster time-to-first-token. Enterprise leaders must evaluate build vs buy infrastructure decisions based on actual production benchmarks.
March 22, 2026AI InfrastructureGimlet Labs Raises $80M to Solve AI's Biggest Waste Problem
Stanford founder's multi-silicon cloud delivers 3-10x faster AI inference by orchestrating workloads across CPUs, GPUs, and specialized chips. Already running at top frontier labs and hyperscalers with 8-figure revenue.
March 23, 2026CloudflareCloudflare Dynamic Workers Run AI Agent Code 100x Faster Than Containers
Cloudflare Dynamic Workers Run AI Agent Code 100x Faster Than Containers. For enterprise decision-makers: strategic analysis, cost implications, and implemen...
March 25, 2026GoogleGoogle TurboQuant Cuts AI Memory 6x With Zero Accuracy Loss and 8x Speedup
Google TurboQuant Cuts AI Memory 6x With Zero Accuracy Loss and 8x Speedup. For enterprise decision-makers: strategic analysis, cost implications, and implem...
March 25, 2026ArmArm Enters AI Chip Manufacturing With $15B Revenue Target as Citi Calls It 'Most Significant Shift' in Company History
Arm Enters AI Chip Manufacturing With $15B Revenue Target as Citi Calls It 'Most Significant Shift' in Company History. For enterprise decision-makers: strat...
March 26, 2026AI HardwareAnvil's $5.5M Bet: Why 'Legos for Robots' Is the AWS of Physical AI
Most companies spend 6+ months building robot prototypes. Anvil ships custom robots in 48 hours for $5K-$10K. Here's why modular hardware just became the bottleneck-breaker for enterprise AI.
April 2, 2026