AI coding agents: which company is ahead?

In our agentic AI market deck, you will find everything you need to understand the market
SUMMARY
Cursor is clearly ahead in the AI coding agent race today. Cognition is the strongest startup challenger, Replit leads natural-language application creation, and Factory remains the most credible candidate to disrupt the current top three.
Cursor’s lead is primarily commercial, not just technical. Its approximately $4 billion annualized revenue is roughly eight times Cognition’s latest run rate and around seventeen times Replit’s last achieved sales figure.
The company also covers more of the software-development workflow than any startup rival. Cursor now combines interactive coding, local and cloud agents, code review, command-line tools, mobile access, enterprise controls, proprietary models, and multi-agent workspaces.
Cognition is the most dangerous challenger because it combines two previously separate products. Windsurf supports developers while they work, while Devin handles longer tasks that can be delegated and completed remotely.
Replit is competing in a slightly different race. It is strongest when a customer wants to turn an idea into a working and deployed application without separately managing an editor, database, infrastructure provider, and hosting stack.
Factory remains difficult to rank precisely. Its long-running Missions, persistent agent environments, model routing, enterprise customers, and reported monthly revenue doublings look impressive, but the company still does not disclose enough revenue or deployment-depth data.
Independent evidence suggests that task selection matters almost as much as product selection. Agents perform much better on documentation and bug fixes than on complex new features, and a large share of agent-generated pull requests still gets rejected.
Cursor has the strongest enterprise evidence because its claims are supported by unusually large deployments and quantified customer outcomes. Stripe, Nvidia, Salesforce, and Amplitude provide a more convincing picture than a broad Fortune 500 usage percentage alone.
The market’s strongest moats are increasingly based on workflow ownership and usage data rather than exclusive model access. Frontier models can be copied or replaced quickly, but deeply embedded development environments, repository context, enterprise policies, and accumulated agent feedback are harder to reproduce.
Claude Code, Codex, and GitHub are serious threats to every startup in the ranking. Claude Code may already rival Cursor in developer influence, while GitHub can become the neutral control layer through which companies use several competing agents.
The ranking is most reliable at the top. Cursor, Cognition, and Replit have enough commercial, product, and customer evidence to support a clear order; Factory, Augment Code, and Poolside remain harder to separate because their public disclosures are thinner.
The current hierarchy is therefore clear but not permanent: Cursor leads overall, Cognition leads pure task delegation, Replit leads AI-native application creation, and Factory has the clearest path to an upset if its reported momentum converts into visible commercial scale.

This market map, featured in our agentic AI market deck, highlights top companies and startups in the agentic AI market
Which AI coding agent startups belong in this race?
We compare Cursor, Cognition, Replit, Factory, Augment Code, and Poolside because these six companies are currently making the most serious startup bets on agents that can complete substantial software work.
An AI coding agent must be able to understand a repository or application, change multiple files, use development tools, run code or tests, and carry a task beyond simple autocomplete. The six companies below meet most or all of those conditions, although they target different customers.
Cursor started as an AI-native code editor and has since expanded into local agents, cloud agents, code review, command-line tools, automations, and multi-agent workspaces. Cognition combines Devin, which handles delegated software tasks, with the Windsurf development environment. Replit concentrates on building and deploying new applications from natural-language instructions. Factory sells enterprise agents called Droids across the software development lifecycle. Augment Code focuses heavily on large repositories and organizational context. Poolside is building its own coding models alongside agent products.
We exclude Lovable, Bolt, Base44, and similar app builders because they mostly compete for fast greenfield application creation, often among users with limited coding experience. Replit overlaps with that group, but its long history as a development environment and its growing enterprise business justify its inclusion.
Magic has disclosed too little current product adoption to support a useful comparison, while CodeRabbit is primarily a code-review company. Anthropic, OpenAI, Microsoft, and GitHub are analyzed later because they are major threats to the startups rather than startups themselves.
Cursor also requires a small qualification. SpaceX agreed in June to acquire Anysphere, Cursor’s parent company, for $60 billion in stock, with the transaction expected to close later in 2026. We still include Cursor because the deal remains pending and the company built its current position independently.
Funding totals should be treated as approximate because databases sometimes classify secondary transactions and extensions differently.
| Startup | Main AI coding agent position | Approximate cumulative funding |
|---|---|---|
| Cursor | AI-native development workspace spanning interactive and autonomous coding | About $3.4B |
| Cognition | Devin autonomous agent plus the Windsurf development environment | About $2.1B |
| Replit | Natural-language application creation, development, and deployment | About $922M |
| Poolside | Proprietary coding models, agents, and supporting infrastructure | About $626M |
| Augment Code | Enterprise agents built around large-codebase and company context | About $252M |
| Factory | Model-flexible enterprise software-development agents | About $220M |
Is Cursor clearly ahead of other AI coding agent startups today?
Cursor is clearly ahead of the other AI coding agent startups overall, although Cognition, Replit, and Factory each lead a narrower part of the market.
The largest gap appears in revenue. Forbes and The Wall Street Journal reported that Cursor passed approximately $4 billion in annualized revenue in early June. Cognition reached $492 million, while Replit’s latest achieved figure was about $240 million. Cursor was therefore generating roughly eight times Cognition’s run rate and about seventeen times Replit’s last disclosed sales figure.
Cursor also covers more of the normal development workflow than it did a year ago. Developers can work directly with an agent inside the editor, send tasks to cloud environments, run agents through the command line, review pull requests with Bugbot, manage several agents at once, or start work from a phone. Recent enterprise additions include centralized organizations, budget controls, security settings, and separate environments for cloud agents.
Cognition poses the clearest challenge because it combines two forms of AI coding that were previously sold separately. Windsurf helps developers while they work, while Devin can take a task and continue working asynchronously. Cognition can therefore serve both supervised coding and full task delegation.
Replit leads when the customer wants to create a new application quickly without configuring an editor, database, hosting provider, or deployment pipeline. Factory is building a credible enterprise alternative around model choice and long-running workflows. Augment and Poolside are technically ambitious, but their commercial positions remain far less visible.
The market has one broad startup leader and several narrower specialists. Cursor’s lead is large. It just does not extend evenly across every coding workflow.
If you want more recent data on this point, please see our latest agentic AI market report.

As this chart shows, and as featured in our agentic AI market deck, search interest in AI agents has been rising rapidly
Which AI coding agent has the strongest commercial traction?
Cursor has the strongest commercial traction in AI coding by a wide margin, and Cognition is the only startup currently operating in the same broad financial league.
Cursor crossed $500 million in annual recurring revenue in June 2025 and $1 billion by November. It then doubled again to roughly $2 billion around February before reaching the latest reported figure. The climb from $500 million to approximately $4 billion took about one year, an eightfold increase from an already substantial base.
Cognition’s growth also stands out. Devin reportedly moved from $1 million in annual recurring revenue in September 2024 to $73 million by June 2025. After acquiring Windsurf, the combined company reached a $492 million annualized revenue run rate. The acquisition gave Cognition existing enterprise customers, a development environment, and a larger sales organization.
Replit generated about $240 million in annual sales when its chief executive discussed the business in October 2025. The company currently says it is on track to reach a $1 billion run rate before the end of 2026, but that figure remains a forecast.
Factory says its revenue doubled every month for six consecutive months, which mathematically represents a sixty-fourfold increase. Without a starting or ending figure, though, we cannot tell whether Factory remains relatively small or has already reached major scale.
Augment and Poolside have not released current revenue figures that can be compared meaningfully with Cursor, Cognition, or Replit.
| Company | Latest useful revenue evidence | Approximate scale versus Cursor | What we can reasonably conclude |
|---|---|---|---|
| Cursor | About $4B annualized | 100 | Clear commercial leader |
| Cognition | $492M annualized | 12 | Strong second place, but still about eight times smaller |
| Replit | About $240M achieved; $1B targeted | 6 | Large business, although the newest headline is still a target |
| Factory | Undisclosed | Unknown | Very fast reported growth from an unknown base |
| Augment Code | Undisclosed | Unknown | Enterprise positioning is clearer than commercial scale |
| Poolside | Undisclosed | Unknown | Funding and technical ambition exceed visible adoption |
Which AI coding agent startup is growing fastest and has the strongest momentum now?
Cognition currently has the strongest challenger momentum, while Cursor is adding far more revenue and expanding across more parts of the development workflow.
Cognition reached $73 million in annual recurring revenue before completing the Windsurf acquisition and later reported $492 million. That is roughly a 6.7-fold increase in less than a year, although part of the jump came from acquiring Windsurf rather than organic Devin sales alone. Cognition also said enterprise Devin usage grew 50% month over month for six consecutive months. If applied consistently to the same usage measure, that would represent an increase of more than eleven times.
Recent product and financing moves reinforce that growth. Cognition raised more than $1 billion at a $26 billion valuation, launched Devin Desktop, added a productivity guarantee, expanded into government deployments, and tightened the integration between Devin and Windsurf. It is also using consulting firms such as Infosys and Cognizant to reach large companies faster.
Cursor’s percentage growth was slightly less dramatic during its latest stretch, but the dollars were much larger. The company moved from more than $1 billion in annualized revenue in November to approximately $2 billion around February, then doubled again. It has also launched cloud-agent environments, enterprise organizations, mobile access, visual prompting, a cheaper Bugbot, an agent SDK, and continual improvements to Composer.
Factory may be growing faster than both in percentage terms. Six consecutive monthly revenue doublings would be extraordinary at almost any starting level. Its recent Series C, persistent Droid Computers, Missions product, model routing, desktop agent, and new chief revenue officer suggest that Factory is preparing for a much larger commercial push. The missing revenue figure still prevents a direct comparison.
Replit experienced the most extreme earlier acceleration, moving from $2.8 million to more than $150 million in annualized revenue before reaching approximately $240 million. Its recent momentum is strongest in application creation, with Agent 4, parallel task execution, custom instructions, and a large new financing round.
Cognition’s $492 million run rate remains far behind Cursor. Still, it is currently the startup most likely to narrow the gap, while Factory is the smaller challenger with the greatest upside.
If you want more recent data on this point, please see our latest agentic AI market report.

This chart, included in our agentic AI market deck, illustrates yearly VC funding for agentic AI startups
Which AI coding agent is most mature and autonomous for professional software teams?
Cursor currently offers the most complete AI coding agent product for professional software teams, while Cognition’s Devin remains the strongest option for handing off an entire task.
Cursor’s product has grown around the way developers already work. A developer can ask questions, edit several files, run a local agent, send longer jobs to cloud agents, compare parallel attempts, review generated code, or continue the work from another device. Cursor 3 brought local and remote agents into one workspace instead of making users manage separate tools and windows.
The company has also built the administrative features that large customers eventually demand. Cursor Enterprise supports organization-level administration, separate team budgets, security controls, usage visibility, and managed development environments. Recent pricing changes added heavier-use seats and better alerts after customers complained that agent consumption had become difficult to predict.
Cognition is more focused on delegation. A user gives Devin a task, and the agent creates a plan, works inside a development environment, uses tools, runs tests, and returns a pull request or completed output. Cognition has also added Devin Desktop, Devin Review, government products, MCP connections, and Windsurf integration, allowing teams to move between supervised coding and autonomous work.
Factory is pursuing longer and more complex autonomous workflows. Its Missions product coordinates several Droids across work that Factory says may take days or weeks. Droid Computers preserve files, credentials, services, and running processes between sessions, so the agents do not have to recreate their environment every time.
Replit is the most mature option for building a new application from an idea. Agent 4 can design, develop, test, and deploy inside the same cloud platform. Parallel tasks and automatic conflict handling allow several parts of an application to be built at once, and Replit says its system resolves merge conflicts automatically about 90% of the time.
Cursor is also becoming more autonomous through cloud agents, automations, parallel workspaces, and its agent SDK. Its advantage is that teams can increase delegation gradually rather than moving immediately to a fully remote-agent workflow.
Full autonomy remains less common than product demos suggest. A recent study of agent-generated fixes found that 46.41% of pull requests from Copilot, Devin, Cursor, and Claude were rejected. Common reasons included incorrect approaches, failed tests, incomplete implementations, and lost sessions.
Cursor is the most mature overall platform. Devin is better for pure task delegation, Factory has the boldest long-horizon architecture, and Replit is strongest for autonomous greenfield application creation.
Which AI coding agent performs best on real coding tasks?
No startup wins every type of coding task, but Cursor currently has the strongest real-world record for bug fixes, while Devin shows the clearest improvement over time.
A 2026 study examined 7,156 pull requests created by Codex, GitHub Copilot, Devin, Cursor, and Claude Code. Cursor achieved an 80.4% acceptance rate on fix tasks. Claude Code performed best on documentation and feature work. Codex was consistently strong across categories, while Devin was the only agent whose acceptance rate improved steadily, rising by about 0.77 percentage points per week over thirty-two weeks.
Task type mattered as much as the agent. Documentation pull requests reached an 82.1% acceptance rate across the dataset, compared with 66.1% for new features. Choosing suitable work for an agent may therefore create a larger gain than switching from one leading product to another.
The wider AIDev dataset contains 932,791 agent-authored pull requests across more than 116,000 repositories. That volume confirms that AI coding agents are being used in real projects, although it does not prove that one product creates more valuable software because repositories, users, and review standards differ sharply.
Company benchmarks need more care. Factory reported a 58.75% score on Terminal-Bench and held first place when the result was published. Cursor’s Composer models have also posted strong results in Cursor’s own harness. Newer OpenAI models later reported much higher results on updated benchmark versions, showing how quickly a lead can disappear.
Cursor currently has the best independent result for bug fixes. Devin has the strongest improvement trend, Replit is better suited to complete new applications, and Factory looks competitive on terminal-heavy work.

This chart, included in our agentic AI market deck, shows how Cognition is positioned in agentic AI
Which AI coding agent startup has won the strongest enterprise customers?
Cursor has the strongest enterprise customer base among AI coding agent startups, and its best deployments go well beyond small experiments.
Cursor says more than 70% of the Fortune 500 now uses its platform. That percentage alone is broad, but several customer examples show deeper adoption.
Stripe rolled Cursor out to roughly 3,000 engineers. Nvidia reported three times more committed code across 30,000 developers. Salesforce reported a velocity increase above 30%, while Amplitude said it shipped three times more production code. Cursor has also published examples from Dropbox, PayPal, Coinbase, National Australia Bank, Wayfair, Box, and Faire.
Cognition’s customers include Goldman Sachs, Citi, Dell, Cisco, Palantir, Mercedes-Benz, NASA, Santander, Nubank, and Mercado Libre. Partnerships with Infosys and Cognizant could matter even more because they give Cognition a route into many large corporate accounts. The company has not disclosed enough seat counts or contract values to show how deep most deployments are.
Factory says hundreds of thousands of developers use Droids daily at companies including Nvidia, Adobe, EY, Palo Alto Networks, and Adyen. The announcement does not show how much usage comes from paid contracts or how many customers have expanded.
Replit says people from 85% of the Fortune 500 have built with its platform and names customers including Atlassian, LabCorp, PayPal, Zillow, Adobe, and UKG. Its strongest position often sits outside central engineering teams, where employees create internal tools, prototypes, and smaller applications.
| Company | Best available enterprise evidence | How convincing is it? |
|---|---|---|
| Cursor | Large deployments at Stripe and Nvidia, plus repeated quantified case studies | Strongest evidence of wide and deep enterprise adoption |
| Cognition | Major banks, technology companies, manufacturers, NASA, Infosys, and Cognizant | High-quality customer list, but fewer disclosed seat and contract figures |
| Factory | Hundreds of thousands of reported daily users at named large enterprises | Strong emerging position, with limited contract detail |
| Replit | Users from 85% of the Fortune 500 and several named business customers | Very broad reach, although “users from” may include small pockets of adoption |
| Augment Code | Enterprise-focused product and selected customer stories | Credible, but insufficient aggregate data |
| Poolside | Few comparable public customer disclosures | Commercial position remains difficult to judge |
Which AI coding agent gives customers the best value?
Cursor offers the safest value for established development teams, while Replit can produce better economics for new applications because it bundles coding, infrastructure, and deployment.
Subscription prices reveal only part of the cost. Agent usage also creates model charges, cloud-environment costs, failed attempts, and review work. A cheaper agent can become expensive if engineers spend hours correcting its output.
Cursor currently sells several individual, team, and enterprise tiers. Its newer Teams structure separates normal users from heavy agent users and gives administrators better spending controls. Cursor’s growing use of its own models may also reduce its dependence on expensive external inference, although the company has not published gross margins by model.
Devin uses plans and metered consumption, so customers pay more as autonomous work increases. That structure aligns price with usage but makes complex tasks harder to budget, especially when the agent spends time understanding an environment or retrying failed work.
Factory’s paid self-serve plans begin at $20 per month, with larger business and enterprise arrangements. Its Router can choose between models according to quality, reliability, and cost, which could help enterprises avoid using the most expensive model on every task.
Replit’s value comes from bundling. A customer can generate the application, store data, collaborate, run the software, and publish it through one platform. Business Insider reported that Replit’s overall gross margin was about 23% in mid-2025, compared with approximately 80% for enterprise business. That suggests heavy consumer agent use was expensive, while larger contracts were much healthier.
Independent research remains mixed. A meta-analysis of 23 studies found a moderate positive effect on programming productivity overall, with smaller gains in open-source and enterprise settings. Another Microsoft study found that engineers using command-line agents merged roughly 24% more pull requests, while acknowledging that pull-request volume does not equal business value.
Cursor is the safest broad purchase for existing engineering teams. Replit has the clearest value for greenfield applications. Devin and Factory make more sense when a company has repeatable work that genuinely justifies heavier autonomous usage.
If you want more recent data on this point, please see our latest agentic AI market report.

This chart, included in our agentic AI market deck, illustrates yearly funding for agentic AI startups
Which AI coding agent company has the funding and infrastructure to scale?
Cursor has already shown that it can support the largest AI coding agent deployments, while Cognition has the capital to challenge it.
Cursor has raised approximately $3.4 billion, including a $2.3 billion Series D at a $29.3 billion valuation. It now serves millions of developers and large engineering organizations while operating local products, cloud agents, code review, enterprise administration, and proprietary model training. The pending SpaceX deal could give Cursor even greater access to computing infrastructure.
Cursor has also converted its capital into a much larger business than its startup rivals. Its revenue run rate is now comparable to, or larger than, all the private funding it raised. That does not prove profitability, but it shows unusually strong commercial output for a company investing heavily in model training and cloud agents.
Cognition has raised approximately $2.1 billion and was valued at $26 billion after its latest round. That gives it the resources to operate remote development environments, pay for model usage, support major enterprises, and integrate Devin with Windsurf. Its early period was particularly efficient: Devin reached $73 million in recurring revenue while Cognition said historical net burn remained below $20 million.
Replit has raised about $922 million and already runs a broad cloud platform for more than fifty million users. It must support development environments, databases, hosting, deployment, collaboration, and agent inference together. That infrastructure is expensive, but it also lets Replit control the whole path from idea to working application.
Factory has raised about $220 million, including a $150 million Series C. Its persistent Droid Computers and model-routing layer are designed for enterprise use, but the company has not yet shown several deployments at Cursor-like scale or disclosed enough revenue to measure capital efficiency.
Poolside has raised about $626 million while taking on expensive model and infrastructure work. The Financial Times reported that a major CoreWeave data-center agreement collapsed after delays, disrupting Poolside’s planned Texas project and a proposed $2 billion fundraising round. The episode showed the risk of building heavy infrastructure before commercial adoption is established.
Cursor currently combines the strongest infrastructure, customer scale, and capital conversion. Cognition has the balance sheet and enterprise channels to close part of the gap, while Replit has already proved it can operate at very large user volume.
What can the leading AI coding agent startups defend from copycats?
Cursor currently has the strongest defensible position because competitors would need to copy its workflow, distribution, enterprise integrations, product data, and model work together.
The editor gives Cursor frequent daily use. Once a company has indexed repositories, created rules, connected tools, trained employees, set budgets, and approved security policies, switching platforms becomes inconvenient enough to support retention.
Cursor also learns from millions of agent sessions. That data can reveal which suggestions survive, where agents fail, how developers interrupt them, and which workflows lead to accepted code. The company has used that experience to train Composer and improve its agent system rather than relying entirely on outside models.
Cognition’s advantage comes from long-running delegated work. Devin creates data about planning, environment setup, retries, tool use, failed sessions, and human review. Windsurf adds interactive developer behavior and codebase navigation.
Replit sees the whole application cycle, from the first prompt through deployment and later edits. Factory is building around model-independent orchestration, while Augment’s Context Engine aims to preserve knowledge across large repositories and teams.
Anthropic, OpenAI, GitHub, and Microsoft can copy many individual features. The durable advantage comes from owning an important workflow and learning faster from real use. Cursor currently has the strongest combination of both.
If you want more recent data on this point, please see our latest agentic AI market report.

This chart, included in our agentic AI market deck, compares the main business model options for autonomous AI agent platforms
Can Claude Code, Codex, and GitHub overtake the AI coding agent startups?
Claude Code, OpenAI Codex, and GitHub can overtake individual AI coding agent startups, and Claude Code may already rival Cursor on developer influence.
Anthropic reported that Claude Code passed $2.5 billion in run-rate revenue after more than doubling from the start of 2026. Weekly active users doubled over the same period, business subscriptions quadrupled, and enterprise customers generated more than half of Claude Code revenue. Anthropic also cited an external estimate that Claude Code produced 4% of public GitHub commits worldwide.
Claude Code is therefore already much larger than Cognition and Replit financially. It gives developers direct access to a frontier model without requiring a separate AI-native editor, which pushed Cursor to invest more heavily in Composer.
OpenAI is attacking the same market through a dedicated Codex application, the command line, an IDE extension, and the web. OpenAI says GPT-5.3-Codex reached 56.8% on SWE-Bench Pro and 77.3% on Terminal-Bench 2.0 while running 25% faster than the previous version. Company-run benchmarks require caution, but they show how quickly model providers can erase a startup’s technical lead.
GitHub may have the strongest distribution position. Agent HQ lets customers run agents from Anthropic, OpenAI, Google, Cognition, and other providers inside GitHub and Visual Studio Code. GitHub can become the control layer through which several agents receive tasks, access repositories, and submit pull requests.
Cursor, Cognition, Replit, Factory, and Augment remain strongest where they provide a better workflow than direct model access. Their position weakens whenever Claude, Codex, or GitHub can offer the same experience through software customers already use.
Can we trust the numbers used to rank AI coding agent startups?
We can trust the broad ranking of AI coding agent startups, but several precise gaps remain uncertain because private companies report different kinds of numbers.
Revenue is the most useful metric when a credible publication provides the value and reporting period. Even then, most figures are annualized run rates rather than audited yearly revenue. A strong month multiplied by twelve may overstate what the company will actually collect.
User statistics are harder to compare. “Millions of developers,” “hundreds of thousands of daily users,” “used by 70% of the Fortune 500,” and “users from 85% of the Fortune 500” describe different levels of adoption. None reveals paid seats, retention, customer concentration, or how much code reaches production.
Company case studies prove that strong outcomes can occur, but they rarely show average performance. Independent studies provide a useful correction: one pull-request analysis found large differences by task type and no universal winner, while another study found that 46.41% of agent-generated fixes were rejected.
Benchmark results also age quickly. Factory once led Terminal-Bench with 58.75%, while newer models later reported much higher results on updated versions. Changes in models, prompts, agent harnesses, test-time compute, and benchmark design can all move the ranking.
Our confidence is highest on the first three positions. Several independent indicators place Cursor ahead, Cognition second, and Replit third. Factory, Augment, and Poolside remain harder to order because they disclose fewer comparable commercial results.

This chart, featured in our agentic AI market deck, shows the share of revenue generated by each customer segment in the agentic AI market
Which AI coding agent startups are actually ahead?
Cursor is currently the clear overall leader in AI coding agents, Cognition is the strongest challenger, and Replit leads the adjacent race to let anyone build complete applications.
Cursor takes first place because it combines the largest business, a mature daily product, major enterprise deployments, cloud autonomy, code review, and increasingly strong proprietary models. No other startup has demonstrated that entire combination at a similar scale.
Cognition ranks second because Devin remains the clearest autonomous software engineer, while Windsurf gives the company a strong interactive product. Cognition is growing quickly and has excellent enterprise customers, but Cursor still generates roughly eight times more annualized revenue.
Replit ranks third because it has built a genuinely large business around a different customer need. It is currently the strongest startup for turning a plain-language idea into a deployed application. Replit would move closer to the top two by reaching its $1 billion target and proving that Agent can handle complex established codebases as well as greenfield applications.
Factory ranks fourth. Its model-flexible architecture, long-running Missions, persistent computers, strong enterprise names, and rapid reported growth give it a credible path upward. Actual revenue, retention, and deployment-depth figures would be needed to place it above Replit.
Augment ranks fifth because its work on organizational context addresses a real limitation of coding agents, particularly in very large repositories. Cosmos remains too new, and the company has disclosed too little adoption, to support a higher ranking.
Poolside ranks sixth. Owning coding models could become valuable, but Poolside has taken on enormous infrastructure and financing risk while revealing little commercial traction.
Cursor’s approximately $4 billion run rate creates a commercial gap that smaller product announcements cannot erase. Cognition could narrow it by sustaining its current growth, while Replit and Factory still need to prove their newest ambitions at larger commercial scale.
Today, Cursor is ahead overall. Cognition is the closest startup challenger and the leader in pure task delegation. Replit leads AI-native application creation. Factory has the clearest chance to disrupt that top three.
| Rank | Startup | Why it ranks here |
|---|---|---|
| 1 | Cursor | Clear overall leader in revenue, product breadth, enterprise deployment, distribution, and proprietary agent development |
| 2 | Cognition | Strongest autonomous-engineering specialist, with Devin, Windsurf, rapid growth, and major enterprise customers |
| 3 | Replit | Leader in building and deploying new applications through natural language, supported by a large user and revenue base |
| 4 | Factory | Fast-rising enterprise challenger with long-running agents, model flexibility, and strong reported adoption |
| 5 | Augment Code | Credible large-codebase and organizational-context strategy, but limited evidence of current commercial scale |
| 6 | Poolside | Ambitious proprietary-model approach with substantial capital, yet weaker proof of adoption and operational execution |
If you want more recent data on this point, please see our latest agentic AI market report.
OUR METHODOLOGY
The answer to which AI coding agent company is ahead is not obvious because leadership can mean different things: generating the most revenue, growing the fastest, completing the most autonomous work, winning major enterprises, performing well on real coding tasks, or building the strongest long-term position. We therefore broke the main question into separate analytical dimensions rather than relying on product reputation or general market perception.
For each dimension, we reviewed recent commercial results, product releases, customer deployments, funding events, technical evaluations, independent research, and competitive developments. We aggregated the most relevant evidence and assessed it point by point, giving more weight to achieved revenue, quantified deployments, and independent studies than to forecasts, broad usage claims, or company-run benchmarks.
We also kept category leadership separate from overall leadership. Replit can lead natural-language application creation and Cognition can lead autonomous task delegation without either company ranking first across the complete market. The final order reflects the combined pattern across commercial scale, product maturity, enterprise adoption, technical evidence, infrastructure, and defensibility.
Private-company numbers require some judgment. Most revenue figures are annualized run rates rather than audited yearly revenue, customer percentages do not reveal deployment depth, and technical benchmarks can become outdated quickly. We used them as comparative evidence, not as perfectly standardized measurements.
Because the market moves unusually fast, we emphasized the freshest credible evidence available. Older benchmarks, product impressions, and valuation assumptions carried less weight when newer commercial or technical results showed that the competitive position had changed.
Key sources used for this analysis include: Forbes on Cursor’s approximately $4 billion annualized revenue, the U.S. Securities and Exchange Commission filing for the SpaceX–Cursor merger agreement, Cursor’s funding and commercial-scale announcement, Cursor’s Stripe deployment, Cursor’s Nvidia deployment, Cognition’s latest financing and revenue update, Cognition’s Windsurf acquisition announcement, Cognition’s Devin Desktop release, Replit’s introduction of Agent 4, Factory’s Series C and growth announcement, Augment Code’s Cosmos announcement, and Augment Code’s Context Engine evaluation.
Independent technical evidence came from the AIDev dataset of 932,791 agent-authored pull requests, the comparative study of coding-agent performance by task type, the empirical study of accepted and rejected agent-authored pull requests, and Microsoft’s study of command-line coding agents and development output. We also used Anthropic’s Claude Code adoption and revenue update and OpenAI’s GPT-5.3-Codex release to assess pressure from major AI platforms.

This chart, included in our agentic AI market deck, shows how autonomous AI agent platform technology has evolved over time
Related blog posts
- Who is winning AI coding now?
- Vibe coding: which startup is ahead?
- Which customer service agent startup has the edge?
Who is the author of this content?
NEW MARKET PITCH TEAM
We track new markets so founders and investors can move fasterWe build living "market pitch" documents for emerging markets: AI, synthetic biology, new proteins, and more. Instead of outdated PDFs or hallucinated LLM answers, our clients get a clean, visual, always-updated view of what's really happening: key players, deals, regulations, and signals that matter. Learn more about us.