AI Notebook · Living reference · Updated monthly

AI Agent Statistics

Last updated 2026-08-29 · First published 2026-08-29 · Every number verified against its source

AI agents crossed from pilots to production in 2026. Among large enterprises (over $1 billion in revenue), 40% now report scaling AI agents in at least one business function, up from 27% a year earlier (McKinsey, published Aug 25, 2026), and 57.3% of organizations surveyed by LangChain have agents running in production. The money is keeping pace: Salesforce's Agentforce passed $1.5 billion in ARR, up over 240% year over year (Aug 26, 2026), and Anthropic's Claude Code passed a $2.5 billion revenue run rate. The caution flag is just as concrete: Gartner predicts over 40% of agentic AI projects will be canceled by the end of 2027.

What happened in August 2026

August 2026 was the month the personal AI assistant became a venture category of its own. On August 26, the Wall Street Journal reported that Instinct, the invite-only assistant built by 23-year-old former Sierra researcher Noah Shinn, closed a $250 million Series B co-led by Index Ventures and Benchmark, bringing total funding to $350 million at a $2.5 billion valuation. Forbes noted the same day that the valuation had jumped roughly fivefold in a matter of weeks, from just over $500 million in early August. One day later, Newcomer reported that Town, the enterprise assistant startup founded by ex-Plaid CTO Jean-Denis Greze, was in talks to raise at a $1 billion valuation in a round led by Index, the same firm backing Instinct. Earlier in the month, on August 12, Bloomberg-sourced reporting said Cognition was already in talks to raise at a valuation of at least $40 billion, less than three months after closing its round at $26 billion post-money.

The enterprise data caught up with the hype in the same week. McKinsey's State of AI survey, published August 25, found 40% of large organizations (over $1 billion in revenue) now scaling AI agents in at least one function, up from 27% a year earlier, while smaller firms sit flat at 22%. Deloitte's August 12 study of 501 US organizations put a finer point on the maturity gap: 85% are testing or expanding agents, but only 15% have reached scaled, orchestrated multi-agent adoption. And on August 26, Salesforce reported Agentforce ARR exceeding $1.5 billion, up over 240% year over year, alongside a new usage metric: 7 billion Agentic Work Units delivered to date, 3.2 billion of them in the most recent quarter.

Capability benchmarks kept moving underneath it all. As of the Vals AI leaderboard update on August 26, Claude Opus 5 tops SWE-bench Verified at 97.00%, with seven of 86 evaluated models now at 95% or better, which is why the field's attention is shifting to harder tests like OSWorld 2.0, where the best agent completes just 20.6% of long-horizon computer tasks. The gap between saturated coding benchmarks and unsolved real-world autonomy is now the clearest single picture of where AI agents actually stand.

How many companies are actually using AI agents?

StatisticValueAs ofKindSource
Large organizations (over $1B annual revenue) scaling AI agents in one or more functions40%, up from 27% last yearQ2 2026 (survey fielded May 4 - June 8, 2026; published 2026-08-25)surveyMcKinsey, The State of AI in 2026: On the road to ROI (n=1,719 across 97 nations)
Smaller organizations (under $1B revenue) scaling AI agents, flat year over year22%Q2 2026 (survey fielded May 4 - June 8, 2026)surveyMcKinsey, The State of AI in 2026
Organizations reaching the scaling phase with AI agents across the enterprise (all sizes); coding agents similar; chatbots remain the most widely scaled AI toolAbout 2 in 10 for AI agents; about 2 in 10 for software coding agents (31% at larger enterprises); 47% for chatbotsQ2 2026 (survey fielded May 4 - June 8, 2026)surveyMcKinsey, The State of AI in 2026
Respondents whose organizations decided against buying at least one software product because the functionality could be built in-house with agentic coding tools32% (nearly half of AI high performers vs 31% of other respondents)Q2 2026 (survey fielded May 4 - June 8, 2026)surveyMcKinsey, The State of AI in 2026
AI high performers (about 6% of respondents) vs others on scaling agents; cost constraints on AI useMore than 3x as likely to be scaling agents in most business functions; 2x for software coding agents; 2.7x for other agentic AI. About 20% of all respondents say AI operating costs (including tokens) constrained AI use; about 1 in 10 report cost-constrained use for each of chatbots, AI agents, and coding agentsQ2 2026 (survey fielded May 4 - June 8, 2026)surveyMcKinsey, The State of AI in 2026
Respondents with AI agents in production, plus share actively developing agents with concrete deployment plans57.3% in production (up from 51% the prior year); 30.4% actively developing. By size: 67% of 10k+ employee orgs vs 50% of sub-100 orgs have agents in productionSurvey fielded Nov 18 - Dec 2, 2025 (n=1,340)surveyLangChain, State of Agent Engineering report
US enterprise leaders ($1B+ revenue orgs) deploying AI agents, and share orchestrating multiple agents across workflows53% deploying agents (vs 55% the prior quarter); 18% orchestrating multiple agents, doubled from 9%Q2 2026 (survey fielded April 28 - May 25, 2026; published 2026-06-24; n=204)surveyKPMG, AI Quarterly Pulse Survey Q2 2026
US organizations' AI agent maturity: testing, expanding, and scaled orchestrated multi-agent adoption42% testing/small deployments; 43% expanding across more than one function; only 15% have scaled, orchestrated multi-agent adoptionSurvey fielded April - June 2026 (n=501); published 2026-08-12surveyDeloitte, AI agents are only the beginning: The path to agentic transformation
Organizations with a mature governance model for agentic AI, and expected agent use by 2027Only 21% report mature agentic AI governance; 74% expect their companies to use AI agents at least 'moderately' by 2027 (23% 'extensively', 5% fully integrated)Survey fielded Aug - Sep 2025 (n=3,235, 24 countries); analysis published 2026-04-24surveyDeloitte, State of AI in the Enterprise 2026
Senior US executives saying AI agents are already being adopted; depth of adoption; budget intent79% adopting agents; of adopters, 35% adopting broadly and 17% fully adopted in almost all workflows; 88% plan to increase AI-related budgets in the next 12 months due to agentic AI; 66% of adopters report increased productivity; 18% not using agents at allSurvey fielded April 22-28, 2025 (n=308 US business executives)surveyPwC, AI Agent Survey
Enterprise IT leaders planning to expand AI agent use in the next 12 months, and share who had already implemented agents96% plan to expand agent use (about half aiming for significant, organization-wide expansion); 57% implemented AI agents in the past two years (21% within the last year)Released 2025-04-16 (n=nearly 1,500 enterprise IT leaders, 14 countries)surveyCloudera, The Future of Enterprise AI Agents
Overall enterprise AI context from the same McKinsey 2026 survey: regular AI use and enterprise-wide scalingNearly 9 in 10 use AI regularly in at least one function; 44% report AI scaling across the enterprise (up from 38%); 54% of $1B+ orgs scaling enterprise-wide vs one-third of smaller orgs; 37% attribute at least some EBIT impact to AI (flat YoY)Q2 2026 (survey fielded May 4 - June 8, 2026)surveyMcKinsey, The State of AI in 2026

How much money is flowing into AI agents?

Funding rounds and revenue figures are hard events; market-size projections are analyst models and are labeled as such.

StatisticValueAs ofKindSource
Instinct (AI personal assistant) raising a Series B co-led by Index Ventures and Benchmark; WSJ reports the startup began testing Instinct in private beta in February 2026$250M Series B at $2.5B valuation2026-08-26fundingThe Wall Street Journal (Kate Clark, exclusive), confirmed by direct fetch
Instinct's cumulative funding after the Series B; founder Noah Shinn is 23 (per TechCrunch) and a former Sierra researcher (per Newcomer); corroborated independently by Newcomer$350M total raised at $2.5B valuation2026-08-26fundingTechCrunch (Aug 26, 2026); corroborated by Newcomer (Aug 27, 2026)
Instinct's valuation jumped roughly fivefold in weeks: $75M Series A led by Mamoon Hamid of Kleiner Perkins valued it over $500M in early August 2026; weeks later the Benchmark/Index round valued it over $2.5Bover $500M (early Aug 2026) to over $2.5B (Aug 26, 2026)2026-08-26fundingForbes (Iain Martin and Rashi Shrivastava)
Town (enterprise personal-assistant startup, founded by ex-Plaid CTO Jean-Denis Greze and ex-Google applied-AI product director Tony Vincent) in talks to raise at a round led by Index Ventures; reported deal-in-progress, not closedraising at $1B valuation (in talks)2026-08-27fundingNewcomer (Madeline Renbarger and Eric Newcomer)
Town's prior round: Series A led by Andreessen Horowitz with Forerunner Ventures, First Round, Alt Capital, and Conviction participating; Town was approaching 10,000 users and claimed 99% two-month retention among users who built at least one custom automation$55M Series A2026-06-03fundingFortune Term Sheet (Lily Mae Lazarus, exclusive)
Sierra (Bret Taylor's AI customer-agent company) Series E led by Tiger Global and GV; CNBC, Yahoo Finance, and Tech Startups put the valuation at $15.8B; Taylor's own announcement says 'over $15 billion'$950M round; post-money above $15B ($15.8B per CNBC)2026-05-04fundingTechCrunch; cross-checked against CNBC and Tech Startups
Sierra customer base and revenue: more than 40% of the Fortune 50 as customers; ARR of $100M in late November 2025 rising to $150M by early February 2026 (company-disclosed)>40% of Fortune 50 customers; $100M ARR (Nov 2025) to $150M ARR (Feb 2026)2026-02company claimSierra company claims via TechCrunch
Cognition (maker of AI software engineer Devin) raised more than $1B led by Lux Capital, General Catalyst, and 8VC; more than doubled its $10.2B post-money valuation from September 2025>$1B raised at $25B pre-money / $26B post-money2026-05-27fundingTechCrunch (Julie Bort)
Cognition revenue: annualized run-rate with enterprise usage of Devin growing 50% month-over-month for six months; customers include Mercedes-Benz, NASA, Goldman Sachs, and Santander$492M annualized revenue run-rate2026-05company claimCognition company claims via TechCrunch
Cognition reportedly already in talks for another round, predicated on reaching a $1B annualized revenue run rate, per sources cited by Bloomberg; reported talks, not a closed roundat least $40B valuation (reported talks)2026-08-12fundingTechCrunch (citing Bloomberg)
Cognition acquired The Interaction Company of California, maker of consumer texting agent Poke (launched March 2026; works over iMessage, SMS, Telegram, and WhatsApp in select markets), per co-founder Marvin von Hagen'low nine figures' acquisition2026-07-24fundingTechCrunch, 'Why Cognition bought Poke'
Macro context: global venture funding hit a record in H1 2026, surpassing the $440B invested in all of 2025; more than 70% of Q2 2026 startup capital went to AI companies, up from just under 50% a year earlier; OpenAI and Anthropic alone took $217B (43% of H1)$510B global VC in H1 2026; >70% of Q2 capital to AI; OpenAI+Anthropic $217B (43% of H1)H1 2026fundingCrunchbase News (Gene Teare, July 2, 2026)
SpaceX confirmed intent to acquire Anysphere (maker of Cursor), described by Crunchbase as the largest startup acquisition ever; confirmed intent, not yet a completed transaction$60B acquisition (largest startup M&A on record)Q2 2026fundingCrunchbase News (July 2, 2026)
Grand View Research estimate for the global AI agents market; North America held a 39.63% revenue share in 2025; treat as a vendor market estimate, definitions vary widely$182.97B by 2033 (49.6% CAGR, 2026-2033)2026-03market estimateGrand View Research press release
MarketsandMarkets AI agents market estimate; note the large spread vs Grand View's figure, these are vendor estimates with differing scopes, cite with attribution, not as consensus$7.84B (2025) to $52.62B (2030), 46.3% CAGR2025 (report TC 9168 PR published April 23, 2025)market estimateMarketsandMarkets press release

How big are the agent platforms?

Most of these are companies reporting their own numbers. That does not make them false, but the label says what they are.

StatisticValueAs ofKindSource
Salesforce Agentforce deals closed since launch (Oct 2024)over 29,000 deals, up 50% Q/QQ4 FY26 (quarter ended 2026-01-31, announced 2026-02-25)company claimSalesforce Q4 FY26 earnings press release
Agentforce ARR at end of fiscal 2026$800 million ARR, up 169% Y/Y (Agentforce + Data 360 combined: exceeds $2.9 billion, up over 200% Y/Y, a figure that includes $1.1 billion Informatica Cloud ARR)Q4 FY26 (announced 2026-02-25)company claimSalesforce Q4 FY26 earnings press release
Agentforce ARR two quarters later; note Salesforce broadened the definition this quarter ('Effective Q2 FY27, Agentforce ARR includes our AI offerings, Slackbot and Headless 360')exceeded $1.5 billion, up over 240% Y/Y (Agentforce + Data 360: nearly $3.9 billion, up over 210% Y/Y)Q2 FY27 (quarter ended 2026-07-31, announced 2026-08-26)company claimSalesforce Q2 FY27 earnings press release
Agentic Work Units (Salesforce's metric for tasks accomplished by an AI agent) delivered across Agentforce and Slack7.0 billion AWUs delivered to date, with 3.2 billion in Q2 alone, growing 97% Q/QQ2 FY27 (announced 2026-08-26)company claimSalesforce Q2 FY27 earnings press release
Claude Code run-rate revenue (agentic coding tool, GA May 2025)over $2.5 billion run-rate revenue; more than doubled since the beginning of 20262026-02-12company claimAnthropic Series G funding announcement
Claude Code user and business adoption growthweekly active users doubled since January 1, 2026; business subscriptions quadrupled since the start of 2026; enterprise use is over half of all Claude Code revenue2026-02-12company claimAnthropic Series G funding announcement
Combined users of OpenAI's Codex and ChatGPT Work, roughly double the count from two weeks earlier (ChatGPT Work launched 2026-07-09); OpenAI did not define the activity window10 million users2026-07-21company claimOpenAI (Codex lead Tibo Sottiaux on X; also shared with Bloomberg), reported by Unite.AI and 9to5Mac
OpenAI Codex weekly active users, standalone; up more than 6x since the desktop app launched in February 2026, with knowledge workers about 20% of usersmore than 5 million weekly active users2026-06-02company claimOpenAI, 'Codex is becoming a productivity tool for everyone'
Microsoft 365 Copilot paid seats; was more than 20 million as of April 2026, so roughly 10 million seats added in one quarterover 30 million paid seatsQ4 FY26 (quarter ended 2026-06-30, announced 2026-07-29)company claimMicrosoft FY26 Q4 earnings, reported by CNBC (Jordan Novet)
GitHub Copilot total users (Satya Nadella on the FY26 Q4 earnings call); up from 15 million developers at Build 202550 million users2026-07-29company claimMicrosoft FY26 Q4 earnings call via CNBC; 15M baseline verified at Microsoft's Build 2025 blog
Fortune 500 companies using active AI agents built with low-code/no-code tools, per Microsoft first-party telemetry (active = deployed to production with real activity in the past 28 days, Nov 2025)more than 80% of Fortune 500 companies2025-11 (published 2026-02-10)company claimMicrosoft Security Blog
Organizations that have used Microsoft Copilot Studio to build AI agents and automationsmore than 230,000 organizations, including 90% of the Fortune 5002025-05-19 (Build 2025)company claimMicrosoft Official Blog, Build 2025
Google Gemini Enterprise (agent platform) paid monthly active user growth40% growth in paid monthly active users quarter-over-quarter (Q1 2026)Q1 2026 (stated 2026-04-22 at Google Cloud Next '26)company claimGoogle Cloud Blog, 'Welcome to Google Cloud Next 26'
Per-customer agent fleets Google cited on Gemini Enterprise: Virgin Voyages managing 1,000+ specialized agents; GE Appliances 800+ across manufacturing, logistics, and supply chain; WPP thousands; KPMG 100+ in its first month1,000+ (Virgin Voyages), 800+ (GE Appliances), thousands (WPP), 100+ in first month (KPMG)2026-04-22 (Google Cloud Next '26)company claimGoogle Cloud Blog
MCP (Model Context Protocol) server count listed on PulseMCP, a daily-updated public directory; most listed servers are community-built and other registries list far fewer, so treat as an upper-bound directory count21,989 servers2026-08-29 (live count re-fetched and confirmed)market estimatePulseMCP server directory

How capable are AI agents right now?

StatisticValueAs ofKindSource
Top SWE-bench Verified score (bash-only mini-swe-agent harness): Claude Opus 5 leads all 86 evaluated models97.00%2026-08-26benchmarkVals AI, SWE-bench Verified leaderboard
SWE-bench Verified is saturating: seven of 86 evaluated models score 95% or better, with the leader 3 points from perfect7 of 86 models at >=95%2026-08-26benchmarkVals AI, SWE-bench Verified leaderboard
Best open-weight model on SWE-bench Verified: DeepSeek V4 Pro 0813, second overall, 0.60 points behind the closed leader96.40%2026-08-26benchmarkVals AI, SWE-bench Verified leaderboard
2024 baseline for the same benchmark: upgraded Claude 3.5 Sonnet set the then-state-of-the-art on SWE-bench Verified, up from 33.4% for its predecessor (vs 97% today, a ~2x improvement in 22 months)49.0%2024-10-22company claimAnthropic announcement
OSWorld 2024 baseline: in the original paper the best model completed 12.24% of real computer tasks vs a 72.36% human baseline12.24% (human: 72.36%)2024-04benchmarkOSWorld paper / official project site (arXiv 2404.07972, NeurIPS 2024)
Current OSWorld-Verified best: Claude Mythos Preview, pass@1 on 361 tasks at 100 max steps (5-run avg, Anthropic's revised harness, self-reported in the Claude 5 system card); agents now exceed the 72.4% human baseline on OSWorld 1.0-era tasks85.4%2026-06company claimSteel.dev OSWorld leaderboard (tracking the Claude 5 system card)
OSWorld 2.0 (released June 2026, 108 long-horizon tasks across 31 self-hosted websites, ~1.6h median human completion time) resets the ladder: the best agent, Claude Opus 4.8 with maximum thinking and batched tool calls, completes only 20.6% of tasks at 500 steps; GPT-5.5 plateaus near 14%20.6% binary / 54.8% partial2026-06benchmarkOSWorld 2.0 official site / paper (arXiv 2606.29537), XLang Lab HKU
WebArena current best tracked result: WebTactix system using DeepSeek v3.2, 594 of 812 tasks correct (74.34% as reported by the project page, which excludes 13 tasks with no valid evaluation outcome; strictly 594/812 = 73.2%) vs the original 2023 GPT-4 baseline of 14.41% and a 78.24% human baseline74.34% (594 tasks correct)2026-02benchmarkSteel.dev WebArena leaderboard, citing the WebTactix project page
GAIA current best: OPS-Agentic-Search (Alibaba Cloud), an official GAIA leaderboard submission using a multi-model ensemble, tied at 92.36% with openJiuwen-deepagent, vs OpenAI Deep Research at 67.36% pass@1 in Feb 202592.36%2026-03benchmarkSteel.dev GAIA leaderboard, tracking the official Hugging Face GAIA leaderboard
METR's original finding: the length of software tasks (measured by how long they take human professionals) that frontier AI agents can complete with 50% reliability has been 'doubling approximately every 7 months for the last 6 years'; Claude 3.7 Sonnet's 50% time horizon was approximately one hourdoubling ~every 7 months (2019-2025)2025-03-19benchmarkMETR, 'Measuring AI Ability to Complete Long Tasks' (arXiv 2503.14499)
METR Time Horizon 1.1 update: overall P50 doubling time of 196.5 days across the full 2019-2026 trend, but post-2023 progress is faster at 130.8 days [CI 107-161] (~4.3 months) on the expanded 228-task suite; Claude Opus 4.5's 50% horizon measured at 320 minutes [CI 170-729]doubling every ~131 days post-20232026-01-29benchmarkMETR blog, 'Time Horizon 1.1'
Longest measured 50% time horizon to date: Claude Mythos Preview (early snapshot, Inspect harness) at ~1,045 minutes [CI 509-3,304 min], with Claude Opus 4.6 at ~719 minutes; METR cautions that 'measurements above 16 hrs are unreliable with our current task suite'~1,045 minutes (~17.4 hours)2026-05-08benchmarkMETR, Task-Completion Time Horizons data page

What do the forecasts say?

Every row here is a prediction, not a measurement. Forecasts from the same firms have missed before.

StatisticValueAs ofKindSource
Over 40% of agentic AI projects will be canceled by the end of 2027, due to escalating costs, unclear business value or inadequate risk controlsover 40% canceled by end of 20272025-06-25 (announced); target end of 2027forecastGartner press release (confirmed verbatim on gartner.com)
Agent washing: Gartner estimates only about 130 of the thousands of agentic AI vendors are real (the rest rebrand AI assistants, RPA and chatbots without substantial agentic capabilities)only about 130 of thousands of vendors2025-06-25market estimateGartner press release (confirmed verbatim)
At least 15% of day-to-day work decisions will be made autonomously through agentic AI by 2028, up from 0% in 2024at least 15% by 2028, up from 0% in 20242025-06-25 (announced); target 2028forecastGartner press release (confirmed verbatim)
33% of enterprise software applications will include agentic AI by 2028, up from less than 1% in 202433% by 2028, up from <1% in 20242025-06-25 (announced); target 2028forecastGartner press release (confirmed verbatim)
40% of enterprise applications will be integrated with task-specific AI agents by the end of 2026, up from less than 5% in 202540% by end of 2026, up from <5% in 20252025-08-26 (announced, updated 2025-09-05); target end of 2026forecastGartner press release (confirmed verbatim)
Gartner's best-case scenario: agentic AI could drive approximately 30% of enterprise application software revenue by 2035; same release predicts one-third of agentic AI implementations will combine agents with different skills by 2027 and at least 50% of knowledge workers will develop new skills to work with, govern or create AI agents by 2029~30% of enterprise app software revenue, surpassing $450 billion, by 2035 (up from 2% in 2025)2025-08-26 (announced); target 2035forecastGartner press release (confirmed verbatim)
Up to $234 billion of enterprise application spending is exposed to 'agentic arbitrage' between now and 2030 (the 'Saaspocalypse' thesis: agents complete tasks across systems, breaking the seat-license model)up to $234 billion exposed by 2030 (~20% of enterprise app SaaS spend)2026-07-01 (announced); horizon through 2030forecastGartner press release (confirmed verbatim)
At least 80% of governments will deploy AI agents to automate routine decision-making by 2028; by 2029, 70% of government agencies will require explainable AI and human-in-the-loop mechanisms for all automated decisions impacting citizen service deliveryat least 80% of governments by 2028; 70% requiring XAI/HITL by 20292026-03-17 (announced); targets 2028/2029forecastGartner press release (confirmed verbatim)
By 2028, AI agents will outnumber sellers by 10 times, yet fewer than 40% of sellers will say AI agents have improved productivityagents outnumber sellers 10:1 by 2028; <40% of sellers report productivity gains2026-07-28 (announced); target 2028forecastGartner press release (full body fetched and confirmed verbatim)
By 2026, 40% of all G2000 job roles will involve working with AI agents; by 2026, 70% of G2000 CEOs will focus AI ROI on growth, aiming to boost revenue and reinvent business models without growing headcount40% of G2000 job roles by 2026; 70% of G2000 CEOs2025-10-23 (announced); target 2026forecastIDC FutureScape 2026 press release (confirmed verbatim)
IDC: by 2030, 45% of organizations will orchestrate AI agents at scale; by 2028 pure seat-based pricing will be obsolete, forcing 70% of vendors to refactor their value proposition; by 2030 up to 20% of G1000 organizations will have faced lawsuits, substantial fines and CIO dismissals from inadequate agent controls and governance45% of organizations orchestrating agents at scale by 20302025-10-23 (announced); targets 2028-2030forecastIDC FutureScape 2026 press release (confirmed verbatim)
As AI's hype fades, enterprises will defer a quarter of their planned AI spend into 2027, with fewer than one-third of decision-makers able to tie the value of AI to their organization's financial growth25% of planned AI spend deferred into 2027; <1/3 tie AI to financial growth2025-10-28 (announced); prediction for 2026-2027forecastForrester 2026 Technology & Security Predictions (confirmed verbatim)
McKinsey forecasts agentic commerce (AI agents that shop, negotiate and transact on behalf of humans) could orchestrate as much as $1 trillion in US retail revenue by 2030; the global estimate is independently corroborated by McKinsey's own LinkedIn post and Retail Diveup to $1 trillion US; $3-5 trillion global by 20302025-10 (research published); target 2030forecastDigital Commerce 360 reporting McKinsey agentic commerce research
Stanford WORKBank audit of the US workforce: share of tasks where the workers performing them express a positive attitude toward AI agent automation (1,500 domain workers, 104 occupations, 844 tasks, 52 AI experts); share of Y Combinator AI companies mapped to tasks workers do not want automated46.1% of tasks; 41.0% of YC companies in low-desire zonessurvey fielded Jan-May 2025 (paper last revised 2026-02-01)surveyStanford SALT Lab, Future of Work with AI Agents (arXiv:2506.06576)
Anthropic Economic Index 'Cadences' report (usage data Apr 10-Jun 10, 2026 plus survey of ~9,700 Claude users): expectations of AI capability vs perceived job risk, skill value, productivity, and agentic session structure (median blog-post-producing agentic Claude Code session contains a single human prompt vs 13 rounds of back-and-forth in chat)over 1/3 expect AI to do most/nearly all their tasks next year; 10% rate own job loss likely; 57% skills more valuable; 86% speed gains2026-04-10 to 2026-06-10 (data period); published 2026-06-26surveyAnthropic Economic Index: Cadences (confirmed verbatim)

Who is building what?

StatisticValueAs ofKindSource
IBM's citable definition of an AI agent: 'An artificial intelligence (AI) agent is a system that autonomously performs tasks by designing workflows with available tools.' (verbatim, by Anna Gutowska, AI Engineer, Developer Advocate, IBM)definition (verbatim quote; confirmed word-for-word on live page)2026-08-29 (page live, undated)company claimIBM Think, 'What are AI agents?'
Anthropic's definition distinguishes workflows ('systems where LLMs and tools are orchestrated through predefined code paths') from agents: 'systems where LLMs dynamically direct their own processes and tool usage, maintaining control over how they accomplish tasks' (verbatim)definition (both quotes confirmed verbatim on live page)2024-12-19 (publication date; live as of 2026-08-29)company claimAnthropic, 'Building Effective Agents'
OpenAI's definition, page 4 of its agents guide: 'Agents are systems that independently accomplish tasks on your behalf.' The guide adds that applications that integrate LLMs but don't use them to control workflow execution (simple chatbots, single-turn LLMs, sentiment classifiers) are not agentsdefinition (verbatim; confirmed by reading the PDF directly)2025-04 guide (PDF re-verified 2026-08-29)company claimOpenAI, 'A Practical Guide to Building Agents' (PDF)
LangChain (langchain-ai/langchain, self-described 'The agent engineering platform.') GitHub stars145,255 stars (24,243 forks)2026-08-29 (re-verified via live API call)benchmarkGitHub API, langchain-ai/langchain
LangGraph (langchain-ai/langgraph) GitHub stars; PyPI downloads of the langgraph package in the trailing month40,685 stars; 66,035,075 downloads last month (11,259,434 last week)2026-08-29 (re-verified via live API calls)benchmarkGitHub API + PyPI Stats (pypistats.org)
CrewAI (crewAIInc/crewAI) GitHub stars; PyPI downloads of the crewai package in the trailing month57,803 stars; 29,105,346 downloads last month2026-08-29 (re-verified via live API calls)benchmarkGitHub API + PyPI Stats (pypistats.org)
Microsoft AutoGen (microsoft/autogen, 'A programming framework for agentic AI') GitHub stars60,696 stars (9,164 forks)2026-08-29 (re-verified via live API call)benchmarkGitHub API, microsoft/autogen
OpenAI Agents SDK for Python (openai/openai-agents-python) GitHub stars29,064 stars (4,629 forks)2026-08-29 (re-verified via live API call)benchmarkGitHub API, openai/openai-agents-python
Consumer agent Instinct (operated by Spear Street Technology; Forbes, citing California corporate filings, identifies founder Noah Shinn as a former Sierra researcher) is still in private/invite-only beta, connects to users' apps and devices and communicates via texts and calls. Note: Forbes, published the same day, described the $2.5B round as still in talks$250M Series B; $350M total raised; $2.5B valuation2026-08-26fundingTechCrunch (ex-Sierra detail: Forbes, Aug 26)
Town's assistants are customizable named characters ('Townies', e.g. a bunny, a silver fox, a capybara) whose personality the user shapes; Series A led by Andreessen Horowitz in June 2026, Forerunner Ventures also became an investor$55M Series A led by a16z, June 20262026-07-16 (article date)fundingInc. (Lisa Bonos), 'How Town Became Silicon Valley's New Favorite AI Tool'
In its second month out of beta, Town said its user count had more than quadrupled since early June 2026 (revenue and userbase size undisclosed; many companies Inc. spoke to were still on free trials)users 'more than quadrupled since early June'; $15-$199/mo individual, team plans from $59/seat2026-07-16company claimInc. (Lisa Bonos)
Weeks after its Series A, Town was in talks to raise at a $1 billion valuation in a round led by Index Ventures, the same firm backing Instinct's round, with Newcomer itself noting surprise that Index is backing both startups at the heart of the personal-agent frenzy$1B valuation (round in talks; raise amount undisclosed)2026-08-27fundingNewcomer

How these numbers are chosen

  • Every figure was verified by opening the source and seeing the number. Nothing is quoted from an aggregator or from an AI model's memory.
  • Every figure carries the date it refers to, and a label for what kind of number it is: a survey measured something, a forecast predicts something, a company claim is the company talking about itself, a benchmark is a published leaderboard score, and a market estimate is an analyst firm's model. They are not interchangeable, and this page never pretends they are.
  • Statistics that failed verification are recorded internally and never published, so a future update cannot quietly resurrect a bad number.
  • The page is refreshed monthly; the changelog below records every change.

Changelog

August 2026

  • First published: 81 verified statistics across 6 categories, each checked against its source.

Go deeper