The AI landscape doesn't move in one direction — it lurches. Some techniques leap from experiment to table stakes in a single quarter; others stall against regulatory walls, technical ceilings, or organisational inertia that no amount of hype can dislodge. Knowing which is which is the hard part. The State of Play cuts through the noise with a rigorously maintained index of AI techniques across every major business domain — classified by maturity, evidenced by real-world adoption, and updated daily so you always know where you stand relative to the field. Stop guessing. Start knowing.
A daily newsletter distilling the past two weeks of movement in a domain or two — delivered to your inbox while the index updates in the background.
Each dot marks the weighted maturity of practices within a domain — hover for a brief summary, click for more detail
AI that automates spreadsheet tasks including formula creation, data analysis, pivot tables, and formatting from natural language. Includes Excel/Sheets AI assistants and formula generation; distinct from natural language to SQL which queries databases rather than manipulating spreadsheets.
AI-driven spreadsheet automation entered late 2026 facing a critical bifurcation: elite scaled deployment in finance and bounded workflows (52% accounting adoption, 250% ROI, case studies showing 75% cycle time reduction and +9.4% revenue per rep) versus a widespread adoption/usage paradox that reveals the true mainstream constraint is organizational, not technical. Microsoft's Q3 FY26 earnings report 20 million paid Copilot seats growing 250% YoY, with half of Fortune 500 adopting Copilot Cowork within 6 months—yet 43.7% of organizations report their AI implementations are "implemented but not used," and 95% of enterprise AI pilots fail to deliver measurable P&L impact. All major vendors (Microsoft, Google, OpenAI, Anthropic) have independently converged on spreadsheet agents as core platform strategy, and Copilot Cowork GA demonstrates tested 30-40% cost efficiency gains on multi-step workflows. The capability ceiling is well-established: peer-reviewed benchmarks confirm frontier models "frequently fall short of professional finance standards" on multi-step logic, and independent analysis documents specific competitive displacement (GitHub Copilot lost to Cursor and Claude Code). Practice remains at bleeding-edge: elite deployment in bounded finance and operations workflows with clear ROI; mainstream stalled by organizational adoption barriers (training, manager modeling, governance) not feature maturity, compounded by ROI measurement failure (only 5-8% of enterprises achieve at-scale returns despite $186M average budgets).
August 2026 evidence reinforces the dual reality: platform convergence at scale paired with organizational adoption barriers and competitive displacement risks. Microsoft Q3 FY26 earnings report 20 million paid Microsoft 365 Copilot seats (250% YoY growth, fastest quarterly adds since launch) with monthly engagement now at Outlook levels; Copilot Cowork GA (June 2026) tested 30-40% cost efficiency gains and now comprises 49% of all Copilot tasks (up from 29%), with half of Fortune 500 adopted within 6 months. All major vendors (Google, Microsoft, OpenAI, Anthropic) have GA spreadsheet automation with multilingual support and agent coordination: Google's July 2026 Workspace features include Gemini formula error diagnosis (70.48% accuracy) expanded to 28 languages; Microsoft's work IQ context integration supports 30+ file types up to 50MB; Anthropic's Claude for Excel add-in (GA July 2026) offers multi-sheet reading with formula explanation and tracked changes. Yet specific capability benchmarks remain firm: SpreadsheetBench v2 (321 real workflows) tops at 34.89%; Vals AI Finance Agent capped at 23% on financial modeling; independent research reveals 70-80% accuracy on extraction but only 25-30% on decision-making tasks. Real production failures persist: FP&A teams report disabling Copilot within six weeks of June 2026 release due to unreliable multi-sheet and sign-convention handling.
The adoption/usage paradox is the dominant signal for mainstream stagnation. Deloitte's 3,235-leader survey (24 countries) shows 74% intend AI agent deployment by 2027, but only 21% have mature governance (53-point gap); critically, 43.7% of organizations implementing AI report it is "implemented but not used" due to organizational barriers (training gaps, manager modeling, permission issues)—a signal that features and adoption metrics diverge sharply from realized value. Industry-wide ROI measurement failure dominates: 95% of enterprise AI pilots show zero measurable P&L impact; only 5-8% achieve at-scale returns despite $186M average budgets; the root cause is integration gaps (tools not connected to underlying workflows) and budget misalignment toward sales/marketing instead of back-office automation. Finance sector remains strongest signal: accounting 52% adoption with 250% ROI, AR automation 40% payment acceleration, three-statement modeling 30→4 minutes, case studies showing Cloud Supply Chain 75% cycle time reduction and sales operations +9.4% revenue per rep via agentic automation. Specialized bounded-scope adoption signals emerge: Tracelight AI (error detection, not autonomous generation) deployed with 7 of top 10 management consulting firms and PE funds >$600B AUM. Mainstream adoption stalls: 60% achieve minimal ROI on investment; CSA May 2026 survey documents 82% discovered ungoverned shadow agents, 65% experienced security incidents, 61% reported data exposure. Strategic risks emerging: GitHub Copilot lost market lead to Cursor and Claude Code despite Microsoft's $13B OpenAI investment and distribution dominance—indicating competitive displacement in aligned markets. The practice bifurcates sharply: finance operations with bounded scope and verification gates show sustained ROI at scale; error-detection and review workflows (not autonomous generation) achieve enterprise scaling; mainstream knowledge work adoption blocked by organizational change barriers, not capability maturity. Practice remains at bleeding-edge: elite deployment in finance and operations with proven ROI; mainstream advancement blocked by adoption/usage gaps, ROI measurement failures, and competitive displacement in aligned segments.
— Microsoft internal case study: Cloud Supply Chain pilot deployed 70+ specialized agents reducing cycle time 75%; sales organization achieved +9.4% revenue per rep and -20% deal close cycle via agentic automation.
— Copilot Cowork GA with 30M+ paid seats; multi-step workflows now 49% of all tasks (up from 29%); tested 30-40% cost efficiency gain vs single-model alternatives; half of Fortune 500 adopted within 6 months of preview.
— Strategic analysis: Microsoft's <4.5% conversion (20M of 450M seats) despite $13B OpenAI investment; GitHub Copilot lost market lead to Cursor and Claude Code; spreadsheet automation faces competitive displacement despite distribution dominance.
— Google GA announcements: Gemini-powered formula error diagnosis with one-click debugging and fix suggestions (70.48% accuracy); Fill with Gemini expanded to 28 languages; multi-series scatter and combo charts.
— Microsoft fiscal Q3 2026 earnings: M365 Copilot paid seats surpassed 20 million with 250% YoY growth; monthly engagement now at Outlook levels; GitHub Copilot Enterprise adoption nearly 140,000 orgs (tripled YoY).
— Independent analyst (GCN): 5 million seats added in Q2 (fastest quarterly growth since launch); companies with 50k+ seats quadrupled YoY; Accenture's 743k-seat deployment is largest publicly announced Copilot rollout.
— Enterprise adoption analysis: 43.7% of AI implementations are 'implemented but not used'; barriers are organizational (training gaps, manager modeling, permissions issues) not technical, revealing adoption/usage gap despite high licensing.
— Multi-source synthesis (MIT, BCG, KPMG, McKinsey, 2,100+ execs): 95% of AI pilots fail with zero measurable P&L impact; only 5-8% achieve at-scale ROI; failure rooted in workflow integration gaps and budget misalignment to back-office automation.