Google gemini’s AI juggernaut: 8 million enterprises, $20b cloud boom, and a TPU power play

Fortune

The gist

Google’s Gemini AI, now turbocharged by proprietary TPUs and multimodal power, is rewriting the rules of enterprise cloud dominance—with 8 million paid subscribers and a $20B revenue boom.

What to know

  • Gemini 2.5 and 3’s multimodal AI smarts, natively woven into Google Workspace, Microsoft 365, and Salesforce, have driven 8 million enterprise signups by early 2026.
  • Google’s strategic pivot to in-house TPU hardware slashed AI training costs by up to 50% and is set to explode TPU revenue from $3B to $25B by 2027, eroding Nvidia’s cloud grip.
  • AI-fueled Google Cloud revenue surged 63% YoY to $20B in Q1 2026, with a $460B AI backlog, record $84.75B equity raise, and DeepMind’s AlphaEvolve now reclaiming 0.75% of global data center power.

Gemini’s AI Integration Leap

Google’s rapid evolution from Gemini 2.5 to 3 transformed AI from a feature into the backbone of enterprise workflows, empowering non-technical teams and breaking Nvidia’s hardware monopoly.

Google's Gemini AI models have rapidly evolved to redefine enterprise AI integration, beginning with Gemini 2.5's launch in late 2025 which introduced advanced multimodal capabilities—enabling AI agents to process voice, images, and video beyond traditional text. This innovation underpins Google's strategy to embed AI deeply within core products like Google Search and Workspace, enhancing personalized, proactive user experiences and automating complex tasks across platforms.

The Gemini Enterprise platform, unveiled shortly after Gemini 2.5, acts as a centralized AI hub that unifies workflows across enterprises by integrating seamlessly with major software ecosystems such as Google Workspace, Microsoft 365, Salesforce, and SAP. Its no-code conversational interface democratizes AI access beyond technical teams, empowering roles like marketing and finance to automate tasks like CRM querying and report generation, which has driven rapid adoption—marketers’ usage jumped from 33% to 51% within a year.

By late 2025, Gemini 3 marked a paradigm shift not only in AI model performance—leading leaderboards across text, images, video, and code and scoring 91.9% on PhD-level science benchmarks—but also in infrastructure independence, as Google trained it entirely on proprietary TPUs, breaking reliance on Nvidia’s supply chain. This hardware control enabled Google to scale AI efficiently and extend TPU access to leading AI companies like Anthropic and Midjourney, thereby expanding its AI ecosystem influence.

By early 2026, Gemini Enterprise matured into a comprehensive AI cloud platform that simplifies deployment and reduces integration costs by connecting data from key enterprise systems such as Workday, Palantir, and ServiceNow. It supports industry-specific AI agents tailored for sectors like commerce and security, incorporates proactive security capabilities leveraging Mandiant’s threat intelligence, and fosters a robust partner ecosystem. The platform’s scalable, no-code to low-code agent deployment empowers enterprises to build and manage secure, customizable AI agents at scale, fueling Google’s record earnings growth and positioning it strongly against competitors like Microsoft and AWS.

Sources
Cloud Wars Live with Bob EvansMidnight Signal AITheSequenceAI with AishProduct Growth

TPUs Redefine AI Economics

By vertically integrating custom TPUs, Google slashed AI training costs and seized control of its AI infrastructure, freeing up GPUs for cloud customers and fueling a new era of chip-driven cloud competition.

Google's strategic investment in custom Tensor Processing Units (TPUs) has been pivotal in powering the advanced capabilities of its Gemini 3 AI model, marking a decisive shift away from reliance on Nvidia GPUs. As Mandeep Singh highlighted in late 2025, the entire training process for Gemini 3 was conducted on Google's TPUs, a move that not only showcased TPU's superior architecture but also allowed Google to serve AI workloads at scale without tapping into Nvidia's supply chain. This vertical integration underscores Google's confidence in its TPU design, which is optimized specifically for matrix operations central to neural network training, enabling faster iteration and enhanced debugging compared to generic GPU providers.

By leveraging its proprietary TPU infrastructure internally for both training and inference, Google has significantly reduced its dependence on third-party GPUs, particularly Nvidia's, thereby freeing up GPU capacity for Google Cloud customers. This strategic allocation not only enhances cloud service offerings but also creates a virtuous cycle where increased external TPU usage funds further chip development. As Singh noted, Google continues to purchase Nvidia chips primarily for cloud deployment, which could boost Google Cloud revenue by improving GPU availability, while TPU pods are marketed at competitive rates that undercut GPU instances, reinforcing Alphabet's dual operational and commercial advantage.

Alphabet's decade-long commitment to custom AI silicon since 2015 has yielded substantial cost and performance benefits, with industry analysts estimating training cost reductions of 30-50% compared to off-the-shelf GPUs. This investment enables Google to scale TPU clusters massively—supporting up to one million TPUs in training environments—and tightly integrate them with CPUs to optimize AI workloads. By early 2026, TPU-related infrastructure revenue was forecasted to surge from $3 billion to $25 billion in 2027, reflecting strong market demand for efficient AI compute solutions amid a broader industry shift from training-centric to inference-dominant workloads, where TPUs excel in cost-efficient inference across services like Search, Gmail, and Maps.

Google's TPU strategy exemplifies a broader industry trend toward owning the entire AI compute stack end-to-end, encompassing both hardware and software ecosystems. By optimizing chips specifically for its own AI models and integrating them with frameworks like TensorFlow, Google creates a technical moat that rivals find difficult to breach. This approach not only enhances margins and operational flexibility but also intensifies competition among cloud providers to develop proprietary AI accelerators. Meanwhile, the AI chip market is fragmenting with specialized decoding chips emerging from players like Groq, Microsoft, Amazon, and Facebook, signaling a future where disaggregated, purpose-built hardware solutions reduce power and capital costs, especially in edge environments.

Sources

Enterprise AI: From Hype to Reality

While Gemini’s adoption soared to 8 million enterprises, real-world deployments revealed both the promise and growing pains of AI at scale, with integration challenges and a rapidly expanding agent marketplace shaping the landscape.

Google’s Gemini 2.5 model rapidly gained traction among enterprises, particularly marketers, with adoption rising from 33% to 51% within a year, largely due to its seamless integration into Google Workspace subscriptions. This integration marked a strategic shift from experimental pilots to embedding advanced AI agents directly into core products like Google Search and Workspace, enabling more proactive, personalized, and multimodal experiences that extend beyond text to voice, image, and video interactions.

By early 2026, Gemini Enterprise had amassed 8 million paid enterprise subscribers and over 100 million sign-ups, signaling strong market interest despite a roughly even split in user satisfaction. Challenges such as unreliable agent builders and connectivity issues reflected the nascent state of enterprise AI adoption, a struggle shared by competitors like Microsoft and Salesforce, underscoring the evolving and early-stage nature of this technology in business environments.

As enterprises moved beyond experimentation, leaders like Piyush Saxena of HCLTech emphasized the demand for mature, measurable AI returns through tightly integrated ecosystems combining Gemini Enterprise, Agent Development Kit, Vertex AI, and BigQuery. Real-world deployments, including AskHub and predictive defect detection across manufacturing and financial services, demonstrated tangible efficiency gains, while the rapid expansion of the Google Cloud AI Agent Marketplace—approaching 300 agents—facilitated broader customer engagement and strategic partnerships.

Google’s 2026 launch of the Gemini Enterprise Agent platform marked a pivotal advancement, enabling businesses to build, deploy, and manage customizable, secure AI agents tailored to complex workflows, particularly in the US market. This strategic push, emphasizing trust, scalability, and deep ecosystem integration, fueled record earnings growth and intensified competition with Microsoft and OpenAI. Enterprises are now constructing comprehensive end-to-end agentic workflows that deliver measurable business value, supported by democratized AI development tools like AI Studio and robust frameworks addressing security, observability, and compliance—transforming AI from pilot projects into scalable production operations.

Sources

AI Cloud Profits Break Records

Google Cloud’s full-stack AI dominance—spanning custom chips to advanced models—drove $20B in quarterly revenue and a $460B backlog, outpacing rivals and doubling profitability through tight vertical control.

Google Cloud’s AI-driven revenue surge is deeply rooted in its full-stack AI strategy, which tightly integrates advanced Gemini models with proprietary TPU hardware. This approach not only frees up Nvidia GPUs for external customers, enhancing cloud service capacity, but also delivers frontier AI capabilities with improved efficiency, as seen in the Gemini 3 model’s multimodal reasoning advances. The market rewarded these innovations handsomely, with Alphabet’s shares hitting record highs and adding approximately $140 billion in market capitalization following Gemini 3’s release, underscoring investor confidence in Google’s AI-powered cloud trajectory.

By early 2026, Google Cloud’s AI integration propelled a record 63% year-over-year revenue growth to $20 billion in Q1, surpassing analyst expectations and marking the first quarter exceeding $20 billion. This growth shifted the revenue mix significantly, with AI-powered enterprise solutions—especially Gemini-based products—growing nearly 800% year-over-year and attracting over 2,800 companies and 8 million paid seats. Operating income tripled to $6.6 billion, with margins nearly doubling from 17.8% to 32.9%, reflecting how Google’s control over every AI stack layer—from custom TPUs to models and platforms—creates a unique competitive advantage that translates into cloud profitability.

Alphabet’s strategic commitment to AI is further evidenced by its massive $460 billion AI-related backlog and an unprecedented $84.75 billion equity raise in April 2026, the largest in company history. This capital infusion underscores Google Cloud’s ambition to expand AI infrastructure amid intense competition with AWS and Microsoft Azure. Notably, Google’s AI backlog surpassed AWS for the first time, signaling overwhelming enterprise demand that currently outstrips compute capacity, justifying increased capital expenditures and marking a pivotal shift in cloud market dynamics.

Google’s TPU-centric AI infrastructure not only supports its internal AI workloads and Gemini models but is evolving into a high-margin external revenue stream, with direct chip sales to customers expected to become meaningful by 2027. This hardware strategy reduces AI operational costs by optimizing inference workloads, aligning with the market’s transition from training-heavy to inference-led AI phases. Additionally, ventures like the AI compute partnership with Blackstone diversify revenue channels and strengthen Google Cloud’s competitive positioning against AWS and Microsoft Azure, while enterprise-focused platforms such as Gemini Enterprise Agent emphasize trust, scalability, and deep integration to capture high-value contracts.

Sources

Autonomous AI Drives Efficiency

AlphaEvolve and Google’s self-improving AI agents are revolutionizing enterprise optimization, reclaiming data center capacity and accelerating model training with breakthroughs that surpass human-devised solutions.

By early 2026, Google's autonomous AI research system had pioneered a closed empirical loop leveraging specialist agents to iteratively refine training recipes without human intervention. These agents partition recipe surfaces and share lineage feedback, transforming evaluator outcomes—including failures like crashes and budget overruns—into program-level edits, thereby producing auditable trajectories of code rewrites that enhanced public starting recipes. This approach yielded significant performance gains, such as a 38.7% increase in NanoChat-D12 CORE and a 4.59% reduction in CIFAR-10 Airbench96 wallclock time, validating the system’s ability to optimize complex AI workflows through recursive self-improvement.

Google DeepMind’s AlphaEvolve exemplifies the next evolution of autonomous AI systems by applying evolutionary algorithms and reinforcement learning to enterprise-scale optimization challenges. Demonstrated in mid-2026, AlphaEvolve autonomously discovered a workload scheduling method that reclaimed nearly 0.75% of Google’s global data center computing power and devised a faster matrix multiplication technique that accelerated flagship model training by about 1%. These breakthroughs highlight how AI agents can outperform human experts by running hundreds of experiments—such as Andrej Karpathy’s agent achieving an 11% speedup on already optimized GPT training scripts—underscoring the power of automated, incremental improvements in real-world settings.

With its July 2026 launch, Google Cloud positioned AlphaEvolve not as a niche research tool but as a broadly accessible developer resource integrated into its AI ecosystem, aiming to democratize advanced optimization capabilities across industries. Targeting high-impact sectors like semiconductor chip design, logistics, and medical research, AlphaEvolve addresses critical pain points where even fractional efficiency gains translate into multimillion-dollar savings and competitive advantages. This strategic rollout also signals Google Cloud’s intent to convert DeepMind’s cutting-edge AI research into scalable commercial revenue streams amid intense competition from Microsoft Azure and AWS, marking a pivotal step in embedding autonomous AI-driven optimization into enterprise workflows.

Sources

Ecosystem and Security Power Play

Google Cloud’s partner ecosystem and Mandiant-powered security agents make Gemini Enterprise both deeply integrated and resilient, enabling scalable, compliant AI deployments across critical industries.

By early 2026, Google Cloud had cultivated a powerful partner ecosystem that significantly amplifies the capabilities of Gemini Enterprise beyond Google’s proprietary technologies. This collaborative network acts as a force multiplier, enabling seamless integration across diverse systems of record such as Workday, Salesforce, Palantir, and ServiceNow, while supporting industry- and domain-specific AI agents tailored to unique sector needs like commerce. This strategic ecosystem not only enhances operational management but also ensures that AI deployments are both scalable and finely attuned to enterprise governance and compliance requirements.

Complementing its ecosystem strategy, Google Cloud has fortified Gemini Enterprise’s security posture by embedding advanced, proactive threat detection capabilities derived from its 2022 acquisition of Mandiant. The platform features an agentic security operations center with continuously active security agents that monitor, report, and respond to threats in real time, creating a formidable security footprint tailored to industry-specific demands. This integration underscores Google’s commitment to delivering secure AI cloud services that address the complex governance and compliance challenges enterprises face in deploying AI at scale.

Sources
Cloud Wars Live with Bob Evans

Part of these trends

Get the stories behind the trends

Deep-dive reporting and the weekly brief, in your inbox.