2026 GPU Neocloud Pricing and Capacity Showdown
Reported by MarkTechPost: Best GPU Neoclouds 2026: CoreWeave, Nebius, Lambda, Crusoe, and Groq Ranked by Published Pricing and Contracted Power. Analysis and context written by TickrWire.
Marktechpost compares five leading GPU neocloud providers, detailing Q2 2026 pricing, power contracts, and financial performance.

- CoreWeave and Nebius are the only publicly listed GPU neocloud providers, offering transparent quarterly financials.
- Nebius offers the lowest on‑demand B300 price and the only published B300 rate among the group.
- Lambda provides the cheapest B200 instance but lacks a spot‑tier pricing option.
- Crusoe is the only provider with AMD Instinct GPUs, pricing its H200 at $4.29 per GPU‑hour.
- Groq’s inference service is priced per token, differentiating it from the GPU‑hour models of its rivals.
Marktechpost released a data‑driven comparison of the five biggest GPU neocloud operators as of August 2026. The analysis covers CoreWeave, Nebius, Lambda, Crusoe and Groq, focusing on published rate cards, active and contracted gigawatt capacity, recent financial results and the hardware roadmaps each company follows.
CoreWeave, a Nasdaq‑listed firm, posted Q2 2026 revenue of $2.575 billion, a 112 percent year‑over‑year increase, and a backlog of roughly $104 billion. Its active power stood at 1.5 GW, while contracted power reached about 3.7 GW, with a public claim of over 4.2 GW across 51 data centers. Nebius, also public, reported $582.3 million in group revenue, a 454 percent jump YoY, and an AI‑cloud run‑rate of $3.0 billion. The company announced a contracted‑power target of 5 GW for year‑end and plans to add more than 1 GW annually from 2027 onward. Lambda, a private player, secured a $1.5 billion Series E round in late 2025 and a $1 billion credit facility in mid‑2026, while Crusoe raised $1.375 billion in a Series E at a valuation above $10 billion and is reportedly courting a $3 billion raise at a $30 billion valuation. Groq, after licensing its LPU inference chip to NVIDIA in December 2025, raised $650 million in June and $350 million in August 2026, operating 13 data centers and targeting 200 MW of capacity by 2027.
The market context is a rapid expansion of AI‑intensive workloads that demand both training‑scale and inference‑scale GPU resources. Public filings from CoreWeave and Nebius provide transparent quarterly data, a rarity among cloud providers, and signal that investors are closely watching the scalability of GPU infrastructure. The emergence of “neocloud” terminology reflects a shift toward specialized AI‑focused data centers that differ from traditional hyperscalers in pricing flexibility and contract structures.
Pricing differences are stark. CoreWeave’s H100 rate of $6.16 per GPU‑hour sits about 60 percent above Nebius and Lambda, and SemiAnalysis estimates a 10‑15 percent premium over other managed clusters. Nebius undercuts CoreWeave on Blackwell‑class chips, offering a B200 price of $7.15 and a unique B300 on‑demand rate of $7.85. Lambda provides the cheapest published B200 instance at $6.69, though its 1‑Click Cluster bundles are priced higher for larger GPU counts. Crusoe stands out as the sole AMD option, listing an H200 price of $4.29 and an H100 rate of $3.90, comparable to Nebius and Lambda. Groq does not publish GPU‑hour rates; instead, its GroqCloud service is billed per token, positioning it as a niche inference offering.
The analysis also highlights limits and risks. Spot‑tier pricing is only available from CoreWeave and Nebius, with Nebius pre‑emptible H100 at $2.15 and CoreWeave spot at $2.46 per GPU‑hour, while Lambda lacks a spot market altogether. Several providers, including CoreWeave and Lambda, have not disclosed pricing for the newest Vera Rubin NVL72 hardware, leaving buyers to negotiate sales contracts. Large‑scale capacity deals, such as Nebius’s $12 billion five‑year order from Meta, create floor utilization but also tie up significant capital, potentially affecting flexibility.
Looking ahead, the competitive landscape will be shaped by upcoming capacity expansions and financial milestones. CoreWeave expects active power above 1.85 GW by year‑end and a full‑year revenue range of $12.4‑$13.2 billion. Nebius aims to deploy over 1 GW per year starting in 2027, while Crusoe’s development pipeline exceeds 40 GW across multiple campuses. Groq’s partnership with NVIDIA’s LPX platform suggests future hybrid offerings that blend LPU inference with GPU acceleration. Investors and customers should monitor IPO filings from Lambda and Crusoe, as well as any shifts in pricing strategy that could arise from the growing demand for AI‑specific hardware.
Provides concrete cost data for selecting GPU cloud resources for training and inference workloads.
Shows financial health and capacity commitments of major AI infrastructure providers, informing procurement decisions.
Highlights revenue growth, funding rounds and upcoming IPOs in the AI‑cloud sector.
- neocloud
- A specialized cloud service focused on AI workloads, often using dedicated GPU hardware.
- LPU
- Learning Processing Unit, Groq's custom inference chip sold on a per‑token basis.
- Vera Rubin NVL72
- NVIDIA's Hopper‑based GPU platform, recently validated by CoreWeave and Nebius.
BusinessTrump bought SpaceX shares two weeks after blockbuster IPO
BusinessAmjad Masad, CEO and co-founder of Replit, joins the Disrupt Stage at TechCrunch Disrupt 2026
BusinessHarvard’s $699 startup bootcamp offers AI avatars of its instructors
BusinessNvidia partners with data center developer Cloverleaf
BusinessThe DOJ is investigating a16z. What does this mean for venture capital?
AI ResearchPew study confirms sharp rise of AI-written text on the web since ChatGPT's launch
A Pew Research Center study reveals that over a third of English language web pages published since late 2022 show indicators of machine authorship.
SecurityInstinct’s powerful AI assistant is raising privacy and security concerns
Instinct, a new AI personal assistant, is drawing attention for its powerful features but also raising serious concerns about user privacy, security, and control over personal data.

Advancing price-performance for developers with GPT‑5.6 in Kiro
OpenAI’s GPT‑5.6 is now integrated into Kiro, giving developers higher quality code and an 82% cost reduction on benchmark tests.
HardwareCerebras unveils CS-4 with double the performance on the same chip
Cerebras has launched the CS-4, a rack-scale AI accelerator that doubles performance over its predecessor by optimizing power and cooling for the WSE-3 chip.
AI ResearchKids outlearn AI—and we still don’t know why
A new analysis highlights the stark data gap between children and large language models, showing that children master language with far fewer exposures. Researchers are using the BabyLM competition to probe how limited data can still yield linguistic competence.
AI ResearchWho’s behind the new ‘stealth model’ Ox Alpha?
A mysterious reasoning model named Ox Alpha appeared on OpenRouter, prompting widespread speculation regarding its anonymous creator.