TLDR
- Cerebras shares declined 7-9% following allegations that OpenAI employs Nvidia processors rather than Cerebras technology for its “Ultrafast” AI service tier.
- According to SemiAnalysis, OpenAI’s GPT-6.1 Sol Ultrafast operates on Nvidia GPUs with minimal batch processing, bypassing Cerebras’ Wafer-Scale Engine.
- OpenAI represents Cerebras’ biggest revenue commitment, with both companies having publicly stated in August that Cerebras would support the Ultrafast feature.
- No official confirmation or denial has been issued by either Cerebras or OpenAI regarding these allegations.
- Shares have declined significantly since the company’s May IPO debut at $350 per share.
Shares of Cerebras Systems experienced a significant decline this week following a report that cast doubt on the company’s involvement in supporting OpenAI’s premium speed AI offering. Trading saw the stock tumble as much as 8.87% during Wednesday’s market hours.
The catalyst emerged from semiconductor analyst firm SemiAnalysis, which claimed via social media that OpenAI’s latest GPT-6.1 Sol Ultrafast model operates on conventional Nvidia GPUs rather than Cerebras’ proprietary chip architecture.
This assertion carries substantial weight. When OpenAI unveiled its “Ultrafast” service tier, it touted performance gains reaching eight times normal speeds, delivering approximately 300 tokens per second. Cerebras had been publicly identified as the technology provider enabling those enhanced speeds.
Cerebras constructed its entire value proposition around addressing precisely these performance challenges. The company’s Wafer-Scale Engine represents a colossal integrated circuit that incorporates billions of processing cores and extensive on-chip memory within a single silicon wafer. This architecture eliminates latency bottlenecks associated with transferring data across hundreds of discrete GPU units.
Breaking Down the SemiAnalysis Claims
The critical element lies in a specific technical detail mentioned in the analysis: “low batch size.” In artificial intelligence inference operations, batching refers to the practice of consolidating multiple requests for simultaneous processing to maximize GPU utilization.
Nvidia’s processors typically achieve optimal throughput with large batch configurations. Operating with minimal batch sizes means processing fewer simultaneous requests, which prioritizes response latency over computational efficiency. This represents precisely the use case Cerebras has positioned itself to dominate.
Should OpenAI have successfully optimized Nvidia hardware to deliver premium-tier performance at reduced batch sizes while maintaining cost effectiveness, it would directly undermine a cornerstone of Cerebras’ competitive differentiation. This explains the swift market reaction.
Cerebras stock settled around $180.22 on Wednesday following the news circulation. This represents a dramatic decline from its $350 IPO valuation established in May.
The initial public offering generated considerable investor enthusiasm. Market participants showed strong interest in a semiconductor company claiming superior inference speed capabilities compared to Nvidia for specific computational workloads.
That initial optimism has substantially diminished. Post-IPO financial disclosures have not sustained the momentum witnessed during the company’s market debut.
The Strategic Significance of the OpenAI Relationship
OpenAI represents far more than a routine customer for Cerebras. The AI research company constitutes Cerebras’ most substantial client measured by committed revenue, according to Barron’s.
Any erosion of this partnership, regardless of duration, presents legitimate concerns for shareholders. Cerebras has prominently featured its OpenAI collaboration when defending its market valuation.
Neither organization has issued public commentary addressing the SemiAnalysis claims. This communication void has forced market analysts to operate on speculation rather than verified information.
Alternative explanations remain plausible. Infrastructure capacity limitations or ongoing technical refinement processes could account for Nvidia hardware currently handling these workloads.
Cerebras may ultimately provide the computational backbone for Ultrafast mode once any engineering challenges are resolved. No official statements from either party preclude this possibility.
Currently, the market response reflects ambiguity rather than definitively negative developments. Investors are incorporating the possibility that Nvidia’s extensive CUDA software ecosystem and continuous hardware advancements create formidable barriers for competitors seeking exclusive partnerships with leading AI organizations.
Cerebras has not released any communication verifying or refuting its present status within OpenAI’s technology infrastructure. As of Wednesday’s closing bell, shares remained depressed approximately 8% for the trading session.


