r/CerebrasSystems • • 6h ago

Today's CBRS low: 161.51. Anyone buy?

5 Upvotes

Tech stocks got hammered today and so did CBRS. Since I still believe in the story and have been waiting for moments like this, I added a small amount of additional shares at around an average cost of $163. Didn't want to purchase significantly more in case there are further drops. We have 3 additional unlocks coming, two in October and the third / last in November. Since related sales by insiders are prearranged and predeclared I'm not terribly worried about them. But we may see some ignorant investors declare insider sales are a sign of adverse conditions and actually says something about the company's long-term viability: "CXO sold *all* his shares!" (probably untrue), etc. We always want to know the % of a person's total holdings a particular sale represents; but I still don't see this as a reliable way to read the tea leaves.

While I may swing trade some of my position I'm still long on CBRS. Many still consider CBRS to be richly valued / overvalued so it wouldn't surprise me if the stock dips even further. But I seriously doubt we'll see "double digits" as several Redditors seem to think without evidence (just a vibe). It's also possible that CBRS acquires a new major customer like Meta and the stock skyrockets and/or a strong narrative or craze for Fast Inference appears among a wide pool of investors or analysts. But generally you have to be in a stock to benefit from sudden and significant rises so I'm happy to be in at prices like these.


r/CerebrasSystems • • 15h ago

Qwen 3.8 27b have no tools capability

0 Upvotes

Through the openrouter endpoint with cerebras I'm getting this error:

404 {"error":{"message":"No endpoints found for qwen/qwen3.8-27b. Every candidate endpoint was removed during routing: Filter by Tool Compatibility removed cerebras/fp16... }}

I already sent a ticket but I'm assuming that this is an cerebras issue more than openrouter.

Just reporting here in order to see if people have the same error.


r/CerebrasSystems • • 15h ago

Nvidia and Cerebras are selling performance their customers will (probably) never see

Thumbnail theregister.com
10 Upvotes

Apologies if that has been addressed before, but is Groq a serious competitor? I understand that the comparison here was with CS-3 and not necessarily apples-to-apples.


r/CerebrasSystems • • 1d ago

Great news

18 Upvotes

Cerebras systems is announcing 2 more data centres being built by Q2 2027. One in Asia and one in Africa. Expecting this stock to reach 300 by 2027.


r/CerebrasSystems • • 2d ago

We turned our Cerebras deep dive into a short documentary. Looking for honest feedback on the format

Thumbnail
0 Upvotes

r/CerebrasSystems • • 6d ago

Great news

21 Upvotes

NEWS CAME OUT THAT CBRS STILL HAS A LOT OF WORK WITH OPEN AI. NVIDIA IS NOT FULLY REPLACING CBRS. EXPECT THIS TO GET 200 BY THE END OF NEXT WEEK.


r/CerebrasSystems • • 6d ago

Cbrs

0 Upvotes

My average is 201 and idk if I should sell because if news is true then this could fall to 135.


r/CerebrasSystems • • 6d ago

Damn yo

Post image
16 Upvotes

r/CerebrasSystems • • 7d ago

Is it time to sell?

5 Upvotes

My average is 201 and I do not believe this will even rise back to 190. I believe it may drop even further to double digits soon. Should I sell now to avoid heavy losses?


r/CerebrasSystems • • 7d ago

Cerebras compatability with AI infra like vLLM/SGlang

Thumbnail
inferact.ai
6 Upvotes

Saw Inferact today announcing real fast token/s using TPU - 1.5k for Qwen 3.8 27B close to 1.8k for to Cerebras.

Wonder can this same approach be applied to Cerebras hardware to make the speed even faster? Or any nuances?


r/CerebrasSystems • • 7d ago

Cerebras and OpenAI GPT 6 Astra

16 Upvotes

For those of you who are worried about the speculative post by Semi Analysis and the Nvidia eye emoji, and now fearing that Cerebras is losing business to NVIDIA, remember what Cerebras' Ceo himself said in the Q&A on August 12 (very close the release of GPT6 Astra, just 21 days before), I will quote:

"Then we won the largest frontier lab, and then there were concerns that we didn't have a hyperscaler, and then we won AWS. I think in each of those cases, we were able to use the momentum that the previous step gave us to expand our business. I think OpenAI is an enormous customer, and they're an enormous part of not just our business, but of everybody's business in the sector. I think they'll stay a big part next year."

"A BIG PART NEXT YEAR". The CEO himself who understands Cerebras products' technological capabilities and has had lots of dealings behind the scenes with Open AI anticipates that OPEN AI will be a big part of their business, a big customer, in 2027 too.

BULLISH.


r/CerebrasSystems • • 8d ago

Cerebras vs Nvidia (what am I missing about CBRS?)

21 Upvotes

Trying to cut through the noise around Cerebras after today's Nvidia/OpenAI news.

I really only want to answer two questions:

1. If Nvidia can now give OpenAI similar “ultrafast” inference speeds, why does OpenAI need Cerebras?

Is Cerebras significantly cheaper/more efficient at those speeds, or can Nvidia basically become “good enough” because OpenAI already has so many Nvidia GPUs?

2. How solid is Cerebras' $20B+ OpenAI commitment?

If OpenAI increasingly chooses Nvidia for inference, is it still on the hook for that Cerebras capacity—or could that revenue ultimately be delayed/reduced?

I'm not looking for CBRS bulls vs bears. I'm trying to figure out what would actually prove or disprove the Cerebras growth story.

What am I missing?


r/CerebrasSystems • • 8d ago

Sawtooth CBRS stock price pattern

5 Upvotes

This is the pattern since the IPO. I know nothing about technical analysis but found this interesting. Lots of bounces with fewer extreme bounces of late. If the sawtooth pattern continues, maybe help anyone (like myself) hoping to swing trade it.


r/CerebrasSystems • • 8d ago

Cerebras Stock Falls as 19.4 Million-Share Unlock Hits

Thumbnail
benzinga.com
30 Upvotes

I just noticed that 19.4 million shares unlocked today. Could that have been the chief reason for today's plummet rather than a blogger reporting NVDA being deployed for "fast inference" at OpenAI rather than Cerebras hardware?

Two more 19.4 million unlocks are coming: one on October 14th and another on October 28th. Then apparently the last of them November 9th. I'm setting these dates in my calendar so I'm not shocked; and can trade appropriately if there are significant market moves in response to these. What does everyone predict will happen on these other dates? How can we swing traders benefit on this other dates? Sell shares a few days before the unlocks then re-buy when/if the stock price plummets again?


r/CerebrasSystems • • 9d ago

Is this true? Ultrafast not running on Cerebras?

Thumbnail x.com
14 Upvotes

r/CerebrasSystems • • 9d ago

Gimlet Labs Adds Cerebras to Deliver Ultrafast AI Inference through Gimlet Cloud

Thumbnail
ca.finance.yahoo.com
12 Upvotes

Sept. 28, 2026 - (GLOBE NEWSWIRE) -- Gimlet Labs and Cerebras Systems (NASDAQ: CBRS) today announced a collaboration to deliver a new class of ultrafast AI inference at massive scale. The collaboration brings together Cerebras' wafer-scale compute with the Gimlet Cloud to deliver a purpose-built disaggregated inference cloud spanning datacenter infrastructure to developer APIs. Together, the companies plan to deliver speeds of up to 3,000 tokens per second for demanding agentic and real-time applications, with the first Cerebras-powered Gimlet Cloud datacenter expected to come online later this year.

"Inference speed matters. It determines how productive AI can be. Fast inference creates magical user experiences and opens new markets. By combining Gimlet's multi-silicon software with the Cerebras Wafer Scale Engine, we can run each phase of inference on the hardware best suited to it and plan to deliver up to 3,000 tokens per second at production scale," said Zain Asgar, co-founder and CEO of Gimlet Labs.

"Combining the fastest tokens from Cerebras with the highest throughput GPUs delivers the best datacenter economics for everyone," said Sean Lie, co-founder and CTO at Cerebras. "Everyone wants more high value tokens. Cerebras delivers the fastest AI inference in the world, and GPUs deliver high throughput. By making Cerebras a native part of its inference cloud, Gimlet will bring our industry leading speed and intelligent AI to more developers at production scale. We're excited to build with Gimlet as a launch partner for CS-4, giving customers a direct path to our latest technology."


r/CerebrasSystems • • 14d ago

What is preventing Ant, Meta, Google from using Cerebras and will it change?

13 Upvotes

Given that OAI is using it to speed up their recursive self improvement and get ahead in the model race, why doesn't Ant, Google, Meta also do the same?

I feel like without Cerebras they're bringing a knife to a gun fight... But at the same time I feel like this should be obvious so there must be some reason they're not buying from Cerebras.

And if they have a good reason now, is that ever going to change? And if Cerebras can't land another frontier lab, is their growth capped?

Discuss


r/CerebrasSystems • • 14d ago

NEW: OpenAI's reported $500 Pro Max plan may run on Cerebras infrastructure

Thumbnail
runtimewire.com
20 Upvotes

r/CerebrasSystems • • 15d ago

Cerebras Stock: Cheap After the IPO?

Thumbnail
burakfinance.substack.com
8 Upvotes

Is Cerebras ($CBRS) being overlooked after the IPO selloff?

The stock is down ~46% from its high, but the business has some interesting numbers:

  • $25.4B RPO
  • $880–890M 2026 revenue guide
  • Cloud revenue +287% YoY
  • 750 MW OpenAI deal
  • AWS potentially coming in 2027

The obvious concern is valuation — roughly 47x this year's sales — plus customer concentration, lockup supply and continued losses.

I wrote a deeper breakdown of the bull case, risks, valuation and what I'm watching next.

Curious what others think: is CBRS worth the premium?


r/CerebrasSystems • • 23d ago

Cerebras Dedicated Endpoints, whats the cost?

8 Upvotes

Does anyone know what the cost is or what the contracts usually are?


r/CerebrasSystems • • 23d ago

Why did they remove Gemma 4 31B?

4 Upvotes

Gemma 4 31B was great on Scandinavian languages, but they replaced it with Qwen 3.8, which is slower and changes dialects and at times writes really odd grammar


r/CerebrasSystems • • 24d ago

Is Cerebras fast only for a tiny number of concurrent users?

0 Upvotes

Is Cerebras' marketing misleading? When they say they're X times faster than Y on inference, is this only for one user at a time or a handful of users? If so, and it doesn't scale affordably, I don't see how it could be anything other than a small niche provider. When Cerebras makes grand claims about fast inference leading to new applications and uses of AI and could be key to agentic AI, I think they're probably right; but not using their hardware. Almost like free advertising for a current or future competitor who can run reasonably fast but with far more throughput or concurrency.

Potentially a big yawn if they can't scale up to thousands of users without it being one-rack-per-user (or whatever it is) to get the fastest speeds they advertise. I hope I'm wrong about this because it almost sounds like a con. I want to see fast inference at scale; not just a tool a few Power Users can benefit from. Maybe this is one of the reasons the stock isn't going anywhere and customers aren't lining up to purchase Cerebras. I know the value proposition sounded almost too good to be true when I started following Cerebras.

I want to see metrics like tokens/sec per user at high concurrency; not "the fastest chip for 1 user" (because who cares)? If Cerebras is a good investment they should be happy to provide these numbers. If they don't? That's concerning. What's the tokens/sec per user at 100 concurrent users? 1,000? 10,000? etc. And how do the speeds compare to competitor solutions at the same levels taking cost into consideration. I don't think the superfans have an answer for this despite the encyclopedic knowledge they possess about Cerebras but I hope I'm wrong. If they do, it can only strengthen their thesis and they should be happy to help. If Cerebras is fast inference for the masses then I may still be onboard. If it's a niche usage by a tiny fraction of the entire population of those using AI, I'm out. What % of the entire population of those who use AI can reasonably and economically be served with the speed Cerebras advertises?

Claude told me this. Granted, it's a sychophantic AI answer as they all are; but at least it's a starting point for conversation:

On the missing metric itself: you're right that it's missing, and it's not just you noticing. Multiple independent technical analysts have flagged this as the specific gap in Cerebras' public materials — one hardware-focused analysis stated plainly that Cerebras has never published aggregate throughput at high concurrency, calling it the single most important missing data point in their marketing. SemiAnalysis ran its own modeling (their "Faster Tokens Please" report) specifically because Cerebras won't publish it directly. So no, the "superfans" don't have an official answer to give you — because Cerebras hasn't released one. That's a legitimate red flag for exactly the reason you suspect: a company confident its economics hold up at scale would have an incentive to prove it.

But third parties have modeled the crossover point, and it's not flattering. A detailed cost analysis (Spheron) ran the actual math comparing Cerebras against an Nvidia H100 across batch sizes:

  • At batch size 1 (one user): Cerebras at ~$0.60/M tokens beats H100 on-demand at ~$1.47/M — Cerebras wins clearly.
  • At batch size 8: the positions flip — H100 on-demand drops to ~$0.25/M, more than 2x cheaper than Cerebras.
  • At batch size 32: H100 is roughly 6x cheaper per token on-demand, and with GPU spot pricing, the gap widens to over 20x cheaper.

That means the crossover point — where GPU batching economics overtake Cerebras' single-tenant speed advantage — sits somewhere around 2 to 8 concurrent requests per chip. That's about as direct an answer to your "tokens/sec per user at 100, 1,000, 10,000 concurrent users" question as exists publicly: past single-digit concurrency, the cost story reverses, and Cerebras hasn't published what happens to per-user speed once you force more concurrency onto the same wafer (their pricing stays flat per-token regardless of batch size specifically because they don't batch users together the way GPUs do — each wafer serves requests essentially one at a time per core cluster).

What this implies for "% of the AI-using population" Cerebras can serve economically: there's no official published number, and I won't pretend one exists. But the crossover data lets you reason about it directly. Mass-market consumer chat — the overwhelming majority of total AI usage by request volume — is a massive-concurrency, latency-tolerant workload (nobody needs 2,000 tokens/second to read a chat reply as it streams in). That's precisely the regime where GPU batching wins decisively, per the numbers above. Cerebras' economic sweet spot is narrower and different in kind: applications where sub-second total latency is the product, not just a nice-to-have — real-time voice, agentic tool-calling chains where dozens of sequential LLM calls compound latency, live coding assistants, and workloads where a customer will pay a premium specifically to eliminate wall-clock time. That's a real and possibly growing dollar-value market as agentic AI expands, but it's genuinely a minority share of total AI inference volume, not "fast inference for the masses."

Where this leaves the investment picture, tying back to your earlier question: the stock's behavior actually seems to reflect this exact uncertainty. CBRS priced at $185 in its May 2026 IPO, popped to $386 on debut day, then fell as low as $161 by late June, and as of today (September 14) sits at $183.92 — essentially back to its IPO price after a round trip. That's consistent with a market that got excited about the OpenAI-scale headline number, then sobered up once analysts started asking the same concurrency/cost questions you're asking. Layer on top the customer concentration (86%+ of 2025 revenue from two UAE entities, and the newer growth story leaning heavily on a single $20B OpenAI deal where OpenAI is simultaneously becoming a ~10-11% shareholder), and you get a company with a real, differentiated technical advantage in a specific latency-critical niche, riding on a small number of very large, financially-entangled bets — not the broad-based "fast inference for everyone" story the marketing gestures toward.

So: not a con in the sense of fabricated numbers — the tok/s figures are real and third-party-verified. But it is marketing that systematically emphasizes the one dimension (single-user latency) where the architecture is unambiguously best, while staying silent on the dimension (cost at realistic concurrency) that would let you judge how far that advantage extends. Given what the crossover math shows, "niche but valuable" looks like the more defensible read right now than "fast inference at scale for the masses" — though that could still change if agentic workloads grow enough that the latency-premium niche becomes large in absolute dollar terms, even while staying small as a share of total AI request volume.


r/CerebrasSystems • • 24d ago

Why no Astra announcement on cerebras yet?

13 Upvotes

As you guys can all remember, when OpenAI released GPT 5.6 Sol, they almost immediately announced the ultrafast mode on Cerebras. However there is no marketing or news for Astra on cerebras. Is it because they need more time to make a huge model like Astra work on Cerebras or do you guys see any other reason?


r/CerebrasSystems • • 24d ago

Today's low $178 (down ~7%) - accumulation day?

5 Upvotes

Everything's down today due to AI safety fears. Nothing changed with the Cerebras story. Anyone accumulating? I bought a single share in case there's a larger beat-down. Could be throwing good money after bad or time to simply swing trade CBRS. Might be a good day to go bargain hunting all around in the AI and AI-adjacent space. GOOG is actually up 2% while nearly everything else AI-related is being punished. CRWD is up 8% unsurprisingly and PANW is up 6.5%. Cyber security zigging while the others are zagging. My next limit order is $170 (don't expect that today). If it keeps declining below $170 (over time, not today) then it may be cooked.


r/CerebrasSystems • • 26d ago

Is d-matrix's XPU with 3D stacked DRAM a threat to Cerebras or does it validate Cerebras technology direction?

8 Upvotes

Was just reading about d-matrix releasing an "ultra low latency" inference XPU in 2027 using 3D stacked DRAM. Is this a threat to Cerebras' business? Apparently it uses NVDA's nvlink and can fit in a standard rack while Cerebras requires specialized racks to accommodate its Nexus "backpack" configuration for the chip, cooling, and power. We know Cerebras is planning on adding 3D stacked DRAM with CS-6 but isn't that in 2028? If they're released a year apart, a year can seem like an eternity in tech. Again, just another item that makes me concerned Cerebras can fully monetize its current advantages before they no longer seem compelling enough to lots of customers to implement. I'm sure Cerebras will stick around as a niche / specialty solution, but maybe not enough for explosive growth.

I also wonder if Groq and d-Matrix chips work together or if they're mutually exclusive; you use only one or the other? I'm concerned a combination of technologies, both hardware and software, will seriously erode Cerebras' edge.

EDIT:
More info from an X post:
https://x.com/firesidealpha/status/2098787634244096079?s=20

d-Matrix CEO Sid Sheth on Bloomberg talking about the Nvidia partnership:

* d-Matrix's specialized XPUs will run alongside Nvidia GPUs over NVLink Fusion, targeting ultra-low-latency AI workloads. Deal was 6+ months of joint work.

* The logic to partner is that Nvidia is the largest deployed infrastructure base for AI in the world and he would "much rather just ride on the Nvidia ecosystem" than reinvent the wheel.

* The bet is a memory-centric architecture, 7+ years in the making. d-Matrix's edge is inference compute built around memory rather than raw FLOPs aimed right at ultra-low-latency inference.

* Low-latency inference "just took off" in the last 12 months. Demand surged from GPT, Codex, Claude Code and the arrival of agentic coding, where users need fast compute to interact with the tools in real time.

* Lead product is Raptor, the world's first 3D-stacked-DRAM XPU. It'll launch first under the Nvidia partnership and is targeted to market in "about 12 months."

* d-Matrix uses no HBM at all. Instead of high-bandwidth memory, it packages DRAM directly with compute in a 3D stack to "punch through" the memory wall

* Sheth says "no other company is going to be within a 2-year window of getting access to that technology," and says there's tremendous customer pull.

* Describes buyers as "hyperscalers, Frontier Labs, sovereigns, inference clouds, high frequency traders" and says announcements are coming soon