Nefarious Trading Est 2021
⏱ 9 min read Single Stock · Vol. 01 No. 79 · August 2026
CBRS$202.00 ▼ 48% OFF HIGH RPO$25.4B EV/SALES60x GAAP GM14.2% CORE GM40.6% SHORT INT11.7% OF FLOAT STREET PT$291.64 CBRS$202.00 ▼ 48% OFF HIGH RPO$25.4B EV/SALES60x GAAP GM14.2% CORE GM40.6% SHORT INT11.7% OF FLOAT STREET PT$291.64
Single Stock · AI Silicon · Recent IPO
The Wafer-Scale Bet
NASDAQ: CBRS · $202.00 · 52W $160.81–$386.34
$202
IPO $185 · May 2026

They built a chip the size of a dinner plate. The hard part was never the silicon.

CBRS — The Wafer-Scale Bet | Nefarious Research
§ The Thesis

You are buying speed, from one customer, at a price that assumes 2028.

Cerebras built the fastest inference chip in the world by putting the whole model on a single wafer. That part works — OpenAI runs its fastest model on it in production, and the order book stands at $25.4B. The technology risk is largely retired.

What replaced it is customer risk and price risk. Two customers are 66% of revenue, the backlog is overwhelmingly one counterparty, and at 60x sales this is the most expensive name in AI silicon by nearly three times. Meanwhile Nvidia paid roughly $20B in December to buy the competing version of this architecture and now sells it as a tier inside its own platform.

§ Plain English

What you actually own. The whole filing cabinet on the desk.

1
Wafer = one chip
900k
Cores
44GB
Memory on the chip
750
Tokens per second

Every other chip company slices a 300mm wafer into hundreds of small chips. Cerebras doesn't slice — the entire wafer is one processor.

Why that matters: when an AI writes you an answer, the bottleneck isn't thinking, it's fetching. A normal GPU is a fast worker who keeps walking across the room to a filing cabinet. Cerebras built a worker with the whole cabinet on the desk. Less room for files, but no walking.

The result is speed. A token is about three quarters of a word, so 750 tokens a second is 560 words every second — against the 4 words a second you read at. It writes 142x faster than you can read.

On small jobs that gap is invisible. Scale the job up and it becomes the whole story.

The jobOn CerebrasOn a normal setup
A full essay2 sec25 sec
A long research report18 sec4 min
An entire novel3 min41 min
The Lord of the Rings trilogy14 min3.3 hrs
The full Harry Potter series32 min7.4 hrs
An agent refactoring a large codebase59 min13.7 hrs
An agent reading every 10-K in the S&P 5004.9 hrs68.6 hrs
Look at the bottom two rows

Fourteen hours versus one hour is not a nicer experience — it is a different working day. One means you kick the job off and come back tomorrow. The other means you run it, read it, fix it and run it again before lunch.

And the last row is the one that matters commercially. Sixty-eight hours is nearly three days. Five hours is an afternoon. At that scale speed stops being a feature and becomes whether the job is worth doing at all — and that is the only thing Cerebras is selling.

The catch, in one line

44GB is a small desk. Big models don't fit, so they get split across many machines — and the memory on these chips has stopped getting bigger with each generation.

§ The Business

They stopped selling machines and started renting them.

$180M
Q2 GAAP revenue, +74%
+281%
Cloud revenue growth
-23%
Hardware revenue
$25.4B
Contracted backlog

Two lines: hardware, where they sell you a box, and cloud, where they own the box and bill you for time on it. Hardware fell 23% last quarter while cloud grew 281%. That shift is deliberate — recurring revenue at better margins — and it is why capex is running above $500M a quarter.

But read the pivot the other way too. One semiconductor analyst put it bluntly: running your own token factory with your own hardware is not where a chip company wants to be, and you only do it if people aren't buying the hardware directly. Groq walked the same path into GroqCloud before Nvidia absorbed it.

§ Who Runs It

The SeaMicro band, back together.

Founded around 2015 in Sunnyvale by five people who worked together at SeaMicro. CEO Andrew Feldman co-founded that company and sold it to AMD in 2012 for roughly $334M. CTO Sean Lie is a co-founder. CFO Bob Komin joined in April 2024 — he took Sunrun public and sold three companies to Yahoo, Pandora and Microsoft. An exit CFO, hired two years before the listing.

Feldman holds about 14.1M shares and Lie about 8.4M, almost all Class B carrying 20 votes each. Founder control is intact.

The insider read

Zero open-market buys since the IPO. Every filing is a sale. But Feldman's and Lie's were mandatory tax withholding on vesting shares. The genuinely voluntary sellers were the venture directors — Benchmark sold $15.7M and Foundation $10.9M, both within days of being allowed to.

§ The Margin Trap

The scary number is mostly an accounting artifact.

GAAP gross margin fell from 31.1% to 14.2%. Everyone read that as the business breaking. Here is the actual bridge.

StepRevenueMargin
Core (company's own measure)$209.9M40.6%
Add back pass-through billing$224.4M38.0%
Subtract stock comp in cost$224.6M31.1%
Subtract OpenAI warrant charge$180.1M14.2%

A warrant granted to OpenAI is booked as a reduction to revenue rather than as a cost. That single item is $44.3M in the quarter and explains essentially the whole 17-point drop. On the core basis, margin actually improved 940 basis points.

It is also structurally perverse: the faster OpenAI buys, the bigger the charge and the worse GAAP looks. The asset is $1.1B and runs to October 2031 — roughly 25 more quarters of this.

The problem that is real

Core margin still fell sequentially, from 46.5% to 40.6%, because Cerebras is renting its own machines back from cloud customers — demand arrived faster than they could build data centres. Management says that cost 500bps, that Q3 is the low point, and that Q4 improves as owned capacity comes online.

§ The Customer Problem

The risk moved from Abu Dhabi to San Francisco.

CustomerQ2 2026Q2 2025
Customer A34%70%
Customer B32%under 10%
Customer C10%under 10%
Top two66%

The top customer fell from roughly 87% of revenue in early 2024 to 34%. Genuine progress. But two names are still two thirds of the business, and filings show OpenAI revenue of $56.8M in the quarter — 31.5% of the total.

And OpenAI is not just a customer. It is simultaneously the largest customer, the holder of most of the $25.4B backlog, a $1B lender whose loan is forgiven if repaid in compute rather than cash, and a warrant holder for ~33M shares at $0.00001. No cash has been repaid on that loan — it is being worked off in billing credits.

In plain English

OpenAI lends Cerebras money, Cerebras builds data centres, OpenAI buys the capacity, and the loan gets cancelled out in the process — while OpenAI earns shares for doing it. That is circular financing. If you're uneasy about OpenAI's balance sheet, this is one of the most concentrated ways to own that risk.

§ The Supply Chain

TSMC holds every card.

Cerebras depends on one foundry, on 5nm. The real lock-in isn't the relationship — it's the reticle-stitching process, co-developed with TSMC over roughly a decade, which independent analysts say does not port to another fab. There is no second source at any price.

Worse, the filings state manufacturing is bought on a purchase-order basis with no capacity or volume commitments. Management says they've secured the wafers they need. That is an assertion, not a contract.

The offsetting advantage is real though. Cerebras uses no HBM, no CoWoS packaging and no leading-edge node — the three tightest bottlenecks in AI silicon. As Feldman puts it, the constraints facing the industry don't apply to them.

§ The Demand Side

The order book is enormous and very slow.

When the $25.4B convertsShareApprox
Next 24 months~22%~$5.6B
Months 25–4843%~$10.9B
Thereafter~35%~$8.9B

Backlog is roughly 29x guided revenue. In any normal industry that's extraordinary. But only about a fifth lands inside two years, so it is contracted and a long way from the income statement.

The constraint is Cerebras's own — over 600MW live or contracted, and they're renting systems back to serve demand they can't host. Which means a new customer doesn't unlock revenue. It joins a queue. That is the flaw in the argument that one big deal re-rates this stock.

Where the demand leaks is straightforward: any latency-sensitive workload Cerebras can't serve on time gets served by Nvidia instead.

§ The Competition

Nvidia already bought the other version of this idea.

In December 2025, Nvidia licensed Groq's inference technology and hired its team including founder Jonathan Ross — reported at roughly $20B. Groq survived, re-based at a $3.5B valuation down 49%, and now resells Nvidia GPUs as a cloud partner. It is no longer a chip company.

Nvidia now ships that IP as NVIDIA Groq 3 LPX, the seventh chip of the Vera Rubin platform, claiming 35x higher throughput per megawatt on very large models. It arrives 2H 2026.

Note what Nvidia changed. LPX is not memory-on-chip only — each rack pairs 128GB of fast memory with 12TB of conventional memory and runs alongside GPUs. Nvidia didn't accept the pure approach. It turned it into a tier inside its own platform, which is far harder to compete with than a rival chip.

CerebrasNvidia + LPXQualcomm
Bets onMax speedEverything, at high powerCapacity per watt
StatusShipping2H 20262027–2028
WeaknessMemory capacityPower and costNothing shipping yet

Qualcomm is the 2028 threat, not the 2026 one — and note their Meta deal is a CPU agreement contributing nothing until 2028, which a lot of coverage gets wrong.

And the Meta question everyone is asking.

Meta already works with Cerebras — it powers fast inference inside the Llama API, announced at LlamaCon, hitting about 2,600 tokens/sec on Llama 4 Scout. But read what that is: Meta is reselling Cerebras speed to outside developers, not buying compute for itself. Meta is a distribution channel, not a customer.

The proof is in the filings — OpenAI, G42, MBZUAI and AWS are named. Meta is not. Mizuho thinks Meta could be announced as a full customer in 2H 2026, which would make it the third major cloud customer. That is a forecast, not a signed deal.

The complication: Meta is spending $115-135B in 2026 building four generations of its own inference silicon with Broadcom, and describes that program as inference-first. A large external contract would land in the same window it ramps its own chips.

Scale check, stated honestly: Nvidia holds roughly 75-85% of AI accelerator revenue and hyperscaler in-house chips another 15-20%. Merchant challengers including Cerebras are well under 1%. Cerebras's whole FY26 guide is about 0.35% of Nvidia's trailing revenue.
§ The Bear Case

Three reasons the technology may not be as strong as it looks.

One
The memory doesn't scale with the chip
The high-density memory cell has been stuck at the same size from 5nm through 3nm into 2nm, while competing memory scales toward 1TB+ by 2028. Shrinking the process gets them more logic, not more memory. The gap is widening, not closing — which is exactly why CS-4 is three existing wafers bolted together rather than a new chip.
Two
They're 2-5x more expensive per token
Against a throughput-optimised GPU, latency-optimised Cerebras carries a 2-5x cost penalty. It only becomes competitive when you force the GPU down to the same response time. The customer has to genuinely need the speed — and OpenAI was reportedly looking to move only about 10% of its inference to low-latency hardware.
Three
It was built for training, and training rejected it
Bandwidth was never training's constraint and CUDA was entrenched. Inference arrived later and retroactively justified the architecture. That's good luck meeting good engineering — and luck can turn again if software techniques reduce the value of raw bandwidth.

Worth knowing the precedent. Gene Amdahl tried wafer-scale at Trilogy Systems in 1980, burned roughly $230M, and shipped nothing — killed by yield, lithography and layers delaminating under heat. Cerebras solved that with far finer redundancy: about a million physical cores of which 900,000 are exposed, so bad ones are simply mapped out.

§ Valuation

The most expensive name in AI silicon, by nearly three times.

TickerMkt capEV/SalesType
CBRS$48.0B60.0xHybrid
NBIS$59.4B45.7xNeocloud
MRVL$213B24.4xAI silicon
AVGO$1.73T23.5xAI silicon
NVDA$5.27T20.7xAI silicon
AMD$760B18.2xAI silicon
Comparability note: the neoclouds aren't real comps — they're capex-financed rental businesses. The silicon comps median 22x. Cerebras is a hybrid that sells chips and runs its own cloud, and the one clean comp that existed, Groq, was bought by Nvidia.

On the FY26 guide the multiple is ~46x. On FY27 consensus of $2.95B it's 13.9x. On Morgan Stanley's contracted path to ~$6B by 2028, roughly 6.8x. To sit at the silicon median today, Cerebras needs about $1.85B of revenue against an $885M guide.

You are paying today for 2027 to 2028 execution. This is not a 2026 story at any price. The fair offset: consensus three-year growth is 145% versus 48% for Nvidia, so growth-adjusted it's less absurd than it looks.

Fair value, on FY27 revenue.

FY27 revenue8x12x16x20x
$2.4B miss$111$151$192$232
$2.95B consensus$129$179$229$278
$3.5B beat$148$207$266$325
Calculated: (FY27 revenue × multiple + $7.12B net cash) ÷ 237.56M shares. At $202 the market is paying about 14x FY27 consensus. The hard floor is roughly $30/share in net cash, about 15% of the price.
§ What To Watch

The calendar, in order.

EventWhenWhat settles it
Nvidia LPX ships 2H 2026 First independent speed tests against CS-4. Within 2x on latency and the moat argument weakens badly.
Q3 results ~Nov 2026 Management guided core margin to a 38-40% low point. Below that and the rent-back is worse than admitted.
Lockup expiry Nov 9 or 2 days post-Q3, whichever is earlier ~171M shares free — 5x the IPO size against a 110M float with 11.7% already short. Largest scheduled event on the calendar.
Q4 results ~Feb 2027 The guide needs $267-277M, a 24-28% sequential jump. And core margin above 45% to prove the turn. Under $250M and the 2027 tripling is in doubt.
Q3 and Q4 reporting dates are estimated from the Q2 pattern — the quarter ended June 30 and was reported August 12. Cerebras has not confirmed either date.
§ Technicals

Sitting right on the shelf.

CBRS daily chart showing a rounded base from June through August 2026
CBRS daily. Rounded base off the $160.81 low, price back above both moving averages, and heavy volume on the August advance.
Floor
$160.81
-20%
52-week low, only tested floor
Support
$200
-1%
The shelf. Lose it and structure breaks
Now
$202
Sitting directly on support
Reclaim
$265
+31%
Confirms the base

The structure is a clean rounded bottom and the volume profile supports it — distribution into the June lows, a long quiet middle, then rising volume on the advance. What it hasn't done is reclaim $265. More urgent is the downside: at $202 price is directly on the $200 shelf. Hold it and the base is intact. Lose it and the next tested floor is 20% lower.

Chart is John's own. Levels are read off it, not independently derived.
§ My Take

Expensive, and worth it.

Johnny's verdict

Cerebras is a very expensive company compared to the others — I'm not going to pretend otherwise. But I think it's worth it, because they've done something genuinely special that nobody else has managed to do.

And it isn't theory. The OpenAI deal proves the thing works. The financing could be better, and yes, that ties them to OpenAI's success — that is the honest risk in this one.

But here is where I land. I think Cerebras is the closest company in the world to Nvidia on chips. And the fact that Nvidia went out and bought Groq to put that technology inside its own platform tells you everything you need to know. You don't pay $20 billion to absorb something that isn't a threat.

Not financial advice — please do your own research.

🔒 CBRS Trade Plan
Members only
Entry
• • • •
Position size
• • • •
Add level
• • • •
Take profit 1
• • • •
Take profit 2
• • • •
Stop
• • • •
Unlock plan
• • • •
Other positions
• • • •

Want the entries, stops, and sizing?

Join the Discord to find out! →
discord.gg/nfrs · @Nefarioustrading
Nefarious Trading
Equity research and trading commentary — AI infrastructure, semiconductors, aerospace, energy, commodities.
AuthorJohnny Li
Sources
Cerebras Q2 2026 results release and non-GAAP reconciliation tables (Aug 12, 2026) · Form 10-Q for the quarter ended June 30, 2026 · Q2 2026 earnings call · IPO pricing release (May 13, 2026) · Form 4 filings under CIK 0002021728 · SUPERNOVA event release and CS-4 launch (Aug 18, 2026) · Callosum partnership release (Aug 20, 2026) · Groq–Nvidia non-exclusive licensing release (Dec 24, 2025) and subsequent CNBC, Tom's Hardware and TrendForce coverage · NVIDIA Groq 3 LPX product page and developer blog · Qualcomm Dragonfly roadmap and Investor Day (June 24, 2026) and AI200/AI250 release (Oct 27, 2025) · Meta MTIA custom silicon announcement (March 2026) and Meta–Cerebras Llama API release · Mizuho, Morgan Stanley, UBS, Citi, Rosenblatt and Needham notes via TheFly and TipRanks · peer market caps, enterprise values and multiples via stockanalysis.com / S&P Global Market Intelligence (Aug 20, 2026) · SRAM scaling and wafer-scale bottleneck analysis via Vikram Sekar / SemiExponent (May 12, 2026) · Trilogy Systems history via contemporaneous accounts of the 1980–1984 period · CBRS price of $202 supplied by John on Aug 21, 2026; the embedded chart is the Aug 19 close of $215.69. Fair value grid and net cash per share are calculated, with inputs shown inline. Customer identities are anonymised in Cerebras filings — any mapping to named entities is inference and labelled as such. Q3 and Q4 reporting dates are estimated from the Q2 pattern and are not company-confirmed.
One trader's view — not investment advice. Do your own research. HONA $168.51 intraday Aug 7, 2026; quotes move and this will be stale by the time you read it. HONA has traded publicly only since June 29, 2026, so historical data is extremely limited and the 52-week range may include when-issued or pre-separation pricing. Scenario prices are illustrative constructions from management guidance and peer multiples, not forecasts. Enterprise value is approximate, calculated as market cap plus long-term debt without netting cash. Forward multiples rely on management guidance that was already revised downward once. Casting supplier names reflect industry structure and are not confirmed HONA vendors, which the company does not disclose. Peer book-to-bill figures are segment-level and not comparable to HONA's company-wide figure. Rolls-Royce figures are GBP-converted and approximate. The 2030 normalisation timeline is third-party industry analysis, not company guidance; management guides to meaningful improvement in 2027. © 2026 Nefarious Trading.