S&P 5007,785.76▼0.2% Nasdaq26,729.16▼0.3% Dow53,732.41▼0.2% Russell 2K3,068.42▲0.5% 10-Yr4.70%+6bp VIX14.25−0.38 WTI$82.40▲1.4% Gold$4,432.00▲1.6% EUR/USD1.157▲0.4% BTC$64,996▲0.5% Nikkei68,309▲1.2%
At close · Fri, Aug 14, 2026
Daily Market Updates.

Earnings

HomeEarningsAnalyst RatingsCerebras targets faster AI inference with wafer-scale…

Cerebras targets faster AI inference with wafer-scale chips

The company says embedding 44GB of on-chip SRAM helps eliminate data shuttling that slows token generation, aiming for up to 14x faster output than conventional setups.

AI spending is shifting from training clusters toward real-time inference, where how quickly systems generate tokens can determine whether products work for users, according to MarketBeat Ratings.

Cerebras Systems is positioned around that change with a wafer-scale architecture designed to address the “memory wall” seen in traditional chips, which rely on repeated moves of data between compute and external memory during the decode phase.

The company says its approach embeds 44GB of static random-access memory directly on the chip and uses a single-silicon-wafer design, with the goal of delivering tokens up to 14 times faster than conventional setups.

Cerebras also points to infrastructure partnerships and an inference approach that separates prefill and decode, including an AMD alliance linking Helios rack platforms to Cerebras hardware.

More like this

Sources

Get the close, explained.

One email every trading day: what moved, why it moved, and what's on deck tomorrow. Read in 3 minutes.

Free. Unsubscribe anytime.