TechCrunch · Kate Park ·

Xcena, whose MX1 chip performs data orchestration and KV cache management directly within memory modules, raised a $135M Series B at a $570M valuation

Every time you ask ChatGPT a question, your request triggers a data relay race. Information leaves memory, passes through a CPU for preprocessing …

In this story MX1 Xcena
Xcena, whose MX1 chip performs data orchestration and KV cache management directly within memory modules, raised a $135M Series B at a $570M valuation

Lead Source

How this story grew

Coverage · 0 Discussion · 0
May 29May 30

More

SiliconANGLE: SiliconANGLE
Pulse 2.0: Pulse 2.0
Crypto Briefing: Crypto Briefing
Tech in Asia: Tech in Asia

Discussion

TechSnif Coverage

Xcena Raises $135M to Kill the AI Memory Bottleneck

Xcena's MX1 chip handles data orchestration inside memory modules, cutting out the middleman in AI inference.

Every time you fire off a prompt to ChatGPT, your data goes on a relay race — bouncing from memory to CPU and back before anything useful happens. Xcena wants to end that nonsense.

The startup just closed a $135M Series B at a $570M valuation. Its secret weapon: the MX1 chip, which performs data orchestration and KV cache management directly within memory modules instead of shuttling everything through a processor first.

KV cache is the mechanism that lets large language models remember context during a conversation. Managing it efficiently is a massive bottleneck in AI inference workloads. Xcena's approach embeds that logic right where the data already lives.

It's a clever architectural bet. If AI inference demand keeps scaling the way everyone expects, eliminating memory-to-CPU round trips could matter a lot.