AMD announced the deal on Thursday, after the market closed, and did not say what it paid. Taalas, founded in 2023, builds what it calls model-specific chips. Instead of loading weights from memory, it bakes them straight into the silicon. AMD says the technology will join its accelerator roadmap, working alongside its Instinct GPUs.

“I’m a big believer that there’s no one-size-fits-all as it comes to chips,” AMD chief executive Lisa Su said at a July event. That line is the whole logic of the deal. The industry spent four years buying general-purpose GPUs. AMD is betting the next phase rewards something far narrower.

What Taalas actually built

A normal AI chip keeps memory and compute apart, then burns huge effort shuttling data between them. That gap is why modern systems need stacked memory, exotic packaging, and liquid cooling. Taalas merges the two. Its co-founder, Ljubisa Bajic, says the result needs none of it: no high-bandwidth memory, no 3D stacking, no liquid cooling.

The speed claims are startling. Taalas’s first chip runs Meta’s Llama 3.1 model at about 17,000 tokens a second per user, which it says is many times faster than a leading GPU, at a fraction of the cost and power. Those are the company’s own numbers. The first version also leans on aggressive compression that dents output quality.