AMD
AMD announced it has acquired AI chip startup Taalas, which develops model-specific integrated circuits that etch model weights directly into silicon to boost inference performance.
Why it matters for sellers
M&A integration = tooling and consolidation needs
Signal details
- Counterparty
- Taalas
- Reported
- August 6, 2026
- Source
- theregister.com
From the coverage · theregister.com
AI and ML Early tech demos show model-specific integrated circuits churning out up to 17,000 tokens a second In AMD’s latest bid to upset Nvidia's dominance in AI hardware, the House of Zen has acquired AI chip company Taalas, which bakes model weights directly into silicon in a process that promises to boost inference performance by an order of magnitude or more. The deal, announced at market close on Thursday, appears to be framed in much the same context as Nvidia’s $20 billion licensing deal with Groq last December: make high-performance “premium” inference services prized for AI agents, like code assistants, faster and cheaper to run.
AMD didn’t disclose the terms of the deal, but from what we understand, this is an actual acquisition rather than an acquihire. Founded in 2023 and based in Toronto, Taalas’ approach to inference is radically different from conventional GPUs or the dataflow architectures that underpin Groq LPUs or Cerebras' waferscale accelerators. The startup’s chips don’t rely on HBM to store the model weights but rather etch them directly into the silicon. In a sense, Taalas’ chips are really model-specific integrated circuits or MSICs. Perhaps more importantly, Taalas’ tech isn’t just conceptual.
In February, the startup revealed its first test chip fabbed on TSMC’s 6nm process tech, which it called the HC1. 5x faster than Cerebras' accelerators. 1 is ancient by today’s standards, having made its debut all the way back in mid 2024, the reticle-sized chip was really intended to prove the concept. Taalas has been incredibly secretive about how its chips actually work, but we know its processors are comprised of two main regions: the mask-ROM recall fabric where model weights are etched, and the SRAM recall fabric where KV caches and fine-tuning adapters are stored.
For its second-gen HC2 chip due out this summer, Taalas aims to boost parameter count to 20 billion parameters. That might not sound like much, but just like with GPUs for larger models, weights are simply distributed across multiple accelerators using pipeline parallelism. At 20 billion parameters per chip, you’d need just 50 accelerators to support a trillion-parameter model, and AMD just so happens to have a rack-scale compute platform and in-house system design team that can comfortably accommodate that. That’s quite a bit more space and power efficient than Nvidia’s recently unveiled LPX systems, which would need a few dozen GPUs and at least 2,000 Groq LPUs to serve the same model.
From what we understand, AMD intends to pair its Instinct-based Helios racks with chips based on Taalas’ tech, which implies a disaggregated architecture where compute-heavy prompt processing is done on GPUs while token generation is offloaded to Taalas-based accelerators. It’s also possible that AMD could adopt a sort of tick-tock cadence in which customers initially deploy and validate models on Instinct accelerators and, once they’re satisfied with them, transition to Taalas accelerators. We can only speculate at this point, but here’s what AMD’s SVP of AI, Vamsi Boppana, had to say about it in a canned statement: “AMD is building a full-stack AI platform that gives customers the flexibility to deploy the right compute solutions for every AI workload."
While the tech is blazing fast, if you hadn’t already figured it out, it comes with a pretty substantial downside. Once the chips are deployed you’re stuck with that model. Any change bigger than something like a LoRA adapter is going to require a re-spin of the chips, which is not only expensive but time-consuming. Nearly four years into the AI boom, new models are rolling out on a nearly monthly basis. In order to benefit from Taalas’ tech, AMD’s customers are going to have to be really sure about their choice of models, which will be easier for some than others.
However, if the startup is to be believed, the situation isn’t quite as bad as it sounds. While new models will require a re-spin, it doesn’t require starting over from scratch.
Get signals like this for every account you sell to
Our engine detects, verifies, and deduplicates thousands of buying signals every day across 50M+ companies — delivered via API, GCS push, or flat file.
More from the last 24 hours
Microsoft
New LocationMicrosoft has opened its fourth cloud region in India, located in Hyderabad, as part of a $20.5 billion investment in the country.
Airbnb
EarningsAirbnb reported a 17% increase in second-quarter revenue, reaching $3.6 billion, surpassing analysts' expectations.
Lyft
EarningsLyft's second-quarter revenue increased by 16% to $1.84 billion, surpassing analysts' expectations of $1.81 billion.
Adobe
Product LaunchAdobe has launched a comprehensive ChatGPT plugin that integrates all 70 of its creative and productivity tools, including Photoshop, Premiere, and Acrobat.