Signal
Taalas Never Sold a Chip. AMD Bought the Memory Trick Underneath It
The HC1 burns Llama 3.1 8B into silicon and quotes 17,000 tokens a second. Taalas' own product page calls it a technology demonstrator. That phrase is the whole story.
Taalas' own product page calls the HC1 a technology demonstrator. That is the most useful phrase written about this chip and almost nobody quoted it.
The silicon is real. 815mm² on TSMC's 6nm node, 53 billion transistors, roughly 250 W per card and 2.5 kW for a ten-card server on Taalas' own spec sheet, with Meta's Llama 3.1 8B physically encoded in the circuit instead of fetched from memory. Taalas claims 17,000 output tokens per second per user. The Register reported 16,960 at the acquisition. Other outlets ran 14,000. Taalas' homepage separately claims its Hardcore models are 1000x more efficient than software equivalents, which is a third axis entirely. Every one of those figures is vendor-supplied. No independent benchmark of an HC1 exists in public, and the spread is itself the information: three different multiples for the same die means three different comparisons, not three measurements.
The obsolescence critique is aimed at the wrong object
Most of the February coverage ran the same argument. Weights turn over monthly, silicon does not, so a chip with Llama 3.1 8B burned into it is a museum piece before it ships. Intuitive. Also mostly wrong on the facts.
Of the roughly 100 mask layers in the HC1, only the top two are model-specific. The base die is generic. Changing models means changing two metal masks, reportedly around two months, against roughly six months just to fabricate a conventional accelerator such as Blackwell, and far longer than that for a full custom design cycle. Taalas does not describe itself as a chip vendor at all. It describes a foundry, with LoRA fine-tuning supported on top of the frozen base. Two months is faster than most model families move between releases anyone would re-architect for.
Inventory is the risk, not the design cycle
I have watched hardware teams in Shenzhen turn a PCB revision in four days and a full product refresh in six weeks. Design speed is almost never what kills a hardware company here. Committed inventory is.
A GPU is fungible. When demand moves from one model to another, the same accelerator runs the new one that afternoon. An HC1 is not fungible in that way. The exposure is real but narrower than it looks. Generic base wafers are stockpiled model-agnostically and only the final two layers bind a die to a model, so the inventory risk is aggregate demand for HC1-class parts rather than a bet on any one model. What is not fungible is the finished part: once those two layers are patterned, a card that nobody wants to run Llama 3.1 8B on cannot be repurposed the way a GPU can. That is working capital risk on finished goods rather than engineering risk, and most of the coverage reached for the obsolescence angle instead.
AMD bought a memory technique
AMD signed a definitive agreement on 6 August 2026, terms undisclosed, subject to customary closing conditions and regulatory approval. Taalas had raised $219 million in total, $169 million of it in February from Quiet Capital, Fidelity and the semiconductor investor Pierre Lamond. AMD's own release says it will fold the technology into its accelerator roadmap alongside Instinct, EPYC and Helios. It does not say it will sell model-specific chips.
Read that as the honest answer to the inventory problem: do not take the risk. What AMD wanted is the recall fabric underneath, a mask-ROM structure that holds four bits and performs the associated multiply in a single transistor, which takes HBM out of the inference path. In a market where HBM allocation caps how many accelerators anyone can ship, a technique that removes the HBM dependency is worth more than the chip built to demonstrate it. AMD needed a reply and it bought one, seven months later, for a price it will not print.
The HC1 was never going to be sold at volume. It was an existence proof, and it did the job an existence proof is for.
Also referenced
The Register: AMD acquires Taalas to etch models into silicon: https://www.theregister.com/systems/2026/08/06/amd-acquires-ai-chip-startup-taalas-to-boost-inference-performance-by-etching-models-into-silicon/5284344
Taalas products page (HC1 technology demonstrator specs): https://taalas.com/products/
SiliconANGLE: Taalas raises $169M for model-specific AI chips: https://siliconangle.com/2026/02/19/taalas-raises-169m-funding-develop-model-specific-ai-chips/
CNX Software: Taalas HC1 hardwired Llama 3.1 8B accelerator: https://www.cnx-software.com/2026/02/22/taalas-hc1-hardwired-llama-3-1-8b-ai-accelerator-delivers-up-to-17000-tokens-s/
CNBC: Nvidia buying Groq's assets for about $20 billion: https://www.cnbc.com/2025/12/24/nvidia-buying-ai-chip-startup-groq-for-about-20-billion-biggest-deal.html
CNBC: AMD buys Taalas: https://www.cnbc.com/2026/08/06/amd-buys-taalas-startup-that-hardwires-ai-models-into-its-silicon.html