AMD has acquired AI chip startup Taalas to integrate model-specific integrated circuits directly into silicon. Early demonstrations indicate these custom chips can achieve inference speeds of up to 17,000 tokens per second. This move aims to significantly boost performance for specific AI workloads by hardcoding model architectures.
- AMD targets inference speed gains by integrating Taalas' custom silicon technology.
- Early demos show model-specific chips reaching 17,000 tokens per second throughput.
- Hardcoding models into hardware reduces latency compared to general-purpose accelerators.
- Acquisition signals AMD's push into specialized AI inference hardware solutions.