Toronto-based Taalas develops specialized silicon designed to reduce computing and memory bottlenecks in AI inference, the process of running trained AI models to generate responses or predictions.

Financial terms or expected timeline for completion were not disclosed

Early tech demos show model-specific integrated circuits churning out up to 17,000 tokens a second