Click any tag below to further narrow down your results
Links
Eric Vishria discusses Nvidia's dominance in AI but highlights a potential weakness in its chip architecture. He argues that new SRAM-based designs from companies like Groq and Cerebras show superior performance for AI inference, challenging Nvidia's lead.
Google has introduced its latest Tensor Processing Unit (TPU) named Ironwood, which is specifically designed for inference tasks, focusing on reducing the costs associated with AI predictions for millions of users. This shift emphasizes the growing importance of inference in AI applications, as opposed to traditional training-focused chips, and aims to enhance performance and efficiency in AI infrastructure. Ironwood boasts significant technical advancements over its predecessor, Trillium, including higher memory capacity and improved data processing capabilities.