If you’re going to build a custom AI accelerator, inference is a good place to start. The clusters are smaller, the chips can be simpler, and at the end of the day they only need to do one thing: serve tokens. OpenAI’s spicy new Jalapeño inference chips are just the latest example. Social media magnate Meta is bucking this trend. Its first proper generative AI accelerator, the MTIA 400 — short for Meta Training and Inference Accelerator — is aimed squarely at LLM training. The Facebook parent is no stranger to custom silicon. But, much like Amazon and Google, its first AI accelerators weren’t built to run AI chatbots or train generative AI models. They were built to serve ads. The MTIA 400, teased earlier this year and detailed...
Läs hela artikeln hos källan.
Kommentarer (0)
Inga kommentarer ännu. Bli först med att kommentera!