
ZML, a French AI startup, releases ZML/LLMD, an inference-performance software that enables open-source large language models to run on multiple chip architectures, aiming to break silos and improve AI efficiency. The startup collaborates with European chipmakers to co-design silicon, offering flexibility and cost savings for enterprises. ZML focuses on creating a versatile inference platform that can leverage different chip architectures, emphasizing the growing importance of inference in AI deployment and the "inference gold rush."
Microsoft is simplifying Copilot by combining its consumer and business apps, and dropping AI-generated podcasts, Group Chats, Deep Research, and its Mico character.

Nvidia has a plan to make sure its GPUs won't lose value. It wants to convince a new crop of financiers to keep lending for AI buildouts.

The tech giant has considered a nine-figure budget for the payments, according to the WSJ.
