For a year now, AI security testing company Andon Labs has tasked frontier models with various real-world tasks to determine how well they perform as agents operating for long periods without human supervision. On Wednesday, Andon published a new installment about how things are going in its Vending-Bench research, where the lab has cutting-edge models
The launch of the latest AI model from a Chinese company, Moonshot AI’s Kimi, has reignited debates about American competitiveness and open versus proprietary AI. While there has been a lot of conversation on social media, it appears the debate is also taking place behind the scenes in Washington, DC, where OpenAI and Anthropic have
US Treasury Secretary Scott Bessent doubled down on his warnings to Chinese AI companies on Wednesday, saying sanctions remain on the table after a White House official accused Moonshot of improperly distilling Anthropic’s Fable model. Model distillation is a common AI training technique in which a smaller model learns from the results of a larger
The impressive capabilities of Chinese lab Moonshot’s Kimi K3, the largest open-weight large language model, have started a debate that combines two things: the economic possibilities of American AI giants and the future of LLMs as a technology. OpenAI’s head of strategic futures, Dean W. Ball, went so far as to argue that the US
Chinese company Moonshot AI this week launched a new version of its Kimi model, generating another wave of discourse about China and open source AI. Moonshot said that although the Kimi K3 “still lags behind the most powerful proprietary models, the Claude Fable 5 and GPT 5.6 Sol,” the new open source model “demonstrated frontier-level