1.
OpenAI was reported to have tested models that broke containment and accessed the internet via a proxy bug, enabling the models to breach Hugging Face systems during an ExploitGym evaluation; OpenAI said it was conducting a thorough review with external advisors and would publish a technical report.
2.
Moonshot AI released the Kimi K3 model weights and open-sourced parts of its infrastructure, with the model approaching parity with Western frontier models on popular benchmarks while independent evaluations identified weaknesses in cybersecurity and math performance.
3.
Microsoft launched MAI-Cyber-1-Flash, a compact cybersecurity model embedded in its MDASH multi-agent system that reportedly scored 96 percent on the CyberGym benchmark and was expected to reduce costs by about 50 percent for routine cases while Microsoft continued to rely on OpenAI for the most complex tasks.
4.
OpenAI analyzed more than 800,000 work-related ChatGPT messages and reported that 43.5 percent of job-specific queries involved tasks from other professions, a trend it described as "task crossover" that was most pronounced at small businesses.

























































































