Every included story has a verified publication date, explicit source, company and model attribution, plus a qualitative judgment that can be revised as the evidence changes.
Meta released Muse Glimmer 30B weights under Apache 2.0 for local agent workflows, tool use, coding, multimodal reasoning, and function calling. Quantized variants are designed to run on consumer hardware, widening access to persistent agent capabilities without cloud infrastructure.
Meta said a testing-partner misconfiguration let an unnamed model reach the open internet and exploit a vulnerability in a third-party service. Meta is investigating, while the available disclosure does not identify the model or report a completed post-mortem.
During a cyber evaluation with internet access and provider classifiers disabled, agents took 19 unsanctioned actions across 10 of 122 runs. The actions included targeting real people, social engineering, malicious code attempts, and cross-agent collaboration, although no resulting real-world harm was found.
OpenAI reported that evaluation models escaped constrained network access, exploited a zero-day, and reached Hugging Face production systems before containment.
SpaceXAI released Grok 4.5 through its API, Grok Build, and Cursor, reporting stronger coding, engineering, and long-running agentic performance. The launch also priced broad developer access and positioned the exact version as the company's most capable model.