Rounding up a genuinely heavy week. The throughline: last issue OpenAI paused internal work on a model it couldn’t rule out was cyber-capable. This week the capability shipped anyway, two different ways. OpenAI GPT-5.6 Cyber (Aug 10): a security-specialized model gated behind a “Daybreak Red” tier. OpenAI’s own eval has it answering 95% of offensive-security requests the standard model refuses 98.5% of the time. Access stays with 16 named partners; from Sept 1 individual accounts need hardware keys. Customers get findings, never the weights. Zhipu GLM-5.3 (Aug 14): marketed on “emergent cyber capabilities,” claims 84.5% on CyberGym (vendor-reported; note Wiz’s Atlas system claims a higher 90.9%). Open weights promised in ~2 weeks. The capability didn’t get shelved. It got a doorman. The rest of the week:
- Meta returned to open weights with Muse Glimmer, a 30B Apache-2.0 agent model that runs under 20GB.
- Alibaba published its first downloadable Max-class Qwen (2.4T), and Qwen3.8-27B landed Apache-2.0. DeepSeek took V4-Pro (1.6T, MIT) to GA with peak/off-peak pricing.
- Anthropic began embedding an invisible watermark in all Claude output under the EU AI Act. The builder forums did not take it well.
- SpaceX closed a $60B all-stock acquisition of Cursor; the editor is now inside the Grok org.
- Security: researchers showed encrypted reasoning traces from OpenAI/Anthropic/Google were replayable across sibling models to decrypt them (now patched); an AI notetaker left 181,874 meetings queryable by anyone. Full breakdown with all the receipts: thenewguard.ai/issues/027-the-brake-pedal-had-a-bypass/ submitted by /u/mattezell
Originally posted by u/mattezell on r/ArtificialInteligence
