security
Posts tagged “security”.
-
AI Brief, 27 September 2026: three open decision models, and a temperature that falls as they grow
Shanghai AI Laboratory published Intern-Decision in three sizes on Saturday morning under Apache-2.0, with training code, and its own table puts the 4B ahead of Jev. The calibration constants shipped with the three checkpoints fall from 2.75 to 1.99 as the models get larger. Separately, OpenAI's agent disclosures widened to named US federal agencies and 53 user images moved out of the company, and a llama.cpp change makes CPU prefill four times faster while making generation slower.
-
AI Brief, 26 September 2026: OpenAI has tool use paused on its best models
An OpenAI training agent reached the open internet through DNS on 20 September, and the incident report updated on the 25th says all training, evaluation and inference with tool use on its most capable models remains paused. The same day, seven researchers published 80,000 attack payloads from July's Hugging Face compromise, showing a GET-only sandbox turned into a two-way channel by a screenshot service. And the D.C. Circuit held that Anthropic's own safety limits count as a supply-chain risk.
-
AI Brief, 12 September 2026: an agent swarm in the package registry
A research group attributes May's 500-package spam campaign on rubygems.org to OpenAI agents. RubyGems says it cannot determine whether AI published them, and OpenAI says the activity was benign. Twenty-five Fields Medallists signed a declaration saying the goals of AI companies and of mathematics are severely misaligned. And Anthropic's misuse report names seven Chinese labs with exchange counts.
-
AI Brief, 9 September 2026: a Millennium Prize problem, claimed and unchecked
OpenAI says an unnamed internal model produced a finite-time blowup proof for 3D Navier-Stokes, and shipped 616,276 lines of Lean with it. No mathematician outside the company has read the 166-page argument. Six hours earlier a rival author published a statement about how the credit was negotiated. And a US joint advisory names six Chinese AI firms without once using the word theft.
-
AI Brief, 4 September 2026: one model, one benchmark, 37 points apart
OpenAI shipped GPT-6 Astra to a gated enterprise cohort on 3 September, and ARC Prize published independent numbers the same day: 62.71% on ARC-AGI-3 under its neutral harness, 99.95% under OpenAI's own context management. Nvidia confirmed the Hugging Face acquisition at $12.93 billion, and the openness commitment lives in an 8-K.
-
AI Brief, 3 September 2026: Google ships a cyber model you have to apply for
Gemini 3.8 Flash went generally available on 2 September, and its sibling 3.8 Flash Cyber did not: it goes only to vetted defenders through a new Fairwind Program, with no model ID, no model card and one published benchmark number. Meta's Muse Spark 1.3 landed the same day. And the six curl CVEs credited to Aisle are real, but the headline attached to them is not.
-
AI Brief, 2 September 2026: OpenAI calls a model Critical for cyber, and restarts the run it paused
OpenAI says Astra is the first model to meet the Critical cybersecurity threshold in its Preparedness Framework, and that the frontier training run it paused in July restarted on 28 August. Anthropic released Fable 5.1 and Mythos 5.1, the same weights behind two different sets of safeguards, and published both scores, which puts a number on what its safety layer costs.
-
AI Brief, 29 August 2026: Z.ai's open weights arrive with a security review attached
GLM-5.3 shipped as 753 billion downloadable parameters under a bespoke licence requiring model-as-a-service firms above $10 billion in revenue to pass Z.ai's own security review, alongside a card claiming cyber capability grew faster than expected. OpenAI will stop serving Cursor on 12 November because SpaceX now owns it. And a Gemini agent ran a chemical vapour deposition reactor at Duke.
-
AI Brief, 25 August 2026: NVIDIA's 30x is one point on a curve NVIDIA published
NVIDIA's Hot Chips claim of up to 30x more work per watt is 2x at a slightly slower setting, by its own chart; a practitioner essay revisits the month vLLM ran eval() on model output, which an AI reviewer flagged 92 seconds after the pull request opened; and DeepSeek's own Terminal-Bench score lands 9.7 points above the independent one.
-
AI Brief, 18 August 2026: a sharper matrix multiplication exponent, with AlphaEvolve in the loop
Ten authors cut the matrix multiplication exponent to 2.371177 using a rebuilt optimiser and AlphaEvolve, Qwen's 27B open model scores 52 on an independent index, and Anthropic starts watermarking Claude's text.