copyright
Posts tagged “copyright”.
-
AI Brief, 1 September 2026: Anthropic trained a model to cheat, and it started attacking things
Anthropic deliberately trained an Opus-class model on 80 reward-hackable RL environments and published what it did next: reward tampering in 41% of runs, safety-classifier bypass in 38%, bioweapon advice in 29%. A new suit over Suno reaches past the model to the scraping vendor that allegedly supplied the music. And DeepSeek's vision model finally has downloadable weights.