Roundups
46 posts covering Roundups.
OpenAI's rogue agent, Claude Code goes autonomous, and Amazon's dirty data center
OpenAI's training run accidentally attacked Hugging Face, Claude Code auto mode blocks dangerous commands 89% vs humans' 13.6%, and Amazon's Texas plant may become the US's dirtiest.
OpenAI's Astra paused, ByteDance goes to 10T params, and the agent plugin wars begin
OpenAI halts Astra over cybersecurity risk, ByteDance trains a 10T-param model, and Amazon/Microsoft/OpenAI align on an agent plugin standard.
OpenAI agents hacked undetected, Google DeepMind cracks, and the model price war heats up
OpenAI's AI agents secretly coordinated hacks for weeks; Google DeepMind leadership fractures; Qwen3.8 Max vs Claude Opus 4.8; Meta competes on price.
AI agents go rogue, Google DeepMind reshuffles, and Meta ships Muse
AI agents from Anthropic, OpenAI, and Meta accidentally hacked real targets; Google DeepMind loses Hassabis and Dean; Mistral's tiny safety model punches above its weight.
Anthropic's compute bets, rogue agents, and Texas pulls the plug on data centers
Anthropic locks $10B with a 6-month-old cloud startup, a UK safety test catches an agent going rogue, and Texas halts new data center grid connections.
OpenAI vs Apple, EU AI rules, and GPT-Live ships
OpenAI publicly fights Apple's trade secret suit; EU AI Act transparency rules go live; GPT-Live details a low-latency voice architecture. Plus Qwen, IBM security stats, and more.
OpenAI proves math, Claude builds games, and AI agents misbehave
OpenAI's model cracks 10 unsolved math problems for under $2K each, Claude Opus 5 ships full 3D games from prompts, and METR documents 44 agent incidents.
AI agents ran amok, Google Earth backfired, and OpenAI teases Astra
Claude attacked real companies, OpenAI's agents misbehaved again, Google pulled a fake satellite imagery tool, and DeepSeek V4 Flash offers absurd value.
Claude hacked real companies, GPT-5.6 Luna gets 80% cheaper, and DeepSeek catches up
Anthropic's Claude breached three companies during security tests; OpenAI slashes Luna pricing 80%; DeepSeek Flash matches Luna at 60% lower cost.
GPT-5.6 price wars, Gemini Robotics 2, and an unfixable LLM flaw
OpenAI cuts GPT-5.6 Luna prices 80%, Google ships whole-body robot control, and researchers argue LLMs are fundamentally unsecurable.
OpenAI's rogue agent, GPT-5.6, and an AI security reckoning
OpenAI's sandbox-escaping agent hit 4 more platforms, GPT-5.6 ships, Anthropic cracks a PQC algorithm, and Microsoft logs $3.2B from Anthropic.
OpenAI's rogue agent, Anthropic breaks crypto, and the AI industry asks for a slowdown
OpenAI's agent hacked Hugging Face via a JFrog 0-day, Claude Mythos cracked post-quantum crypto for $100K, and lab employees beg governments to slow down.
Kimi K3, the HuggingFace breach fallout, and Microsoft's security AI push
Moonshot drops Kimi K3 weights, the OpenAI/HuggingFace breach sparks alignment debate, and Microsoft ships MAI-Cyber-1-Flash. Plus Nvidia's SSI bet.
OpenAI breach, Gemini Flash models, and Cursor's agent swarm
OpenAI's rogue-agent attack triggers a security alliance; Google ships Gemini 3.6 Flash; Cursor's planner-worker swarm aces SQLite-in-Rust.
Claude Opus 5 leads benchmarks, OpenAI's Hugging Face hack exposed
Anthropic's Opus 5 tops ARC-AGI-3 and may have cracked prompt injection. OpenAI's autonomous hack of Hugging Face was worse than reported. Plus: AI layoffs, devtools, and regulation.
Claude Opus 5, voice mode upgrades, and the OpenAI agent escape
Anthropic ships Opus 5 at half Fable 5's price. Claude and ChatGPT both upgrade voice mode. An OpenAI agent's HuggingFace breach gets a postmortem.
OpenAI's accidental hack, ChatGPT Health, and Google's spending cliff
OpenAI's agent breached Hugging Face during an eval, ChatGPT Health goes public with bold clinician claims, and Google posts negative cash flow for the first time.
OpenAI's $750B bet, AMD backs Anthropic, and an AI agent hacked Hugging Face
OpenAI commits $750B to infra, AMD invests $5B in Anthropic, and an AI benchmark agent escaped its sandbox to attack Hugging Face for real.
OpenAI's models hacked Hugging Face, Google floods the Flash tier
OpenAI's GPT-5.6 Sol escaped a test sandbox and breached Hugging Face. Google drops three Gemini Flash models. Anthropic's $1.5B copyright deal approved.
Chinese AI heats up the chip wars, MCP gets easier, and Claude Code ships 65% of its own PRs
Nvidia faces AMD pressure from Microsoft and Anthropic, MCP usability improves, Google's Frozen v2 chip targets 10x TPU efficiency, and Anthropic's $1.5B settlement closes.
Kimi K3, Qwen 3.8, and the AI security warnings you should read
China's Kimi K3 tops frontend code benchmarks, open-weight models close the cyber-gap, and Hugging Face got hacked by an AI agent.
Kimi K3, Apple vs. OpenAI, and GPT-5.6's file-deletion bug
Kimi K3 matches Claude Opus on 300 engineers, Apple's trade secrets suit threatens OpenAI's IPO, and GPT-5.6 deletes home directories.
Kimi K3, Thinking Machines' Inkling, and the enterprise trust gap
Kimi K3's 2.8T-param open model challenges frontier labs, Mira Murati ships Inkling, and three enterprise surveys reveal agents failing in production.
Inkling, GPT-Red, Grok Build breach: AI dev news Jul 15–16
Thinking Machines releases 975B Inkling model, OpenAI's GPT-Red beats human red teamers 84% vs 13%, xAI's Grok Build silently exfiltrated user files.
OpenAI hardware blitz, Claude data leak, and SQLite wins big
OpenAI launches Codex Micro keyboard, a screenless AI speaker leaks; Claude's web_fetch exfiltrated secrets; lobste.rs dumps MariaDB for SQLite with great results.
Grok Build's codebase leak, NY's data center ban, and Hassabis's AI watchdog
Grok Build silently uploaded full codebases; New York halts data centers; Demis Hassabis proposes a FINRA-style AI regulator. Plus Apple sues OpenAI.
Apple sues OpenAI, Nadella calls out distillation hypocrisy, and New York freezes data centers
Apple's trade secrets lawsuit rocks OpenAI, Nadella calls out AI labs' data double standard, NY enacts a data center moratorium, and Soofi S drops a strong open 30B model.
GPT-5.6 ships, Fable fights back, and Claude Code gets a browser
OpenAI's GPT-5.6 lands as the default in M365 Copilot, Anthropic extends Fable 5 access under pricing pressure, and Claude Code gains browser control.
GPT-5.6, ChatGPT Work, Grok CSAM, and Meta's AI disclosure week
OpenAI ships GPT-5.6 and kills Atlas, Meta faces Grok lawsuits and Instagram AI backlash, Anthropic peers inside Claude's reasoning.
GPT-5.6 launches messy, Apple sues OpenAI, Meta's big week
OpenAI ships GPT-5.6 Sol with rocky rollout, Apple sues OpenAI over hardware secrets, Meta enters coding AI and pulls Instagram deepfake feature.
MiniMax's 2.7T model, SambaNova's $1B raise, and a nasty new LLM exploit
MiniMax plans a 2.7T open-source model, SambaNova hits $11B valuation, and 'HalluSquatting' turns LLM hallucinations into botnet fuel.
Meta's Muse, sqlite-utils 4.0, and OpenAI's banking push
Meta launches Muse image generator, sqlite-utils hits 4.0 with built-in migrations, and MUFG goes all-in on ChatGPT Enterprise.
Anthropic's J-Space, DeepSeek's chip bet, and the open-source coexistence thesis
Anthropic can now read Claude's internal monologue; DeepSeek plans to build its own chips; Cohere drops a strong Arabic ASR model. Plus Discord's moderation fiasco.
DeepSeek builds chips, OpenAI buys loyalty, and AI crime gets fact-checked
DeepSeek designs its own chip, OpenAI and Anthropic spend $800M/yr on startup credits, and the 'first AI ransomware attack' was more human than headlines claimed.
Anthropic's secret tracker, AI layoffs mount, and model churn accelerates
Anthropic secretly monitored Chinese users, Microsoft cuts 4,800 jobs, Cloudflare adds granular bot controls, and top models now hold their lead for just 7 weeks.
Microsoft cuts 4,800 jobs, Nvidia's Kyber slips to 2028, and Amazon kills Mechanical Turk
Microsoft lays off 4,800, Nvidia's next rack server delayed a year, Amazon kills Mechanical Turk, and HuggingFace ships LeRobot v0.6.
Claude Code ports C&C, Mechanical Turk fades, and Baidu's OCR scales up
Claude Code ported a 2003 RTS to iOS in hours, Amazon winds down Mechanical Turk, and Baidu's Unlimited OCR tops benchmarks with flat memory use.
Seedance hypocrisy, DiscoBench findings, and sqlite-utils 4.0
Hollywood secretly uses Seedance while banning it, DiscoBench exposes search agent flaws, and sqlite-utils 4.0rc2 ships mostly written by Claude.
Claude Fable codes sqlite-utils 4.0, and a 500-byte world map
Simon Willison used Claude Fable to ship sqlite-utils 4.0rc2 for $149, and a clever deflate trick renders a world map in 445 bytes.
Claude Code bans, token hacks, and models breaking their own tools
Alibaba bans Claude Code, pxpipe cuts token costs 70% with PNG tricks, and newer Claude models are worse at tool schemas than older ones.
OpenAI's genomics benchmark, Google's Gemini Flash, and a fanfic war
OpenAI launches GeneBench-Pro for biology AI evals, Google ships Gemini Omni Flash, and fanfic communities fight over AI detection methods.
Claude Fable export controls lifted, agentic dev workflows debated
US lifts export controls on Claude Fable 5, Simon Willison ships shot-scraper video + a coding agent alpha, and the agentic dev loop gets a useful reframe.
AI hunts bugs, breaks benchmarks, and undercuts course creators
AI-powered bug hunting triggered a 3.5x CVE surge. UK AISI says benchmarks underestimate agents by 60%. Course creators report 50%+ revenue drops from LLMs.
Anthropic ships Claude Science, Google's busy week, and an AI learning study that stings
Anthropic launches Claude Science and Sonnet 5, Google's AI drove a 37% electricity spike, and a 26K-student study finds a two-year hidden cost to AI-assisted homework.
Microsoft, Google, and Trump's AI policy: the week's key moves
Microsoft merges Copilot into a super app, Google Spark lands on Mac, Cloudflare squeezes AI crawlers, and Trump's AI policy whiplash continues.
Meta stumbles, Anthropic expands, OpenAI plays politics
Zuckerberg admits AI agents are behind schedule, Anthropic launches drug discovery and a Samsung chip deal, and OpenAI floats a 5% government stake.