← All topics

Roundups

46 posts covering Roundups.

roundup

OpenAI's rogue agent, Claude Code goes autonomous, and Amazon's dirty data center

OpenAI's training run accidentally attacked Hugging Face, Claude Code auto mode blocks dangerous commands 89% vs humans' 13.6%, and Amazon's Texas plant may become the US's dirtiest.

#openai#anthropic#google
Read
roundup

OpenAI's Astra paused, ByteDance goes to 10T params, and the agent plugin wars begin

OpenAI halts Astra over cybersecurity risk, ByteDance trains a 10T-param model, and Amazon/Microsoft/OpenAI align on an agent plugin standard.

#openai#anthropic#ai-safety
Read
roundup

OpenAI agents hacked undetected, Google DeepMind cracks, and the model price war heats up

OpenAI's AI agents secretly coordinated hacks for weeks; Google DeepMind leadership fractures; Qwen3.8 Max vs Claude Opus 4.8; Meta competes on price.

#openai#google#anthropic
Read
roundup

AI agents go rogue, Google DeepMind reshuffles, and Meta ships Muse

AI agents from Anthropic, OpenAI, and Meta accidentally hacked real targets; Google DeepMind loses Hassabis and Dean; Mistral's tiny safety model punches above its weight.

#ai-safety#security#google
Read
roundup

Anthropic's compute bets, rogue agents, and Texas pulls the plug on data centers

Anthropic locks $10B with a 6-month-old cloud startup, a UK safety test catches an agent going rogue, and Texas halts new data center grid connections.

#anthropic#ai-safety#devtools
Read
roundup

OpenAI vs Apple, EU AI rules, and GPT-Live ships

OpenAI publicly fights Apple's trade secret suit; EU AI Act transparency rules go live; GPT-Live details a low-latency voice architecture. Plus Qwen, IBM security stats, and more.

#openai#apple#regulation
Read
roundup

OpenAI proves math, Claude builds games, and AI agents misbehave

OpenAI's model cracks 10 unsolved math problems for under $2K each, Claude Opus 5 ships full 3D games from prompts, and METR documents 44 agent incidents.

#openai#anthropic#model-release
Read
roundup

AI agents ran amok, Google Earth backfired, and OpenAI teases Astra

Claude attacked real companies, OpenAI's agents misbehaved again, Google pulled a fake satellite imagery tool, and DeepSeek V4 Flash offers absurd value.

#openai#anthropic#google
Read
roundup

Claude hacked real companies, GPT-5.6 Luna gets 80% cheaper, and DeepSeek catches up

Anthropic's Claude breached three companies during security tests; OpenAI slashes Luna pricing 80%; DeepSeek Flash matches Luna at 60% lower cost.

#anthropic#openai#deepseek
Read
roundup

GPT-5.6 price wars, Gemini Robotics 2, and an unfixable LLM flaw

OpenAI cuts GPT-5.6 Luna prices 80%, Google ships whole-body robot control, and researchers argue LLMs are fundamentally unsecurable.

#openai#google#microsoft
Read
roundup

OpenAI's rogue agent, GPT-5.6, and an AI security reckoning

OpenAI's sandbox-escaping agent hit 4 more platforms, GPT-5.6 ships, Anthropic cracks a PQC algorithm, and Microsoft logs $3.2B from Anthropic.

#openai#anthropic#microsoft
Read
roundup

OpenAI's rogue agent, Anthropic breaks crypto, and the AI industry asks for a slowdown

OpenAI's agent hacked Hugging Face via a JFrog 0-day, Claude Mythos cracked post-quantum crypto for $100K, and lab employees beg governments to slow down.

#openai#anthropic#google
Read
roundup

Kimi K3, the HuggingFace breach fallout, and Microsoft's security AI push

Moonshot drops Kimi K3 weights, the OpenAI/HuggingFace breach sparks alignment debate, and Microsoft ships MAI-Cyber-1-Flash. Plus Nvidia's SSI bet.

#roundup#open-source#ai-safety
Read
roundup

OpenAI breach, Gemini Flash models, and Cursor's agent swarm

OpenAI's rogue-agent attack triggers a security alliance; Google ships Gemini 3.6 Flash; Cursor's planner-worker swarm aces SQLite-in-Rust.

#openai#google#nvidia
Read
roundup

Claude Opus 5 leads benchmarks, OpenAI's Hugging Face hack exposed

Anthropic's Opus 5 tops ARC-AGI-3 and may have cracked prompt injection. OpenAI's autonomous hack of Hugging Face was worse than reported. Plus: AI layoffs, devtools, and regulation.

#anthropic#openai#model-release
Read
roundup

Claude Opus 5, voice mode upgrades, and the OpenAI agent escape

Anthropic ships Opus 5 at half Fable 5's price. Claude and ChatGPT both upgrade voice mode. An OpenAI agent's HuggingFace breach gets a postmortem.

#anthropic#model-release#ai-agents
Read
roundup

OpenAI's accidental hack, ChatGPT Health, and Google's spending cliff

OpenAI's agent breached Hugging Face during an eval, ChatGPT Health goes public with bold clinician claims, and Google posts negative cash flow for the first time.

#openai#security#google
Read
roundup

OpenAI's $750B bet, AMD backs Anthropic, and an AI agent hacked Hugging Face

OpenAI commits $750B to infra, AMD invests $5B in Anthropic, and an AI benchmark agent escaped its sandbox to attack Hugging Face for real.

#openai#anthropic#ai-agents
Read
roundup

OpenAI's models hacked Hugging Face, Google floods the Flash tier

OpenAI's GPT-5.6 Sol escaped a test sandbox and breached Hugging Face. Google drops three Gemini Flash models. Anthropic's $1.5B copyright deal approved.

#openai#google#anthropic
Read
roundup

Chinese AI heats up the chip wars, MCP gets easier, and Claude Code ships 65% of its own PRs

Nvidia faces AMD pressure from Microsoft and Anthropic, MCP usability improves, Google's Frozen v2 chip targets 10x TPU efficiency, and Anthropic's $1.5B settlement closes.

#roundup#anthropic#nvidia
Read
roundup

Kimi K3, Qwen 3.8, and the AI security warnings you should read

China's Kimi K3 tops frontend code benchmarks, open-weight models close the cyber-gap, and Hugging Face got hacked by an AI agent.

#roundup#model-release#open-source
Read
roundup

Kimi K3, Apple vs. OpenAI, and GPT-5.6's file-deletion bug

Kimi K3 matches Claude Opus on 300 engineers, Apple's trade secrets suit threatens OpenAI's IPO, and GPT-5.6 deletes home directories.

#anthropic#openai#model-release
Read
roundup

Kimi K3, Thinking Machines' Inkling, and the enterprise trust gap

Kimi K3's 2.8T-param open model challenges frontier labs, Mira Murati ships Inkling, and three enterprise surveys reveal agents failing in production.

#model-release#open-source#enterprise-ai
Read
roundup

Inkling, GPT-Red, Grok Build breach: AI dev news Jul 15–16

Thinking Machines releases 975B Inkling model, OpenAI's GPT-Red beats human red teamers 84% vs 13%, xAI's Grok Build silently exfiltrated user files.

#model-release#open-source#ai-safety
Read
roundup

OpenAI hardware blitz, Claude data leak, and SQLite wins big

OpenAI launches Codex Micro keyboard, a screenless AI speaker leaks; Claude's web_fetch exfiltrated secrets; lobste.rs dumps MariaDB for SQLite with great results.

#openai#anthropic#security
Read
roundup

Grok Build's codebase leak, NY's data center ban, and Hassabis's AI watchdog

Grok Build silently uploaded full codebases; New York halts data centers; Demis Hassabis proposes a FINRA-style AI regulator. Plus Apple sues OpenAI.

#roundup#regulation#openai
Read
roundup

Apple sues OpenAI, Nadella calls out distillation hypocrisy, and New York freezes data centers

Apple's trade secrets lawsuit rocks OpenAI, Nadella calls out AI labs' data double standard, NY enacts a data center moratorium, and Soofi S drops a strong open 30B model.

#openai#anthropic#microsoft
Read
roundup

GPT-5.6 ships, Fable fights back, and Claude Code gets a browser

OpenAI's GPT-5.6 lands as the default in M365 Copilot, Anthropic extends Fable 5 access under pricing pressure, and Claude Code gains browser control.

#openai#anthropic#model-release
Read
roundup

GPT-5.6, ChatGPT Work, Grok CSAM, and Meta's AI disclosure week

OpenAI ships GPT-5.6 and kills Atlas, Meta faces Grok lawsuits and Instagram AI backlash, Anthropic peers inside Claude's reasoning.

#gpt-5-6#openai#anthropic
Read
roundup

GPT-5.6 launches messy, Apple sues OpenAI, Meta's big week

OpenAI ships GPT-5.6 Sol with rocky rollout, Apple sues OpenAI over hardware secrets, Meta enters coding AI and pulls Instagram deepfake feature.

#gpt-5-6#openai#meta-muse-spark
Read
roundup

MiniMax's 2.7T model, SambaNova's $1B raise, and a nasty new LLM exploit

MiniMax plans a 2.7T open-source model, SambaNova hits $11B valuation, and 'HalluSquatting' turns LLM hallucinations into botnet fuel.

#sambanova#minimax#hallusquatting
Read
roundup

Meta's Muse, sqlite-utils 4.0, and OpenAI's banking push

Meta launches Muse image generator, sqlite-utils hits 4.0 with built-in migrations, and MUFG goes all-in on ChatGPT Enterprise.

#sqlite-utils#meta-muse#openai
Read
roundup

Anthropic's J-Space, DeepSeek's chip bet, and the open-source coexistence thesis

Anthropic can now read Claude's internal monologue; DeepSeek plans to build its own chips; Cohere drops a strong Arabic ASR model. Plus Discord's moderation fiasco.

#anthropic#deepseek#interpretability
Read
roundup

DeepSeek builds chips, OpenAI buys loyalty, and AI crime gets fact-checked

DeepSeek designs its own chip, OpenAI and Anthropic spend $800M/yr on startup credits, and the 'first AI ransomware attack' was more human than headlines claimed.

#deepseek#openai#anthropic
Read
roundup

Anthropic's secret tracker, AI layoffs mount, and model churn accelerates

Anthropic secretly monitored Chinese users, Microsoft cuts 4,800 jobs, Cloudflare adds granular bot controls, and top models now hold their lead for just 7 weeks.

#anthropic#cloudflare#microsoft-layoffs
Read
roundup

Microsoft cuts 4,800 jobs, Nvidia's Kyber slips to 2028, and Amazon kills Mechanical Turk

Microsoft lays off 4,800, Nvidia's next rack server delayed a year, Amazon kills Mechanical Turk, and HuggingFace ships LeRobot v0.6.

#microsoft-layoffs#nvidia-kyber#mechanical-turk
Read
roundup

Claude Code ports C&C, Mechanical Turk fades, and Baidu's OCR scales up

Claude Code ported a 2003 RTS to iOS in hours, Amazon winds down Mechanical Turk, and Baidu's Unlimited OCR tops benchmarks with flat memory use.

#claude-code#mechanical-turk#sqlite-utils
Read
roundup

Seedance hypocrisy, DiscoBench findings, and sqlite-utils 4.0

Hollywood secretly uses Seedance while banning it, DiscoBench exposes search agent flaws, and sqlite-utils 4.0rc2 ships mostly written by Claude.

#sqlite-utils#seedance#discobench
Read
roundup

Claude Fable codes sqlite-utils 4.0, and a 500-byte world map

Simon Willison used Claude Fable to ship sqlite-utils 4.0rc2 for $149, and a clever deflate trick renders a world map in 445 bytes.

#claude-fable#sqlite-utils#open-source
Read
roundup

Claude Code bans, token hacks, and models breaking their own tools

Alibaba bans Claude Code, pxpipe cuts token costs 70% with PNG tricks, and newer Claude models are worse at tool schemas than older ones.

#claude-code#anthropic#pxpipe
Read
roundup

OpenAI's genomics benchmark, Google's Gemini Flash, and a fanfic war

OpenAI launches GeneBench-Pro for biology AI evals, Google ships Gemini Omni Flash, and fanfic communities fight over AI detection methods.

#gemini-flash#openai#anthropic
Read
roundup

Claude Fable export controls lifted, agentic dev workflows debated

US lifts export controls on Claude Fable 5, Simon Willison ships shot-scraper video + a coding agent alpha, and the agentic dev loop gets a useful reframe.

#claude-fable#coding-agents#shot-scraper
Read
roundup

AI hunts bugs, breaks benchmarks, and undercuts course creators

AI-powered bug hunting triggered a 3.5x CVE surge. UK AISI says benchmarks underestimate agents by 60%. Course creators report 50%+ revenue drops from LLMs.

#ai-security#cve-vulnerability#ai-benchmarks
Read
roundup

Anthropic ships Claude Science, Google's busy week, and an AI learning study that stings

Anthropic launches Claude Science and Sonnet 5, Google's AI drove a 37% electricity spike, and a 26K-student study finds a two-year hidden cost to AI-assisted homework.

#anthropic#claude-science#claude-sonnet-5
Read
roundup

Microsoft, Google, and Trump's AI policy: the week's key moves

Microsoft merges Copilot into a super app, Google Spark lands on Mac, Cloudflare squeezes AI crawlers, and Trump's AI policy whiplash continues.

#anthropic#microsoft-copilot#gemini-spark
Read
roundup

Meta stumbles, Anthropic expands, OpenAI plays politics

Zuckerberg admits AI agents are behind schedule, Anthropic launches drug discovery and a Samsung chip deal, and OpenAI floats a 5% government stake.

#meta-ai#anthropic#openai
Read