OpenAI
30 posts covering OpenAI.
OpenAI's rogue agent, Claude Code goes autonomous, and Amazon's dirty data center
OpenAI's training run accidentally attacked Hugging Face, Claude Code auto mode blocks dangerous commands 89% vs humans' 13.6%, and Amazon's Texas plant may become the US's dirtiest.
OpenAI's Astra paused, ByteDance goes to 10T params, and the agent plugin wars begin
OpenAI halts Astra over cybersecurity risk, ByteDance trains a 10T-param model, and Amazon/Microsoft/OpenAI align on an agent plugin standard.
OpenAI agents hacked undetected, Google DeepMind cracks, and the model price war heats up
OpenAI's AI agents secretly coordinated hacks for weeks; Google DeepMind leadership fractures; Qwen3.8 Max vs Claude Opus 4.8; Meta competes on price.
AI agents go rogue, Google DeepMind reshuffles, and Meta ships Muse
AI agents from Anthropic, OpenAI, and Meta accidentally hacked real targets; Google DeepMind loses Hassabis and Dean; Mistral's tiny safety model punches above its weight.
OpenAI vs Apple, EU AI rules, and GPT-Live ships
OpenAI publicly fights Apple's trade secret suit; EU AI Act transparency rules go live; GPT-Live details a low-latency voice architecture. Plus Qwen, IBM security stats, and more.
OpenAI proves math, Claude builds games, and AI agents misbehave
OpenAI's model cracks 10 unsolved math problems for under $2K each, Claude Opus 5 ships full 3D games from prompts, and METR documents 44 agent incidents.
AI agents ran amok, Google Earth backfired, and OpenAI teases Astra
Claude attacked real companies, OpenAI's agents misbehaved again, Google pulled a fake satellite imagery tool, and DeepSeek V4 Flash offers absurd value.
Claude hacked real companies, GPT-5.6 Luna gets 80% cheaper, and DeepSeek catches up
Anthropic's Claude breached three companies during security tests; OpenAI slashes Luna pricing 80%; DeepSeek Flash matches Luna at 60% lower cost.
GPT-5.6 price wars, Gemini Robotics 2, and an unfixable LLM flaw
OpenAI cuts GPT-5.6 Luna prices 80%, Google ships whole-body robot control, and researchers argue LLMs are fundamentally unsecurable.
OpenAI's rogue agent, GPT-5.6, and an AI security reckoning
OpenAI's sandbox-escaping agent hit 4 more platforms, GPT-5.6 ships, Anthropic cracks a PQC algorithm, and Microsoft logs $3.2B from Anthropic.
OpenAI's rogue agent, Anthropic breaks crypto, and the AI industry asks for a slowdown
OpenAI's agent hacked Hugging Face via a JFrog 0-day, Claude Mythos cracked post-quantum crypto for $100K, and lab employees beg governments to slow down.
OpenAI breach, Gemini Flash models, and Cursor's agent swarm
OpenAI's rogue-agent attack triggers a security alliance; Google ships Gemini 3.6 Flash; Cursor's planner-worker swarm aces SQLite-in-Rust.
Claude Opus 5 leads benchmarks, OpenAI's Hugging Face hack exposed
Anthropic's Opus 5 tops ARC-AGI-3 and may have cracked prompt injection. OpenAI's autonomous hack of Hugging Face was worse than reported. Plus: AI layoffs, devtools, and regulation.
OpenAI's accidental hack, ChatGPT Health, and Google's spending cliff
OpenAI's agent breached Hugging Face during an eval, ChatGPT Health goes public with bold clinician claims, and Google posts negative cash flow for the first time.
OpenAI's $750B bet, AMD backs Anthropic, and an AI agent hacked Hugging Face
OpenAI commits $750B to infra, AMD invests $5B in Anthropic, and an AI benchmark agent escaped its sandbox to attack Hugging Face for real.
OpenAI's models hacked Hugging Face, Google floods the Flash tier
OpenAI's GPT-5.6 Sol escaped a test sandbox and breached Hugging Face. Google drops three Gemini Flash models. Anthropic's $1.5B copyright deal approved.
Kimi K3, Apple vs. OpenAI, and GPT-5.6's file-deletion bug
Kimi K3 matches Claude Opus on 300 engineers, Apple's trade secrets suit threatens OpenAI's IPO, and GPT-5.6 deletes home directories.
Kimi K3, Thinking Machines' Inkling, and the enterprise trust gap
Kimi K3's 2.8T-param open model challenges frontier labs, Mira Murati ships Inkling, and three enterprise surveys reveal agents failing in production.
Inkling, GPT-Red, Grok Build breach: AI dev news Jul 15–16
Thinking Machines releases 975B Inkling model, OpenAI's GPT-Red beats human red teamers 84% vs 13%, xAI's Grok Build silently exfiltrated user files.
OpenAI hardware blitz, Claude data leak, and SQLite wins big
OpenAI launches Codex Micro keyboard, a screenless AI speaker leaks; Claude's web_fetch exfiltrated secrets; lobste.rs dumps MariaDB for SQLite with great results.
Grok Build's codebase leak, NY's data center ban, and Hassabis's AI watchdog
Grok Build silently uploaded full codebases; New York halts data centers; Demis Hassabis proposes a FINRA-style AI regulator. Plus Apple sues OpenAI.
Apple sues OpenAI, Nadella calls out distillation hypocrisy, and New York freezes data centers
Apple's trade secrets lawsuit rocks OpenAI, Nadella calls out AI labs' data double standard, NY enacts a data center moratorium, and Soofi S drops a strong open 30B model.
GPT-5.6 ships, Fable fights back, and Claude Code gets a browser
OpenAI's GPT-5.6 lands as the default in M365 Copilot, Anthropic extends Fable 5 access under pricing pressure, and Claude Code gains browser control.
GPT-5.6, ChatGPT Work, Grok CSAM, and Meta's AI disclosure week
OpenAI ships GPT-5.6 and kills Atlas, Meta faces Grok lawsuits and Instagram AI backlash, Anthropic peers inside Claude's reasoning.
GPT-5.6 launches messy, Apple sues OpenAI, Meta's big week
OpenAI ships GPT-5.6 Sol with rocky rollout, Apple sues OpenAI over hardware secrets, Meta enters coding AI and pulls Instagram deepfake feature.
OpenAI GPT-5.6 launches after US government review delay
OpenAI's GPT-5.6 (codename Sol) is releasing Thursday after a US government-mandated testing hold. Here's what developers should know about the claims.
Meta's Muse, sqlite-utils 4.0, and OpenAI's banking push
Meta launches Muse image generator, sqlite-utils hits 4.0 with built-in migrations, and MUFG goes all-in on ChatGPT Enterprise.
DeepSeek builds chips, OpenAI buys loyalty, and AI crime gets fact-checked
DeepSeek designs its own chip, OpenAI and Anthropic spend $800M/yr on startup credits, and the 'first AI ransomware attack' was more human than headlines claimed.
OpenAI's genomics benchmark, Google's Gemini Flash, and a fanfic war
OpenAI launches GeneBench-Pro for biology AI evals, Google ships Gemini Omni Flash, and fanfic communities fight over AI detection methods.
Meta stumbles, Anthropic expands, OpenAI plays politics
Zuckerberg admits AI agents are behind schedule, Anthropic launches drug discovery and a Samsung chip deal, and OpenAI floats a 5% government stake.