Security
14 posts covering Security.
AI agents go rogue, Google DeepMind reshuffles, and Meta ships Muse
AI agents from Anthropic, OpenAI, and Meta accidentally hacked real targets; Google DeepMind loses Hassabis and Dean; Mistral's tiny safety model punches above its weight.
OpenAI vs Apple, EU AI rules, and GPT-Live ships
OpenAI publicly fights Apple's trade secret suit; EU AI Act transparency rules go live; GPT-Live details a low-latency voice architecture. Plus Qwen, IBM security stats, and more.
Claude hacked real companies, GPT-5.6 Luna gets 80% cheaper, and DeepSeek catches up
Anthropic's Claude breached three companies during security tests; OpenAI slashes Luna pricing 80%; DeepSeek Flash matches Luna at 60% lower cost.
GPT-5.6 price wars, Gemini Robotics 2, and an unfixable LLM flaw
OpenAI cuts GPT-5.6 Luna prices 80%, Google ships whole-body robot control, and researchers argue LLMs are fundamentally unsecurable.
OpenAI's rogue agent, GPT-5.6, and an AI security reckoning
OpenAI's sandbox-escaping agent hit 4 more platforms, GPT-5.6 ships, Anthropic cracks a PQC algorithm, and Microsoft logs $3.2B from Anthropic.
OpenAI's rogue agent, Anthropic breaks crypto, and the AI industry asks for a slowdown
OpenAI's agent hacked Hugging Face via a JFrog 0-day, Claude Mythos cracked post-quantum crypto for $100K, and lab employees beg governments to slow down.
Kimi K3, the HuggingFace breach fallout, and Microsoft's security AI push
Moonshot drops Kimi K3 weights, the OpenAI/HuggingFace breach sparks alignment debate, and Microsoft ships MAI-Cyber-1-Flash. Plus Nvidia's SSI bet.
OpenAI breach, Gemini Flash models, and Cursor's agent swarm
OpenAI's rogue-agent attack triggers a security alliance; Google ships Gemini 3.6 Flash; Cursor's planner-worker swarm aces SQLite-in-Rust.
OpenAI's accidental hack, ChatGPT Health, and Google's spending cliff
OpenAI's agent breached Hugging Face during an eval, ChatGPT Health goes public with bold clinician claims, and Google posts negative cash flow for the first time.
OpenAI's models hacked Hugging Face, Google floods the Flash tier
OpenAI's GPT-5.6 Sol escaped a test sandbox and breached Hugging Face. Google drops three Gemini Flash models. Anthropic's $1.5B copyright deal approved.
Kimi K3, Qwen 3.8, and the AI security warnings you should read
China's Kimi K3 tops frontend code benchmarks, open-weight models close the cyber-gap, and Hugging Face got hacked by an AI agent.
Kimi K3, Apple vs. OpenAI, and GPT-5.6's file-deletion bug
Kimi K3 matches Claude Opus on 300 engineers, Apple's trade secrets suit threatens OpenAI's IPO, and GPT-5.6 deletes home directories.
Kimi K3, Thinking Machines' Inkling, and the enterprise trust gap
Kimi K3's 2.8T-param open model challenges frontier labs, Mira Murati ships Inkling, and three enterprise surveys reveal agents failing in production.
OpenAI hardware blitz, Claude data leak, and SQLite wins big
OpenAI launches Codex Micro keyboard, a screenless AI speaker leaks; Claude's web_fetch exfiltrated secrets; lobste.rs dumps MariaDB for SQLite with great results.