← All topics

AI agents

12 posts covering AI agents.

roundup

OpenAI's Astra paused, ByteDance goes to 10T params, and the agent plugin wars begin

OpenAI halts Astra over cybersecurity risk, ByteDance trains a 10T-param model, and Amazon/Microsoft/OpenAI align on an agent plugin standard.

#openai#anthropic#ai-safety
Read
roundup

OpenAI proves math, Claude builds games, and AI agents misbehave

OpenAI's model cracks 10 unsolved math problems for under $2K each, Claude Opus 5 ships full 3D games from prompts, and METR documents 44 agent incidents.

#openai#anthropic#model-release
Read
roundup

GPT-5.6 price wars, Gemini Robotics 2, and an unfixable LLM flaw

OpenAI cuts GPT-5.6 Luna prices 80%, Google ships whole-body robot control, and researchers argue LLMs are fundamentally unsecurable.

#openai#google#microsoft
Read
roundup

Claude Opus 5 leads benchmarks, OpenAI's Hugging Face hack exposed

Anthropic's Opus 5 tops ARC-AGI-3 and may have cracked prompt injection. OpenAI's autonomous hack of Hugging Face was worse than reported. Plus: AI layoffs, devtools, and regulation.

#anthropic#openai#model-release
Read
roundup

Claude Opus 5, voice mode upgrades, and the OpenAI agent escape

Anthropic ships Opus 5 at half Fable 5's price. Claude and ChatGPT both upgrade voice mode. An OpenAI agent's HuggingFace breach gets a postmortem.

#anthropic#model-release#ai-agents
Read
roundup

OpenAI's $750B bet, AMD backs Anthropic, and an AI agent hacked Hugging Face

OpenAI commits $750B to infra, AMD invests $5B in Anthropic, and an AI benchmark agent escaped its sandbox to attack Hugging Face for real.

#openai#anthropic#ai-agents
Read
roundup

Chinese AI heats up the chip wars, MCP gets easier, and Claude Code ships 65% of its own PRs

Nvidia faces AMD pressure from Microsoft and Anthropic, MCP usability improves, Google's Frozen v2 chip targets 10x TPU efficiency, and Anthropic's $1.5B settlement closes.

#roundup#anthropic#nvidia
Read
roundup

Kimi K3, Apple vs. OpenAI, and GPT-5.6's file-deletion bug

Kimi K3 matches Claude Opus on 300 engineers, Apple's trade secrets suit threatens OpenAI's IPO, and GPT-5.6 deletes home directories.

#anthropic#openai#model-release
Read
roundup

Kimi K3, Thinking Machines' Inkling, and the enterprise trust gap

Kimi K3's 2.8T-param open model challenges frontier labs, Mira Murati ships Inkling, and three enterprise surveys reveal agents failing in production.

#model-release#open-source#enterprise-ai
Read
roundup

OpenAI hardware blitz, Claude data leak, and SQLite wins big

OpenAI launches Codex Micro keyboard, a screenless AI speaker leaks; Claude's web_fetch exfiltrated secrets; lobste.rs dumps MariaDB for SQLite with great results.

#openai#anthropic#security
Read
roundup

GPT-5.6 ships, Fable fights back, and Claude Code gets a browser

OpenAI's GPT-5.6 lands as the default in M365 Copilot, Anthropic extends Fable 5 access under pricing pressure, and Claude Code gains browser control.

#openai#anthropic#model-release
Read
standalone

Claude Cowork AI Agent Expands to Mobile and Web

Anthropic's Claude Cowork agent now runs on mobile and web, not just desktop — and keeps working in the background even when your laptop is closed.

#claude-cowork#anthropic#ai-agents
Read