The archive
AI Research & Insights
Daily digests of what's actually happening in AI — from research breakthroughs to new model releases, minus the hype.

AI Model Comparisons
GPT-6.1 Sol vs Claude Opus 5.5: Price & Benchmarks
Compare GPT-6.1 Sol and Claude Opus 5.5 on price, benchmarks, retries, and real task cost before choosing a production model.
AI Security
MCP Vulnerability: How AI Agent Trust Enables Attacks

Agent Memory Systems
AI Agent Memory: Learning From Past Failures Safely

AI Detection Analysis
AI-Generated Text Detection: Can It Spot Opus 5.5?

AI Agent News
Caddy AI Agent Launches on iMessage and Android RCS

AI Security
OpenAI AI Agents Australia Breach: Security Lessons for Teams

AI Agent Engineering
MCP AI Agents: Reliable APIs and Enterprise Workflows

AI Agent Architecture
Long-Running AI Agents: Build Systems That Finish Reliably

AI Agent Research
Self-Improving Coding Agents: MIT’s Cheaper SIFT Method

AI Agent Security
OpenAI Jev Oversight for Swarming AI Agents Explained

AI Safety Research
Trillium Labs AI Research: The Frontier Openness Test

AI Model Launches
Claude Sonnet 5.5: Anthropic Claims 30% Lower Agent Costs

