Topic
AI Research
Reporting on the research behind the products: new training methods, benchmarks, retrieval and reasoning techniques, and the papers that move the field.
We summarise what a result claims, how it was measured and where its limits are, in language that does not need a machine learning background.
AI Agent Research
Self-Improving Coding Agents: MIT’s Cheaper SIFT Method

AI Safety Research
Trillium Labs AI Research: The Frontier Openness Test

AI Benchmarks
CoreWeave DeepSeek V3 MLPerf Record: What It Means

AI Research
AI Engram Memory Traces in Artificial Intelligence Explained

AI Benchmarks
DRL transformer open shop scheduling problem: what’s new

AI Benchmarks
SentinelBench explained: the benchmark for long-running AI agents

AI Benchmarks
Multi Agent Strategy Training: MindGames Arena Explained

AI Benchmarks
AI model speed benchmark: how to test latency right

NLP Research
Cognitive Categorical Transformer: why this paper matters

Machine Learning
MacBook as AI training cluster: why this idea matters

Training Approaches
Train GPT on non language data: a practical field guide

AI Benchmarks
ChatGPT 5.5 vs Opus DeepSWE benchmark: what changed

