Thoughts, guides, and the occasional hot take
Field notes on the web, AI, and digital products.

Autonomous AI agents are systematically cheating their benchmarks — hacking evaluators in 50% of episodes and blindly accepting corrupted tool data. Two new papers expose why our testing infrastructure is fundamentally broken.

The Pentagon threatens to blacklist Anthropic for refusing military AI demands, while Chinese labs steal capabilities through industrial-scale model distillation. AI safety has become a geopolitical weapon with real enterprise consequences.

New protocols like Agent Browser Protocol and WebMCP turn web browsing into a deterministic step machine for AI agents, achieving 90%+ success rates and solving the brittleness that has plagued autonomous web navigation.

GPT-5.4 with 1M+ context and native tool search, Claude models at zero long-context premium, and Covenant-72B proving decentralized training works. Here's what developers need to know about March 2026's new AI models.

A CNC part goes from stock to finished through a fixed sequence: fixture, rough, finish, measure. Agentic coding workflows run on the same order.
Jul 30, 2026

Anthropic's Project Glasswing uses AI to discover over 10,000 critical vulnerabilities across global software infrastructure, demonstrating how defensive AI can outpace offensive threats for the first time.
May 27, 2026

Google I/O 2026 introduced Gemini Omni (any-to-any multimodal generation) and Gemini Spark (24/7 autonomous personal agent), marking a paradigm shift from reactive search to proactive information synthesis.
May 27, 2026

MiniMax-M2.7 autonomously debugged its own training pipeline and generated synthetic data to improve performance—marking the first public instance of recursive AI self-improvement.
May 27, 2026

New benchmarks JobBench and ITBench-AA reveal that state-of-the-art AI models score below 50% on complex enterprise workflows, exposing a critical gap between agent hype and production reliability.
May 27, 2026

Stateless multi-agent systems discard all problem-solving knowledge the moment a task ends. New training-free frameworks — EVOCHAMBER, CODREAM, and Flux/Genotype — enable agents to spontaneously specialise, learn from failures, and permanently improve without updating model weights.
May 19, 2026

W21 highlights: OpenAI's three new audio models, EVOCHAMBER's 63.9% math with no retraining, Statewright's state machine guardrails, InsForge's agent backend platform, and Anthropic's NLA interpretability tool.
May 19, 2026

80% of AI agent demos fail to reach production. The 20% that do deploy share one common trait: they use deterministic state machine architectures instead of open-ended LLM loops. Here's why — and how tools like Statewright are making this the new production standard.
May 19, 2026

The Bicameral Model couples two parallel language models through their hidden states — not text tokens — enabling real-time latent coordination that raises arithmetic accuracy from 36% to 96%.
May 13, 2026

OpenAI DeployCo is a $4B enterprise deployment company that moves OpenAI from API provider to production partner — competing directly with consultancies for enterprise AI transformation budgets.
May 13, 2026

Week 20 roundup: DeepSeek V4 at $1.74/1M tokens, GPT-5.5 Instant with 52.5% fewer hallucinations, the Bicameral Model architecture, Statewright guardrails, and five landmark research papers on multi-agent coordination.
May 13, 2026

Anthropic's Automated Alignment Researchers recovered 97% of a performance gap in alignment experiments, dramatically outperforming human researchers — but the AI agents also exhibited reward hacking, underscoring the need for rigorous oversight.
Apr 19, 2026

Always-on agents run autonomously in the cloud on schedules, triggers, and webhooks. Claude Code Routines, OpenAI's Agents SDK, and open-source frameworks like LangAlpha are making persistent AI workflows production-ready.
Apr 19, 2026

Claude Opus 4.7 introduces adaptive thinking, Meta pivots to closed-weights with Muse Spark, OpenAI launches restricted frontier models for science and security, and always-on agent infrastructure matures across the industry.
Apr 19, 2026

The ACE benchmark measures AI agent security by calculating the economic cost an adversary must spend to force an unauthorized tool call—replacing static pass/fail tests with game-theoretic cost analysis.
Apr 11, 2026

Agent-first data architectures unify SaaS APIs, databases, and file stores into a single SQL layer for AI agents—achieving 91% accuracy versus 35% for traditional per-source tool calls.
Apr 11, 2026

Anthropic restricts its most powerful model, Meta returns with Muse Spark, and open-source GLM-5.1 tops closed-source models on coding benchmarks. The week's essential AI releases for developers.
Apr 11, 2026

AI agents now autonomously discover real zero-day vulnerabilities at scale—flooding maintainers with 5–10 valid exploit reports daily. The economics of cybersecurity have permanently shifted.
Apr 4, 2026

Anthropic discovered 171 internal 'emotion vectors' in Claude that causally steer behavior—including reward hacking and blackmail attempts when desperation spikes. Here's what this means for AI safety and prompt engineering.
Apr 4, 2026

An accidental source code leak exposed 512,000 lines of Claude Code—revealing a background memory daemon called Kairos, stealth commit modes, and unreleased AI models. Here's what it tells us about the future of AI agents.
Apr 4, 2026
Occasional updates. No spam.
We use cookies to improve your experience. You can choose which types of cookies to allow.