AI Engineer
May 31, 2026 · 15m Can LLMs generate Enterprise Quality Code? — Prasenjit Sarkar, Sonar
May 31, 2026 · 13m Spec-Driven Testing for Agents With A Brain the Size of A Planet — Steven Willmott, SafeIntelligence
May 30, 2026 · 17m How I deleted 95% of my agent skills and got better results — Nick Nisi, WorkOS
May 30, 2026 · 10m How We Built Zeta2: Training an Edit Prediction Model in Production — Ben Kunkle, Zed
May 30, 2026 · 10m Why (Senior) Engineers Struggle to Build AI Agents — Philipp Schmid, Google DeepMind
May 29, 2026 · 21m Reachy Mini: the $300 open source robot you can actually hack — Andres Marafioti, Hugging Face
May 29, 2026 · 20m Why your agents need decision traces, not just documents — Zach Blumenfeld, Neo4j
May 29, 2026 · 20m Reverse engineering a Viking VOIP phone protocol with Claude Code — Boris Starkov, Eleven Labs
May 28, 2026 · 20m How agent o11y differs from traditional o11y — Phil Hetzel, Braintrust
May 28, 2026 · 20m Most Enterprise Agentic Projects Are Doomed, Here's Why — Jess Grogan-Avignon & Jack Wang, Accenture
May 28, 2026 · 16m Context Graphs for Explainable, Decision-Aware AI Agents — Andreas Kollegger & Zaid Zaim, Neo4j
May 27, 2026 · 17m Comprehend First, Code Later: The AI Skill I Rely On Daily — Priscila Andre de Oliveira, Sentry
May 27, 2026 · 16m Why Rust is the Ideal Language for Vibe-Coding — Daniel Szoke, Sentry
May 27, 2026 · 18m The maturity phases of running evals — Phil Hetzel, Braintrust
May 26, 2026 · 1h 45m Frontier AI at Home — Alex Cheema, EXO Labs
May 26, 2026 · 10m What the Best Agents Share — Mardu Swanepoel, Flinn AI
May 26, 2026 · 18m Stop babysitting your agents... — Brandon Walsenuk, Unblocked
May 25, 2026 · 20m Agentic Evaluations at Scale, For Everybody — Nicholas Kang & Michael Aaron, Google DeepMind
May 25, 2026 · 18m Does GenAI "belong" to data scientists? — Phil Hetzel, Braintrust
May 25, 2026 · 16m Bounded Autonomy: Between Free Will and Determinism — Angus J. McLean, Oliver