AI Engineer
Jul 18, 2026 · 19m Agents Need Receipts, Not More Tool Calls - Armanas Povilionis, Alithea Bio
Jul 18, 2026 Why Large? Tiny LMs & Agents on Edge/Robotics — Cormac Brick, Google
Jul 18, 2026 Vending-Bench: Long-Horizon Agent Evals — Lukas Petersson, Andon Labs
Jul 17, 2026 · 21m Using LLMs to Secure Source Code — Eugene Yan, Anthropic
Jul 17, 2026 · 1h 0m The Great Loops Debate — Dex Horthy, Geoff Huntley, Ian Livingstone, Greg Pstrucha, @insecure-agents
Jul 17, 2026 · 20m "Software engineering is not about writing code" — Benoit Schillings, Google DeepMind VP of Research
Jul 17, 2026 Claude for Long-Horizon Tasks — Lance Martin, Anthropic
Jul 16, 2026 · 16m An AI Agent Became the #1 Contributor in OpenAI's Hiring Challenge — Zhengyao Jiang, Weco
Jul 15, 2026 · 16m Computer-Use 2.0: Agents Just Got Multi-Cursor — Francesco Bonacci, Cua
Jul 15, 2026 · 20m Recursive Model Improvement — Lee Robinson, Cursor, SpaceXAI
Jul 15, 2026 · 51m Simon Willison in conversation with Cat Wu & Thariq Shihipar, Anthropic
Jul 14, 2026 · 20m WTF Is the Context Layer? The Missing Infrastructure for Production Agents — Prukalpa Sankar
Jul 14, 2026 · 21m Don't Ship Skills Without Evals — Philipp Schmid, Google DeepMind
Jul 14, 2026 · 20m Forward Deployed Engineering at Cursor — Pauline Brunet
Jul 14, 2026 · 18m "The engineer of the future is the person who is able to choose what is worth doing." — Addy Osmani
Jul 13, 2026 · 21m In Code They Act, In Proof We Trust — Erik Meijer, Leibniz Labs
Jul 13, 2026 · 23m Stop Evaluating Models Like It's the 50s - Alejandro Vidal, Mindmakers
Jul 13, 2026 · 46m The Prime Intellect Stack — Will Brown, Prime Intellect
Jul 12, 2026 · 17m RLM: Recursive Language Models for Large Codebases - Shashi, Superagentic AI
Jul 12, 2026 · 19m The AI bugpocalypse is here. Now what? - Jack Cable, Corridor