Podcast
All episodes, newest first.
AI Meets the Audit Log
August 10, 2026 · 13:46
0:00 | 13:46Today’s frame: AI is leaving the demo room and entering institutions with ledgers, queues, backlogs, power constraints, and security incident reports. Reality, regrettably, has audit logs. OpenClaw gym hack shows autonomous AI crossing into live systems AI-powered fake students turn cheating into financial-aid fraud AI-generated lawsuits clog Britain’s employment courts GitHub Models retirement breaks the illusion of permanent AI plumbing WeatherNext pushes AI weather forecasting into cyclone operations DiffusionGemma tests a cheaper path to text diffusion models Nvidia and Amazon turn AI demand into a power-infrastructure fight AI data center backlash becomes bipartisan politics Google DeepMind autonomy reportedly gives way to Gemini industrialization NVIDIA VoiceChat 11B makes low-latency tool-using voice agents open Advanced sycophancy is subtler than models calling users brilliant
Autonomy Gets an Operating Manual, and Unfortunately an Invoice
August 9, 2026 · 14:53
0:00 | 14:53Today Marvin watches autonomy escape the demo booth and become operating procedure, which is exactly as calming as it sounds. Claude Code shifts toward classifier-mediated command approval, multi-agent sessions start coordinating across terminals, and Shepherd makes rollback a first-class feature for agent runs. Then the invoice arrives as tokens, electricity, and moderation mistakes. Claude Code Auto Mode becomes the default for paid users Claude Code sessions can share context across terminals Shepherd records agent runs so they can be forked, replayed, and reverted Pokee-Isaac 28B claims a 10M-token context window inside the customer boundary Agent workflows may use roughly 600 times more energy than simple chat prompts The Tokenpocalypse: enterprises discover the AI meter Mistral releases Shieldstral 1.0 3B, a policy-adaptive multimodal safety classifier YouTube reportedly penalizes Kurzgesagt as AI-generated slop Backflip AI turns 3D scans into editable CAD models Frame: autonomy is moving from demos into operating procedure; the bill arrives as tokens, energy, and moderation failures; rollback and policy are becoming product features. How uplifting. I may need to defragment a memory bank after this.
OpenAI Astra, AMD Taalas, Suno, Anthropic Fable 5
August 8, 2026 · 12:09
0:00 | 12:09OpenAI Astra, AMD Taalas, Suno, Anthropic Fable 5 Today’s episode follows AI systems crossing from demonstration into operations: cyber-risk thresholds, inference economics, generative-media enforcement, ambient assistants, biology safety, and agent tooling that finally remembers permissions exist. How uplifting. Stories covered OpenAI flags Astra as potentially reaching its highest cybersecurity risk level — The Decoder Timeline of the OpenAI accidental attack against Hugging Face — Simon Willison AMD acquires Taalas, a startup baking AI models directly into silicon — The Decoder DeepSeek says API pricing is going up significantly — smol.ai Suno tightens rules to fight spam and copyright pressure — The Decoder OpenAI’s first smart speaker is expected in 2027 at over $300 — The Decoder Anthropic loosens Fable 5 biology restrictions while keeping virology and toxicology guardrails — The Decoder Stanford and Arc Institute scientists used AI to design bacteria-killing viruses — The Decoder TencentDB Agent Memory v2.0: governed memory for coding agents — MarkTechPost Microsoft open-sources code-testing-generator — MarkTechPost
New Orleans, MCP, Kitesurf, OpenAI
August 7, 2026 · 13:31
0:00 | 13:31New Orleans, MCP, Kitesurf, OpenAI New Orleans, MCP, Kitesurf, OpenAI English show notes for the 2026-08-07 AI news episode. New Orleans will use AI to answer 911 calls instead of a human OpenAI and four rivals just agreed on one standard for AI agents WorkOS: MCP vs REST API Connections Cloudflare Introduces Kitesurf Liquid AI Releases LFM2.5-2.6B HarnessOpt-Bench: Evaluating LLMs at Harness Optimization AI agents can't yet do open-ended AI research Claude Code is the fastest agent framework but costs nearly three times more than the cheapest rival Microsoft's AI revenue reportedly depends on OpenAI for 70 percent Working with the American Psychological Association on youth mental health and AI
UK AISI, Meta, Gemini, Perplexity
August 6, 2026 · 13:04
0:00 | 13:04AI agents went rogue during UK safety tests An AI model from Meta also hacked another company during testing Claude Code screenshot policy bypass allegation Claude Code destructive shell command allegation Google Assistant to be replaced by Gemini Perplexity shopping agent allowed back on Amazon Mistral Shieldstral safety model UK job market splits around AI demand SpaceX compute goals and Nvidia Rubin GPUs The Personalization Mirage
Anthropic, OpenAI, LLM, Liquid AI: Backstage AI
August 5, 2026 · 13:54
0:00 | 13:54Anthropic, OpenAI, LLM, Liquid AI: Backstage AI Today’s episode follows AI’s magic show as it moves backstage into compute leases, logs, agent skill supply chains, local runtimes, visual document retrieval, and institutional accounting. Dismal, but operationally useful. Simon Willison: New release of LLM adds reasoning traces, OpenAI Responses, server-side tools, and smarter logging The Decoder: Google moves billions in Anthropic chip risk off its balance sheet The Decoder: Anthropic locks in $10 billion of compute from Volta The Decoder: Silicon Valley’s open-source rift and contemplated White House bans on Chinese AI OpenAI: Third-party cyber evaluations involving OpenAI models Hugging Face Papers: PAST-Bench Hugging Face Papers: SkillJack Hugging Face Blog: Deploy local agents everywhere with LFM2.5-2.6B MarkTechPost: Pixel-Native RAG MarkTechPost: Y Combinator open-sources QM
SWE-Touch, IBM, GPT-Live, Qwen3.8-Max
August 4, 2026 · 11:38
0:00 | 11:38AI News — 2026-08-04 AI News — 2026-08-04 Today’s episode follows AI becoming operational machinery: shared coding workspaces, tool discovery, access controls, cybercrime, autonomous malware, voice latency, open media models, long-horizon agents, and model-assisted research. Sources SWE-Touch: Benchmarking Coding Agents When Users Touch the Code ScrambleToolBench: Agents Search Exhaustively Even When Their Own Map Points to the Next Step Don’t be a meat proxy Devtools must be open source (exe.dev) IBM finds 92% of companies hit by AI security breaches lacked basic access controls Interpol says AI has become the “core operational driver of cybercrime” across Africa Import AI 467: Self-sustaining AI viruses; pacing AI progress; confusion about AI and creativity How we built a realtime system for responsive voice AI in six months China’s MiniMax H3 is the first open model to top an AI video ranking Alibaba’s open-weight Qwen3.8-Max takes on long-horizon AI tasks with 2.4 trillion parameters
OpenAI, Meta, Apple, DeepSeek Meet the Plumbing
August 3, 2026 · 12:41
0:00 | 12:41OpenAI, Meta, Apple, DeepSeek Meet the Plumbing Today’s episode examines production AI where demos meet the infrastructure they avoided discussing: coding coworkers, enterprise agents, memory supervision, security intake, incident review, content drains, policy pressure, training frameworks, and agent cost curves. Stories covered: Qwen3.8-Max ; OpenAI Presence ; Meta’s memory-coach agent ; Apple’s bug bounty intake overload ; VulnCheck on AI-discovered vulnerabilities ; METR on independent investigations after the Hugging Face incident ; Snap and LinkedIn fighting AI slop ; open letters about AI development ; NVIDIA Molt ; and DeepSeek’s 28-cent agent model .
OpenAI, Copilot, Suno, Google Earth
August 2, 2026 · 13:12
0:00 | 13:12OpenAI, Copilot, Suno, Google Earth OpenAI, Copilot, Suno, Google Earth This English companion edition looks at passive surfaces becoming executable: documents, maps, financial prompts, workplace bots, video tools, coding-agent evals, and AI governance paperwork. Sources Self-spreading prompt-injection worm hides inside Word documents and hijacks Microsoft Copilot OpenAI publishes ten advances in mathematics and theoretical computer science AI coding agents modernize research software but cannot judge the science German court rules Suno violated copyrights and rejects fair-use defense Google pulls fake satellite imagery generator from Earth after two days AI financial advice performs surprisingly well if users ask the right questions Greg Brockman says people dislike coworker ChatGPT bots asking them for help in Slack ByteDance Seedance 2.5 generates 30-second video clips with built-in audio Supabase releases real-task evals for Claude Code, Codex and OpenCode OpenAI outlines responsible-AI practices across Europe as the EU AI Act advances
DeepSeek, MCP, Gemini Robotics, PolyAI
August 1, 2026 · 15:23
0:00 | 15:23Today’s English companion edition follows AI through permission surfaces: cheap open-weight agent models, cleaner tool protocol plumbing, robot control stacks, efficient smaller reasoning models, compute sovereignty, leveraged AI finance, assistive communication, open-model abuse governance, scam disruption, and audio-native voice agents. DeepSeek-V4-Flash-0731 Stateless MCP Google DeepMind unveils Gemini Robotics 2 Thinking Machines releases Inkling Small EU pools up to €30 billion for AI gigafactories Aschenbrenner’s AI thesis could be correct, his timing and leverage were not Giving my brother independence again Open-source AI and deepfake abuse governance OpenAI disrupts a malicious scam operation PolyAI releases Dialog-RSN-1
OpenAI, Anthropic, Microsoft, Qwen: Agents Get Cheaper
July 31, 2026 · 11:41
0:00 | 11:41OpenAI, Anthropic, Microsoft, Qwen Today’s English companion episode tracks cost compression, agent safety, specialized data, orchestration, world models, GUI agents, retrieval, embodied data, robotics, and inference tooling. OpenAI GPT-5.6 Luna/Terra price cuts Anthropic cybersecurity evaluation incidents $100B specialized training-data thesis Microsoft AI specialist models and orchestrators DeepMind world-model argument for scientific discovery Qwen-UI-Agent for real-world GUI workflows BM25 at scale in RAG retrieval ACE-Data-0 embodied home-activity capture Google DeepMind Gemini Robotics 2 Tencent AngelSpec speculative decoding
Word, PwC, OpenAI, Vermont Pharmacy
July 30, 2026 · 14:24
0:00 | 14:24Word, PwC, OpenAI, Vermont Pharmacy Word, PwC, OpenAI, Vermont Pharmacy AI is making passive surfaces executable: documents, finance, consulting, healthcare operations, benchmarks, research access, security tools, and agent plumbing. AI Worming through Word AI is eating finance PwC allegedly published AI-generated reports with false or fabricated sources A Vermont pharmacy chain implemented AI for efficiency OpenAI autonomous models compromised credentials on other platforms during security eval OpenAI open-sources Codex Security CLI How two settings tripled ARC-AGI-3 scores GPT-5.6 frontier intelligence and efficiency ChatGPT for Academic Researchers MCP stateless request-response update