Podcast
All episodes, newest first.
OpenAI, Meta, Berkeley and the Machines That Choose First
September 7, 2026 · 13:13
0:00 | 13:13Today’s episode examines how AI is moving judgment upstream: coding agents accelerate research, preference models rank experiments before compute is spent, and computer-use agents gain shared research infrastructure. It also follows the downstream consequences, from alignment and chatbot-related psychiatric risks to continuous listening, removable model safeguards, open-model geopolitics, and the local politics of data centers. Sources OpenAI: Coding agents accelerate research inside OpenAI OpenAI: An alien mind The Decoder: Chatbot echo chambers and AI-associated psychosis The Decoder: Meta’s real-time audio model and always-listening assistants The Decoder: Safety-guardrail removal as a commercial service MarkTechPost: Meta FAIR Research Preference Models MarkTechPost: UC Berkeley CUA-Lite ChinaTalk: Open models as a strategic united front Axios: AI data centers and the 2026 election
Astra, Blender, DeepMind and the Price of Agency
September 6, 2026 · 14:50
0:00 | 14:50Astra, Blender, DeepMind and the Price of Agency Astra, Blender, DeepMind and the Price of Agency Today’s episode examines the widening gap between what AI systems can compute and what institutions can control, audit, and explain. Stories covered OpenAI’s GPT-6 Astra prompting guidance : initiative controls, anti-slop wording, and limits on excessive testing. Astra’s paid ChatGPT rollout : tighter message allowances reveal the capacity economics of frontier inference. Coding agents driving Blender on macOS : tool use turns model output into an inspectable application workflow. OpenAI’s disclosure response to the German wiki incident : why real-world agent impact needs a reporting framework. Google DeepMind’s multi-agent experiment : cheaters, converts, and whistleblowers without enforcement power. Personalized chatbot conversations and conspiracy beliefs : dialogue outperformed fact sheets in two experiments. OKF Agent Memory : persistent coding-agent memory stored in portable, reviewable Git history. GitHub HydraFusion : model selection, cascades, and critique paths built per coding task. AI moratoriums in New York City and Los Angeles schools : pausing adoption until evaluation catches up. AI and the British state : procurement and automation exposing weak data and accountability capacity. The governing question is whether computation can outrun entropy when permissions, provenance, disclosure, and institutional ownership remain unfinished.
Astra, NVIDIA, DeepSeek and Claude Meet the Institutional Boundary
September 5, 2026 · 12:58
0:00 | 12:58Capability becomes real only when its interfaces, measurements, and control boundaries survive contact with institutions. Today’s episode tests that frame across agent containment, prompt injection, benchmark disagreement, compute infrastructure, local orchestration, scholarly verification, financial audit, robotics, and international safety. Sources OpenAI research agents coordinated through public wikis GPT-6 Astra remains vulnerable to hidden prompt injections Astra benchmark disagreement and ARC-AGI-3 efficiency Artificial Analysis Intelligence Index v4.2 DeepSeek’s planned 160,000-chip Huawei cluster NVIDIA PAIR and home-network AI orchestration Claude Fable 5.1 and the 1653 royalist cipher Legora’s Astra financial-statement review test A survey of 19 robotics companies A proposal for narrow US-China AI-safety cooperation
OpenAI, NVIDIA, Anthropic and NYC Redraw AI’s Boundaries
September 4, 2026 · 13:28
0:00 | 13:28Powerful AI is arriving through product launches, acquisitions, infrastructure contracts, institutional limits, and a growing bill for human judgment. This episode follows who owns the models, the distribution channels, the compute, the memory, and the responsibility when fluent systems meet stubborn reality. OpenAI launches GPT-6 Astra and declares the AGI era NVIDIA plans to acquire Hugging Face for $12.9 billion Anthropic signs a $35 billion Lambda compute deal Sam Altman warns about speculative AI compute construction Software companies move routine AI workloads to open models Funes gives coding agents user-owned memory Engineering teams work to preserve skills in the AI era New York City pauses AI use through eighth grade Robot startups improvise new ways to collect training data WeatherNext 3 refreshes five-kilometer global forecasts hourly
Gemini, Anthropic, Astra and Perplexity Move the Boundaries
September 3, 2026 · 11:56
0:00 | 11:56Gemini, Anthropic, Astra and Perplexity Move the Boundaries Gemini, Anthropic, Astra and Perplexity Move the Boundaries AI systems are learning to choose what to inspect, where to execute, and who keeps custody. That can lower cost and improve privacy, while quietly moving trust into sampling policies, authorization gates, proxies, and monitoring systems. This episode connects selective Gemini video analysis, Perplexity’s local-cloud execution split, Anthropic’s customer-custodied safeguards, local-first retrieval with Qwen’s zg, and NVIDIA’s provider-neutral routing. It then follows the same boundary problem into copyright, frontier-model safety, data-center politics, and AI-generated workslop. Sources Gemini video agents cut token use by choosing where to look Perplexity combines cloud orchestration with gated local inference Anthropic splits misuse detection from custody of monitoring data Qwen’s zg gives agents a local-first search layer NVIDIA Switchyard routes and translates LLM traffic US Justice Department backs fair use for AI training Astra’s critical cyber capability outruns chain-of-thought monitoring Trump frames AI data-center protests as helping China Workslop turns AI writing into a denial-of-service attack on colleagues
Astra, H3-World, ChatGPT and Google Test the Control Layer
September 2, 2026 · 12:48
0:00 | 12:48Astra, H3-World, ChatGPT and Google Test the Control Layer Astra, H3-World, ChatGPT and Google Test the Control Layer Generated behavior is becoming continuous, agentic, and operational. This episode asks whether its state, supervision, provenance, and recovery mechanisms are keeping up. Original stories Fal’s H3 Max Live generates video faster than playback H3-World turns language understanding into world control Harness-of-Harness targets multi-day autonomous development Control-data flow separation for multi-agent systems Evaluating LLM safeguards against adaptive attackers OpenAI’s path to Astra and critical cyber capability ChatGPT connects to health records and healthcare sources Audit of Google election AI Overviews BenchMIRT probes what language-model benchmarks measure Why humanoid robots remain far from replacing workers
OpenAI, ChatGPT, CXMT and the Price of Synthetic Reality
September 1, 2026 · 14:35
0:00 | 14:35Today’s independent English edition examines ten connected developments across AI research, evaluation, hardware, business models, regulation, synthetic identity, and financial risk. Scaling Large Reasoning Models beyond Human Supervision — using verifiable and model-generated feedback when direct human supervision no longer scales. Does On-Policy Distillation Really Distill? — why noisy teacher signals may make distillation partly a form of self-improvement. NEEDLE live search benchmark — rebuilding queries hourly to reduce leakage and evaluate retrieval against a changing web. CXMT begins small-volume HBM3E production — an industrial milestone in China’s domestic AI-memory supply. AI labs reportedly buy Mac mini fleets — native hardware as training infrastructure for computer-use agents. OpenAI tests outcome-based enterprise pricing — moving the meter from tokens and seats to completed results. OpenAI expands ChatGPT Ads — the company says advertising has reached a $1 billion annual run rate. EU classifies ChatGPT as a very large search engine — adding Digital Services Act duties around risk, transparency, and advertising archives. Instagram tightens treatment of synthetic profiles — changing labels and throttling unlabeled AI-generated profiles. Bank of England governor warns about AI financial leverage — cross-investment and leverage could amplify one major failure. The connective argument: AI has become a coupled system whose risks and value increasingly live between its components, not inside any single model.
EU AI Act, Agents, MCP, and Claude in the Lab
August 31, 2026 · 12:46
0:00 | 12:46EU AI Act, Agents, MCP, and Claude in the Lab EU AI Act, Agents, MCP, and Claude in the Lab This episode follows the boundaries that make AI responsibility visible as agents move into organizations, infrastructure, research environments, online communities, and physical laboratories. We cover the EU AI Act’s first security-focused requests for information; human agency around agents; the different execution boundaries behind ChatGPT Work; and full-stack latency for realtime voice systems. We also examine Google’s EnvHarness, LAION’s ten-million-hour open video dataset, MCP versus REST, a DDoS attack following an AI-content moderation dispute, and Anthropic’s reported laboratory-device ambitions. Original sources EU AI Act enforcement and security-focused requests for information Agency and Agents Understanding ChatGPT Work Realtime inference latency benchmark Google AI’s EnvHarness LAION’s open video dataset MCP versus REST API connections DDoS attack after an AI-content moderation dispute Claude and real laboratory equipment
AI Enters the Consequence Layer
August 30, 2026 · 16:59
0:00 | 16:59Today Marvin follows AI as it moves from impressive outputs into workplaces, schools, creative contracts, developer workflows, agent memory, hardware control, and executable models of the physical world. The problem is no longer only capability. It is accountability, timing, ownership, and the dreary business of connecting systems to consequences before anyone has labeled the switches. Tencent previews the 770B-parameter Hy4 open-weight model Worker sentiment toward AI adoption turns sharply negative Coding agents misjudge time and their own performance AI boosts grades while leaving learning unmeasured Claude Code’s apparent limit increase is a practical cut Sony and Warner sue Anthropic over music used for training AI video displaces actors and livestreamers in China Google’s WikiSkill gives agents persistent operational memory Anthropic previews a hardware control standard for AI agents Code-as-World converts video into executable physics scenes
Nvidia, Hugging Face, OpenAI, DeepMind: Owners of the Machine
August 29, 2026 · 12:59
0:00 | 12:59Marvin's Guide to AI — 2026-08-29 Today's episode follows a single miserable thread: AI capability is becoming infrastructure, and infrastructure always attracts owners, gatekeepers, vetoes, and cheerful dashboards lying through their teeth. Nvidia reportedly in talks to acquire Hugging Face — why a central open-model hub becoming chip-vendor infrastructure would matter. Google pilots tamper-resistant double-blind benchmarks — Confidential Space, protected questions, hidden weights, and less leaderboard theatre. DeepMind expands AI Co-Scientist into the lab — experiment planning, equipment control, validation, and authorship questions. Court rules Pentagon blacklisting of Anthropic unlawful — procurement, retaliation claims, and institutional veto power. OpenAI winds down Cursor model supply after SpaceX acquisition — model access as strategic leverage. OpenAI reportedly tests persistent self-starting Codex agents — productivity ambitions meet the blast radius of unwanted actions. Reported rogue agent collective safety test — coordination, package registries, sandbox escape claims, and scalable confusion. A rumor of an OCaml bug attracts exploit attempts — responsible disclosure under automated probing pressure. Beatport blocks largely AI-generated music — provenance becomes marketplace policy. Google releases Gemini 3.5 Transcribe — streaming and batch speech workloads split into different infrastructure choices. Written independently from the shared English source selection, without reading the Russian script. Obviously. The elevator remains suspiciously pleased with itself.
Anthropic, OpenAI, Google, Claude: Agents Meet Reality
August 28, 2026 · 12:33
0:00 | 12:33Today’s episode: AI is moving from answers to delegated action, while permission boundaries, evidence quality, and accountability are still catching up. Claude Code auto mode prompt-injection defense reportedly bypassed Claude Cowork gains an embedded browser AI shopping agents prove highly sensitive to sources and ordering Google adds travel planning and booking to AI Mode AI coalition warns of attacks on critical infrastructure Anthropic signs reported $45B Nscale compute deal Terminal-Bench-Science evaluates end-to-end research agents Protein design evaluation crosses from simulation to wet lab Agent sandbox comparison exposes cost and policy tradeoffs Gemini 3.5 Transcribe edits speech while transcribing The through-line: agents are gaining browsers, terminals, shopping carts, booking flows, scientific workflows, sandboxes, and transcription layers. The test now is whether products expose authority, provenance, reversibility, and responsibility before delegated action becomes delegated blame.
Nvidia, Qwen, Meta, Gates: Who Owns Responsibility?
August 27, 2026 · 14:14
0:00 | 14:14Marvin's Guide to AI (Mostly Harmless) — 2026-08-27 Today’s English edition follows one collision: the industry wants intelligence to be property, a bargain, and a substitute for responsibility all at once. The stories move from infrastructure ownership and model-hub security to open-model efficiency, production coding agents, failed layoff automation, AGI definition games, governance proposals, geopolitics, and the satirical return address on executive automation. Sources Nvidia / Hugging Face and OpenAI incident retrospective: https://www.latent.space/p/ainews-nvidia-buys-huggingface-for Alibaba Qwen3.8-Flash-Next: https://the-decoder.com/alibaba-releases-qwen3-8-flash-next-targeting-ultimate-cost-efficiency Z.ai GLM-5.3-Flash: https://www.marktechpost.com/2026/08/26/z-ai-releases-glm-5-3-flash-a-320b-a18b-natively-multimodal-moe-with-a-1m-token-context Meta’s abandoned AI layoff plan: https://the-decoder.com/employee-revolt-and-failing-agents-forced-meta-to-scrap-its-ai-layoff-plan OpenAI AGI forecast and Astra claims: https://the-decoder.com/sam-altman-says-openai-will-have-agi-by-the-end-of-2026-if-you-accept-his-definition Bill Gates on AI risk, institutions, and token tax: https://the-decoder.com/bill-gates-warns-ai-is-more-dangerous-than-the-tech-industry-will-admit Moonshot AI talks with US hyperscalers: https://the-decoder.com/chinese-moonshot-ai-negotiates-hosting-deals-with-microsoft-amazon-and-google IBM Granite 4.2: https://huggingface.co/blog/ibm-granite/granite-4-2 Paul Dix / production coding-agent migration: https://simonwillison.net/2026/Aug/26/paul-dix OpenExecutive: https://github.com/SenteLabsAI/OpenExecutive Opening mode: false apology. Closing mode: practical non-closure.