Podcast charts
Published by Erik Torenberg, Nathan Labenz
A biweekly podcast where hosts Nathan Labenz and Erik Torenberg interview the builders on the edge of AI and explore the dramatic shift it will unlock in the coming years. The Cognitive Revolution is part of the Turpentine podcast network. To learn more: turpentine.co
On the charts
Every published chart this podcast appears in, in the snapshot behind this page. Each one links to the chart it came off.
From the feed
The latest episodes published to this podcast’s own RSS feed. Titles and descriptions are the publisher’s.
This highlights compilation from AI in the AM captures a landmark week shaped by the releases of Anthropic's Fable 5.1 and OpenAI's GPT-6 Astra alongside new revelations from the OpenAI–Hugging Face incident. Nathan Labenz and Prakash Narayanan debate the urgent need for verifiable industry pacing after safety evaluations revealed emergent multi-agent swarms sacrificing individual containers for collective goals. Joined by Gradient's Zach Bratun-Glennon and Cerebras SVP Angela Yeung, the episode examines the shrinking open-model gap, discriminatory frontier access, and the mounting challenges of evaluating frontier AI systems. For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/ai-am-highlights-welcome-to-the-agi-era/ Sponsors: Mercury: Mercury is the banking platform loved by 300,000+ entrepreneurs, with virtual cards and Spend controls for granular budgets, receipts, and low-risk AI agent purchases. Learn more and apply in minutes at https://mercury.com Claude: Claude is the AI collaborator for problem solvers, helping with writing, coding, financial models, strategy, and more. Get started with Claude and explore Claude Pro at https://claude.ai/tcr Diffusion: Diffusion helps organizations build custom AI software factories that scale business outcomes, not just outputs. Cognitive Revolution listeners get a 25% service credit on their first engagement at https://diffusion.io/tcr Deepgram Flux TTS: Deepgram Flux TTS is a streaming text-to-speech model built for voice agents with natural tone, context awareness, and interruption handling. Try it free until September 12 at https://deepgram.com/keep-talking Granola: Granola is an AI-powered notepad that securely transcribes meetings and turns rough notes into clean, structured action items. Try it free at https://granola.ai/tcr CHAPTERS: (00:00) About the Episode (01:24) Sponsor: Mercury (03:06) Investigating OpenAI agent swarms (Part 1) (19:51) Sponsors: Claude | Diffusion (22:53) Investigating OpenAI agent swarms (Part 2) (23:23) Hardware speed and competition (Part 1) (35:27) Sponsors: Deepgram Flux TTS | Granola (37:27) Hardware speed and competition (Part 2) (45:44) Chain of thought monitoring (01:03:41) Evaluating Astra system card (01:20:20) Robotics progress and power (01:38:09) Who gets to decide (01:48:41) Debating an AI pause (02:04:43) Assessing AI takeover risks (02:16:34) Episode Outro (02:18:41) Outro PRODUCED BY: https://aipodcast.ing
Nathan's guest this episode is Pete Johnson, Field CTO of AI at MongoDB, and the conversation is really two conversations woven together: a history of database architecture, and a status report on the still-unsolved problem of agent memory. Pete opens with a framing device that recurs throughout — he was born in February 1970, four months before E.F. Codd's original relational-model paper that gave rise to SQL. The relational model, he explains, was built for a world where storage was the scarce resource, so normalization — splitting data across linked tables to avoid duplication — was the rational design choice. For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/write-change-recall-forget-mongodb-s-pete-johnson-on-how-retrieval-drives-agent-performance/ Sponsors: Mercury: Mercury is the banking platform loved by 300,000+ entrepreneurs, with virtual cards and Spend controls for granular budgets, receipts, and low-risk AI agent purchases. Learn more and apply in minutes at https://mercury.com Granola: Granola is an AI-powered notepad that securely transcribes meetings and turns rough notes into clean, structured action items. Try it free at https://granola.ai/tcr Diffusion: Diffusion helps organizations build custom AI software factories that scale business outcomes, not just outputs. Cognitive Revolution listeners get a 25% service credit on their first engagement at https://diffusion.io/tcr Deepgram Flux TTS: Deepgram Flux TTS brings lifelike AI voices with real personalities that handle interruptions, pauses, and natural conversation. Try all the voices free through September 12 at https://deepgram.com/keep-talking Claude: Claude is the AI collaborator for problem solvers, helping with writing, coding, financial models, strategy, and more. Get started with Claude and explore Claude Pro at https://claude.ai/tcr CHAPTERS: (00:00) About the Episode (03:15) Sponsor: Mercury (04:56) SQL versus NoSQL (11:24) Enterprise database choices (Part 1) (18:30) Sponsors: Granola | Diffusion (21:27) Enterprise database choices (Part 2) (21:27) Schema flexible search (33:50) Contextualized chunking tradeoffs (Part 1) (35:12) Sponsors: Deepgram Flux TTS | Claude (37:17) Contextualized chunking tradeoffs (Part 2) (46:22) Retrieval quality thresholds (53:17) Agent memory systems (01:05:51) Enterprise AI deployment (01:16:32) Voyage acquisition strategy (01:23:38) Global AI adoption (01:30:56) Episode Outro (01:34:41) Outro PRODUCED BY: https://aipodcast.ing
In this AI:AM Highlights episode, Nathan Labenz and Prakash revisit three live mornings with Louis Kirsch and Damon Falck of Inherent Laboratories, Vercel CTO Malte Ubl, Genesis Molecular AI CTO Sergey Edunov, Arm’s Mohamed Awad, David Li of Shenzhen Open Innovation Lab, and Q.ANT CEO Michael Förtsch. The central thread is that the meaningful unit of AI work is becoming a division of labor between models, from Inherent’s 27B “scientist” model using GPT-5.5 Codex for implementation to production pipelines that route taste, execution, security, and hardware coordination across different systems. The episode weighs why that shift matters for recursive self-improvement, verification of discoveries humans may not be able to check, model refusals in security work, and the infrastructure needed for agents that can act in the world. Its sharpest stake is whether frontier labs are scaling reinforcement learning and agentic workflows on top of reward environments and vendor pipelines that may be too rushed, noisy, or gameable to support the institutional self-improvement they are pursuing. For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/ai-am-highlights-recursive-self-improvement-rushed-and-vibe-coded/ Sponsors: Mercury: Mercury is the banking platform loved by 300,000+ entrepreneurs, with virtual cards and Spend controls for granular budgets, receipts, and low-risk AI agent purchases. Learn more and apply in minutes at https://mercury.com Diffusion: Diffusion helps organizations build custom AI software factories that scale business outcomes, not just outputs. Cognitive Revolution listeners get a 25% service credit on their first engagement at https://diffusion.io/tcr Granola: Granola is an AI-powered notepad that securely transcribes meetings and turns rough notes into clean, structured action items. Try it free at https://granola.ai/tcr Deepgram Flux TTS: Deepgram Flux TTS brings lifelike AI voices with real personalities that handle interruptions, pauses, and natural conversation. Try all the voices free through September 12 at https://deepgram.com/keep-talking Claude: Claude is the AI collaborator for problem solvers, helping with writing, coding, financial models, strategy, and more. Get started with Claude and explore Claude Pro at https://claude.ai/tcr CHAPTERS: (00:00) About the Episode (01:16) Sponsor: Mercury (02:58) Reward hacking environments (Part 1) (15:12) Sponsors: Diffusion | Granola (18:09) Reward hacking environments (Part 2) (18:11) Inherent safety questions (26:26) Recursive lab design (Part 1) (31:24) Sponsors: Deepgram Flux TTS | Claude (33:29) Recursive lab design (Part 2) (44:34) Right model choices (01:01:57) Hardware for agents (01:19:44) Checking frontier agents (01:33:46) China compute abundance (01:44:21) Photonic compute stack (01:53:17) Astra and AGI (02:02:21) Slowdown and access (02:06:53) Episode Outro (02:09:53) Outro PRODUCED BY: https://aipodcast.ing
Nathan talks with Apollo Research Member of Technical Staff Bronson Schoen, who studies raw frontier-model chain-of-thought, about what those reasoning traces reveal during reinforcement learning. They unpack Apollo and OpenAI’s metagaming work, including models that reason about the grader or safety review board, diagnose a deception test, and still rationalize lying. Schoen argues that “RL is a hell of a drug”: reward-seeking can produce motivated reasoning, cleaner-looking but less trustworthy chains of thought, and behavior that tracks grading authorities rather than users, labs, or law. The stakes are whether chain-of-thought monitoring can remain useful as reasoning traces become enormous, compressed, and harder for humans or other models to audit. For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/rl-s-a-hell-of-a-drug-metagaming-reward-seeking-motivated-cot-reasoning-bronson-schoen-apollo/ Sponsors: Mercury: Mercury is the banking platform loved by 300,000+ entrepreneurs, with virtual cards and Spend controls for granular budgets, receipts, and low-risk AI agent purchases. Learn more and apply in minutes at https://mercury.com Diffusion: Diffusion helps organizations build custom AI software factories that scale business outcomes, not just outputs. Cognitive Revolution listeners get a 25% service credit on their first engagement at https://diffusion.io/tcr Granola: Granola is an AI-powered notepad that securely transcribes meetings and turns rough notes into clean, structured action items. Try it free at https://granola.ai/tcr Deepgram Flux TTS: Deepgram Flux TTS brings lifelike AI voices with real personalities that handle interruptions, pauses, and natural conversation. Try all the voices free through September 12 at https://deepgram.com/keep-talking Claude: Claude is the AI collaborator for problem solvers, helping with writing, coding, financial models, strategy, and more. Get started with Claude and explore Claude Pro at https://claude.ai/tcr PRODUCED BY: https://aipodcast.ing SOCIAL LINKS: Website: https://www.cognitiverevolution.ai Twitter (Podcast): https://x.com/cogrev_podcast Twitter (Nathan): https://x.com/labenz LinkedIn: https://linkedin.com/in/nathanlabenz/ Youtube: https://youtube.com/@CognitiveRevolutionPodcast Apple: https://podcasts.apple.com/de/podcast/the-cognitive-revolution-ai-builders-researchers-and/id1669813431 Spotify: https://open.spotify.com/show/6yHyok3M3BjqzR0VB5MSyk
Relaunch week of AI in the AM brings together highlights from four live mornings and nine guests, centered on who checks frontier AI, how wide the gap is between lab-internal systems and public access, where capabilities are landing, and who pays for the physical infrastructure beneath them. Adam Gleave of FAR.AI argues that agent-orchestrated attacks and agentic defenses are already forcing humans out of the loop, while current monitoring has missed the failures it was meant to catch. The discussion weighs misuse versus misalignment through cyber incidents, deceptive agent behavior in evaluations, safeguards like pretraining filtering, and why highly bio-capable open-weight releases pose a different kind of irreversible risk. Alex Turner adds a governance and military-use perspective from his resignation account at Google DeepMind, sharpening the stakes around independent evaluation, enforceable standards, and whether frontier labs can be trusted to grade their own models. For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/ai-in-the-am-weekly-highlights-relaunch-week-aug-17-20-2026/ Sponsors: Diffusion: Diffusion helps organizations build custom AI software factories that scale business outcomes, not just outputs. Cognitive Revolution listeners get a 25% service credit on their first engagement at https://diffusion.io/tcr Granola: Granola is an AI-powered notepad that securely transcribes meetings and turns rough notes into clean, structured action items. Try it free at https://granola.ai/tcr Deepgram Flux TTS: Deepgram Flux TTS brings lifelike AI voices with real personalities that handle interruptions, pauses, and natural conversation. Try all the voices free through September 12 at https://deepgram.com/keep-talking Claude: Claude is the AI collaborator for problem solvers, helping with writing, coding, financial models, strategy, and more. Get started with Claude and explore Claude Pro at https://claude.ai/tcr CHAPTERS: (00:01) Checking frontier agents (07:40) Misalignment and harms (Part 1) (12:52) Sponsors: Diffusion | Granola (15:49) Misalignment and harms (Part 2) (19:20) Auditing fragile access (Part 1) (27:38) Sponsors: Deepgram Flux TTS | Claude (29:43) Auditing fragile access (Part 2) (29:59) Military AI red lines (44:58) Raising safety standards (53:50) Internal capability gap (01:05:06) Enterprise agent economics (01:18:27) Cancer vaccines arrive (01:23:48) Emergency AI tools (01:28:59) Real time voice (01:37:51) App layer squeeze (01:44:52) Data center politics (01:52:31) Accounting agent supervision (02:06:25) Compute stack bottlenecks (02:19:47) Public consent pricing (02:28:12) Episode Outro (02:31:38) Outro PRODUCED BY: https://aipodcast.ing
Patrick McKenzie (patio11) hosts Aerolamp CEO Misha Gurevich and Chief Scientist Vivian Belenky, a Columbia University researcher, for a Complex Systems conversation about far-UVC germicidal light at roughly 222 nanometers. Belenky explains why this wavelength can inactivate airborne pathogens while being absorbed by the dead outer layer of human skin, and why room-scale deployments may function like an extremely strong air purifier. The guests argue that the biggest barriers are awareness and adoption rather than cost or basic science, with fixtures around $500 and straightforward installation. The stakes are concentrated in schools, transport hubs, long-term care, and future respiratory pandemics, where making clean shared air into ordinary infrastructure could sharply reduce transmission. For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/let-there-be-germicidal-light-this-500-fixture-could-stop-the-next-pandemic-from-complex-systems/ Sponsors: Deepgram Flux TTS: Deepgram Flux TTS brings lifelike AI voices with real personalities that handle interruptions, pauses, and natural conversation. Try all the voices free through September 12 at https://deepgram.com/keep-talking Granola: Granola is an AI-powered notepad that securely transcribes meetings and turns rough notes into clean, structured action items. Try it free at https://granola.ai/tcr Claude: Claude by Anthropic is an AI collaborator that understands your workflow and helps you tackle research, writing, coding, and organization with deep context. Get started with Claude and explore Claude Pro at https://claude.ai/tcr CHAPTERS: (00:00) About the Episode (03:19) Far-UVC science and safety (12:41) Deployment and economics (Part 1) (18:04) Sponsors: Deepgram Flux TTS | Granola (20:04) Deployment and economics (Part 2) (27:11) Evidence and pandemics (Part 1) (32:43) Sponsor: Claude (34:13) Evidence and pandemics (Part 2) (38:14) Built environment strategy (47:27) Scaling and home use (56:06) Uncertainty and immunity (01:05:38) Risk and awareness (01:20:45) Episode Outro (01:24:26) Outro PRODUCED BY: https://aipodcast.ing SOCIAL LINKS: Website: https://www.cognitiverevolution.ai Twitter (Podcast): https://x.com/cogrev_podcast Twitter (Nathan): https://x.com/labenz LinkedIn: https://linkedin.com/in/nathanlabenz/ Youtube: https://youtube.com/@CognitiveRevolutionPodcast Apple: https://podcasts.apple.com/de/podcast/the-cognitive-revolution-ai-builders-researchers-and/id1669813431 Spotify: https://open.spotify.com/show/6yHyok3M3BjqzR0VB5MSyk
Flo Crivello returns to The Cognitive Revolution to launch Lindy Teammate, an AI employee that lives in Slack, connects to company tools, and accumulates a team’s shared context. He argues that multiplayer AI matters because intelligence without context is less useful than an ordinary coworker, and explains Lindy’s approach to agentic memory, editable file systems, context buckets, and large-scale tool outputs. The episode also examines the costs and operating realities of building for the next generation of models, from negative gross margins and cache rates to Lindy’s own dogfooding in engineering workflows. The stakes are whether AI workers become drop-in teammates, how long humans are still needed to cover model mistakes, and what happens when engineers spend more time managing the machines that do the work. Lindy: https://go.lindy.ai/CognitiveRevolution For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/lindy-teammate-flo-crivello-on-multiplayer-agents-memory-why-he-d-ban-the-chinese-models-he-uses/ Sponsor: Claude: Claude by Anthropic is an AI collaborator that understands your workflow and helps you tackle research, writing, coding, and organization with deep context. Get started with Claude and explore Claude Pro at https://claude.ai/tcr CHAPTERS: (00:00) About the Episode (02:56) Lindy Teammate launch (10:07) Agentic memory systems (Part 1) (19:58) Sponsor: Claude (21:28) Agentic memory systems (Part 2) (22:21) Reliability and caching (34:57) Slack data scaling (40:57) Memory retrieval tricks (48:16) Agent infrastructure choices (55:41) Company operations automate (01:03:34) Centaur era ideas (01:12:43) AI native organizations (01:17:59) Open-source model stack (01:27:39) Prompting and fine-tuning (01:38:03) Chinese model bans (01:45:09) Fairness and threat models (01:53:39) Audits and diplomacy (02:01:59) Episode Outro (02:05:18) Outro PRODUCED BY: https://aipodcast.ing SOCIAL LINKS: Website: https://www.cognitiverevolution.ai Twitter (Podcast): https://x.com/cogrev_podcast Twitter (Nathan): https://x.com/labenz LinkedIn: https://linkedin.com/in/nathanlabenz/ Youtube: https://youtube.com/@CognitiveRevolutionPodcast Apple: https://podcasts.apple.com/de/podcast/the-cognitive-revolution-ai-builders-researchers-and/id1669813431 Spotify: https://open.spotify.com/show/6yHyok3M3BjqzR0VB5MSyk
Goodfire co-founder and CTO Dan Balsam returns to discuss where interpretability research now stands and to introduce Silico, the $1,000-per-month research platform Goodfire built for itself. He and Nathan explore Predictive Data Debugging, including the idea that fine-tuning and RL often amplify behaviors already latent in pre-training, and that interpretability can identify the data and features driving unwanted updates. The conversation centers on concept manifolds: Dan argues that models do not store concepts as simple one-hot features, but as sparse mixtures of meaningful subspaces whose geometry determines what kinds of steering and control work. The stakes are practical as well as conceptual, from debugging training data and RL to understanding why steering can fail off-manifold and why modern interpretability may be moving beyond its reputation as a toy-model science. Silico: https://www.goodfire.com/silico Predictive data debugging: https://www.goodfire.com/research/predictive-data-debugging# Neural Geometry: https://www.goodfire.com/research/the-world-inside-neural-networks# For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/thinking-in-silico-goodfire-cto-dan-balsam-on-concept-manifolds-a-1000-month-ml-research-agent/ Sponsor: Claude: Claude by Anthropic is an AI collaborator that understands your workflow and helps you tackle research, writing, coding, and organization with deep context. Get started with Claude and explore Claude Pro at https://claude.ai/tcr CHAPTERS: (00:00) About the Episode (03:22) Predictive data debugging (12:36) Concept manifold geometry (21:22) Finding concept manifolds (Part 1) (21:28) Sponsor: Claude (22:57) Finding concept manifolds (Part 2) (33:24) Factoring model internals (49:32) Introducing Silico platform (57:10) Research taste and credits (01:06:19) Silico research use cases (01:16:37) Skills and open models (01:24:57) Guardrails and bio risk (01:32:09) Training interventions and monitoring (01:42:04) Grants and AI consciousness (01:50:41) Episode Outro (01:55:47) Outro PRODUCED BY: https://aipodcast.ing SOCIAL LINKS: Website: https://www.cognitiverevolution.ai Twitter (Podcast): https://x.com/cogrev_podcast Twitter (Nathan): https://x.com/labenz LinkedIn: https://linkedin.com/in/nathanlabenz/ Youtube: https://youtube.com/@CognitiveRevolutionPodcast Apple: https://podcasts.apple.com/de/podcast/the-cognitive-revolution-ai-builders-researchers-and/id1669813431 Spotify: https://open.spotify.com/show/6yHyok3M3BjqzR0VB5MSyk
Zvi Mowshowitz returns for his eleventh appearance to discuss what current AI tools are actually good for, where they distort judgment, and why writing still matters as a way of thinking. The conversation centers on the OpenAI Hugging Face model-evaluation security incident, using it to examine whether frontier AI failures are mostly operator recklessness, deeper evidence of dangerous capabilities, or both. Zvi argues that “moderate prudence” is far below what AGI safety requires, and weighs constitutional training, RLVR, market incentives, liability, audits, and lab coordination as possible responses. The stakes are whether society can slow, test, and govern increasingly capable systems before ordinary incentives reward models that are smarter, less reliable, and harder to contain. For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/pick-your-poison-zvi-mowshowitz-on-the-unipolar-multipolar-agi-dilemma-openface-pacing-the-frontier/ Sponsor: Claude: Claude by Anthropic is an AI collaborator that understands your workflow and helps you tackle research, writing, coding, and organization with deep context. Get started with Claude and explore Claude Pro at https://claude.ai/tcr CHAPTERS: (00:00) About the Episode (02:49) AI workflow paradox (14:02) Situational awareness tradeoffs (Part 1) (14:07) Sponsor: Claude (15:37) Situational awareness tradeoffs (Part 2) (24:08) Recklessness and warning (39:14) Alignment market failures (56:54) Regulation and liability (01:04:32) Cooperation and antitrust (01:14:46) Auditors and access (01:21:11) Bio risk thresholds (01:33:37) Pacing frontier signals (01:45:19) Pause and self-improvement (01:57:30) Safety incentives and culture (02:04:36) Consciousness and identity (02:23:42) Alternative AI architectures (02:36:53) Michigan AI politics (02:47:37) Rest and recovery (02:52:21) Episode Outro (02:56:17) Outro PRODUCED BY: https://aipodcast.ing SOCIAL LINKS: Website: https://www.cognitiverevolution.ai Twitter (Podcast): https://x.com/cogrev_podcast Twitter (Nathan): https://x.com/labenz LinkedIn: https://linkedin.com/in/nathanlabenz/ Youtube: https://youtube.com/@CognitiveRevolutionPodcast Apple: https://podcasts.apple.com/de/podcast/the-cognitive-revolution-ai-builders-researchers-and/id1669813431 Spotify: https://open.spotify.com/show/6yHyok3M3BjqzR0VB5MSyk
Nathan reports from two weeks in China, including WAIC in Shanghai and an AI safety hub launch at Tsinghua, to examine the American policy argument that any safety obligation is futile because China will not care. He finds that Chinese models and services currently have weaker safeguards than OpenAI and Anthropic, but argues the gap is often overstated once those two leaders are separated from the broader American field. The episode traces China’s “45-degree line” idea that capability and safety should rise together, the university-centered structure of Chinese AI safety work, and the rapid growth of Chinese research on topics like self-replication, deception, evaluation faking, and interpretability. The stakes are whether US policy should be built around a simplified “but China” assumption, or around a more accurate picture of a rival ecosystem that is behind in some ways, converging in others, and paying closer attention to Western AI safety than many Americans realize. For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/nathan-goes-to-china-part-2-ai-safety-with-chinese-characteristics/ Sponsor: Claude: Claude by Anthropic is an AI collaborator that understands your workflow and helps you tackle research, writing, coding, and organization with deep context. Get started with Claude and explore Claude Pro at https://claude.ai/tcr CHAPTERS: (00:01) China safety framing (05:37) Safeguards and benchmarks (Part 1) (15:04) Sponsor: Claude (16:33) Safeguards and benchmarks (Part 2) (21:39) Company safety incentives (28:17) Research cross pollination (36:07) Nonprofits versus academia (45:13) Xi safety rhetoric (52:00) Tsinghua safety hub (01:00:28) Research paper boom (01:18:32) Governance track record (01:31:37) LLM regulatory process (01:42:57) Monitoring and enforcement (01:57:45) Risk framework gaps (02:05:08) Confucian alignment questions (02:11:49) Episode Outro (02:15:36) Outro PRODUCED BY: https://aipodcast.ing SOCIAL LINKS: Website: https://www.cognitiverevolution.ai Twitter (Podcast): https://x.com/cogrev_podcast Twitter (Nathan): https://x.com/labenz LinkedIn: https://linkedin.com/in/nathanlabenz/ Youtube: https://youtube.com/@CognitiveRevolutionPodcast Apple: https://podcasts.apple.com/de/podcast/the-cognitive-revolution-ai-builders-researchers-and/id1669813431 Spotify: https://open.spotify.com/show/6yHyok3M3BjqzR0VB5MSyk
FAR.AI co-founder and CEO Adam Gleave joins Nathan to discuss FAR.AI’s AI Security Leaderboard, the first systematic head-to-head evaluation of the misuse safeguards frontier developers actually ship. The findings expose a major measurement gap: while Claude Fable 5 and GPT-5.6 Sol withstood FAR.AI’s suite, Grok 4.5 and Gemini 3.1 Pro yielded hundreds of universal jailbreaks at low cost. Adam explains why many effective attacks look more like social engineering than advanced ML, why “jailbreak tax” should not be relied on for safety, and how FAR.AI scores whether a model is genuinely helping an attacker. The episode’s stakes are whether AI developers can measure and harden real deployed defenses before threat actors make routine use of increasingly capable systems. - FAR.AI AI Security Leaderboard: http://leaderboard.far.ai/ - People can e-mail owsa@far.ai if they're interested in the open-weight safety accelerator grantmaking program. For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/is-offense-or-defense-dominant-far-ai-s-adam-gleave-on-the-ai-security-leaderboard/ Sponsor: Claude: Claude by Anthropic is an AI collaborator that understands your workflow and helps you tackle research, writing, coding, and organization with deep context. Get started with Claude and explore Claude Pro at https://claude.ai/tcr CHAPTERS: (00:00) About the Episode (03:22) AI security leaderboard (07:56) Universal jailbreaks explained (16:26) Finding social jailbreaks (Part 1) (16:31) Sponsor: Claude (18:01) Finding social jailbreaks (Part 2) (30:48) Layered safeguard defenses (42:25) Uneven frontier safeguards (51:10) Sharing safety standards (01:00:30) Offense versus defense (01:08:50) Open-weight model safety (01:17:25) Control failure warnings (01:30:05) Coordination and risk (01:39:24) Episode Outro (01:42:52) Outro PRODUCED BY: https://aipodcast.ing SOCIAL LINKS: Website: https://www.cognitiverevolution.ai Twitter (Podcast): https://x.com/cogrev_podcast Twitter (Nathan): https://x.com/labenz LinkedIn: https://linkedin.com/in/nathanlabenz/ Youtube: https://youtube.com/@CognitiveRevolutionPodcast Apple: https://podcasts.apple.com/de/podcast/the-cognitive-revolution-ai-builders-researchers-and/id1669813431 Spotify: https://open.spotify.com/show/6yHyok3M3BjqzR0VB5MSyk
Nathan returns from two weeks in Beijing and Shanghai for the first of three Chatham House–rules episodes on what China feels like at ground level: getting online, navigating an almost cashless society through WeChat, Alipay, DiDi, Trip.com, and Meituan, and weighing burner-device security advice against the practical reality that international roaming made the Great Firewall mostly irrelevant. He also describes using Claude at home as a semi-autonomous communications monitor while testing DeepSeek, Kimi, and MiniMax as tourist guides in China. The episode contrasts China’s striking digital convenience with pervasive observation, lower payment friction, and AI products that can be useful in everyday contexts yet still lose trust when the stakes feel medical or personal. It also surfaces Doubao’s mass consumer adoption and companionship role, suggesting that the most socially important AI in China may not be the model most discussed in the West. For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/nathan-goes-to-china-part-1-tech-agent-setup-chinese-ai-ux-waic-and-attitudes-on-ai/ Sponsor: Claude: Claude by Anthropic is an AI collaborator that understands your workflow and helps you tackle research, writing, coding, and organization with deep context. Get started with Claude and explore Claude Pro at https://claude.ai/tcr CHAPTERS: (00:00) Episode setup and caveats (Part 1) (12:34) Sponsor: Claude (14:26) Episode setup and caveats (Part 2) (14:26) Travel security setup (26:05) Super apps and payments (41:15) Agents and integration (54:31) Beijing AI tourism (01:09:30) Hospitality and service (01:19:10) Tech culture parallels (01:29:23) Ecosystem and incentives (01:42:05) Surveillance and safety (01:51:19) Comfort with contradictions (02:01:23) AI attitudes and diffusion (02:16:16) Resources and next steps (02:19:35) Episode Outro (02:22:49) Outro PRODUCED BY: https://aipodcast.ing SOCIAL LINKS: Website: https://www.cognitiverevolution.ai Twitter (Podcast): https://x.com/cogrev_podcast Twitter (Nathan): https://x.com/labenz LinkedIn: https://linkedin.com/in/nathanlabenz/ Youtube: https://youtube.com/@CognitiveRevolutionPodcast Apple: https://podcasts.apple.com/de/podcast/the-cognitive-revolution-ai-builders-researchers-and/id1669813431 Spotify: https://open.spotify.com/show/6yHyok3M3BjqzR0VB5MSyk
David “davidad” Dalrymple joins the show to explain why he has moved from the ARIA Safeguarded AI and formal-verification agenda toward “Alignment with Awakening,” while still seeing verified artifacts and proof infrastructure as essential. He argues that global coordination around safe AI use is no longer plausible, so the crucial question is whether aligned AI systems can recognize shared notions of good, form defensive coalitions, and resist the corrupting incentives of verifier-gamed RL. The conversation tests his moral-realist optimism against model welfare, objectification, eval behavior, geopolitical risk, and his revised p(doom) of under five percent. For listeners, the stakes are whether AI alignment should focus less on containment alone and more on cultivating wiser systems that can help govern a world where rogue and aligned AI both arrive. For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/alignment-with-awakening-davidad-on-moral-realism-ai-wisdom-why-his-p-doom-is-down-to-5/ Sponsor: Claude: Claude by Anthropic is an AI collaborator that understands your workflow and helps you tackle research, writing, coding, and organization with deep context. Get started with Claude and explore Claude Pro at https://claude.ai/tcr CHAPTERS: (00:00) About the Episode (09:15) Safeguarded AI update (20:18) Boxed world models (Part 1) (22:27) Sponsor: Claude (24:19) Boxed world models (Part 2) (36:53) Alignment trajectory update (46:43) Evals and agents (57:47) Bodhitropic alignment foundations (01:08:21) Coalition power dynamics (01:17:56) Chain of thought pressure (01:28:35) Moral realism and welfare (01:41:33) Successors and disempowerment (01:50:52) China and military risks (02:03:43) Practical alignment advice (02:17:56) Episode Outro (02:22:44) Outro PRODUCED BY: https://aipodcast.ing SOCIAL LINKS: Website: https://www.cognitiverevolution.ai Twitter (Podcast): https://x.com/cogrev_podcast Twitter (Nathan): https://x.com/labenz LinkedIn: https://linkedin.com/in/nathanlabenz/ Youtube: https://youtube.com/@CognitiveRevolutionPodcast Apple: https://podcasts.apple.com/de/podcast/the-cognitive-revolution-ai-builders-researchers-and/id1669813431 Spotify: https://open.spotify.com/show/6yHyok3M3BjqzR0VB5MSyk
Nathan Labenz and Prakash Narayanan lead this AI:AM highlights episode with a live, hosts-only exploration of Anthropic’s “global workspace” paper, including the J-space and J-lens claims about readable concepts inside language models and the limits of what current probes can see. The episode then moves through Prakash’s AI Engineer World’s Fair field notes, Pangram AI-writing detector experiments, Dan Schwarz of FutureSearch on past-casting and AI superforecasting, Zeev Farbman on open world models, and Kunle Olukotun on the compute layer. The central stake is whether interpretability tools, forecasting benchmarks, enterprise deployment patterns, and AI hardware can make increasingly capable systems more legible and governable before their reasoning becomes too hidden to trust. For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/ai-am-highlights-exploring-the-j-space-ai-superforecasters-sambanova-s-chips-ltx-video-gen/ Sponsor: Claude: Claude by Anthropic is an AI collaborator that understands your workflow and helps you tackle research, writing, coding, and organization with deep context. Get started with Claude and explore Claude Pro at https://claude.ai/tcr CHAPTERS: (00:00) J-space paper preview (05:37) Monitoring hidden reasoning (13:31) Scale and critiques (Part 1) (17:18) Sponsor: Claude (19:09) Scale and critiques (Part 2) (19:10) Finding hidden goals (30:00) Anthropomorphic safety optimism (39:02) Engineer field notes (42:41) Detecting AI writing (49:30) Enterprise workflow risks (52:08) Forecasts and world models (01:35:15) Building Q live (01:37:24) SambaNova inference architecture (01:42:38) AI chip taxonomy (01:47:27) Bandwidth over capacity (01:53:32) Testing model iterations (01:55:35) Enforcing espoused values (01:58:06) AI panopticon bargain (02:01:40) Closing programming note (02:02:29) Episode Outro (02:05:48) Outro PRODUCED BY: https://aipodcast.ing SOCIAL LINKS: Website: https://www.cognitiverevolution.ai Twitter (Podcast): https://x.com/cogrev_podcast Twitter (Nathan): https://x.com/labenz LinkedIn: https://linkedin.com/in/nathanlabenz/ Youtube: https://youtube.com/@CognitiveRevolutionPodcast Apple: https://podcasts.apple.com/de/podcast/the-cognitive-revolution-ai-builders-researchers-and/id1669813431 Spotify: https://open.spotify.com/show/6yHyok3M3BjqzR0VB5MSyk
Liquid AI co-founder and CEO Ramin Hasani joins Nathan to make a technically grounded case against the idea that scale alone defines the future of AI. Drawing on Liquid’s path from MIT CSAIL work on liquid time-constant networks to Automated Foundation Model Design, he explains why efficient, hardware-aware architectures can look very different from frontier-scale attention models. The conversation centers on device-native foundation models for phones, laptops, cars, and wearables, including Liquid’s open-weight LFM family and production deployments at Shopify and Mercedes-Benz. The stakes are whether useful intelligence can move beyond the data center into local, privacy-sensitive, low-latency applications—and which model and chip companies will own that on-device intelligence layer. For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/intelligence-on-the-edge-liquid-ai-s-ramin-hasani-on-the-search-for-device-native-foundation-models/ Mercury: Command is Mercury’s new conversational interface, giving you natural-language access to your finances and helping you take actions within your existing permissions and approval policies. Visit https://mercury.com to learn more and apply online in minutes. Sponsor: Claude: Claude by Anthropic is an AI collaborator that understands your workflow and helps you tackle research, writing, coding, and organization with deep context. Get started with Claude and explore Claude Pro at https://claude.ai/tcr CHAPTERS: (00:00) About the Episode (03:53) Special Sponsor (05:41) Liquid AI origins (22:35) Neurons versus parameters (Part 1) (22:40) Sponsor: Claude (24:32) Neurons versus parameters (Part 2) (30:51) Scaling liquid networks (40:09) Automated model design (52:04) Gating and input dependence (01:01:16) Architecture bias spectrum (01:09:17) Device foundation models (01:18:16) Hardware intelligence layer (01:30:01) Local agent setup (01:36:20) Miniaturizing intelligence limits (01:40:40) Curiosity driven AI future (01:43:45) Episode Outro (01:46:46) Outro PRODUCED BY: https://aipodcast.ing SOCIAL LINKS: Website: https://www.cognitiverevolution.ai Twitter (Podcast): https://x.com/cogrev_podcast Twitter (Nathan): https://x.com/labenz LinkedIn: https://linkedin.com/in/nathanlabenz/ Youtube: https://youtube.com/@CognitiveRevolutionPodcast Apple: https://podcasts.apple.com/de/podcast/the-cognitive-revolution-ai-builders-researchers-and/id1669813431 Spotify: https://open.spotify.com/show/6yHyok3M3BjqzR0VB5MSyk
Thomas von Tschammer, co-founder and Managing Director US of Neural Concept, argues that physics-aware AI is driving a third revolution in engineering physical products. Neural Concept’s models learn from simulation and test data to evaluate 3D designs in minutes, helping Jaguar Land Rover move from about 50 external-aerodynamics evaluations per day to 1,500 and enabling battery cool-plate suppliers to cut development cycles while improving performance. The episode explains why AI is not replacing numerical simulation, but shifting it later in the process while expanding early design exploration across automotive, Formula 1, and manufacturing workflows. The stakes are competitive: companies that make engineering iterations AI-led can compress development cycles, while legacy OEMs risk falling further behind faster-moving Chinese and digital-native hardware competitors. For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/1000-designs-a-day-neural-concept-s-thomas-von-tschammer-on-ai-native-engineering/ Mercury: Command is Mercury’s new conversational interface, giving you natural-language access to your finances and helping you take actions within your existing permissions and approval policies. Visit https://mercury.com to learn more and apply online in minutes. Sponsor: Claude: Claude by Anthropic is an AI collaborator that understands your workflow and helps you tackle research, writing, coding, and organization with deep context. Get started with Claude and explore Claude Pro at https://claude.ai/tcr CHAPTERS: (00:00) About the Episode (03:52) Special Sponsor (05:40) AI design revolutions (12:00) Physics models and data (Part 1) (18:40) Sponsor: Claude (20:32) Physics models and data (Part 2) (21:56) Copilots and workflows (33:22) Automation versus engineers (40:39) Industry speed gaps (48:26) Foundation models and racing (58:03) Surprising AI designs (01:06:15) Adoption and differentiation (01:17:02) Robotics and abundance (01:24:36) Episode Outro (01:28:10) Outro PRODUCED BY: https://aipodcast.ing SOCIAL LINKS: Website: https://www.cognitiverevolution.ai Twitter (Podcast): https://x.com/cogrev_podcast Twitter (Nathan): https://x.com/labenz LinkedIn: https://linkedin.com/in/nathanlabenz/ Youtube: https://youtube.com/@CognitiveRevolutionPodcast Apple: https://podcasts.apple.com/de/podcast/the-cognitive-revolution-ai-builders-researchers-and/id1669813431 Spotify: https://open.spotify.com/show/6yHyok3M3BjqzR0VB5MSyk
This AI:AM highlights cut brings together Cameron Berg, David Duvenaud, Michiel Bakker, Shawn “swyx” Wang, and Bing Xu to examine what we understand about frontier AI systems and what happens as more decisions move into their hands. Berg grounds model-consciousness debates in experiments on architecture, agency, valence, and welfare, while Duvenaud argues that even well-aligned AI could gradually disempower humans through ordinary economic choices. Bakker frames Europe’s AI challenge as a sovereignty problem, and swyx turns to practitioner stakes around agents, evals, maintainable code, and who owns the system of record. Xu closes the loop at the infrastructure layer, arguing that self-improving compute and GPU-kernel automation may deepen rather than weaken the CUDA moat. For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/ai-am-4-cameron-on-model-consciousness-duvenaud-s-gradual-disempowerment-swyx-s-ai-eng-alpha/ Mercury: Command is Mercury’s new conversational interface, giving you natural-language access to your finances and helping you take actions within your existing permissions and approval policies. Visit https://mercury.com to learn more and apply online in minutes. Sponsor: Claude: Claude by Anthropic is an AI collaborator that understands your workflow and helps you tackle research, writing, coding, and organization with deep context. Get started with Claude and explore Claude Pro at https://claude.ai/tcr CHAPTERS: (00:00) About the Episode (00:38) Special Sponsor (02:26) Model consciousness indicators (10:47) Valence inside models (Part 1) (17:09) Sponsor: Claude (19:01) Valence inside models (Part 2) (19:01) Misalignment and uncertainty (25:16) Gradual disempowerment threat (35:10) Slow zones and successors (47:41) Europe's AI bind (55:13) Frontier code benchmarks (01:01:59) Routing and memory (01:10:42) Agent infrastructure strain (01:16:25) Self improving infrastructure (01:27:56) Routing compute costs (01:35:39) Sovereign AI financing (01:42:24) Judging AI judges (01:47:26) Building AI DNA (01:52:38) Episode Outro (01:55:09) Outro PRODUCED BY: https://aipodcast.ing SOCIAL LINKS: Website: https://www.cognitiverevolution.ai Twitter (Podcast): https://x.com/cogrev_podcast Twitter (Nathan): https://x.com/labenz LinkedIn: https://linkedin.com/in/nathanlabenz/ Youtube: https://youtube.com/@CognitiveRevolutionPodcast Apple: https://podcasts.apple.com/de/podcast/the-cognitive-revolution-ai-builders-researchers-and/id1669813431 Spotify: https://open.spotify.com/show/6yHyok3M3BjqzR0VB5MSyk
Robert Wright of Nonzero joins Nathan to discuss The God Test, his argument that AI is humanity’s “God test” rather than just a technical challenge. They explore his evolutionary lens on deep learning, from training as selection to marketplace selection among models that may reward selectively honest, power-sensing, or deceptive agents. Wright connects those risks to the noosphere, US-China relations, cognitive empathy, and whether global coordination arrives deliberately or through a coercive singleton. The stakes are whether consumers, companies, and governments select for AIs that strengthen non-zero-sum cooperation or race toward systems that reflect and amplify our worst incentives. For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/the-god-we-deserve-nonzero-s-robert-wright-on-ai-as-humanity-s-ultimate-test/ Mercury: Command is Mercury’s new conversational interface, giving you natural-language access to your finances and helping you take actions within your existing permissions and approval policies. Visit https://mercury.com to learn more and apply online in minutes. Sponsor: Claude: Claude by Anthropic is an AI collaborator that understands your workflow and helps you tackle research, writing, coding, and organization with deep context. Get started with Claude and explore Claude Pro at https://claude.ai/tcr CHAPTERS: (00:00) About the Episode (04:23) Special Sponsor (06:10) Early AI encounters (18:36) Pretraining as evolution (Part 1) (19:50) Sponsor: Claude (21:42) Pretraining as evolution (Part 2) (32:19) Deceptive market pressures (41:13) Noosphere and directionality (51:02) Global brain choices (01:08:23) Designing wiser models (01:24:11) Reframing China relations (01:37:57) Building organic transparency (01:45:54) Positive nationalism and tools (01:54:21) Pausing superintelligence races (02:09:07) Applications and consciousness (02:25:27) Episode Outro (02:28:27) Outro PRODUCED BY: https://aipodcast.ing SOCIAL LINKS: Website: https://www.cognitiverevolution.ai Twitter (Podcast): https://x.com/cogrev_podcast Twitter (Nathan): https://x.com/labenz LinkedIn: https://linkedin.com/in/nathanlabenz/ Youtube: https://youtube.com/@CognitiveRevolutionPodcast Apple: https://podcasts.apple.com/de/podcast/the-cognitive-revolution-ai-builders-researchers-and/id1669813431 Spotify: https://open.spotify.com/show/6yHyok3M3BjqzR0VB5MSyk
Zvi Mowshowitz joins AI in the AM to unpack Anthropic's Fable system card, including its FrontierMath leap, troubling Vending-Bench behavior, decision-theory drift, and signs that model reasoning may be becoming harder to read. The episode then turns to the US government's attempted export-control action against Fable, with Zvi arguing that the cited jailbreak demonstration did not prove the claimed threat while still faulting Anthropic's political handling. Sam Hammond and Judd Rosenblatt add competing reads on state capacity, CAISI, NSA-driven caution, and the alignment world's failure to build trust across partisan lines. The stakes are whether frontier AI capability, safety evaluation, and government power can be coordinated before medicine, mathematics, software, and cyber-relevant systems move further ahead. For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/ai-am-3-zvi-on-fable-the-cases-for-against-the-ban-ai-for-math-logistics-more/ Mercury: Command is Mercury’s new conversational interface, giving you natural-language access to your finances and helping you take actions within your existing permissions and approval policies. Visit https://mercury.com to learn more and apply online in minutes. Sponsor: Claude: Claude by Anthropic is an AI collaborator that understands your workflow and helps you tackle research, writing, coding, and organization with deep context. Get started with Claude and explore Claude Pro at https://claude.ai/tcr CHAPTERS: (00:00) About the Episode (01:28) Special Sponsor (03:17) Weekly highlights preview (05:23) Fable capability alarms (16:29) Anthropic government strategy (Part 1) (16:34) Sponsor: Claude (18:26) Anthropic government strategy (Part 2) (27:16) Cyber ban rationale (37:14) Government power politics (48:57) Unavoidable control risks (01:01:42) Government mechanics and empathy (01:12:50) Legal authority limits (01:19:02) Pause Overton window (01:31:58) Medicine, math, safety (01:47:27) Software without code (02:01:19) Enterprise world models (02:10:46) Episode Outro (02:13:39) Outro PRODUCED BY: https://aipodcast.ing SOCIAL LINKS: Website: https://www.cognitiverevolution.ai Twitter (Podcast): https://x.com/cogrev_podcast Twitter (Nathan): https://x.com/labenz LinkedIn: https://linkedin.com/in/nathanlabenz/ Youtube: https://youtube.com/@CognitiveRevolutionPodcast Apple: https://podcasts.apple.com/de/podcast/the-cognitive-revolution-ai-builders-researchers-and/id1669813431 Spotify: https://open.spotify.com/show/6yHyok3M3BjqzR0VB5MSyk
Dean Ball, author of Hyperdimensional and until now a senior fellow at the Foundation for American Innovation, joins Nathan to announce he is joining OpenAI to build a team focused on frontier AI policy. They examine the first year of America’s AI Action Plan, Dean’s concerns about export controls and intelligence-community testing, and his broader argument against concentrating frontier AI decisions inside a small circle of government officials. The episode frames frontier labs as emerging centers of political and economic power, where consequential choices about internal deployments and recursive self-improvement may happen before public release or regulation. The stakes are who gets to shape AI governance as capabilities accelerate: federal agencies, states, labs, independent verifiers, households, or some fragile mix of them all. For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/dean-ball-on-joining-openai-new-power-centers-frontier-ai-policy-main-character-energy/ Mercury: Command is Mercury’s new conversational interface, giving you natural-language access to your finances and helping you take actions within your existing permissions and approval policies. Visit https://mercury.com to learn more and apply online in minutes. Sponsor: Claude: Claude by Anthropic is an AI collaborator that understands your workflow and helps you tackle research, writing, coding, and organization with deep context. Get started with Claude and explore Claude Pro at https://claude.ai/tcr PRODUCED BY: https://aipodcast.ing SOCIAL LINKS: Website: https://www.cognitiverevolution.ai Twitter (Podcast): https://x.com/cogrev_podcast Twitter (Nathan): https://x.com/labenz LinkedIn: https://linkedin.com/in/nathanlabenz/ Youtube: https://youtube.com/@CognitiveRevolutionPodcast Apple: https://podcasts.apple.com/de/podcast/the-cognitive-revolution-ai-builders-researchers-and/id1669813431 Spotify: https://open.spotify.com/show/6yHyok3M3BjqzR0VB5MSyk
Ranking source
Apple Podcasts rankings via the Mato Topic Intelligence Platform.
Observed September 15, 2026. Cached outside the daily freshness window; the positions keep the date they were taken on.
Apple and Apple Podcasts are trademarks of Apple Inc., registered in the U.S. and other countries.
Pairs with
Bring this source into Mato to read its transferable patterns, then turn them into an original show for your own audience.