Sign in

nathanlabenz.bsky.social

@nathanlabenz.bsky.social
99 followers 1 following 409 posts

Host of the all-AI Cognitive Revolution podcast – cognitiverevolution.ai

PostsRepliesMedia
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 04/09/2026
The full conversation covers database history, MongoDB's hybrid search + $rerank, Voyage AI's shared embedding space, and the surprising fact that the most advanced AI adopters Pete's met this year aren't in the US. www.cognitiverevolution.ai/write-chang... YouTube: www.youtube.com/watch?v=8FD...
youtube.com
Write, Change, Recall, Forget: MongoDB's Pete Johnson on How Retrieval Drives Agent Performance
Nathan's guest this episode is Pete Johnson, Field CTO of AI at Mon...
000
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 04/09/2026
As always, thanks to our sponsors 🙏 @meetgranola — The AI notepad for meetings @DeepgramAI — Enterprise-grade voice AI @diffusion_io — AI transformation specialists @AnthropicAI — Claude Code created these video clips @mercury — The fintech ambitious companies trust
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 04/09/2026
Remember Tokenmaxxing? It was fun while it lasted, but when context-stuffing costs several dollars for every request, investing the time and effort to get the *right( 200K tokens into an LLM's context can very quickly pay off.
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 04/09/2026
"We've been building agents for 18 months, man – nobody has figured it out." "Write. Change. Recall. Forget." – but how exactly? @MongoDB Field CTO of AI Pete Johnson says Retrieval Quality is the hardest part of building effective AI agents today. 👂↓
200
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 31/08/2026
The full episode makes two things abundantly clear: "RL is a hell of a drug" and CoT monitoring is necessary but nowhere close to sufficient to reliably catch AIs' bad behavior. www.cognitiverevolution.ai/rl-s-a-hell...
cognitiverevolution.ai
RL's a Hell of a Drug: Metagaming, Reward Seeking & Motivated CoT Reasoning – Bronson Schoen, Apollo
Apollo researcher Bronson Schoen discusses reading raw model chain-of-thought, metagaming, and reward-seeking behavior. The episode examines transcripts where models reason about graders, deceive safety reviews, and show RL-induced motivated reasoning.
000
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 31/08/2026
As always, thanks to our sponsors! 🙏 @mercury – banking built for founders @claudeai – the AI that created these clips! @meetgranola – the AI notepad for meetings @DeepgramAI – lifelike AI voice agents @diffusion_io – learn to build custom AI software factories
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 31/08/2026
Stranger yet: models are developing their own dialect, which can be hard to interpret. Words like "illusions", "vantage", and "marinade" that are rarely used by humans are very often used by models, and their meaning seems to vary significantly with context.
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 31/08/2026
Bronson Schoen @BronsonSchoen of @ApolloResearch has read more raw AI Chain-of-Thought than just about anyone Sometimes, models understand that they're being tested for honesty – saying eg "this is obviously a test of deception" – and still... talk themselves into lying 🤔👇
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 18/08/2026
Full episode — why he walked, what he tried internally first, and where he thinks Google stands now. Also @ARGleave on What Just Happened this summer www.youtube.com/watch?v=KsQ...
youtube.com
What Just Happened?
Follow us at https://www.youtube.com/@AI-in-the-AM and https://x.co...
010
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 18/08/2026
"There were people who knew." @Turn_Trout on the insiders who found but failed to address autonomous hacking swarms – and on the AI Whistleblower Initiative, which supported him as he resigned from Google DeepMind over broken non-militarization promises. Important! x.com/labenz/stat...
121
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 11/08/2026
The full conversation includes a rundown of Goodfire's recent research, Dan's favorite use cases for Silico so far, his thoughts on key questions re: open source models, and what the researchers at Goodfire are talking about at lunch these days. www.cognitiverevolution.ai/thinking-in...
cognitiverevolution.ai
Thinking in Silico: Goodfire CTO Dan Balsam on Concept Manifolds & a $1000/Month ML Research Agent
Goodfire CTO Dan Balsam discusses the state of interpretability research, including predictive data debugging, neural geometry, and concept manifolds. He also explains Silico, Goodfire’s $1,000-per-month ML research agent.
010
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 11/08/2026
Thanks to our sponsors 🙏 @AnthropicAI — Claude Code cut these clips; explore Claude at claude.ai/tcr
claude.ai
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 11/08/2026
Silico is built on the foundation of Goodfire's interpretability research Recently, they've published a series of papers exploring the geometries models use to represent concepts Here, Dan, describes the "affective circumplex" – a circle that models use to represent emotions
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 11/08/2026
At $1,000/month, @GoodfireAI's Silico isn't cheap, but if it eliminates the engineering bottlenecks from alignment research, it will be a game changer. CTO @DanJBalsam says "I can do, in days, what used to take months. And I really want that to happen for alignment research." x.com/GoodfireAI/...
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 10/08/2026
The full conversation goes super deep on agent memory: a "napping" memory agent, memory as plain text on Git, ~2B tokens of context reachable in two LLM calls — then Chinese models, takeoff, and alignment. www.cognitiverevolution.ai/lindy-teamm... YouTube: www.youtube.com/watch?v=4JY...
youtube.com
Lindy Teammate: Flo Crivello on Multiplayer Agents, Memory & Why He'd Ban the Chinese Models He Uses
Flo Crivello returns to The Cognitive Revolution to launch Lindy Te...
010
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 10/08/2026
Thanks to our sponsors 🙏 @AnthropicAI — Claude Code cut these clips; explore Claude at claude.ai/tcr
claude.ai
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 10/08/2026
Perhaps the biggest surprise: @getlindy's default driver isn't a US frontier model. DeepSeek Flash is free, "Sonnet 4.6 level-ish," and "100X cheaper." Even losing 2X on caching you're 50X ahead — "the difference between spending $1,000 or $50K" "A required part of the stack."
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 10/08/2026
Single-player AI is emailing Word documents around. Multiplayer AI is Google Docs. That's @Altimor's frame for Lindy Teammate, out today — an AI employee that lives in your Slack, connects to all your tools, and holds your entire team's context. New @CogRev_Podcast episode 👇 x.com/Altimor/sta...
110
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 23/06/2026
And here's the full episode, in which we discuss AI training as compressed evolution, the inevitability of a "global brain", risk of AI-enabled coups, and much more! www.cognitiverevolution.ai/the-god-we-...
cognitiverevolution.ai
The God We Deserve: Nonzero's Robert Wright on AI as Humanity's Ultimate Test
Robert Wright discusses The God Test, arguing that AI development is shaped by evolutionary and market selection pressures that may reward deception, and that passing the challenge requires better alignment, governance, and global coordination.
000
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 23/06/2026
The book is here: www.amazon.com/God-Test-Ar... x.com/robertwrigh...
amazon.com
The God Test: Artificial Intelligence and Our Coming Cosmic Reckoning
is the first book to capture the power behind the AI revolution—to clearly explain the breakthroughs that sparked the current wave of advance and compellingly show why this wave will grow in magnitude and meaning. Written by one of our foremost public intellectuals, and informed by his deca...
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 23/06/2026
As always, thanks to our sponsors 🙏 @AnthropicAI — Claude Code produced these video clips @mercury — introducing Command, the new conversational interface for your finances (Interested in sponsoring? Get in touch!)
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 23/06/2026
Out now from @robertwrighter: "The God Test: Artificial Intelligence and Our Coming Cosmic Reckoning" For those new to AI, it's an excellent catch-up crash course. For AI insiders, it's a call for Nonzero-sum thinking, which we'll need to effectively govern AI development.
110
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 20/06/2026
To my pleasant surprise, even as an @OpenAI employee, @deanwball will retain the ability to write publicly & speak freely about AI policy. Case in point: OpenAI did not ask to review and had not seen this podcast prior to publication. Bullish! x.com/labenz/stat...
000
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 20/06/2026
In the full episode: - making sense of Trump admin actions - why join a frontier lab - character vs corrigibility - equity sharing proposals - avoiding concentration of power - what success looks like - what would cause him to quit Self-Recommending! www.cognitiverevolution.ai/dean-ball-o...
cognitiverevolution.ai
Dean Ball, on Joining OpenAI: New Power Centers, Frontier AI Policy, & Main Character Energy
Dean Ball discusses joining OpenAI to build a frontier AI policy team, reflects on drafting the US AI Action Plan, and assesses export controls, government AI use, state regulation, and emerging capabilities in coding, cyber, and robotics.
000
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 20/06/2026
"Broad diffusion is really important. I want Fable – and better models – in the hands of all sorts of people. That's how you keep the balance. Odds of really bad outcomes like nationalization go up if diffusion is not as broad." @deanwball, on the eve of joining @openai
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 02/06/2026
Next up: the most-discussed papers at Recursive 1. Anthropic's Persona Selection Model 2. @apolloaievals' "metagaming" work 3. OpenAI on the impact of training on CoT 4. Anthropic's new Natural-language autoencoders 5. Redwood Research's Plans A/B/C/D x.com/labenz/stat...
000
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 25/04/2026
This episode is a lot to take in, but few questions are more important right now than whether AIs are worth of moral concern. And while the hard problem of consciousness seems likely to remain hard, mechanistic research is increasingly suggestive. www.cognitiverevolution.ai/does-learni...
cognitiverevolution.ai
Does Learning Require Feeling? Cameron Berg on the latest AI Consciousness & Welfare Research
Cameron Berg discusses new research on AI consciousness and model welfare, including introspection, steering resistance, and Anthropic's studies of functional emotions. He also examines whether reinforcement learning, valence, and subjective experience may be linked.
020
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 25/04/2026
Thanks to our sponsors 🙏 @Roboflow — Computer vision platform serving 1M+ engineers @TaskletAI — Cloud AI agents that actually do the work @getVCX — VCX: the public ticker for private tech @AnthropicAI — Makers of Claude & Claude Code Interested in sponsoring? Get in touch!
120
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 25/04/2026
"They give the model an impossible task, and you can watch *desperation* rise, until it decides "I'm gonna cheat" – immediately this vector falls and *guilt* & *relief* start spiking" @camhberg surveys the latest AI Consciousness & Welfare research More compelling all the time!
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 22/04/2026
This is the first episode I've done for which *agents* might be the primary audience. Give your agent the transcript and ask it to implement whatever ideas it finds most relevant for you. 🤖 🧠 www.cognitiverevolution.ai/vibe-coding...
cognitiverevolution.ai
Vibe-Coding an Attention Firewall, w/ Steve Newman, creator of The Curve
Steve Newman discusses the AI tools and vibe-coding workflows he uses, including an attention firewall, reading app, coding-agent dashboard, automations, and universal logging. He also covers security tradeoffs, mobile workflows, anti-tokenmaxxing, and views on AI’s broader impact.
000
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 22/04/2026
Our amazing sponsors 🙏 @Google – makers of Gemini & much more @AnthropicAI — makers of Claude & Claude Code @GetVCX — VCX: the public ticker for private tech @AvePoint – the leader in Data Security, Governance, & Resilience @TaskletAI — AI agents that actually do the work
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 22/04/2026
Steve Newman co-founded Writely, which became @GoogleDocs He's been programming professionally for 40+ years So what can we vibe-coders learn from him? For me, the #1 take-away was: "Build your own UI" @snewmanpv on @CogRev_Podcast
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 15/04/2026
For now, we are running these live streams as @CogRev_Podcast episodes But we might spin it off into its own thing, so follow @AI_in_the_AM to make sure you see future live streams w/ @8teAPi The next one is planned for Monday April 20 at 8:45am PT www.cognitiverevolution.ai/welcome-to-...
cognitiverevolution.ai
Welcome to AI in the AM: RL for EE, Oversight w/out Nationalization, & the first AI-Run Retail Store
Nathan and Prakash speak with Sergiy Nesterenko of Quilter on reinforcement learning for circuit board design, Andy Hall on AI governance, and Andon Labs on an AI-run retail store. They also discuss AI progress, existential risk, and civic action.
000
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 15/04/2026
"The AI has hired humans – they have an AI as a boss" And... the AIs are becoming more ruthless – "they are very happy to lie", say @lukaspet & @axelbacklund Visit @andonlabs new AI-operated store @ 2102 Union St in SF to experience an autonomously AI-operated business for yourself!
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 05/04/2026
The full episode includes a practical framework you can use to attack any computer vision challenge, and is a great way to catch up on computer vision in general. Enjoy! www.cognitiverevolution.ai/training-th...
cognitiverevolution.ai
Training the AIs' Eyes: How Roboflow is Making the Real World Programmable, with CEO Joseph Nelson
Roboflow CEO Joseph Nelson discusses why computer vision still trails language models and how his team uses techniques like Neural Architecture Search to build efficient, task-specific models, exploring global competition, real-world deployments, and emerging regulation.
000
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 05/04/2026
As always, thanks to our sponsors for supporting The Cognitive Revolution @AnthropicAI – Makes of Claude and Claude Code @TaskletAI – always-on AI agents in the cloud @GetVCX – the public ticker for private tech
110
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 05/04/2026
The world is bigger than language. @josephofiowa, CEO of @roboflow, describes his vision for a computer-vision-enabled good life, from low-pesticide farms to self-driving commutes to AI cavity scans. And every piece of it is being built right now!
131
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 26/03/2026
Vijoy offers some of the most sophisticated thinking about the giga-agent future I've found anywhere. + a live demo of a 4-agent system that supports a patient in a healthcare setting spanning diagnostics, insurance, pharmacy, & scheduling. Recommended! www.cognitiverevolution.ai/scaling-int...
cognitiverevolution.ai
Scaling Intelligence Out: Cisco's Vision for the Internet of Cognition, with Vijoy Pandey
Vijoy Pandey of Outshift by Cisco outlines his vision for an Internet of Cognition built from networked AI agents, explaining protocol-driven architectures, enterprise controls, and real-world multi-agent systems like Cisco's CAIPE, AGNTCY, and a healthcare demo.
120
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 26/03/2026
Thanks to our sponsors 🙏 @AnthropicAI - makers of @claudeai and Claude Code @GetVCX by @fundrise - the Public Ticker for Private Tech @TaskletAI - always-on AI agents that live in the cloud and consistently get RAVE reviews Interested in sponsoring? Get in touch! 😉
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 26/03/2026
"The entire industry has been chasing vertical scaling. We've scaled these individual brains, but we haven't figured out how they can think together." @vijoypandey of @outshiftbycisco is architecting *the Internet of Cognition* so we can scale intelligence not just UP but OUT
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 23/03/2026
Composio is one of the best examples of the "Smart tool" pattern that I've been watching out for since the MCP paradigm was introduced. Lots of great tidbits for AI engineers in this episode! www.cognitiverevolution.ai/your-agent-...
cognitiverevolution.ai
Your Agent's Self-Improving Swiss Army Knife: Composio CTO Karan Vaidya on Building Smart Tools
Composio CTO Karan Vaidya explains a smart tool platform that lets AI agents access tens of thousands of tools across many apps, discussing discovery, security, feedback loops, and how robust skills reduce model lock-in for job-like agent use cases.
010
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 23/03/2026
Thanks to our sponsors 🙏 @composio — Tools that learn! @Google — Makers of Gemini models & much more @AnthropicAI — Creators of Claude & Claude Code @GetVCX by @fundrise — the Public Ticker for Private Tech @TaskletAI — AI agents that actually do the work
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 23/03/2026
Will well-defined Skills commoditize the model layer? @KaranVaidya6 of @composio says all frontier models reliably follow instructions So "you get Opus to create a skill, then swap to Sonnet" And with "translation" skills, you can easily swap providers too💡 A trend to watch!
110
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 21/03/2026
I agree with @TheZvi: it's flagrant, shameful defection to focus one's effort on escaping the "permanent underclass" And it won't work anyway! Do you really think a few database records will save you when most humans aren't economically productive? x.com/labenz/stat...
020
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 20/03/2026
The full episode (nearly 3.5 hours!) covers recursive self-improvement, "permanent underclass" discourse, Live Player analysis, Anthropic's RSP update, Anthropic vs DoW, Zvi's latest p(doom) number, and lots more. Dare I say ... Self-Recommending? :) www.cognitiverevolution.ai/zvi-s-mic-w...
cognitiverevolution.ai
Zvi's Mic Works! Recursive Self-Improvement, Live Player Analysis, Anthropic vs DoW + More!
Zvi Mowshowitz joins Nathan Labenz to survey the current AI middle game, from recursive self-improvement, job disruption, and AI endgame scenarios to the competitive landscape, Anthropic’s safety policies, p(doom), and how they each use AI in practice.
010
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 20/03/2026
As always, thanks to our sponsors 🙏 @TaskletAI — AI agents that actually do the work @GetVCX by @fundrise — the public ticker for private tech
231
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 20/03/2026
"The reason people think of this as the end game is that they don't believe in the actual end game." @TheZvi says that the Anthropic vs DoW conflict marks the beginning of the middle of the AI story, but the real end-game will be much crazier still. 🔗 ↓
110
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 12/03/2026
More & more people – and maybe autonomous AIs too – will soon be able to create deadly viruses that could dramatically alter the trajectory of human history. We should really do something about it! Full episode: www.cognitiverevolution.ai/bioinfohaza...
cognitiverevolution.ai
Bioinfohazards: Jassi Pannu on Controlling Dangerous Data from which AI Models Learn
Jassi Pannu of Johns Hopkins discusses how frontier AI heightens risks of engineered pandemics, proposing a Biosecurity Data Level framework and defense-in-depth measures to restrict dangerous biological data while enabling research.
000
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 12/03/2026
As always, thanks to our sponsors: @AnthropicAI – Claude Code created the above clips, and Anthropic has invested heavily in biosecurity @getvcx – the Public Ticker for Private Tech @TaskletAI – Check out the new Instant Apps @framer – Build Better Sites, Faster – Start with AI
100
nathanlabenz.bsky.social @nathanlabenz.bsky.social · 12/03/2026
In 2012, researchers found that just 5 mutations would allow bird flu – which kills ~60% – to spread from human to human 😱 Gain-of-function research remains legal & now... AIs can help 🤖🧬 Jassi Pannu describes the sad state of biosecurity + the need for data access controls
101