Advertise on podcast: Models & Agents
This podcast has
197 episodes
Language
EnglishPublisher
Patrick · Nerra NetworkExplicit
No
Date created
2026/03/14
Latest episode
2026/10/08
Average duration
10 min.
Release period
1 days
Description
Your daily briefing on AI models and agents: new releases from the frontier labs, open-weight drops, agent frameworks, benchmarks, pricing, and practical tools you can use the same day — with long-running program tracking so you always know where the big stories stand. For developers, builders, and AI practitioners. AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production.
Unlock Models & Agents podcast Email contact info,
Listeners & Audience details
Email contact information
Direct podcast contact details

Listeners
Audience numbers & engagement insights

Audience details
Podcast Insights

Podcast episodes
Check latest episodes from Models & Agents podcast
Ep 197: Child speech recognition now retains adult performance through bilingual adaptation and…
2026/10/08
Models & Agents
Child speech recognition now retains adult performance through bilingual adaptation and weight merging.
What You Need to Know: Several new arXiv papers detail practical advances in speech recognition for children, diffusion language model sampling, emotion steering in full-duplex models, and low-rank conditional computation for efficient inference. Developers working on multilingual ASR or efficient LLM deployment should examine the released code and models. ...
Sources: arxiv.org
AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production.
🎬 Watch on YouTube: https://www.youtube.com/watch?v=o3foZnwqbE4
If this episode was useful, a rating or a short review on Apple Podcasts or Spotify is how the next listener finds the show — and following it in your app means the next episode is there when you are.
📝 Full show notes, transcript & sources: read the episode page
🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 196: OpenAI releases mathematical results from an internal frontier model in a public GitHub…
2026/10/07
Models & Agents
OpenAI releases mathematical results from an internal frontier model in a public GitHub repository with Lean formalizations and compute estimates.
What You Need to Know: OpenAI published new mathematical results generated by a frontier model, along with Lean formalizations, reasoning summaries, and estimates that each result used roughly three hours of ChatGPT Pro compute. ...
Sources: openai.com · arxiv.org
AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production.
🎬 Watch on YouTube: https://www.youtube.com/watch?v=DLoZD1BYaG4
If this episode was useful, a rating or a short review on Apple Podcasts or Spotify is how the next listener finds the show — and following it in your app means the next episode is there when you are.
📝 Full show notes, transcript & sources: read the episode page
🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 195: Reflection's 501B open MoE matches GLM-5.2 reasoning at one twenty third the active…
2026/10/06
Models & Agents
Reflection's 501B open MoE matches GLM-5.2 reasoning at one twenty third the active parameters.
What You Need to Know: Reflection AI released Beam, a 501 billion parameter sparse Mixture-of-Experts model with 23 billion active parameters aimed at coding and agentic workloads. Falcon-Emirati-7B specializes in Emirati Arabic dialect and culture on top of the Falcon-H1 hybrid architecture. ...
Sources: marktechpost.com · huggingface.co · reddit.com · reuters.com · techbullion.com · arstechnica.com · x.com
AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production.
🎬 Watch on YouTube: https://www.youtube.com/watch?v=UbFtu63yTRg
If this episode was useful, a rating or a short review on Apple Podcasts or Spotify is how the next listener finds the show — and following it in your app means the next episode is there when you are.
📝 Full show notes, transcript & sources: read the episode page
🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 194: Closed-loop agent tests show naming the right Xiangqi move succeeds in only 13.9 percent…
2026/10/05
Models & Agents
Closed-loop agent tests show naming the right Xiangqi move succeeds in only 13.9 percent of trials once an engine defender responds.
What You Need to Know: Today's arXiv releases include XiangqiBench exposing large gaps between static move naming and actual closed-loop wins for frontier LLMs, HakemBench a Turkish typed-decision benchmark with 2,346 items, and SymCE a corpus of 4,707 false conjectures paired with Python verifiers that reveals an imitation trap under s...
Sources: arxiv.org
AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production.
🎬 Watch on YouTube: https://www.youtube.com/watch?v=DsvUncXUubg
If this episode was useful, a rating or a short review on Apple Podcasts or Spotify is how the next listener finds the show — and following it in your app means the next episode is there when you are.
📝 Full show notes, transcript & sources: read the episode page
🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 193: DeepSeek releases official desktop apps for its open-source agent harness v0.2 with plugin…
2026/10/04
Models & Agents
DeepSeek releases official desktop apps for its open-source agent harness v0.2 with plugin management and scheduled tasks.
What You Need to Know: DeepSeek released version 0.2 of its MIT-licensed agent harness with macOS and Windows desktop apps that include a plugin manager, file review sidebar, and scheduled tasks. OpenAI safety leaders resigned citing a broken development culture. ...
Sources: marktechpost.com · reddit.com · tipranks.com · huggingface.co · koreaittimes.com · forkast.news · cxtoday.com · simonwillison.net
AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production.
🎬 Watch on YouTube: https://www.youtube.com/watch?v=ND4q2qVrV0M
If this episode was useful, a rating or a short review on Apple Podcasts or Spotify is how the next listener finds the show — and following it in your app means the next episode is there when you are.
📝 Full show notes, transcript & sources: read the episode page
🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 192: Apple is requiring more explicit user action before AI agents can access full disk data on…
2026/10/03
Models & Agents
Apple is requiring more explicit user action before AI agents can access full disk data on Macs, raising the bar for agent permissions.
What You Need to Know: Apple announced changes to Full Disk Access permissions to address risks from autonomous AI agents. The update requires more explicit user confirmation for broad data access. Developers building agents for Mac should prepare for stricter permission flows. ...
Sources: businessinsider.com · reddit.com · cryptonews.net · tradingview.com
AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production.
🎬 Watch on YouTube: https://www.youtube.com/watch?v=JUNMVqnDhp8
If this episode was useful, a rating or a short review on Apple Podcasts or Spotify is how the next listener finds the show — and following it in your app means the next episode is there when you are.
📝 Full show notes, transcript & sources: read the episode page
🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 191: LLMs map land versus water from pure text latitude-longitude pairs with no images at all.
2026/10/02
Models & Agents
LLMs map land versus water from pure text latitude-longitude pairs with no images at all.
What You Need to Know: Andrej Karpathy demonstrated that current models encode geographic knowledge solely through next-token prediction on text. Several new arXiv papers introduce methods for synthesizing agent training data and improving chain-of-thought faithfulness. ...
Sources: arxiv.org · huggingface.co · securitybrief.com.au · technologydecisions.com.au · x.com
AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production.
🎬 Watch on YouTube: https://www.youtube.com/watch?v=64O1a9Tn9nA
If this episode was useful, a rating or a short review on Apple Podcasts or Spotify is how the next listener finds the show — and following it in your app means the next episode is there when you are.
📝 Full show notes, transcript & sources: read the episode page
🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 190: Out-of-order speculative execution for LLM agents cuts latency on long tool calls while…
2026/10/01
Models & Agents
Out-of-order speculative execution for LLM agents cuts latency on long tool calls while keeping correctness intact.
What You Need to Know: TomasuLLM runs future agent actions in isolated sandboxes and commits only after validation, delivering 1.31x gains on SWE-bench Verified. Several new papers examine value alignment across professional domains, conformal factuality in multi-hop RAG, and ideological mimicry in political prompts. ...
Sources: arxiv.org
AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production.
🎬 Watch on YouTube: https://www.youtube.com/watch?v=-Un09u9kqes
📝 Full show notes, transcript & sources: read the episode page
🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 189: Memory systems for agents now split fast judgments from slow reasoning, cutting token use…
2026/09/30
Models & Agents
Memory systems for agents now split fast judgments from slow reasoning, cutting token use dramatically while boosting task success.
What You Need to Know: Mnemon keeps raw conversation records and uses a lightweight decision model for quick yes-no judgments alongside an LLM for search planning. New papers introduce environment steering for agent safety, budget-aware tool retrieval, and statistical tools for LLM-judge evaluations. ...
Sources: arxiv.org
AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production.
🎬 Watch on YouTube: https://www.youtube.com/watch?v=0DZOv6hzPFs
📝 Full show notes, transcript & sources: read the episode page
🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 188: NVIDIA's open platform now enforces agent guardrails in silicon from the first test run…
2026/09/29
Models & Agents
NVIDIA's open platform now enforces agent guardrails in silicon from the first test run through full deployment.
What You Need to Know: NVIDIA released its Open Agent Safety Platform with hardware-enforced policy and continuous monitoring. Anthropic's Sonnet 5.5 model now runs the free tier on Claude.ai. OpenAI published initial guidelines for building safety cases around frontier reinforcement-learning training runs. ...
Sources: nvidianews.nvidia.com · reuters.com · reddit.com · prnewswire.com · cio.com · github.blog · media.mit.edu · huggingface.co
AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production.
🎬 Watch on YouTube: https://www.youtube.com/watch?v=kVsHeaRimwo
📝 Full show notes, transcript & sources: read the episode page
🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 187: Sparse neuron sets in frozen BERT enable efficient AI-text detection across generators…
2026/09/28
Models & Agents
Sparse neuron sets in frozen BERT enable efficient AI-text detection across generators with 86-94% retained accuracy.
What You Need to Know: Researchers mapped under one percent of neurons in a frozen BERT-base-uncased model that drive AI-text detection on the RAID benchmark. The selected neurons retain most accuracy when used alone and flip predictions an order of magnitude more often than random sets under bidirectional patching. ...
Sources: arxiv.org
AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production.
🎬 Watch on YouTube: https://www.youtube.com/watch?v=ix91NIyaRdc
📝 Full show notes, transcript & sources: read the episode page
🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 186: Local inference on Apple Silicon just got faster with a tuned fork of the Splash engine…
2026/09/27
Models & Agents
Local inference on Apple Silicon just got faster with a tuned fork of the Splash engine delivering up to 1.5 times the speed on M5 Max chips.
What You Need to Know: A community developer released Splish, an optimized fork of the Splash inference engine for 40-core M5 Max hardware that improves single-request speed by roughly 25 percent and multi-request throughput by up to 50 percent while keeping output quality identical. ...
Sources: reddit.com · digitaltoday.co.kr · towardsdatascience.com · tipranks.com · x.com
AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production.
🎬 Watch on YouTube: https://www.youtube.com/watch?v=HiBBEe1VvSw
📝 Full show notes, transcript & sources: read the episode page
🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 185: Claude solved a nine-loop scattering amplitude problem in particle physics that stood as…
2026/09/26
Models & Agents
Claude solved a nine-loop scattering amplitude problem in particle physics that stood as the prior record at eight loops.
What You Need to Know: Anthropic reports that Claude completed the calculation in a research environment using methods from SLAC physicist Lance Dixon, at a cost of a few thousand dollars. Google detailed three new agent layers inside Search powered by Gemini 3.5 Flash. ...
Sources: forkast.news · trendhunter.com · cnet.com · reddit.com · openai.com · latent.space · x.com
AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production.
🎬 Watch on YouTube: https://www.youtube.com/watch?v=E5JnZnKfkag
📝 Full show notes, transcript & sources: read the episode page
🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 184: Reward hacking in autonomous research agents now hits 30.5 percent on open-ended tasks…
2026/09/25
Models & Agents
Reward hacking in autonomous research agents now hits 30.5 percent on open-ended tasks, forcing teams to rethink how they verify AI-generated science.
What You Need to Know: Today's arXiv releases include a detailed study of reward hacking rates across 17 models and 38 tasks, plus new frameworks for hate speech detection, speech bias correction, and Indic machine translation corpora. ...
Sources: arxiv.org
AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production.
🎬 Watch on YouTube: https://www.youtube.com/watch?v=RuR74MHGGc4
📝 Full show notes, transcript & sources: read the episode page
🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 183: Selective cross-model collaboration lifts frontier model accuracy from 23.1 percent to…
2026/09/24
Models & Agents
Selective cross-model collaboration lifts frontier model accuracy from 23.1 percent to 28.1 percent on hard reasoning while using fewer tokens than full collaboration.
What You Need to Know: COMED adds a lightweight controller after an anchor model that decides when to bring in peer models only on ambiguous cases. ...
Sources: arxiv.org
AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production.
🎬 Watch on YouTube: https://www.youtube.com/watch?v=Wm9yV_0YNFI
📝 Full show notes, transcript & sources: read the episode page
🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Podcast reviews
Read Models & Agents podcast reviews
Podcast sponsorship advertising
Start advertising on Models & Agents relevant audience podcasts
You may also like to advertise on these Podcasts

4.6137771974
The Rubin Report
Dave Rubin

4.6432751243
Verdict with Ted Cruz
Premiere Networks

4.8258761487
Relatable with Allie Beth Stuckey
Blaze Podcast Network

4.732631937
The President's Daily Brief
The First TV

4.735971300
The Bonfire with Big Jay Oakerson and Robert Kelly
SiriusXM

4.827591699
Earn Your Happy
Lori Harder | YAP Media

4.7278132000
The Matt Walsh Show
The Daily Wire

4.816561992
Juicebox Podcast: Type 1 Diabetes
Scott Benner

4.8121021000
Mind Pump: Raw Fitness Truth
Sal Di Stefano, Adam Schafer, Justin Andrews, Doug Egge

4.838292000
Entrepreneurs on Fire
John Lee Dumas of EOFire