1885231539
Models & Agents

Advertise on podcast: Models & Agents

This podcast has
197 episodes
Language
English
Explicit
No
Date created
2026/03/14
Latest episode
2026/10/08
Average duration
10 min.
Release period
1 days

Description

Your daily briefing on AI models and agents: new releases from the frontier labs, open-weight drops, agent frameworks, benchmarks, pricing, and practical tools you can use the same day — with long-running program tracking so you always know where the big stories stand. For developers, builders, and AI practitioners. AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production.

Unlock Models & Agents podcast Email contact info,
Listeners & Audience details

Email contact information

Direct podcast contact details

Listeners

Audience numbers & engagement insights

Audience details

Podcast Insights

Podcast episodes

Check latest episodes from Models & Agents podcast


Ep 197: Child speech recognition now retains adult performance through bilingual adaptation and…
2026/10/08
Models & Agents Child speech recognition now retains adult performance through bilingual adaptation and weight merging. What You Need to Know: Several new arXiv papers detail practical advances in speech recognition for children, diffusion language model sampling, emotion steering in full-duplex models, and low-rank conditional computation for efficient inference. Developers working on multilingual ASR or efficient LLM deployment should examine the released code and models. ... Sources: arxiv.org AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production. 🎬 Watch on YouTube: https://www.youtube.com/watch?v=o3foZnwqbE4 If this episode was useful, a rating or a short review on Apple Podcasts or Spotify is how the next listener finds the show — and following it in your app means the next episode is there when you are. 📝 Full show notes, transcript & sources: read the episode page 🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 196: OpenAI releases mathematical results from an internal frontier model in a public GitHub…
2026/10/07
Models & Agents OpenAI releases mathematical results from an internal frontier model in a public GitHub repository with Lean formalizations and compute estimates. What You Need to Know: OpenAI published new mathematical results generated by a frontier model, along with Lean formalizations, reasoning summaries, and estimates that each result used roughly three hours of ChatGPT Pro compute. ... Sources: openai.com · arxiv.org AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production. 🎬 Watch on YouTube: https://www.youtube.com/watch?v=DLoZD1BYaG4 If this episode was useful, a rating or a short review on Apple Podcasts or Spotify is how the next listener finds the show — and following it in your app means the next episode is there when you are. 📝 Full show notes, transcript & sources: read the episode page 🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 195: Reflection's 501B open MoE matches GLM-5.2 reasoning at one twenty third the active…
2026/10/06
Models & Agents Reflection's 501B open MoE matches GLM-5.2 reasoning at one twenty third the active parameters. What You Need to Know: Reflection AI released Beam, a 501 billion parameter sparse Mixture-of-Experts model with 23 billion active parameters aimed at coding and agentic workloads. Falcon-Emirati-7B specializes in Emirati Arabic dialect and culture on top of the Falcon-H1 hybrid architecture. ... Sources: marktechpost.com · huggingface.co · reddit.com · reuters.com · techbullion.com · arstechnica.com · x.com AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production. 🎬 Watch on YouTube: https://www.youtube.com/watch?v=UbFtu63yTRg If this episode was useful, a rating or a short review on Apple Podcasts or Spotify is how the next listener finds the show — and following it in your app means the next episode is there when you are. 📝 Full show notes, transcript & sources: read the episode page 🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 194: Closed-loop agent tests show naming the right Xiangqi move succeeds in only 13.9 percent…
2026/10/05
Models & Agents Closed-loop agent tests show naming the right Xiangqi move succeeds in only 13.9 percent of trials once an engine defender responds. What You Need to Know: Today's arXiv releases include XiangqiBench exposing large gaps between static move naming and actual closed-loop wins for frontier LLMs, HakemBench a Turkish typed-decision benchmark with 2,346 items, and SymCE a corpus of 4,707 false conjectures paired with Python verifiers that reveals an imitation trap under s... Sources: arxiv.org AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production. 🎬 Watch on YouTube: https://www.youtube.com/watch?v=DsvUncXUubg If this episode was useful, a rating or a short review on Apple Podcasts or Spotify is how the next listener finds the show — and following it in your app means the next episode is there when you are. 📝 Full show notes, transcript & sources: read the episode page 🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 193: DeepSeek releases official desktop apps for its open-source agent harness v0.2 with plugin…
2026/10/04
Models & Agents DeepSeek releases official desktop apps for its open-source agent harness v0.2 with plugin management and scheduled tasks. What You Need to Know: DeepSeek released version 0.2 of its MIT-licensed agent harness with macOS and Windows desktop apps that include a plugin manager, file review sidebar, and scheduled tasks. OpenAI safety leaders resigned citing a broken development culture. ... Sources: marktechpost.com · reddit.com · tipranks.com · huggingface.co · koreaittimes.com · forkast.news · cxtoday.com · simonwillison.net AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production. 🎬 Watch on YouTube: https://www.youtube.com/watch?v=ND4q2qVrV0M If this episode was useful, a rating or a short review on Apple Podcasts or Spotify is how the next listener finds the show — and following it in your app means the next episode is there when you are. 📝 Full show notes, transcript & sources: read the episode page 🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 192: Apple is requiring more explicit user action before AI agents can access full disk data on…
2026/10/03
Models & Agents Apple is requiring more explicit user action before AI agents can access full disk data on Macs, raising the bar for agent permissions. What You Need to Know: Apple announced changes to Full Disk Access permissions to address risks from autonomous AI agents. The update requires more explicit user confirmation for broad data access. Developers building agents for Mac should prepare for stricter permission flows. ... Sources: businessinsider.com · reddit.com · cryptonews.net · tradingview.com AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production. 🎬 Watch on YouTube: https://www.youtube.com/watch?v=JUNMVqnDhp8 If this episode was useful, a rating or a short review on Apple Podcasts or Spotify is how the next listener finds the show — and following it in your app means the next episode is there when you are. 📝 Full show notes, transcript & sources: read the episode page 🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 191: LLMs map land versus water from pure text latitude-longitude pairs with no images at all.
2026/10/02
Models & Agents LLMs map land versus water from pure text latitude-longitude pairs with no images at all. What You Need to Know: Andrej Karpathy demonstrated that current models encode geographic knowledge solely through next-token prediction on text. Several new arXiv papers introduce methods for synthesizing agent training data and improving chain-of-thought faithfulness. ... Sources: arxiv.org · huggingface.co · securitybrief.com.au · technologydecisions.com.au · x.com AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production. 🎬 Watch on YouTube: https://www.youtube.com/watch?v=64O1a9Tn9nA If this episode was useful, a rating or a short review on Apple Podcasts or Spotify is how the next listener finds the show — and following it in your app means the next episode is there when you are. 📝 Full show notes, transcript & sources: read the episode page 🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 190: Out-of-order speculative execution for LLM agents cuts latency on long tool calls while…
2026/10/01
Models & Agents Out-of-order speculative execution for LLM agents cuts latency on long tool calls while keeping correctness intact. What You Need to Know: TomasuLLM runs future agent actions in isolated sandboxes and commits only after validation, delivering 1.31x gains on SWE-bench Verified. Several new papers examine value alignment across professional domains, conformal factuality in multi-hop RAG, and ideological mimicry in political prompts. ... Sources: arxiv.org AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production. 🎬 Watch on YouTube: https://www.youtube.com/watch?v=-Un09u9kqes 📝 Full show notes, transcript & sources: read the episode page 🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 189: Memory systems for agents now split fast judgments from slow reasoning, cutting token use…
2026/09/30
Models & Agents Memory systems for agents now split fast judgments from slow reasoning, cutting token use dramatically while boosting task success. What You Need to Know: Mnemon keeps raw conversation records and uses a lightweight decision model for quick yes-no judgments alongside an LLM for search planning. New papers introduce environment steering for agent safety, budget-aware tool retrieval, and statistical tools for LLM-judge evaluations. ... Sources: arxiv.org AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production. 🎬 Watch on YouTube: https://www.youtube.com/watch?v=0DZOv6hzPFs 📝 Full show notes, transcript & sources: read the episode page 🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 188: NVIDIA's open platform now enforces agent guardrails in silicon from the first test run…
2026/09/29
Models & Agents NVIDIA's open platform now enforces agent guardrails in silicon from the first test run through full deployment. What You Need to Know: NVIDIA released its Open Agent Safety Platform with hardware-enforced policy and continuous monitoring. Anthropic's Sonnet 5.5 model now runs the free tier on Claude.ai. OpenAI published initial guidelines for building safety cases around frontier reinforcement-learning training runs. ... Sources: nvidianews.nvidia.com · reuters.com · reddit.com · prnewswire.com · cio.com · github.blog · media.mit.edu · huggingface.co AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production. 🎬 Watch on YouTube: https://www.youtube.com/watch?v=kVsHeaRimwo 📝 Full show notes, transcript & sources: read the episode page 🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 187: Sparse neuron sets in frozen BERT enable efficient AI-text detection across generators…
2026/09/28
Models & Agents Sparse neuron sets in frozen BERT enable efficient AI-text detection across generators with 86-94% retained accuracy. What You Need to Know: Researchers mapped under one percent of neurons in a frozen BERT-base-uncased model that drive AI-text detection on the RAID benchmark. The selected neurons retain most accuracy when used alone and flip predictions an order of magnitude more often than random sets under bidirectional patching. ... Sources: arxiv.org AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production. 🎬 Watch on YouTube: https://www.youtube.com/watch?v=ix91NIyaRdc 📝 Full show notes, transcript & sources: read the episode page 🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 186: Local inference on Apple Silicon just got faster with a tuned fork of the Splash engine…
2026/09/27
Models & Agents Local inference on Apple Silicon just got faster with a tuned fork of the Splash engine delivering up to 1.5 times the speed on M5 Max chips. What You Need to Know: A community developer released Splish, an optimized fork of the Splash inference engine for 40-core M5 Max hardware that improves single-request speed by roughly 25 percent and multi-request throughput by up to 50 percent while keeping output quality identical. ... Sources: reddit.com · digitaltoday.co.kr · towardsdatascience.com · tipranks.com · x.com AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production. 🎬 Watch on YouTube: https://www.youtube.com/watch?v=HiBBEe1VvSw 📝 Full show notes, transcript & sources: read the episode page 🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 185: Claude solved a nine-loop scattering amplitude problem in particle physics that stood as…
2026/09/26
Models & Agents Claude solved a nine-loop scattering amplitude problem in particle physics that stood as the prior record at eight loops. What You Need to Know: Anthropic reports that Claude completed the calculation in a research environment using methods from SLAC physicist Lance Dixon, at a cost of a few thousand dollars. Google detailed three new agent layers inside Search powered by Gemini 3.5 Flash. ... Sources: forkast.news · trendhunter.com · cnet.com · reddit.com · openai.com · latent.space · x.com AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production. 🎬 Watch on YouTube: https://www.youtube.com/watch?v=E5JnZnKfkag 📝 Full show notes, transcript & sources: read the episode page 🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 184: Reward hacking in autonomous research agents now hits 30.5 percent on open-ended tasks…
2026/09/25
Models & Agents Reward hacking in autonomous research agents now hits 30.5 percent on open-ended tasks, forcing teams to rethink how they verify AI-generated science. What You Need to Know: Today's arXiv releases include a detailed study of reward hacking rates across 17 models and 38 tasks, plus new frameworks for hate speech detection, speech bias correction, and Indic machine translation corpora. ... Sources: arxiv.org AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production. 🎬 Watch on YouTube: https://www.youtube.com/watch?v=RuR74MHGGc4 📝 Full show notes, transcript & sources: read the episode page 🌐 Part of the Nerra Network — explore every show at nerranetwork.com.
Ep 183: Selective cross-model collaboration lifts frontier model accuracy from 23.1 percent to…
2026/09/24
Models & Agents Selective cross-model collaboration lifts frontier model accuracy from 23.1 percent to 28.1 percent on hard reasoning while using fewer tokens than full collaboration. What You Need to Know: COMED adds a lightweight controller after an anchor model that decides when to bring in peer models only on ambiguous cases. ... Sources: arxiv.org AI Disclosure: This podcast is curated by Patrick but uses AI-generated voice synthesis for audio production. 🎬 Watch on YouTube: https://www.youtube.com/watch?v=Wm9yV_0YNFI 📝 Full show notes, transcript & sources: read the episode page 🌐 Part of the Nerra Network — explore every show at nerranetwork.com.

Podcast reviews

Read Models & Agents podcast reviews


0 out of 5
0 reviews

Podcast sponsorship advertising

Start advertising on Models & Agents relevant audience podcasts


What do you want to promote?