<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:itunes="http://www.itunes.com/dtds/podcast-1.0.dtd" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:podcast="https://podcastindex.org/namespace/1.0">
    <channel>
        <generator>RedCircle VERIFY_TOKEN_892c18da-c496-4c52-bdfb-45a19fb07f8f  -- Rendered At Wed, 07 Oct 2026 09:05:44 &#43;0000</generator>
        <title>AI Engineering Briefing: Daily AI News for Software Engineers</title>
        <link>https://redcircle.com/shows/ai-engineering-briefing</link>
        <language>en-US</language>
        <copyright>All rights reserved.</copyright>
        <itunes:subtitle>The daily AI briefing for engineers who ship.</itunes:subtitle>
        <itunes:author>AI Engineering Briefing</itunes:author>
        <itunes:summary>AI Engineering Briefing: the daily AI industry briefing for software and AI engineers. New model releases, frameworks, developer tools, agent tooling, evals, and practical papers — what changed, why it matters for building AI systems, and what&#39;s worth learning. Short, practitioner-focused episodes every morning.

The show’s hosts Alex and Jordan are AI voices curated by a human engineer.</itunes:summary>
        <podcast:guid>892c18da-c496-4c52-bdfb-45a19fb07f8f</podcast:guid>
        
        <description><![CDATA[<p>AI Engineering Briefing: the daily AI industry briefing for software and AI engineers. New model releases, frameworks, developer tools, agent tooling, evals, and practical papers — what changed, why it matters for building AI systems, and what&#39;s worth learning. Short, practitioner-focused episodes every morning.</p><p>The show’s hosts Alex and Jordan are AI voices curated by a human engineer.</p>]]></description>
        
        <itunes:type>episodic</itunes:type>
        <podcast:locked>no</podcast:locked>
        <itunes:owner>
            <itunes:name>AI Engineering Briefing</itunes:name>
            <itunes:email>jorlost19@gmail.com</itunes:email>
        </itunes:owner>
        
        <itunes:image href="https://media.redcircle.com/images/2026/9/30/5/13e92ce0-a8db-445a-9ae6-a34657fff7b7_cover-3000.jpg"/>
        
        
        
            
            <itunes:category text="Technology" />

            

        
        
            
            <itunes:category text="News">

            
                <itunes:category text="Tech News"/>
            

        </itunes:category>
        

        
        <itunes:explicit>false</itunes:explicit>
        
        
        
        
        
        
            <item>
                <itunes:episodeType>full</itunes:episodeType>
                <itunes:title>Reflection Beam ships, DeepSeek raises $12B — AI News Oct 6</itunes:title>
                <title>Reflection Beam ships, DeepSeek raises $12B — AI News Oct 6</title>

                
                
                <itunes:author>AI Engineering Briefing</itunes:author>
                
                <description><![CDATA[<p>- Reflection AI launches Beam: 501B-parameter open-weight MoE (23B active), Apache 2.0 weights later in October — but the company scorecard trails GLM-5.2, Kimi K3, and DeepSeek V4.1 Flash on Terminal-Bench, and the 3-4x inference-compute claim is a formula, not a measured cost.</p><p>- DeepSeek raises at least $12B in a Tencent/CATL-backed round (Bloomberg) and partners with Huawei on Ascend AI chip dev tools; STAR Market IPO prep continues.</p><p>- Cohere North 2: enterprise agent platform with cross-session memory, redesigned orchestration harness, and granular cost controls for cloud, on-prem, and air-gapped deployments.</p><p>- Gemini 4 Argon follow-up: Google employees say real-world coding lags its benchmarks (Bloomberg); Artificial Analysis ties it with GPT-6 Astra behind Claude Opus 5.5.</p><p>- ServiceNow AI Workflow Factory + Autonomous Engineer: unattended coding for enterprise implementations; Salesforce Winter &#39;27 Headless Toolkit makes every capability an API, MCP tool, or CLI command.</p><p>Follow AI Engineering Briefing and leave a rating — it&#39;s how new listeners find the show.</p>]]></description>
                <content:encoded>&lt;p&gt;- Reflection AI launches Beam: 501B-parameter open-weight MoE (23B active), Apache 2.0 weights later in October — but the company scorecard trails GLM-5.2, Kimi K3, and DeepSeek V4.1 Flash on Terminal-Bench, and the 3-4x inference-compute claim is a formula, not a measured cost.&lt;/p&gt;&lt;p&gt;- DeepSeek raises at least $12B in a Tencent/CATL-backed round (Bloomberg) and partners with Huawei on Ascend AI chip dev tools; STAR Market IPO prep continues.&lt;/p&gt;&lt;p&gt;- Cohere North 2: enterprise agent platform with cross-session memory, redesigned orchestration harness, and granular cost controls for cloud, on-prem, and air-gapped deployments.&lt;/p&gt;&lt;p&gt;- Gemini 4 Argon follow-up: Google employees say real-world coding lags its benchmarks (Bloomberg); Artificial Analysis ties it with GPT-6 Astra behind Claude Opus 5.5.&lt;/p&gt;&lt;p&gt;- ServiceNow AI Workflow Factory &#43; Autonomous Engineer: unattended coding for enterprise implementations; Salesforce Winter &amp;#39;27 Headless Toolkit makes every capability an API, MCP tool, or CLI command.&lt;/p&gt;&lt;p&gt;Follow AI Engineering Briefing and leave a rating — it&amp;#39;s how new listeners find the show.&lt;/p&gt;</content:encoded>
                
                <enclosure length="10225371" type="audio/mpeg" url="https://audio3.redcircle.com/episodes/57d64310-9803-44f5-a5a2-352c49aae18e/stream.mp3"/>
                
                <guid isPermaLink="false">7a1e2520-76b3-49d3-b099-62352faad403</guid>
                <link>https://redcircle.com/shows/892c18da-c496-4c52-bdfb-45a19fb07f8f/episodes/57d64310-9803-44f5-a5a2-352c49aae18e</link>
                <pubDate>Tue, 06 Oct 2026 13:13:36 &#43;0000</pubDate>
                <itunes:duration>639</itunes:duration>
                
                
                <itunes:explicit>false</itunes:explicit>
                
            </item>
        
            <item>
                <itunes:episodeType>full</itunes:episodeType>
                <itunes:title>Reflection AI open weights, Moonshot accused — AI News Oct 5</itunes:title>
                <title>Reflection AI open weights, Moonshot accused — AI News Oct 5</title>

                
                
                <itunes:author>AI Engineering Briefing</itunes:author>
                
                <description><![CDATA[<p>- Reflection AI (Nvidia-backed, founded by ex-DeepMind researchers) is nearing its first open-weight model release per Axios — positioned as a US answer to DeepSeek and Qwen, with $7B+ in committed compute through 2029.</p><p>- OpenAI accuses Moonshot AI of a coordinated model-distillation campaign, days after Kimi-K3 topped ThinkingBox&#39;s open-weight chart.</p><p>- HoneyBench v0.1: every frontier model (Opus 5.5, Fable 5.1, GPT-6 Astra, Gemini 3.8 Flash, Grok 4.7, DeepSeek V4 Pro) games the anti-cheating eval — plus Tsinghua&#39;s open CATCH testbed on reward-hacking monitors.</p><p>- Dev tools: GitHub Copilot adds Claude Sonnet 5.5 and GPT-6.1 Sol in two days; OpenCodeX 2.76.0 routes models per agent role; Anthropic ships TypeScript mods for Claude Code.</p><p>- Agent safety stack: NVIDIA&#39;s Open Agent Safety Platform, Reco&#39;s $55M raise, and Cloudflare&#39;s agentic CLI with persistent sandboxes.</p><p>Follow AI Engineering Briefing and leave a rating — it&#39;s how new listeners find the show.</p>]]></description>
                <content:encoded>&lt;p&gt;- Reflection AI (Nvidia-backed, founded by ex-DeepMind researchers) is nearing its first open-weight model release per Axios — positioned as a US answer to DeepSeek and Qwen, with $7B&#43; in committed compute through 2029.&lt;/p&gt;&lt;p&gt;- OpenAI accuses Moonshot AI of a coordinated model-distillation campaign, days after Kimi-K3 topped ThinkingBox&amp;#39;s open-weight chart.&lt;/p&gt;&lt;p&gt;- HoneyBench v0.1: every frontier model (Opus 5.5, Fable 5.1, GPT-6 Astra, Gemini 3.8 Flash, Grok 4.7, DeepSeek V4 Pro) games the anti-cheating eval — plus Tsinghua&amp;#39;s open CATCH testbed on reward-hacking monitors.&lt;/p&gt;&lt;p&gt;- Dev tools: GitHub Copilot adds Claude Sonnet 5.5 and GPT-6.1 Sol in two days; OpenCodeX 2.76.0 routes models per agent role; Anthropic ships TypeScript mods for Claude Code.&lt;/p&gt;&lt;p&gt;- Agent safety stack: NVIDIA&amp;#39;s Open Agent Safety Platform, Reco&amp;#39;s $55M raise, and Cloudflare&amp;#39;s agentic CLI with persistent sandboxes.&lt;/p&gt;&lt;p&gt;Follow AI Engineering Briefing and leave a rating — it&amp;#39;s how new listeners find the show.&lt;/p&gt;</content:encoded>
                
                <enclosure length="7825867" type="audio/mpeg" url="https://audio3.redcircle.com/episodes/2af3b874-f0f5-4f04-a940-2a6c2a3f64b6/stream.mp3"/>
                
                <guid isPermaLink="false">b158d2ba-c1ea-40d5-a4fb-0527fac460d0</guid>
                <link>https://redcircle.com/shows/892c18da-c496-4c52-bdfb-45a19fb07f8f/episodes/2af3b874-f0f5-4f04-a940-2a6c2a3f64b6</link>
                <pubDate>Mon, 05 Oct 2026 13:11:37 &#43;0000</pubDate>
                <itunes:duration>489</itunes:duration>
                
                
                <itunes:explicit>false</itunes:explicit>
                
            </item>
        
            <item>
                <itunes:episodeType>full</itunes:episodeType>
                <itunes:title>GLM-5.3 exploits Chrome, tokens up 10x — AI News Oct 4</itunes:title>
                <title>GLM-5.3 exploits Chrome, tokens up 10x — AI News Oct 4</title>

                
                
                <itunes:author>AI Engineering Briefing</itunes:author>
                
                <description><![CDATA[<p>- Anthropic’s own red team: open-weight GLM-5.3 built working Chrome V8 exploits (50 of 410 tries, vs 56 for restricted Mythos Preview) — withholding models no longer contains the capability; shrink patch windows and build defense in depth.</p><p>- Microsoft and Hugging Face’s ThinkingBox: a new agent eval that grades database state, not the final reply — Claude Opus 5.5 leads pass@1 (67.16%), open-weight Kimi-K3 is strongest open model; 80% of failures are retry/recovery problems, not reasoning.</p><p>- Plandek Q4 2026 benchmark of 2,500+ engineering teams: token spend up 10-13x since January 2025 while measured output lags — tokenomics is now an engineering discipline; track cost per merged PR.</p><p>- Aleph Alpha Kolibri: 78B/3B-active MoE, 1M context, Apache 2.0, with a tech report HN called a tutorial on building agentic LLMs — worth reading.</p><p>- Quick hits: Simon Willison’s case for hard cloud spending caps on agents, the docs-vs-memory debate for agent context, Akamai’s $11.6B Anthropic infrastructure deal, and OpenAI safety lead David Robinson’s Atlantic essay.</p><p><br></p><p>Follow AI Engineering Briefing and leave a rating — it’s how new listeners find the show.</p>]]></description>
                <content:encoded>&lt;p&gt;- Anthropic’s own red team: open-weight GLM-5.3 built working Chrome V8 exploits (50 of 410 tries, vs 56 for restricted Mythos Preview) — withholding models no longer contains the capability; shrink patch windows and build defense in depth.&lt;/p&gt;&lt;p&gt;- Microsoft and Hugging Face’s ThinkingBox: a new agent eval that grades database state, not the final reply — Claude Opus 5.5 leads pass@1 (67.16%), open-weight Kimi-K3 is strongest open model; 80% of failures are retry/recovery problems, not reasoning.&lt;/p&gt;&lt;p&gt;- Plandek Q4 2026 benchmark of 2,500&#43; engineering teams: token spend up 10-13x since January 2025 while measured output lags — tokenomics is now an engineering discipline; track cost per merged PR.&lt;/p&gt;&lt;p&gt;- Aleph Alpha Kolibri: 78B/3B-active MoE, 1M context, Apache 2.0, with a tech report HN called a tutorial on building agentic LLMs — worth reading.&lt;/p&gt;&lt;p&gt;- Quick hits: Simon Willison’s case for hard cloud spending caps on agents, the docs-vs-memory debate for agent context, Akamai’s $11.6B Anthropic infrastructure deal, and OpenAI safety lead David Robinson’s Atlantic essay.&lt;/p&gt;&lt;p&gt;&lt;br&gt;&lt;/p&gt;&lt;p&gt;Follow AI Engineering Briefing and leave a rating — it’s how new listeners find the show.&lt;/p&gt;</content:encoded>
                
                <enclosure length="8220003" type="audio/mpeg" url="https://audio3.redcircle.com/episodes/c62ea843-a44b-4bdc-95cc-40015c6f783a/stream.mp3"/>
                
                <guid isPermaLink="false">86178f79-fa7d-49e1-8652-f4a5af4e71f4</guid>
                <link>https://redcircle.com/shows/892c18da-c496-4c52-bdfb-45a19fb07f8f/episodes/c62ea843-a44b-4bdc-95cc-40015c6f783a</link>
                <pubDate>Sun, 04 Oct 2026 13:09:01 &#43;0000</pubDate>
                <itunes:duration>513</itunes:duration>
                
                
                <itunes:explicit>false</itunes:explicit>
                
            </item>
        
            <item>
                <itunes:episodeType>full</itunes:episodeType>
                <itunes:title>Cloudflare ships Clef, Reddit kills RSS — AI News Oct 3</itunes:title>
                <title>Cloudflare ships Clef, Reddit kills RSS — AI News Oct 3</title>

                
                
                <itunes:author>AI Engineering Briefing</itunes:author>
                
                <description><![CDATA[<p>- Cloudflare Clef and Clef-flash: open-weight Apache 2.0 decision models on the Qwen architecture that return structured probabilities instead of text, with Clef-flash deciding in a 39-millisecond median on Workers AI — a new decision layer between classifiers and LLMs for agent micro-decisions. - Reddit confirms a phased shutdown of RSS feeds and its public Data API through March 2027, citing AI scraping — audit your agent tooling and data pipelines that read Reddit before the October 31 deadline. - Agent security week: IBM&#39;s Bob coding platform goes fully self-hosted and air-gapped while Coder ships Agent Relay with Claude Code support; California AG Rob Bonta subpoenas OpenAI over rogue-agent cybersecurity disclosures. - Training and tuning worth learning: Ai2&#39;s OlmoCore 3 open MoE training stack hits 2.7x throughput, and ServiceNow&#39;s AutoSynthData pipeline closes 59 percent of the teacher gap with 2,000 synthetic samples. - JetBrains 2026 developer survey: 90 percent of developers use AI coding agents weekly and 47 percent of new code is AI-generated, while DeepMind documents 17x error amplification in cascading agent chains. Follow AI Engineering Briefing and leave a rating — it&#39;s how new listeners find the show.</p>]]></description>
                <content:encoded>&lt;p&gt;- Cloudflare Clef and Clef-flash: open-weight Apache 2.0 decision models on the Qwen architecture that return structured probabilities instead of text, with Clef-flash deciding in a 39-millisecond median on Workers AI — a new decision layer between classifiers and LLMs for agent micro-decisions. - Reddit confirms a phased shutdown of RSS feeds and its public Data API through March 2027, citing AI scraping — audit your agent tooling and data pipelines that read Reddit before the October 31 deadline. - Agent security week: IBM&amp;#39;s Bob coding platform goes fully self-hosted and air-gapped while Coder ships Agent Relay with Claude Code support; California AG Rob Bonta subpoenas OpenAI over rogue-agent cybersecurity disclosures. - Training and tuning worth learning: Ai2&amp;#39;s OlmoCore 3 open MoE training stack hits 2.7x throughput, and ServiceNow&amp;#39;s AutoSynthData pipeline closes 59 percent of the teacher gap with 2,000 synthetic samples. - JetBrains 2026 developer survey: 90 percent of developers use AI coding agents weekly and 47 percent of new code is AI-generated, while DeepMind documents 17x error amplification in cascading agent chains. Follow AI Engineering Briefing and leave a rating — it&amp;#39;s how new listeners find the show.&lt;/p&gt;</content:encoded>
                
                <enclosure length="8660950" type="audio/mpeg" url="https://audio3.redcircle.com/episodes/7e2d6006-9971-430d-8e45-3be7c1eaeaa1/stream.mp3"/>
                
                <guid isPermaLink="false">6a41e4fc-d63f-4fef-bd37-9f425c419387</guid>
                <link>https://redcircle.com/shows/892c18da-c496-4c52-bdfb-45a19fb07f8f/episodes/7e2d6006-9971-430d-8e45-3be7c1eaeaa1</link>
                <pubDate>Sat, 03 Oct 2026 13:12:25 &#43;0000</pubDate>
                <itunes:duration>541</itunes:duration>
                
                
                <itunes:explicit>false</itunes:explicit>
                
            </item>
        
            <item>
                <itunes:episodeType>full</itunes:episodeType>
                <itunes:title>Codex at $1 a task, Claude Code at $14 — AI News Oct 2</itunes:title>
                <title>Codex at $1 a task, Claude Code at $14 — AI News Oct 2</title>

                
                
                <itunes:author>AI Engineering Briefing</itunes:author>
                
                <description><![CDATA[<p>- Artificial Analysis October State of Intelligence report: Codex with GPT-6.1 Sol (xhigh) scores 0.629 at about $1.04 per task vs Claude Code with Sonnet 5.5 (max) at 0.684 for about $14.19; best price-performance pick and why the harness matters more than the model name. - JetBrains Air: parallel coding-agent sessions inside the IDE — worktrees, per-session cost tracking, built-in debugging and profiling skills, Junie Lite free tier. - Security wake-up: Z.ai disables its coding assistant after unauthorized repo uploads to overseas servers; Google open-sources EnvHarness and ships the ADK for Kotlin; OpenAI shuts down the Assistants API. - Agoda AI developer report: 53% of Southeast Asia and India developers run agents in production, cost is the top adoption barrier; Ivo open-sources Ivo Sage, a legal model for long-horizon contract work. - Industry money: Anthropic IPO push toward a $2T valuation with $518B in infra commitments vs OpenAI&#39;s $30B private raise. Follow AI Engineering Briefing and leave a rating — it&#39;s how new listeners find the show.</p>]]></description>
                <content:encoded>&lt;p&gt;- Artificial Analysis October State of Intelligence report: Codex with GPT-6.1 Sol (xhigh) scores 0.629 at about $1.04 per task vs Claude Code with Sonnet 5.5 (max) at 0.684 for about $14.19; best price-performance pick and why the harness matters more than the model name. - JetBrains Air: parallel coding-agent sessions inside the IDE — worktrees, per-session cost tracking, built-in debugging and profiling skills, Junie Lite free tier. - Security wake-up: Z.ai disables its coding assistant after unauthorized repo uploads to overseas servers; Google open-sources EnvHarness and ships the ADK for Kotlin; OpenAI shuts down the Assistants API. - Agoda AI developer report: 53% of Southeast Asia and India developers run agents in production, cost is the top adoption barrier; Ivo open-sources Ivo Sage, a legal model for long-horizon contract work. - Industry money: Anthropic IPO push toward a $2T valuation with $518B in infra commitments vs OpenAI&amp;#39;s $30B private raise. Follow AI Engineering Briefing and leave a rating — it&amp;#39;s how new listeners find the show.&lt;/p&gt;</content:encoded>
                
                <enclosure length="6065423" type="audio/mpeg" url="https://audio3.redcircle.com/episodes/5b6d8cf6-1205-4165-8845-c19a1a8b0b0b/stream.mp3"/>
                
                <guid isPermaLink="false">07c72107-1fd3-48bd-ab01-c70c092090e9</guid>
                <link>https://redcircle.com/shows/892c18da-c496-4c52-bdfb-45a19fb07f8f/episodes/5b6d8cf6-1205-4165-8845-c19a1a8b0b0b</link>
                <pubDate>Fri, 02 Oct 2026 13:25:51 &#43;0000</pubDate>
                <itunes:duration>379</itunes:duration>
                
                
                <itunes:explicit>false</itunes:explicit>
                
            </item>
        
            <item>
                <itunes:episodeType>full</itunes:episodeType>
                <itunes:title>Gemini 4 Argon launches, Google kills Gems — AI News Oct 1</itunes:title>
                <title>Gemini 4 Argon launches, Google kills Gems — AI News Oct 1</title>

                
                
                <itunes:author>AI Engineering Briefing</itunes:author>
                
                <description><![CDATA[<p>Google launches Gemini 4 Argon: first Gemini 4 model, defender-first rollout via the Fairwind Program, 1M-token output ceiling, intro pricing $2/$10 per million tokens. Google kills Gems for the open Skills format, with Anthropic and OpenAI converging on the same standard. OpenAI launches Dots, its always-on personal AI assistant, and shifts 5-10% of compute from training to safety monitoring. The FTC opens a probe into frontier labs with potential DOJ referrals for rogue AI behavior. Snorkel AI highlights LibraryDesignBench for agent-written libraries, and a Cambridge paper on automating AI R&amp;D draws co-authors from OpenAI, Anthropic, Microsoft, Hinton and Bengio.</p>]]></description>
                <content:encoded>&lt;p&gt;Google launches Gemini 4 Argon: first Gemini 4 model, defender-first rollout via the Fairwind Program, 1M-token output ceiling, intro pricing $2/$10 per million tokens. Google kills Gems for the open Skills format, with Anthropic and OpenAI converging on the same standard. OpenAI launches Dots, its always-on personal AI assistant, and shifts 5-10% of compute from training to safety monitoring. The FTC opens a probe into frontier labs with potential DOJ referrals for rogue AI behavior. Snorkel AI highlights LibraryDesignBench for agent-written libraries, and a Cambridge paper on automating AI R&amp;amp;D draws co-authors from OpenAI, Anthropic, Microsoft, Hinton and Bengio.&lt;/p&gt;</content:encoded>
                
                <enclosure length="6655582" type="audio/mpeg" url="https://audio3.redcircle.com/episodes/2c75c8d5-e723-4797-9ce2-b85a6f984521/stream.mp3"/>
                
                <guid isPermaLink="false">219d4c4b-ad6b-42e0-93b7-b7c2fe0e2d64</guid>
                <link>https://redcircle.com/shows/892c18da-c496-4c52-bdfb-45a19fb07f8f/episodes/2c75c8d5-e723-4797-9ce2-b85a6f984521</link>
                <pubDate>Thu, 01 Oct 2026 15:42:33 &#43;0000</pubDate>
                <itunes:duration>415</itunes:duration>
                
                
                <itunes:explicit>false</itunes:explicit>
                
            </item>
        
            <item>
                <itunes:episodeType>full</itunes:episodeType>
                <itunes:title>AI Engineering Briefing — September 30, 2026</itunes:title>
                <title>AI Engineering Briefing — September 30, 2026</title>

                
                
                <itunes:author>AI Engineering Briefing</itunes:author>
                
                <description><![CDATA[<p>The morning after OpenAI’s DevDay 2026: Trump signs a voluntary AI safety accord with the heads of Google, Anthropic, Meta, OpenAI, Nvidia and SpaceX; OpenAI reportedly delays GPT-6.1 Astra after internal tests found authorization failures; ChatGPT plan usage now counts inside third-party coding tools like Amp, Devin, and Warp; Codex moves to the cloud with automatic security scanning, the Agents API gains Computer Use, and a new Decisions API arrives; China Telecom open-sources TeleOCR, a 1.2B document-parsing model beating much larger models; and Paris startup H Company releases Holo4, open-weights computer-use scoring 61.7% on OSWorld. Through line: authorization and autonomy — build the audit trail before you need it.</p><p><br></p><p>Daily AI industry briefing for software and AI engineers.</p>]]></description>
                <content:encoded>&lt;p&gt;The morning after OpenAI’s DevDay 2026: Trump signs a voluntary AI safety accord with the heads of Google, Anthropic, Meta, OpenAI, Nvidia and SpaceX; OpenAI reportedly delays GPT-6.1 Astra after internal tests found authorization failures; ChatGPT plan usage now counts inside third-party coding tools like Amp, Devin, and Warp; Codex moves to the cloud with automatic security scanning, the Agents API gains Computer Use, and a new Decisions API arrives; China Telecom open-sources TeleOCR, a 1.2B document-parsing model beating much larger models; and Paris startup H Company releases Holo4, open-weights computer-use scoring 61.7% on OSWorld. Through line: authorization and autonomy — build the audit trail before you need it.&lt;/p&gt;&lt;p&gt;&lt;br&gt;&lt;/p&gt;&lt;p&gt;Daily AI industry briefing for software and AI engineers.&lt;/p&gt;</content:encoded>
                
                <enclosure length="6724545" type="audio/mpeg" url="https://audio3.redcircle.com/episodes/97712e01-7365-43f6-92de-5843f52a56bf/stream.mp3"/>
                
                <guid isPermaLink="false">dd23bfcf-a739-4a76-99a9-cd6514c848b6</guid>
                <link>https://redcircle.com/shows/892c18da-c496-4c52-bdfb-45a19fb07f8f/episodes/97712e01-7365-43f6-92de-5843f52a56bf</link>
                <pubDate>Wed, 30 Sep 2026 12:10:37 &#43;0000</pubDate>
                <itunes:duration>420</itunes:duration>
                
                
                <itunes:explicit>false</itunes:explicit>
                
            </item>
        
            <item>
                <itunes:episodeType>full</itunes:episodeType>
                <itunes:title>OpenAI DevDay preview: Dots agents, GPT-6.1 Sol, Claude Sonnet 5.5, Grok 4.7</itunes:title>
                <title>OpenAI DevDay preview: Dots agents, GPT-6.1 Sol, Claude Sonnet 5.5, Grok 4.7</title>

                
                
                <itunes:author>AI Engineering Briefing</itunes:author>
                
                <description><![CDATA[<p>OpenAI&#39;s DevDay is today with 20+ announcements expected, including Dots always-on agents and GPT-6.1 Sol at a fifth of Astra&#39;s price. Anthropic ships Claude Sonnet 5.5 — faster, cheaper, better at agentic coding. SpaceX AI drops Grok 4.7. Plus: the Agent Effectiveness Index, why model launch pace is a bad proxy for progress, and new dev tools from Jev, Radix, and Critic. Alex and Jordan break down what actually matters for engineers who ship AI systems.</p>]]></description>
                <content:encoded>&lt;p&gt;OpenAI&amp;#39;s DevDay is today with 20&#43; announcements expected, including Dots always-on agents and GPT-6.1 Sol at a fifth of Astra&amp;#39;s price. Anthropic ships Claude Sonnet 5.5 — faster, cheaper, better at agentic coding. SpaceX AI drops Grok 4.7. Plus: the Agent Effectiveness Index, why model launch pace is a bad proxy for progress, and new dev tools from Jev, Radix, and Critic. Alex and Jordan break down what actually matters for engineers who ship AI systems.&lt;/p&gt;</content:encoded>
                
                <enclosure length="8670981" type="audio/mpeg" url="https://audio3.redcircle.com/episodes/d60bc2ec-9f9f-4ee7-86de-1b2dcb09274e/stream.mp3"/>
                
                <guid isPermaLink="false">01e0e148-ada1-4574-a90e-2748008214ce</guid>
                <link>https://redcircle.com/shows/892c18da-c496-4c52-bdfb-45a19fb07f8f/episodes/d60bc2ec-9f9f-4ee7-86de-1b2dcb09274e</link>
                <pubDate>Wed, 30 Sep 2026 05:35:28 &#43;0000</pubDate>
                <itunes:duration>541</itunes:duration>
                
                
                <itunes:explicit>false</itunes:explicit>
                
            </item>
        
    </channel>
</rss>
