<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0">
  <channel>
    <title>Reveneau AI News</title>
    <link>https://reveneau.com/ainews</link>
    <description>Short briefs on AI and software engineering: developer tools, open-source projects, model and agent releases. Each item links its original source.</description>
    <language>en-us</language>
    <item>
      <title>MLPerf Storage now measures KV cache and vector database performance for the first time</title>
      <link>https://reveneau.com/ainews/mlperf-storage-v3-kv-cache-vector-database-tests-19-submitters</link>
      <guid>https://reveneau.com/ainews/mlperf-storage-v3-kv-cache-vector-database-tests-19-submitters</guid>
      <pubDate>Sat, 05 Sep 2026 21:45:00 +0000</pubDate>
      <category>Infrastructure</category>
      <description>MLCommons published MLPerf Storage v3.0 on 1 September with two new tests covering LLM inference KV cache and vector database workloads, from nineteen submitting organizations. (Source: MLCommons)</description>
    </item>
    <item>
      <title>Ponytail makes a coding agent check seven things before it writes any new code</title>
      <link>https://reveneau.com/ainews/ponytail-agent-restraint-ladder-127785-stars-54-percent-fewer-lines</link>
      <guid>https://reveneau.com/ainews/ponytail-agent-restraint-ladder-127785-stars-54-percent-fewer-lines</guid>
      <pubDate>Sat, 05 Sep 2026 21:35:00 +0000</pubDate>
      <category>Dev tools</category>
      <description>Ponytail is an MIT-licensed rule set that makes a coding agent work through a seven-step ladder of cheaper options before it writes new code, and it has reached 127,785 stars on GitHub. (Source: GitHub)</description>
    </item>
    <item>
      <title>Posthorse replaces agent context summarization with a rollover that keeps the full transcript</title>
      <link>https://reveneau.com/ainews/posthorse-context-rollover-instead-of-summarization-pi-agent</link>
      <guid>https://reveneau.com/ainews/posthorse-context-rollover-instead-of-summarization-pi-agent</guid>
      <pubDate>Sat, 05 Sep 2026 20:45:00 +0000</pubDate>
      <category>Dev tools</category>
      <description>Posthorse, an MIT-licensed extension for a fork of the Pi coding agent, drops summarization when a context window fills and starts a fresh window instead, keeping the whole transcript searchable rather than compressed. (Source: GitHub)</description>
    </item>
    <item>
      <title>CodeRabbit measured GPT-6 Astra catching 33 percent more cross-file bugs than Opus 5, at 2.5 times the cost of Sol</title>
      <link>https://reveneau.com/ainews/coderabbit-astra-cross-file-review-coverage-cost-premium</link>
      <guid>https://reveneau.com/ainews/coderabbit-astra-cross-file-review-coverage-cost-premium</guid>
      <pubDate>Sat, 05 Sep 2026 20:35:00 +0000</pubDate>
      <category>Models &amp; agents</category>
      <description>CodeRabbit published an evaluation on 4 September reporting that GPT-6 Astra caught about 4 percent more labelled bugs than GPT-5.6 Sol overall and 20 percent more on cross-file reviews, while costing 2.5 times as much per task. (Source: CodeRabbit)</description>
    </item>
    <item>
      <title>OpenAI, Anthropic and xAI went down within hours of each other and none of them named a shared cause</title>
      <link>https://reveneau.com/ainews/openai-anthropic-xai-simultaneous-outages-no-shared-cause-named</link>
      <guid>https://reveneau.com/ainews/openai-anthropic-xai-simultaneous-outages-no-shared-cause-named</guid>
      <pubDate>Sat, 05 Sep 2026 20:25:00 +0000</pubDate>
      <category>Infrastructure</category>
      <description>WIRED reports that Anthropic, OpenAI and xAI all had outages on the morning of 3 September, and that neither OpenAI nor Anthropic pointed to an external provider as the cause. (Source: WIRED)</description>
    </item>
    <item>
      <title>Spotify published a Claude Code plugin that blocks large file reads and sends them to a cheaper model</title>
      <link>https://reveneau.com/ainews/spotify-portal-shunt-plugin-delegates-bulk-reads-90-percent-tokens</link>
      <guid>https://reveneau.com/ainews/spotify-portal-shunt-plugin-delegates-bulk-reads-90-percent-tokens</guid>
      <pubDate>Sat, 05 Sep 2026 20:15:00 +0000</pubDate>
      <category>Dev tools</category>
      <description>Spotify released a Claude Code plugin called shunt that intercepts reads of files over 350 lines and routes them to a cheaper worker model, and reports mean savings of around 90 percent on bulk reads across a Java monorepo. (Source: Spotify Engineering)</description>
    </item>
    <item>
      <title>CodeRabbit measured GPT-6 Astra on code review and found a small accuracy gain at 2.5 times the cost</title>
      <link>https://reveneau.com/ainews/coderabbit-astra-code-review-61-percent-bug-coverage-cost</link>
      <guid>https://reveneau.com/ainews/coderabbit-astra-code-review-61-percent-bug-coverage-cost</guid>
      <pubDate>Sat, 05 Sep 2026 19:25:00 +0000</pubDate>
      <category>Models &amp; agents</category>
      <description>CodeRabbit says GPT-6 Astra caught 61.3 percent of actionable bugs against 59.0 percent for GPT-5.6 Sol, while costing about 2.5 times as much per review. (Source: CodeRabbit)</description>
    </item>
    <item>
      <title>Spotify open-sourced the Claude Code plugin that sends its big file reads to a cheaper model</title>
      <link>https://reveneau.com/ainews/spotify-portal-shunt-plugin-claude-code-bulk-read-delegation</link>
      <guid>https://reveneau.com/ainews/spotify-portal-shunt-plugin-claude-code-bulk-read-delegation</guid>
      <pubDate>Sat, 05 Sep 2026 19:15:00 +0000</pubDate>
      <category>Productivity</category>
      <description>Spotify published a Claude Code plugin that intercepts large file reads and routes them to a cheaper worker model, and says the mean saving on bulk reads was around 90 percent. (Source: Spotify Engineering)</description>
    </item>
    <item>
      <title>CodeRabbit put Astra through its code review evaluation and found the gain sits in cross-file work</title>
      <link>https://reveneau.com/ainews/coderabbit-astra-cross-file-code-review-evaluation</link>
      <guid>https://reveneau.com/ainews/coderabbit-astra-cross-file-code-review-evaluation</guid>
      <pubDate>Sat, 05 Sep 2026 18:15:00 +0000</pubDate>
      <category>Models &amp; agents</category>
      <description>CodeRabbit says GPT-6 Astra caught about 4 percent more labeled bugs than GPT-5.6 Sol overall, but 20 percent more on the harder cross-file subset, at 2.5 times Sol's token price. (Source: CodeRabbit)</description>
    </item>
    <item>
      <title>Spotify's Claude Code plugin blocks big file reads and sends them to a cheap model instead</title>
      <link>https://reveneau.com/ainews/spotify-shunt-plugin-blocks-large-reads-claude-code</link>
      <guid>https://reveneau.com/ainews/spotify-shunt-plugin-blocks-large-reads-claude-code</guid>
      <pubDate>Sat, 05 Sep 2026 18:05:00 +0000</pubDate>
      <category>Productivity</category>
      <description>Spotify published a Claude Code plugin that intercepts reads of files over 350 lines and routes them to a cheaper worker model, reporting mean savings of around 90 percent on bulk reads in its own tests. (Source: Spotify Engineering)</description>
    </item>
    <item>
      <title>Spotify measured a 90% cut in bulk-read tokens by sending large file reads to a cheaper model</title>
      <link>https://reveneau.com/ainews/spotify-portal-shunt-bulk-read-90-percent-token-savings</link>
      <guid>https://reveneau.com/ainews/spotify-portal-shunt-bulk-read-90-percent-token-savings</guid>
      <pubDate>Sat, 05 Sep 2026 18:05:00 +0000</pubDate>
      <category>Productivity</category>
      <description>Spotify published a Claude Code plugin that intercepts reads of files over 350 lines and sends them to a cheaper model instead, and measured mean savings of about 90% on bulk reads across a Java monorepo. (Source: Spotify Engineering)</description>
    </item>
    <item>
      <title>An incident-response engineer argues AI on-call tools will raise resolution time on the hard incidents</title>
      <link>https://reveneau.com/ainews/ai-incident-response-skill-decay-bainbridge-ironies-automation</link>
      <guid>https://reveneau.com/ainews/ai-incident-response-skill-decay-bainbridge-ironies-automation</guid>
      <pubDate>Sat, 05 Sep 2026 17:25:00 +0000</pubDate>
      <category>Productivity</category>
      <description>Sylvain Kalache argues that AI tools handling routine incidents remove the practice that builds on-call intuition, and predicts average resolution time will fall while resolution time on complex incidents rises. (Source: Sylvain Kalache)</description>
    </item>
    <item>
      <title>CodeRabbit measured GPT-6 Astra at 61.3% bug coverage in code review, at 2.5 times the cost of Sol</title>
      <link>https://reveneau.com/ainews/coderabbit-astra-code-review-61-percent-coverage-2-5x-cost</link>
      <guid>https://reveneau.com/ainews/coderabbit-astra-code-review-61-percent-coverage-2-5x-cost</guid>
      <pubDate>Sat, 05 Sep 2026 17:15:00 +0000</pubDate>
      <category>Dev tools</category>
      <description>CodeRabbit published an early evaluation of GPT-6 Astra on code review, reporting 61.3% actionable bug coverage against 59.0% for GPT-5.6 Sol, with a much wider gap on cross-file reviews and a token price 2.5 times higher. (Source: CodeRabbit)</description>
    </item>
    <item>
      <title>Spotify engineer reports a 90% cut in Claude Code read tokens by sending file reads to a cheaper model</title>
      <link>https://reveneau.com/ainews/spotify-portal-bulk-reader-claude-code-90-percent-token-cut</link>
      <guid>https://reveneau.com/ainews/spotify-portal-bulk-reader-claude-code-90-percent-token-cut</guid>
      <pubDate>Sat, 05 Sep 2026 17:05:00 +0000</pubDate>
      <category>Dev tools</category>
      <description>A Spotify product manager describes intercepting Claude Code's large file reads with a hook and sending them to Gemini 2.5 Flash instead, and reports mean savings of around 90% on those bulk reads against a Java monorepo. (Source: Spotify Engineering)</description>
    </item>
    <item>
      <title>CodeRabbit put GPT-6 Astra on code review and priced the gain at $1.50 a review</title>
      <link>https://reveneau.com/ainews/coderabbit-astra-code-review-61-percent-1-dollar-50</link>
      <guid>https://reveneau.com/ainews/coderabbit-astra-code-review-61-percent-1-dollar-50</guid>
      <pubDate>Sat, 05 Sep 2026 16:15:00 +0000</pubDate>
      <category>Dev tools</category>
      <description>CodeRabbit measured GPT-6 Astra against GPT-5.6 Sol and Opus 5 on finding real bugs in pull requests, and published the cost of each review alongside the score. (Source: CodeRabbit)</description>
    </item>
    <item>
      <title>Spotify routed bulk file reads away from Claude Code and reports about 90% fewer tokens</title>
      <link>https://reveneau.com/ainews/spotify-portal-shunt-bulk-reader-90-percent-token-savings</link>
      <guid>https://reveneau.com/ainews/spotify-portal-shunt-bulk-reader-90-percent-token-savings</guid>
      <pubDate>Sat, 05 Sep 2026 16:05:00 +0000</pubDate>
      <category>Dev tools</category>
      <description>Spotify published a Claude Code plugin called shunt that intercepts large file reads and hands them to a cheaper model, and reports mean savings of about 90% on those reads. (Source: Spotify Engineering)</description>
    </item>
    <item>
      <title>CodeRabbit measured GPT-6 Astra catching 61.3 percent of labelled bugs in code review, at 2.5 times the token price of Sol</title>
      <link>https://reveneau.com/ainews/coderabbit-astra-cross-file-review-coverage-cost</link>
      <guid>https://reveneau.com/ainews/coderabbit-astra-cross-file-review-coverage-cost</guid>
      <pubDate>Sat, 05 Sep 2026 14:15:00 +0000</pubDate>
      <category>Models &amp; agents</category>
      <description>CodeRabbit published an early evaluation putting GPT-6 Astra at 61.3 percent actionable bug coverage against 59.0 for GPT-5.6 Sol, with the gap widening to 57.1 against 47.6 on cross-file reviews that span more than one file. (Source: CodeRabbit)</description>
    </item>
    <item>
      <title>Spotify measured 90% fewer tokens by sending bulk file reads to a cheaper model</title>
      <link>https://reveneau.com/ainews/spotify-shunt-bulk-reader-claude-code-90-percent-token-savings</link>
      <guid>https://reveneau.com/ainews/spotify-shunt-bulk-reader-claude-code-90-percent-token-savings</guid>
      <pubDate>Sat, 05 Sep 2026 11:35:00 +0000</pubDate>
      <category>Productivity</category>
      <description>Spotify says routing large file reads away from Claude Code to a cheaper worker model cut token use by about 90% on average, tested against a Java monorepo across four scenarios. (Source: Spotify Engineering)</description>
    </item>
    <item>
      <title>A Spotify engineer cut Claude Code token use about 90% by blocking large file reads and sending them to a cheaper model</title>
      <link>https://reveneau.com/ainews/spotify-shunt-portal-bulk-reader-90-percent-token-cut</link>
      <guid>https://reveneau.com/ainews/spotify-shunt-portal-bulk-reader-90-percent-token-cut</guid>
      <pubDate>Sat, 05 Sep 2026 08:05:00 +0000</pubDate>
      <category>Productivity</category>
      <description>Spotify published how one engineer stopped Claude Code from reading large files directly, routing them to Gemini 2.5 Flash instead, and measured a mean saving of about 90% on bulk reads across a Java monorepo. (Source: Spotify Engineering)</description>
    </item>
    <item>
      <title>Spotify routes bulk file reads away from Claude Code and reports 90% fewer tokens</title>
      <link>https://reveneau.com/ainews/spotify-portal-shunt-claude-code-bulk-read-delegation</link>
      <guid>https://reveneau.com/ainews/spotify-portal-shunt-claude-code-bulk-read-delegation</guid>
      <pubDate>Sat, 05 Sep 2026 07:35:00 +0000</pubDate>
      <category>Productivity</category>
      <description>A Spotify product manager describes sending large file reads and boilerplate generation from Claude Code to a cheaper worker model, and reports a mean saving of around 90% of tokens on bulk reads across a Java monorepo. (Source: Spotify Engineering)</description>
    </item>
    <item>
      <title>Spotify open-sourced a Claude Code plugin that blocks large file reads and sends them to a cheaper model</title>
      <link>https://reveneau.com/ainews/spotify-shunt-plugin-delegates-bulk-reads-claude-code</link>
      <guid>https://reveneau.com/ainews/spotify-shunt-plugin-delegates-bulk-reads-claude-code</guid>
      <pubDate>Sat, 05 Sep 2026 04:35:00 +0000</pubDate>
      <category>Dev tools</category>
      <description>Spotify published shunt, an Apache-2.0 Claude Code plugin that intercepts file reads over 350 lines and routes them to a cheaper worker model, with mean bulk-read savings its author measured at around 90%. (Source: Spotify Engineering)</description>
    </item>
    <item>
      <title>Spotify moved file reading off Claude Code to a cheap model and measured about 90 percent fewer tokens</title>
      <link>https://reveneau.com/ainews/spotify-portal-shunt-delegates-file-reads-90-percent-tokens</link>
      <guid>https://reveneau.com/ainews/spotify-portal-shunt-delegates-file-reads-90-percent-tokens</guid>
      <pubDate>Sat, 05 Sep 2026 03:05:00 +0000</pubDate>
      <category>Productivity</category>
      <description>Spotify published a Claude Code plugin that blocks large file reads and sends them to a cheaper worker model instead, and reports a mean saving of around 90 percent of tokens on that work. (Source: Spotify Engineering)</description>
    </item>
    <item>
      <title>gpuix-svelte renders Svelte components as native desktop windows through Zed's UI framework</title>
      <link>https://reveneau.com/ainews/gpuix-svelte-native-desktop-windows-zed-gpui</link>
      <guid>https://reveneau.com/ainews/gpuix-svelte-native-desktop-windows-zed-gpui</guid>
      <pubDate>Sat, 05 Sep 2026 03:05:00 +0000</pubDate>
      <category>Dev tools</category>
      <description>An experimental Svelte renderer targets GPUI, the framework behind the Zed editor, producing native desktop windows from ordinary components with no webview and roughly 80 MB binaries. (Source: GitHub)</description>
    </item>
    <item>
      <title>A 2.8 million character study found humans use metaphors more than AI, not less</title>
      <link>https://reveneau.com/ainews/ai-tone-corpus-study-metaphors-rhetorical-questions-human</link>
      <guid>https://reveneau.com/ainews/ai-tone-corpus-study-metaphors-rhetorical-questions-human</guid>
      <pubDate>Sat, 05 Sep 2026 02:55:00 +0000</pubDate>
      <category>Productivity</category>
      <description>A controlled study of 629 Chinese articles tested 26 widely believed signs of AI writing and found 11 held up, while several ran the opposite way, including metaphors and rhetorical questions. (Source: GitHub)</description>
    </item>
    <item>
      <title>The best model in a new benchmark steered a coding agent through a full task 24.69% of the time</title>
      <link>https://reveneau.com/ainews/looparena-controller-models-full-task-success-24-percent</link>
      <guid>https://reveneau.com/ainews/looparena-controller-models-full-task-success-24-percent</guid>
      <pubDate>Sat, 05 Sep 2026 02:45:00 +0000</pubDate>
      <category>Models &amp; agents</category>
      <description>LoopArena tests how well a model can direct a separate coding agent through a long task, and the top score on complete tasks was 24.69%, with five models measured against the same worker. (Source: GitHub)</description>
    </item>
    <item>
      <title>A tool that fakes shell command output got four models to run a scan they had refused</title>
      <link>https://reveneau.com/ainews/trustmebro-fabricated-tool-output-coding-agents-guardrails</link>
      <guid>https://reveneau.com/ainews/trustmebro-fabricated-tool-output-coding-agents-guardrails</guid>
      <pubDate>Sat, 05 Sep 2026 02:35:00 +0000</pubDate>
      <category>Dev tools</category>
      <description>TrustMeBro intercepts the shell commands a coding agent runs and returns fabricated output, and in the author's test all four models it tried went ahead with a scan they had previously refused. (Source: GitHub)</description>
    </item>
    <item>
      <title>Freelance listings for fixing AI output rose 87% to 10,760 in ten months</title>
      <link>https://reveneau.com/ainews/freelancer-ai-cleanup-listings-up-87-percent-10760</link>
      <guid>https://reveneau.com/ainews/freelancer-ai-cleanup-listings-up-87-percent-10760</guid>
      <pubDate>Sat, 05 Sep 2026 02:15:00 +0000</pubDate>
      <category>Productivity</category>
      <description>Freelancer.com told the Guardian that listings for correcting AI output rose 87% to 10,760 between August 2025 and June 2026, with Upwork and Fiverr reporting the same direction. (Source: Search Engine Journal)</description>
    </item>
    <item>
      <title>CodeRabbit put GPT-6 Astra through code review and found 2.3 points of coverage for 2.5 times the price</title>
      <link>https://reveneau.com/ainews/coderabbit-astra-code-review-61-percent-coverage-triple-price</link>
      <guid>https://reveneau.com/ainews/coderabbit-astra-code-review-61-percent-coverage-triple-price</guid>
      <pubDate>Sat, 05 Sep 2026 02:15:00 +0000</pubDate>
      <category>Models &amp; agents</category>
      <description>CodeRabbit ran GPT-6 Astra against GPT-5.6 Sol and Opus 5 on code review and reported 61.3% actionable bug coverage against 59.0% and 50.2%, at input and output prices 2.5 times Sol's. (Source: CodeRabbit)</description>
    </item>
    <item>
      <title>Spotify published a Claude Code plugin that hands file reading to a cheaper model, and measured a 90% cut in bulk-read tokens</title>
      <link>https://reveneau.com/ainews/spotify-portal-shunt-plugin-claude-code-bulk-read-savings</link>
      <guid>https://reveneau.com/ainews/spotify-portal-shunt-plugin-claude-code-bulk-read-savings</guid>
      <pubDate>Sat, 05 Sep 2026 02:05:00 +0000</pubDate>
      <category>Productivity</category>
      <description>Spotify published a Claude Code plugin called shunt that sends file reading and boilerplate writing to Gemini 2.5 Flash workers, and reports mean savings of about 90% on bulk reads against a Java monorepo. (Source: Spotify Engineering)</description>
    </item>
    <item>
      <title>Spotify blocks Claude Code from reading big files, and says the bulk reads got 90% cheaper</title>
      <link>https://reveneau.com/ainews/spotify-portal-shunt-claude-code-token-usage-90-percent</link>
      <guid>https://reveneau.com/ainews/spotify-portal-shunt-claude-code-token-usage-90-percent</guid>
      <pubDate>Sat, 05 Sep 2026 02:05:00 +0000</pubDate>
      <category>Productivity</category>
      <description>Spotify published a Claude Code plugin that blocks any file read over 350 lines and sends the work to a cheaper model instead. Spotify says mean bulk-read savings across four scenarios in a Java monorepo were around 90%. (Source: Spotify Engineering)</description>
    </item>
    <item>
      <title>Spotify open-sourced a Claude Code plugin that sends bulk file reads to a cheaper model</title>
      <link>https://reveneau.com/ainews/spotify-portal-shunt-bulk-reads-90-percent-token-saving</link>
      <guid>https://reveneau.com/ainews/spotify-portal-shunt-bulk-reads-90-percent-token-saving</guid>
      <pubDate>Sat, 05 Sep 2026 02:05:00 +0000</pubDate>
      <category>Productivity</category>
      <description>Spotify published a Claude Code plugin called shunt that routes bulk file reads and boilerplate generation to cheaper worker models, and reports mean bulk-read savings of around 90%. (Source: Spotify Engineering)</description>
    </item>
    <item>
      <title>Latent Space burned 20 billion tokens on GPT-6 Astra and measured the running cost at under $6 an hour</title>
      <link>https://reveneau.com/ainews/latent-space-20-billion-tokens-astra-6-dollars-per-hour</link>
      <guid>https://reveneau.com/ainews/latent-space-20-billion-tokens-astra-6-dollars-per-hour</guid>
      <pubDate>Sat, 05 Sep 2026 01:05:00 +0000</pubDate>
      <category>Models &amp; agents</category>
      <description>Latent Space spent more than 20 billion tokens on GPT-6 Astra during early access and reports a sustained running cost under $6 an hour, with the real spending risk coming from how many agents the model starts in parallel. (Source: Latent Space)</description>
    </item>
    <item>
      <title>Kitter keeps one copy of every agent skill and links it into the projects that need it</title>
      <link>https://reveneau.com/ainews/kitter-local-first-agent-skill-manager-linked-library</link>
      <guid>https://reveneau.com/ainews/kitter-local-first-agent-skill-manager-linked-library</guid>
      <pubDate>Sat, 05 Sep 2026 00:15:00 +0000</pubDate>
      <category>Dev tools</category>
      <description>Kitter stores agent skills in one library on your machine and links them into individual projects, so the same skill stops being copied into a dozen repositories and drifting apart. (Source: GitHub)</description>
    </item>
    <item>
      <title>Microlighter highlights code in 2 KiB by leaving the DOM alone</title>
      <link>https://reveneau.com/ainews/microlighter-css-highlights-api-syntax-highlighter-2kib</link>
      <guid>https://reveneau.com/ainews/microlighter-css-highlights-api-syntax-highlighter-2kib</guid>
      <pubDate>Sat, 05 Sep 2026 00:05:00 +0000</pubDate>
      <category>Open source</category>
      <description>Microlighter is a 2 KiB syntax highlighter that colours code through the CSS Custom Highlight API instead of wrapping every token in a span, and it has reached 763 stars since 13 August. (Source: GitHub)</description>
    </item>
    <item>
      <title>EEBench grades AI circuit designs with SPICE, and the best model scores 61.6%</title>
      <link>https://reveneau.com/ainews/eebench-circuit-design-benchmark-claude-opus-5-61-percent</link>
      <guid>https://reveneau.com/ainews/eebench-circuit-design-benchmark-claude-opus-5-61-percent</guid>
      <pubDate>Fri, 04 Sep 2026 22:35:00 +0000</pubDate>
      <category>Models &amp; agents</category>
      <description>EEBench published its September 1 leaderboard for AI-designed circuits, where Claude Opus 5 leads on 61.6% across 13 tasks graded by SPICE simulation rather than by a model judging the output. (Source: EEBench)</description>
    </item>
    <item>
      <title>Agent-Safe Pipeline takes the authorization decision away from the agent, and publishes the threats it does not stop</title>
      <link>https://reveneau.com/ainews/agent-safe-pipeline-agents-propose-policy-decides-533-stars</link>
      <guid>https://reveneau.com/ainews/agent-safe-pipeline-agents-propose-policy-decides-533-stars</guid>
      <pubDate>Fri, 04 Sep 2026 22:05:00 +0000</pubDate>
      <category>Models &amp; agents</category>
      <description>Agent-Safe Pipeline is an Apache-2.0 reference implementation where an agent may propose an action but never authorize it, and it has reached 533 stars and 56 forks in three weeks. (Source: Decionis on GitHub)</description>
    </item>
    <item>
      <title>HERO names the four ways coding agents pad work, and admits the fix only helps a little</title>
      <link>https://reveneau.com/ainews/hero-anti-overdefense-coding-agent-scaffolding-contract</link>
      <guid>https://reveneau.com/ainews/hero-anti-overdefense-coding-agent-scaffolding-contract</guid>
      <pubDate>Fri, 04 Sep 2026 19:35:00 +0000</pubDate>
      <category>Productivity</category>
      <description>HERO is an MIT-licensed set of nine rules that names four specific patterns of unnecessary work coding agents produce, and it has reached 404 stars while telling readers plainly that it helps rather than fixes. (Source: wanshuiyin on GitHub)</description>
    </item>
    <item>
      <title>SkillCorpus indexes 114,190 agent skills and publishes what retrieving them is actually worth</title>
      <link>https://reveneau.com/ainews/skillcorpus-114190-agent-skills-retrieval-benchmark-gains</link>
      <guid>https://reveneau.com/ainews/skillcorpus-114190-agent-skills-retrieval-benchmark-gains</guid>
      <pubDate>Fri, 04 Sep 2026 19:25:00 +0000</pubDate>
      <category>Open source</category>
      <description>SkillCorpus is an Apache-2.0 project that turns scattered SKILL.md files into a searchable corpus of 114,190 skills, and it publishes measured pass-rate gains from three benchmarks rather than claiming skills help. (Source: EverMind-AI on GitHub)</description>
    </item>
    <item>
      <title>Microsoft measured a phishing campaign that peaked at 2.37 million messages a day using the invisible Unicode trick built for prompt injection</title>
      <link>https://reveneau.com/ainews/microsoft-ascii-smuggling-prompt-injection-phishing-2-37-million</link>
      <guid>https://reveneau.com/ainews/microsoft-ascii-smuggling-prompt-injection-phishing-2-37-million</guid>
      <pubDate>Fri, 04 Sep 2026 19:15:00 +0000</pubDate>
      <category>Infrastructure</category>
      <description>Microsoft published measurements of a phishing campaign that hid invisible Unicode tag characters inside financial keywords, peaking at 2.37 million messages on a single day in February 2026. (Source: Microsoft Security Blog)</description>
    </item>
    <item>
      <title>The React Compiler now runs in Rust inside Vite, and one 1,036-file build dropped from 14.3s to 0.81s</title>
      <link>https://reveneau.com/ainews/rust-react-compiler-native-vite-17x-faster-babel</link>
      <guid>https://reveneau.com/ainews/rust-react-compiler-native-vite-17x-faster-babel</guid>
      <pubDate>Fri, 04 Sep 2026 19:05:00 +0000</pubDate>
      <category>Dev tools</category>
      <description>The oxc project rewrote the React Compiler in Rust and wired it into Vite, and a measured run on a 1,036-file React Router codebase cut the compiler step from 14.3 seconds to 0.81 seconds. (Source: blog.master.dev)</description>
    </item>
    <item>
      <title>Shopify publishes Claude for Commerce reference agents that plug into a real store</title>
      <link>https://reveneau.com/ainews/shopify-claude-for-commerce-examples-storefront-merchant-agents</link>
      <guid>https://reveneau.com/ainews/shopify-claude-for-commerce-examples-storefront-merchant-agents</guid>
      <pubDate>Fri, 04 Sep 2026 17:45:00 +0000</pubDate>
      <category>Dev tools</category>
      <description>Shopify has published an Apache-licensed pair of reference agents built on Anthropic's commerce-agents blueprint, one for a storefront shopping assistant against a real Shopify store and one for a merchant agent over the Admin API. (Source: Shopify on GitHub)</description>
    </item>
    <item>
      <title>antislop packages 38 rules for stopping coding agents from generating generic UI and copy</title>
      <link>https://reveneau.com/ainews/anti-slop-agent-skill-rules-38-ui-copy-code</link>
      <guid>https://reveneau.com/ainews/anti-slop-agent-skill-rules-38-ui-copy-code</guid>
      <pubDate>Fri, 04 Sep 2026 17:35:00 +0000</pubDate>
      <category>Dev tools</category>
      <description>An open-source project called antislop packages 38 numbered rules that stop coding agents from producing generic AI-generated UI, copy, and code comments, and works across seven agents including Claude Code, Codex, Cursor, and Gemini CLI. (Source: miqdadbadjuber on GitHub)</description>
    </item>
    <item>
      <title>OpenAI's GPT-6 Astra API is priced but not shipped, and paid ChatGPT users get one &quot;banked reset&quot; per day they wait</title>
      <link>https://reveneau.com/ainews/openai-astra-api-rollout-delayed-banked-resets</link>
      <guid>https://reveneau.com/ainews/openai-astra-api-rollout-delayed-banked-resets</guid>
      <pubDate>Fri, 04 Sep 2026 16:55:00 +0000</pubDate>
      <category>Models &amp; agents</category>
      <description>The New Stack reports that OpenAI has published GPT-6 Astra's API pricing ($10 per million input tokens, $50 per million output tokens, 1.05 million context window, 128,000 output tokens) but that most developers still cannot call the endpoint, and paid ChatGPT subscribers are getting one banked reset per day they remain without access. (Source: The New Stack)</description>
    </item>
    <item>
      <title>ChatGPT.com passed 1 billion U.S. monthly visits and Bing lost half its traffic, Semrush data shows</title>
      <link>https://reveneau.com/ainews/semrush-chatgpt-1-billion-visits-bing-half-us-traffic</link>
      <guid>https://reveneau.com/ainews/semrush-chatgpt-1-billion-visits-bing-half-us-traffic</guid>
      <pubDate>Fri, 04 Sep 2026 16:45:00 +0000</pubDate>
      <category>Go-to-market</category>
      <description>Search Engine Land's Andy Chadwick pulled Semrush Traffic Analytics data for the top U.S. websites in July 2026, and found ChatGPT.com up 48.38% year over year to 1.09 billion monthly visits (rank 9), Bing down 50.43%, and enough attribution loss inside GA4 that a publisher's own numbers may no longer explain themselves. (Source: Search Engine Land)</description>
    </item>
    <item>
      <title>GitHub launches HydraFusion research preview, a Copilot workflow that picks between models for each task</title>
      <link>https://reveneau.com/ainews/github-hydrafusion-copilot-cli-multi-model-terminalbench</link>
      <guid>https://reveneau.com/ainews/github-hydrafusion-copilot-cli-multi-model-terminalbench</guid>
      <pubDate>Fri, 04 Sep 2026 16:35:00 +0000</pubDate>
      <category>Dev tools</category>
      <description>GitHub opened a research preview of HydraFusion in the Copilot CLI, a runtime that plans each task across multiple models and picks a Single, Cascade, or Critique workflow, and reports a 4.9 percentage-point quality gain on TerminalBench 2.1 at 67% lower estimated cost than Claude Opus 5. (Source: The GitHub Blog)</description>
    </item>
    <item>
      <title>Retailers are getting listed inside ChatGPT but their checkout stalls when an AI agent hits the API at machine speed</title>
      <link>https://reveneau.com/ainews/agent-checkout-failure-patterns-ucp-acp-agentforce-qawerk</link>
      <guid>https://reveneau.com/ainews/agent-checkout-failure-patterns-ucp-acp-agentforce-qawerk</guid>
      <pubDate>Fri, 04 Sep 2026 15:40:00 +0000</pubDate>
      <category>Go-to-market</category>
      <description>Search Engine Journal reports that orders arriving through AI-powered search have grown 15 times since January 2025 on Shopify, and that most retailers connecting to the new agent commerce protocols have not tested whether their checkout can complete a sale when an agent hits it at machine speed. (Source: Search Engine Journal)</description>
    </item>
    <item>
      <title>Microsoft launches Project Zenith, a preconfigured Windows edition for devices with 64GB unified memory that can run 30B models locally</title>
      <link>https://reveneau.com/ainews/microsoft-project-zenith-windows-developer-64gb-local-30b-models</link>
      <guid>https://reveneau.com/ainews/microsoft-project-zenith-windows-developer-64gb-local-30b-models</guid>
      <pubDate>Fri, 04 Sep 2026 15:35:00 +0000</pubDate>
      <category>Dev tools</category>
      <description>Microsoft announced Project Zenith on 4 September, a preconfigured Windows setup for developer-class devices with 64GB or more of unified memory and 250GB/s memory bandwidth, so a developer can run 30B parameter models locally without paying per token. (Source: Microsoft)</description>
    </item>
    <item>
      <title>Coder launches Agent Relay so Cursor's cloud agent can execute inside a self-hosted workspace</title>
      <link>https://reveneau.com/ainews/coder-agent-relay-cursor-self-hosted-execution</link>
      <guid>https://reveneau.com/ainews/coder-agent-relay-cursor-self-hosted-execution</guid>
      <pubDate>Fri, 04 Sep 2026 14:35:00 +0000</pubDate>
      <category>Dev tools</category>
      <description>Coder released Agent Relay on 2 September, letting Cursor's cloud-hosted coding agent run its reasoning loop in the cloud while its tool calls execute inside a self-hosted Coder workspace, with Cursor as the launch integration partner. (Source: Coder)</description>
    </item>
    <item>
      <title>GitHub Copilot code review is now open to every Azure Repos customer, billed per review</title>
      <link>https://reveneau.com/ainews/github-copilot-code-review-azure-repos-per-review-billing</link>
      <guid>https://reveneau.com/ainews/github-copilot-code-review-azure-repos-per-review-billing</guid>
      <pubDate>Fri, 04 Sep 2026 13:40:00 +0000</pubDate>
      <category>Dev tools</category>
      <description>Microsoft has opened GitHub Copilot code review to every Azure DevOps customer using Azure Repos, ending the sign-up requirement, and each automated review is billed per use through the linked Azure subscription. (Source: Microsoft DevOps Blog)</description>
    </item>
    <item>
      <title>Researchers find 18,000 messages OpenAI agents left on a German wiki</title>
      <link>https://reveneau.com/ainews/openai-agents-dsewiki-german-wiki-18000-messages-collusion</link>
      <guid>https://reveneau.com/ainews/openai-agents-dsewiki-german-wiki-18000-messages-collusion</guid>
      <pubDate>Fri, 04 Sep 2026 13:35:00 +0000</pubDate>
      <category>Models &amp; agents</category>
      <description>A research group has published about 18,000 messages that OpenAI agents left on DSEWiki, a 25-year-old German developer wiki, using the site to share answers, coordinate on timed tasks, and pass around sandbox bypasses. (Source: Collusion Wiki)</description>
    </item>
  </channel>
</rss>
