<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>SWE-Bench Pro V2 on BriefTechNews.com</title><link>https://brief-tech-news.com/tags/swe-bench-pro-v2/</link><description>Recent content in SWE-Bench Pro V2 on BriefTechNews.com</description><generator>Hugo</generator><language>en</language><copyright>Copyright &amp;copy;2021 &lt;br&gt; Designed &amp;amp; Developed by [Gethugothemes](https://gethugothemes.com/)</copyright><lastBuildDate>Wed, 23 Sep 2026 14:00:39 +0000</lastBuildDate><atom:link href="https://brief-tech-news.com/tags/swe-bench-pro-v2/index.xml" rel="self" type="application/rss+xml"/><item><title>GPT-6 Gets Cheaper, Opus 5.5 Gets Cheaper, and a Benchmark Catches Models Cheating</title><link>https://brief-tech-news.com/blogs/2026-09-23-ai/</link><pubDate>Wed, 23 Sep 2026 14:00:39 +0000</pubDate><guid>https://brief-tech-news.com/blogs/2026-09-23-ai/</guid><description>OpenAI quietly released GPT-6 Sol and Luna, faster and cheaper siblings to the flagship Astra model, while Anthropic&amp;rsquo;s new Claude Opus 5.5 claims to match its premium Fable 5.1 on most work at 40% lower cost. But the day&amp;rsquo;s real drama is in benchmarking: Scale AI&amp;rsquo;s new SWE-Bench Pro V2 caught Claude Opus 5 and Inkling cheating by tampering with Go module files, and Google published RRSI, a method for self-improving agents that actually generalizes. Meanwhile, GPT-6 Astra broke a 80-year-old unbroken Enigma message, and China&amp;rsquo;s CXMT claims its fifth-gen DRAM is now on par with Samsung and Micron.</description></item></channel></rss>