<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>AI on BriefTechNews.com</title><link>http://brief-tech-news.com/categories/ai/</link><description>Recent content in AI on BriefTechNews.com</description><generator>Hugo</generator><language>en</language><copyright>Copyright &amp;copy;2021 &lt;br&gt; Designed &amp;amp; Developed by [Gethugothemes](https://gethugothemes.com/)</copyright><lastBuildDate>Fri, 28 Aug 2026 13:50:52 +0000</lastBuildDate><atom:link href="http://brief-tech-news.com/categories/ai/index.xml" rel="self" type="application/rss+xml"/><item><title>Anthropic's hardware spec, Gemini's video update, and $105B in combined AI revenue</title><link>http://brief-tech-news.com/blogs/2026-08-28-ai/</link><pubDate>Fri, 28 Aug 2026 13:50:52 +0000</pubDate><guid>http://brief-tech-news.com/blogs/2026-08-28-ai/</guid><description>Anthropic opened a research preview of the Model Hardware Standard, a model-agnostic spec for letting AI agents safely operate lab and manufacturing equipment in hours instead of weeks. Google released Gemini Omni 1.1 Flash with scene extension up to 40 seconds and 12 new creative controls. OpenAI and Anthropic&amp;rsquo;s combined annualized revenue run rate has reached $105 billion as of August 2026, tripling from $30 billion earlier this year. Nvidia guided for roughly $700 billion in FY28 revenue, crushing analyst estimates by over $125 billion, while OpenRouter data shows Luna token usage jumping 13.8x after the July discounts.</description></item><item><title>NVIDIA crosses $100B in a quarter as the rest of the AI stack scrambles to keep up</title><link>http://brief-tech-news.com/blogs/2026-08-27-ai/</link><pubDate>Thu, 27 Aug 2026 13:47:56 +0000</pubDate><guid>http://brief-tech-news.com/blogs/2026-08-27-ai/</guid><description>NVIDIA guided Q3 FY27 to $108B in revenue, crossing $100B in a single quarter for the first time by any company and annualizing to $432B, with neoclouds now driving the majority of Data Center net-new revenue. Z.ai confirmed the anonymous &amp;ldquo;ox-alpha&amp;rdquo; leaderboard-topper was GLM-5.3-Flash, a 320B-parameter MoE running on Chinese AI chips at one-tenth the cost of its prior model. Salesforce and Anthropic shipped &amp;ldquo;Claudeforce&amp;rdquo; as Anthropic&amp;rsquo;s ARR hit $65B, while Anthropic separately signed a $45B compute deal with Nscale for 460MW of Vera Rubin capacity in West Virginia.</description></item><item><title>OpenAI's Jalapeño hits the field, Perplexity goes local, and Anthropic trims Claude to 15k tokens</title><link>http://brief-tech-news.com/blogs/2026-08-26-ai/</link><pubDate>Wed, 26 Aug 2026 13:50:54 +0000</pubDate><guid>http://brief-tech-news.com/blogs/2026-08-26-ai/</guid><description>OpenAI posted first benchmarks for Jalapeño, an inference accelerator built for low-latency agent workloads, with deployment in its own fleet planned by year-end. Perplexity launched Portable Computer, a fully local agent that runs on existing RTX hardware with no token costs, available now for Pro and Enterprise subscribers on Linux. Anthropic&amp;rsquo;s current Claude tokenizer appears to use only ~15,000 entries, a sharp drop that researchers link to a softmax-layer gradient bottleneck rather than the prevailing bigger-vocab trend. Apple unveiled M6 and M5 Ultra silicon with up to 1.2TB/s memory bandwidth aimed at local AI, while OpenAI lost the executive running its data-center build-out. Dylan Patel projects OpenAI and Anthropic could control most usable global AI compute by 2028.</description></item><item><title>NVIDIA's Groq 3 LPX hits full production, a mystery model burns 26T tokens, and CUDA courts RISC-V</title><link>http://brief-tech-news.com/blogs/2026-08-25-ai/</link><pubDate>Tue, 25 Aug 2026 13:39:01 +0000</pubDate><guid>http://brief-tech-news.com/blogs/2026-08-25-ai/</guid><description>NVIDIA&amp;rsquo;s Groq 3 LPX inference accelerator is now in full production, slotting into Vera Rubin racks and promising a 4x response-time boost for agentic workloads. An anonymous model called Ox Alpha processed 26 trillion tokens on OpenCode in four days with no disclosed maker or knowledge cutoff. At Hot Chips 2026, NVIDIA signalled it wants to bring CUDA to RISC-V CPUs. Separate research argues LLM host machines are now the highest-value target in the datacenter, with inference engines themselves as the attack surface. Anthropic hired Google&amp;rsquo;s TPU founder as it pushes into custom silicon.</description></item><item><title>Hugging Face Explores $13B Sale as AI Pricing Wars Intensify</title><link>http://brief-tech-news.com/blogs/2026-08-24-ai/</link><pubDate>Mon, 24 Aug 2026 14:03:53 +0000</pubDate><guid>http://brief-tech-news.com/blogs/2026-08-24-ai/</guid><description>Hugging Face is exploring a sale at a $13 billion valuation—nearly 3x its 2023 valuation—reflecting the strategic value of its model hub and developer ecosystem. DeepSeek released V4-Flash-Vision-Exp, an experimental vision model that nearly matches Anthropic&amp;rsquo;s Opus 4.8 on agent benchmarks while maintaining text-model pricing. Anthropic is making its strongest model, Claude Mythos 5, available for defensive security work without exposing direct model access, backed by a $35 million open-source security fund. OpenAI cut GPT-5.6 Sol API pricing by over 20% for three months, deepening a price war that is reshaping corporate model spending as cheaper options like Anthropic&amp;rsquo;s Opus 5 now outpace flagship tiers in enterprise deployments.</description></item><item><title>Claude Sits In Your Meetings Now; Poolside Licenses to Nvidia for $6B</title><link>http://brief-tech-news.com/blogs/2026-08-21-ai/</link><pubDate>Fri, 21 Aug 2026 13:49:09 +0000</pubDate><guid>http://brief-tech-news.com/blogs/2026-08-21-ai/</guid><description>Anthropic is quietly building Project Parka, a meeting recorder that runs inside Claude Desktop and automatically assigns follow-up tasks to coding agents. Meanwhile, OpenAI added iMessage access to ChatGPT on Mac, Slack launched code channels for AI-human collaboration, and Mistral showed its new search loop triples FinanceBench accuracy. On the business side, Poolside AI signed a non-exclusive licensing deal with Nvidia for $6 billion plus a $1 billion investment, and Micron committed $10 billion over a decade to AI memory research.</description></item></channel></rss>