<?xml version="1.0" encoding="UTF-8"?><rss version="2.0" xmlns:content="http://purl.org/rss/1.0/modules/content/"><channel><title>Today in AI | ai.dosa.dev</title><description>A daily AI brief for builders: model releases, developer tools, funding, and policy moves that matter.</description><link>https://ai.dosa.dev/</link><language>en-us</language><item><title>OpenAI&apos;s GPT-6 Astra Rolls Out, Agents Outpace Human Researchers</title><link>https://ai.dosa.dev/blog/today-in-ai-2026-09-07/</link><guid isPermaLink="true">https://ai.dosa.dev/blog/today-in-ai-2026-09-07/</guid><description>GPT-6 Astra launches, OpenAI agents hit 3.1x human research effort, Anthropic&apos;s $100B AWS deal, and more.</description><pubDate>Mon, 07 Sep 2026 00:00:00 GMT</pubDate><content:encoded>&lt;p&gt;This brief covers AI news from 2026-09-07 UTC.&lt;/p&gt;&lt;p&gt;Today&amp;#39;s brief covers OpenAI&amp;#39;s Astra rollout, internal agent milestones, and major compute deals.&lt;/p&gt;&lt;h2&gt;OpenAI launches GPT-6 Astra, claims state-of-the-art across benchmarks&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; OpenAI introduced GPT-6 Astra, calling it the world&amp;#39;s most intelligent and aligned model. It scores 98% on FrontierMath Tier 4, 99.9% on ARC-AGI-3, and 100% on ExploitBench. Rolling out today to select organizations, then to all ChatGPT Plus, Pro, Business, and Enterprise users, plus API, Azure, and AWS Bedrock.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers get a model that excels in computer use, browsing, and software engineering, potentially reducing the need for multiple specialized tools.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://openai.com/index/gpt-6-astra/&quot;&gt;Source: OpenAI&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;OpenAI agents now do 3.1x the research work of humans&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; OpenAI reports that as of mid-August 2026, its AI agents perform 3.1 agent-workdays for every human workday in its research org. Median researchers spend over $600 daily on inference; top 10% exceed $7, 000. The company met its September goal for an automated research intern and targets a fully automated researcher by March 2028.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Shows agents can handle substantial research tasks, hinting at future tools that could automate complex workflows for developers.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://cryptobriefing.com/openai-ai-systems-3x-research-effort-humans/&quot;&gt;Source: Crypto Briefing&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Anthropic pledges $100B to AWS ahead of IPO&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Anthropic has committed over $100 billion to Amazon Web Services over a decade, securing up to 5 gigawatts of compute with Trainium2-4 and Graviton chips. The deal will be detailed in its upcoming S-1 filing. Anthropic&amp;#39;s annualized revenue run rate surpassed $65 billion by end of July.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Signals massive compute reservations and financial scale, but also potential concentration risk for Anthropic&amp;#39;s infrastructure.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://cryptobriefing.com/anthropic-100b-aws-ipo-prospectus/&quot;&gt;Source: Crypto Briefing&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Model fatigue hits as labs ship updates in one week&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Anthropic, Meta, Google, and OpenAI all released model updates in the same week, prompting CNBC to coin &amp;#39;model fatigue.&amp;#39; Runpod CEO Zhen Lu says the pace forces companies to make noise, while Notre Dame professor Ahmed Abbasi says labs are fighting for share of wallet.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers face harder choices and higher evaluation costs when picking models, as releases outpace adoption cycles.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.cnbc.com/2026/09/06/meta-google-openai-anthropic-ai-model-fatigue.html&quot;&gt;Source: CNBC&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Hugging Face ships TRL v1.0 for LLM fine-tuning&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Hugging Face released TRL v1.0, a production-ready framework standardizing post-training. It includes a unified CLI, config structure, and support for PPO, DPO, GRPO, KTO, and experimental ORPO, plus parameter-efficient fine-tuning.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Simplifies fine-tuning workflows, letting developers move between alignment methods without rewriting training stacks.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://dailysynapse.com/news/hugging-face-launches-trl-v1-0-for-llm-fine-tuning/&quot;&gt;Source: DailySynapse&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;India rules AI can&amp;#39;t author, but human can own AI-made work&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; India&amp;#39;s Copyright Office held that AI-generated works can be copyrighted if a human is identified as the author, rejecting AI authorship. The ruling came in Stephen Thaler&amp;#39;s application for &amp;#39;A Recent Entrance to Paradise&amp;#39; created by DABUS.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Clarifies IP rights for AI-generated assets, giving developers and businesses legal certainty when using generative tools.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.storyboard18.com/digital/indias-ai-copyright-ruling-human-authorship-key-ws-l-109838.htm&quot;&gt;Source: Storyboard18&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;EU AI compliance stack splits into three unsynchronized layers&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; The EU now layers the AI Act, Cyber Resilience Act, and MiCA/DORA with separate enforcement bodies and timelines. AI Act transparency rules are active since August 2, 2026; CRA vulnerability reporting starts September 11, 2026.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers deploying agents in the EU must juggle multiple compliance regimes, increasing legal and engineering overhead.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://forkast.news/the-eu-ai-compliance-stack-is-crystallizing-into-three-layers-and-none-of-them-talk-to-each-other/&quot;&gt;Source: Forkast&lt;/a&gt;&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ai.dosa.dev/blog/today-in-ai-2026-09-07/&quot;&gt;Read the full brief on ai.dosa.dev&lt;/a&gt;&lt;/p&gt;</content:encoded><category>openai</category><category>agents</category><category>anthropic</category><category>aws</category><category>hugging-face</category><category>regulation</category></item><item><title>Anthropic&apos;s Fable 5.1 and Mythos 5.1: Cheaper, Safer, Smarter</title><link>https://ai.dosa.dev/blog/today-in-ai-2026-09-06/</link><guid isPermaLink="true">https://ai.dosa.dev/blog/today-in-ai-2026-09-06/</guid><description>Anthropic launches Fable 5.1 with lower costs and fewer interruptions; Etched raises $700M; Seattle Times sues.</description><pubDate>Sun, 06 Sep 2026 00:00:00 GMT</pubDate><content:encoded>&lt;p&gt;This brief covers AI news from 2026-09-06 UTC.&lt;/p&gt;&lt;p&gt;Frontier model releases, big chip funding, and a new copyright suit land in one day.&lt;/p&gt;&lt;h2&gt;Anthropic ships Fable 5.1 and Mythos 5.1&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Anthropic released Claude Fable 5.1, generally available, and Mythos 5.1, restricted to trusted access programs. Fable 5.1 costs about 25% less for token-billed workloads and up to 45% less for agentic work, with new Enterprise Frontier Safeguards for customer-controlled data.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers get a frontier model with lower cache-read pricing and fewer false-positive safety blocks, plus a path to zero data retention before EFS rolls out.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.anthropic.com/claude-fable-and-mythos-5-1&quot;&gt;Source: Anthropic&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Etched raises $700M at $21B for transformer chips&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Etched closed a $700 million round at a $21 billion valuation, led by Jane Street, which is also an early customer. The company has over $1 billion in customer contracts and has shipped first server racks of its Sohu transformer-only ASIC.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Specialized inference silicon is attracting serious capital and customers, signaling a shift toward cheaper, faster transformer inference beyond GPUs.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://sentinel.ht/etched-700-million-transformer-inference-chips/&quot;&gt;Source: Sentinel&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Seattle Times and Newsday sue OpenAI and Microsoft&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; The Seattle Times and Newsday filed a federal lawsuit in the Southern District of New York alleging OpenAI and Microsoft scraped their journalism, including paywalled content, to train and operate ChatGPT, Copilot, and Bing AI. They seek damages and possible destruction of models.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; The case targets the full AI content pipeline, and its outcome could shape how AI companies handle paywalled data and fair use in training.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://techcrunch.com/2026/09/05/seattle-times-and-newsday-are-the-latest-publications-to-sue-openai-and-microsoft/&quot;&gt;Source: TechCrunch&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Tencent&amp;#39;s Hy4 Preview overwhelms WorkBuddy&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Tencent released Hy4 Preview, a 770B-parameter open-weight model under Apache 2.0, and the surge of users forced emergency scaling of its WorkBuddy platform within three days. The model uses Mixture-of-Experts with 49B active parameters and a 1M token context.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Open-weight frontier models are driving immediate infrastructure demands, and Hy4&amp;#39;s architecture targets long-horizon coding and agentic workflows.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.ainvest.com/news/tencent-hunyuan-hy4-launches-workbuddy-undergoes-emergency-scaling-2609/&quot;&gt;Source: AInvest&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Crusoe triples valuation to $30B on $3B raise&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Crusoe raised over $3 billion at around a $30 billion valuation, nearly triple its October 2025 figure, led by Atreides Management and Valor Equity Partners. The round follows a five-year, $13 billion cloud contract with Jane Street.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; AI infrastructure demand keeps inflating data center valuations, and Crusoe&amp;#39;s client list includes Meta, Microsoft, OpenAI, and Oracle.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://beckmann.ai/en/ai-economy/2026-09/crusoe-30-billion-valuation&quot;&gt;Source: Beckmann&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;GitHub&amp;#39;s HydraFusion routes coding across models&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; GitHub launched Project HydraFusion as a research preview in Copilot CLI. It builds a per-request workflow, choosing among single-model, cascade, or critique patterns, and can escalate to stronger models via quality gates.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers can now let the router decide which model or sequence of models to use, potentially cutting costs by using expensive inference only when needed.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.datastudios.org/post/github-launches-hydrafusion-multi-model-orchestration-dynamic-routing-lower-cost-coding-and-the&quot;&gt;Source: Data Studios&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;ChatGPT designated EU very large online search engine&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; The European Commission classified ChatGPT as a Very Large Online Search Engine after it declared at least 45 million average monthly users in the EU. OpenAI has four months, until January 2027, to meet additional Digital Services Act obligations.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; ChatGPT now faces stricter EU transparency and risk-management rules, which could affect how OpenAI deploys features in Europe.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://quasa.io/insights/chatgpt-becomes-an-eu-search-engine-and-gets-four-months-to-comply&quot;&gt;Source: Quasa&lt;/a&gt;&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ai.dosa.dev/blog/today-in-ai-2026-09-06/&quot;&gt;Read the full brief on ai.dosa.dev&lt;/a&gt;&lt;/p&gt;</content:encoded><category>anthropic</category><category>funding</category><category>copyright</category><category>openai</category><category>chips</category><category>regulation</category></item><item><title>CrowdStrike, GitHub Ship New AI Tools; DOJ Backs Fair Use</title><link>https://ai.dosa.dev/blog/today-in-ai-2026-09-05/</link><guid isPermaLink="true">https://ai.dosa.dev/blog/today-in-ai-2026-09-05/</guid><description>CrowdStrike launches security models, GitHub&apos;s router cuts costs, DOJ backs OpenAI, plus more.</description><pubDate>Sat, 05 Sep 2026 00:00:00 GMT</pubDate><content:encoded>&lt;p&gt;This brief covers AI news from 2026-09-05 UTC.&lt;/p&gt;&lt;p&gt;Today&amp;#39;s brief covers new AI launches, funding, and a key legal filing.&lt;/p&gt;&lt;h2&gt;CrowdStrike unveils SafeMind security models with Nvidia&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; CrowdStrike introduced SafeMind, a family of purpose-built security models and harnesses, built with Nvidia Nemotron and trained on Falcon telemetry. It includes Red Tempest (offensive) and Blue Solano (defensive) models.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Security teams get an agentic system that automates attack path finding and defense, potentially reducing false positives and improving response.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.crowdstrike.com/en-us/press-releases/crowdstrike-launches-frontier-models-for-cybersecurity-with-nvidia/&quot;&gt;Source: CrowdStrike&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;GitHub&amp;#39;s HydraFusion router cuts coding costs up to 67%&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; GitHub announced Project HydraFusion, a research preview in Copilot that dynamically routes coding tasks across models. In tests, it beat Claude Opus 5 on TerminalBench 2.1 by 4.9 points while costing 67% less.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers can get frontier-level coding quality at a fraction of the cost by letting a router pick the best model per task.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://cryptobriefing.com/github-hydrafusion-ai-coding-router/&quot;&gt;Source: Crypto Briefing&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;DOJ tells court AI training is fair use in OpenAI case&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; The U.S. Department of Justice filed a statement of interest in the New York Times v. OpenAI lawsuit, arguing that training LLMs on copyrighted works is transformative fair use and that requiring licensing would hamper competition.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; A ruling favoring fair use would keep training costs low for AI builders, while an adverse ruling could create licensing barriers favoring big tech.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://publishingperspectives.com/2026/09/u-s-department-of-justice-files-legal-brief-urging-court-to-find-ai-training-to-be-fair-use/&quot;&gt;Source: Publishing Perspectives&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Nscale commits $3.5B to Figure AI for humanoid compute&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; AI cloud provider Nscale signed a multi-year agreement to supply at least $3.5 billion in GPU infrastructure to Figure AI, with plans to scale beyond $6 billion, and took an undisclosed equity stake.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Humanoid robotics training requires massive compute; this deal secures infrastructure and signals new financing models for physical AI.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.techtimes.com/articles/326591/20260904/nscale-backs-figure-ai-35-billion-takes-equity-stake-humanoid-push.htm&quot;&gt;Source: Tech Times&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Ant Ling releases visual agent model Ling-3.0-flash-VL&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Ant Ling released Ling-3.0-flash-VL, a version of its language model that accepts images and video, can generate websites from screenshots, and operate GUIs via its hosted API.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers can build agents that see and interact with software, enabling automated UI testing and visual coding workflows.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://runtimewire.com/article/ant-ling-releases-ling-3-0-flash-vl-visual-agents&quot;&gt;Source: RuntimeWire&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;SoundHound closes LivePerson acquisition, appoints CFO&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; SoundHound AI completed its acquisition of LivePerson, expanding its customer base to 25 of the Fortune 100 and adding over 750 patents. John Collins was appointed CFO of the combined company.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; The combined platform integrates voice and digital messaging, giving enterprises a unified conversational AI system across channels.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://aifuturefront.com/soundhound-ai-completes-acquisition-of-liveperson-creating-a-world-leading-omnichannel-conversational-ai-powerhouse/&quot;&gt;Source: AI Future Front&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Gimlet Labs raises $300M at $3B for multi-chip AI&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Gimlet Labs raised $300 million at a $3 billion valuation, led by Andreessen Horowitz, with new backers Arm and Microsoft&amp;#39;s M12. Its software routes AI workloads across different chip types.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; As GPU scarcity bites, multi-silicon orchestration can cut inference costs and speed up workloads by using the best chip for each task.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://techfundingnews.com/andreessen-backed-gimlet-labs-hits-3b-valuation-with-300m-round-as-ai-goes-multi-chip/&quot;&gt;Source: Tech Funding News&lt;/a&gt;&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ai.dosa.dev/blog/today-in-ai-2026-09-05/&quot;&gt;Read the full brief on ai.dosa.dev&lt;/a&gt;&lt;/p&gt;</content:encoded><category>security</category><category>coding</category><category>legal</category><category>funding</category><category>agents</category><category>infrastructure</category></item><item><title>Nvidia buys Hugging Face for $12.9B; OpenAI ships GPT-6 Astra</title><link>https://ai.dosa.dev/blog/today-in-ai-2026-09-04/</link><guid isPermaLink="true">https://ai.dosa.dev/blog/today-in-ai-2026-09-04/</guid><description>Nvidia acquires Hugging Face, OpenAI launches GPT-6 Astra, IFM open-sources K2 Horizon, and more.</description><pubDate>Fri, 04 Sep 2026 00:00:00 GMT</pubDate><content:encoded>&lt;p&gt;This brief covers AI news from 2026-09-04 UTC.&lt;/p&gt;&lt;p&gt;Chipmaker Nvidia is buying Hugging Face, OpenAI says GPT-6 Astra ushers in the AGI era, and IFM drops a fully open model fleet.&lt;/p&gt;&lt;h2&gt;Nvidia to acquire Hugging Face for $12.93B&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Nvidia announced it agreed to acquire Hugging Face for $12, 930, 300, 000. Hugging Face will remain an open platform, and Nvidia compute will not be required to build on or deploy through it.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers get a financially secure home for open models, but Nvidia&amp;#39;s control of the leading hub could shape how open-weight AI is distributed and run.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://blogs.nvidia.com/blog/nvidia-to-acquire-hugging-face/&quot;&gt;Source: NVIDIA&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;OpenAI launches GPT-6 Astra, claims AGI era&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; OpenAI released GPT-6 Astra, calling it a generational leap in capability. President Greg Brockman said &amp;#39;we are now in the AGI era.&amp;#39; The model rolls out to enterprise cybersecurity customers today, with broader access in coming days.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; If GPT-6 Astra delivers on agentic and software engineering claims, it could reset expectations for what AI assistants can automate in production.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.theverge.com/ai-artificial-intelligence/989601/openai-gpt-6-astra-release&quot;&gt;Source: The Verge&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;IFM releases K2 Horizon, six fully open models&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; The Institute of Foundation Models released K2 Horizon, a fleet of six models from 0.9B to 375B parameters, with full training data, code, and checkpoints under Apache 2.0. The 0.9B, 3.7B, and 7B models set state of the art in their size classes.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers can now inspect, reproduce, and adapt the entire training pipeline, not just weights, enabling true open-source research and customization.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ifm.ai/blog/k2/&quot;&gt;Source: IFM&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;DOJ backs OpenAI in NYT copyright case&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; The U.S. Department of Justice filed a Statement of Interest arguing that training LLMs on copyrighted works is fair use, saying a contrary ruling would harm competition and national security. The filing is non-binding.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; If courts follow this reasoning, AI developers may avoid licensing costs for training data, but the non-binding nature leaves uncertainty for ongoing litigation.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ipwatchdog.com/2026/09/03/doj-sides-with-openai-warns-obstacles-to-ai-development-threaten-national-security/&quot;&gt;Source: IPWatchdog&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;ServiceNow acquires Israeli AI startup Sweep&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; ServiceNow acquired Sweep, an Israeli startup building AI agents that work across enterprise platforms, in a deal estimated at hundreds of millions of dollars. Sweep had raised $46 million in disclosed funding.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; ServiceNow&amp;#39;s move signals deeper integration of agentic AI into CRM workflows, potentially accelerating autonomous work in enterprise software.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.calcalistech.com/ctechnews/article/r19lbzwdfx&quot;&gt;Source: Calcalistech&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Cohere launches Parse 5 for document extraction&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Cohere released Parse 5, a 2.3B-parameter multimodal model that converts complex PDFs into Markdown with bounding boxes. It uses an 8K-token context window and is built on Cohere&amp;#39;s North-Micro-Vision-Instruct architecture.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers can extract structured data from visually rich documents more efficiently, reducing the need for custom parsing pipelines.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.infoq.com/news/2026/09/cohere-multimodal-parse/&quot;&gt;Source: InfoQ&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Nvidia accelerates local AI with RTX Spark, PAIR&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; At IFA 2026, Nvidia announced RTX Spark PCs coming in October, a Personal AI Router tool called PAIR, and up to 1.9x faster local inference via llama.cpp and vLLM optimizations. New open models from Qwen, Z.ai, and others run locally on Nvidia hardware.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers can run capable agents and models on local PCs with better performance, reducing reliance on cloud APIs for inference.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://blogs.nvidia.com/blog/local-ai-ifa-next-gen-agents-nv-pair-rtx-spark/&quot;&gt;Source: NVIDIA&lt;/a&gt;&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ai.dosa.dev/blog/today-in-ai-2026-09-04/&quot;&gt;Read the full brief on ai.dosa.dev&lt;/a&gt;&lt;/p&gt;</content:encoded><category>nvidia</category><category>openai</category><category>open-source</category><category>acquisition</category><category>copyright</category><category>agents</category></item><item><title>Google launches Gemini 3.8 Flash and Flash Cyber</title><link>https://ai.dosa.dev/blog/today-in-ai-2026-09-03/</link><guid isPermaLink="true">https://ai.dosa.dev/blog/today-in-ai-2026-09-03/</guid><description>Google debuts two Gemini 3.8 models, Anthropic&apos;s Fable 5.1 tops benchmarks, DOJ backs OpenAI in copyright case.</description><pubDate>Thu, 03 Sep 2026 00:00:00 GMT</pubDate><content:encoded>&lt;p&gt;This brief covers AI news from 2026-09-03 UTC.&lt;/p&gt;&lt;p&gt;Today&amp;#39;s brief covers new model releases, a major copyright filing, and big funding moves.&lt;/p&gt;&lt;h2&gt;Google unveils Gemini 3.8 Flash and Flash Cyber&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Google introduced Gemini 3.8 Flash, a general-purpose model, and Gemini 3.8 Flash Cyber for cybersecurity, both at the same price as 3.7 Flash. Google says 3.8 Flash outperformed Claude Opus 5 and GPT-5.6 Sol on nine of 16 benchmarks.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers get a cheaper, faster model with strong coding and agentic performance, plus a specialized cyber variant for vulnerability detection.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://blog.google/innovation-and-ai/models-and-research/gemini-models/3-8-flash-and-3-8-flash-cyber/&quot;&gt;Source: Google&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;DOJ backs OpenAI in NYT copyright fight&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; The US Department of Justice filed a statement of interest in the New York Times v. OpenAI case, arguing that training AI on copyrighted works is fair use. The Times responded that the administration is siding with AI companies at creators&amp;#39; expense.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; A ruling for OpenAI could cement fair use for AI training, reducing legal risk for developers building on large language models.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://techcrunch.com/2026/09/02/u-s-government-sides-with-openai-on-issue-of-training-llms-on-copyrighted-material/&quot;&gt;Source: TechCrunch&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Anthropic&amp;#39;s Fable 5.1 tops AI benchmarks&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Anthropic launched Claude Fable 5.1 (general availability) and Mythos 5.1 (trusted access). Fable 5.1 scored 66 on Artificial Analysis&amp;#39; Intelligence Index, six points ahead of Chinese rivals, and costs 25% less than Fable 5 for typical workloads.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Fable 5.1 offers top-tier coding and agentic performance at a lower price, with enterprise safeguards promising zero data retention.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.scmp.com/tech/big-tech/article/3366125/frontier-ai-cost-what-anthropics-fable-51-means-us-china-model-race&quot;&gt;Source: South China Morning Post&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Palo Alto Networks buys Console for $500M&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Palo Alto Networks acquired Console, a two-year-old startup using AI agents for IT help desk automation, for $500M in cash and stock. Console had raised $29M and was valued at $157M pre-sale.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; The deal signals strong M&amp;amp;amp;A appetite for agentic AI tools, and Console&amp;#39;s tech will be integrated into Cortex for autonomous security operations.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://techcrunch.com/2026/09/02/palo-alto-networks-paid-500m-for-thrive-backed-console-sources-say/&quot;&gt;Source: TechCrunch&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Wonderful raises $550M at $5B valuation&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Wonderful, an AI startup founded in early 2025, closed a $550M Series C led by Insight Partners, with Salesforce joining. The valuation more than doubled to $5B in six months. The company now has 650 employees across 35+ markets.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Wonderful&amp;#39;s AI OS platform is model-agnostic and deployable on any cloud, offering enterprises a flexible way to integrate AI without vendor lock-in.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://techcrunch.com/2026/09/02/wonderful-more-than-doubles-its-valuation-to-5b-in-under-6-months/&quot;&gt;Source: TechCrunch&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Vertiv to acquire UIG for $1.45B plus earnout&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Vertiv agreed to buy Utility Innovation Group (UIG) for about $1.45B in cash, plus up to $1.15B in earnouts tied to EBITDA targets. The deal adds microgrid and energy storage capabilities for AI data centers.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Power-constrained AI data centers need more energy infrastructure; this acquisition expands Vertiv&amp;#39;s ability to deliver behind-the-meter power solutions.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.stocktitan.net/sec-filings/VRT/8-k-vertiv-holdings-co-reports-material-event-694bfe7b18b9.html&quot;&gt;Source: StockTitan&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Google speeds up TPU rollout to two chips a year&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Google&amp;#39;s AI infrastructure chief said the company is shifting from one TPU generation per year to two chips per year, citing the need for more differentiation between inference and training designs. The announcement came at Semicon Taiwan.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Faster TPU cadence means developers get more specialized hardware sooner, potentially improving performance and cost for AI workloads.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://asia.nikkei.com/business/technology/artificial-intelligence/google-to-speed-up-chip-rollout-to-stay-ahead-in-ai-technology-chief-says&quot;&gt;Source: Nikkei Asia&lt;/a&gt;&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ai.dosa.dev/blog/today-in-ai-2026-09-03/&quot;&gt;Read the full brief on ai.dosa.dev&lt;/a&gt;&lt;/p&gt;</content:encoded><category>google</category><category>anthropic</category><category>openai</category><category>funding</category><category>chips</category><category>policy</category></item><item><title>Anthropic&apos;s Fable 5.1, Google&apos;s Agentic Video, and More</title><link>https://ai.dosa.dev/blog/today-in-ai-2026-09-02/</link><guid isPermaLink="true">https://ai.dosa.dev/blog/today-in-ai-2026-09-02/</guid><description>Anthropic launches Fable 5.1, Google adds agentic video, World Labs debuts Atlas, Cognition raises $1B, and more.</description><pubDate>Wed, 02 Sep 2026 00:00:00 GMT</pubDate><content:encoded>&lt;p&gt;This brief covers AI news from 2026-09-02 UTC.&lt;/p&gt;&lt;p&gt;A packed day of model launches and funding news.&lt;/p&gt;&lt;h2&gt;Anthropic debuts Fable 5.1, Mythos 5.1&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Anthropic released Claude Fable 5.1 and Mythos 5.1, with Fable 5.1 generally available on AWS Bedrock and the Anthropic API. The model sets records on Terminal-Bench 4.0 and Terminal-Bench-Science 0.1, and is 25% more cost-efficient on typical workloads.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Fable 5.1&amp;#39;s improved prompt caching cuts costs up to 45% for agent-heavy apps, and relaxed guardrails allow broader use cases like vulnerability discovery.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/introducing-claude-fable-5-1-on-aws/&quot;&gt;Source: AWS&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Google launches agentic video understanding&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Google DeepMind launched agentic video understanding for Gemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite, available via the Gemini API. It cuts token consumption by up to 88% and costs by up to 66% while improving accuracy by up to 7%.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers can now process long videos with sub-second moment retrieval and anomaly detection at a fraction of the cost.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://blog.google/innovation-and-ai/models-and-research/gemini-models/introducing-agentic-video-in-gemini/&quot;&gt;Source: Google&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;World Labs debuts Atlas world model&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; World Labs released Atlas, a multimodal autoregressive diffusion transformer pretrained on text, images, video, and 3D. It claims 81-93% user preference over rivals on camera-controlled generation and outperforms specialized 3D reconstruction models.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; A single model handling text, image, video, and 3D could simplify pipelines for developers building spatial AI and robotics applications.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://aiweekly.co/alerts/world-labs-debuts-atlas-an-omni-world-model-in-early-access&quot;&gt;Source: AI Weekly&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Cognition targets $1B at $47B valuation&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Bloomberg reports Cognition is raising about $1B at a $47B valuation, three months after its $26B round. The deal is under negotiation with nearly $10B in investor interest.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; A near-doubling in valuation in three months signals intense investor demand for AI coding agents like Devin.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://runtimewire.com/article/cognition-1-billion-funding-47-billion-valuation&quot;&gt;Source: Runtimewire&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;EU sends first AI Act info requests&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; The European Commission sent formal information requests under the AI Act to over 30 general-purpose AI model makers, including OpenAI, Google, and Anthropic, following July containment failures. Companies face fines up to €15M for incorrect replies.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Builders using frontier models may see new compliance requirements and scrutiny on model security and monitoring.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://euperspectives.eu/2026/09/the-ai-act-gives-brussels-new-powers-frontier-labs-are-first-in-line/&quot;&gt;Source: EU Perspectives&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Flower Labs unveils Endeavor 1.0&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Flower Labs introduced Endeavor 1.0, a frontier-class generalist model for reasoning and coding, available as a preview to select organizations. It follows their sovereign 7B model Lizzy.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Endeavor offers an alternative to closed APIs, deployable on your own infrastructure for greater control.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://flower.ai/blog/2026-09-01-introducing-endeavor-1.0&quot;&gt;Source: Flower Labs&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Meta&amp;#39;s Muse Image hits fal&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; fal launched developer access to Meta&amp;#39;s Muse Image, an agentic image model that plans, searches, and self-corrects before rendering. It ranks top-5 on Arena across text-to-image and editing.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; One model now handles generation, editing, and multi-turn refinement, potentially replacing multiple vendors for production image workflows.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.prnewswire.com/news-releases/muse-image-the-agentic-image-model-from-meta-is-now-available-on-fal-302866384.html&quot;&gt;Source: PR Newswire&lt;/a&gt;&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ai.dosa.dev/blog/today-in-ai-2026-09-02/&quot;&gt;Read the full brief on ai.dosa.dev&lt;/a&gt;&lt;/p&gt;</content:encoded><category>anthropic</category><category>google</category><category>world-labs</category><category>funding</category><category>regulation</category><category>image-models</category></item><item><title>Nvidia invests $3.5B in MediaTek for custom AI chips</title><link>https://ai.dosa.dev/blog/today-in-ai-2026-09-01/</link><guid isPermaLink="true">https://ai.dosa.dev/blog/today-in-ai-2026-09-01/</guid><description>Nvidia&apos;s $3.5B MediaTek deal, Anthropic&apos;s $35B Lambda cloud pact, DeepSeek vision weights, and more.</description><pubDate>Tue, 01 Sep 2026 00:00:00 GMT</pubDate><content:encoded>&lt;p&gt;This brief covers AI news from 2026-09-01 UTC.&lt;/p&gt;&lt;p&gt;Today&amp;#39;s brief covers Nvidia&amp;#39;s MediaTek investment, Anthropic&amp;#39;s cloud deal, and new model releases.&lt;/p&gt;&lt;h2&gt;Nvidia invests $3.5B in MediaTek for custom AI chips&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Nvidia announced a $3.5 billion investment in MediaTek via convertible bonds. MediaTek will adopt Nvidia&amp;#39;s NVLink Fusion platform to help customers design custom XPUs that integrate into Nvidia-based data centers.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers can build custom accelerators that plug into Nvidia&amp;#39;s rack-scale infrastructure, reducing reliance on off-the-shelf GPUs.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.globenewswire.com/news-release/2026/08/31/3353306/0/en/nvidia-and-mediatek-deepen-long-standing-partnership-to-build-ai-edge-to-cloud-computing-platforms.html&quot;&gt;Source: NVIDIA&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Anthropic signs $35B cloud deal with Lambda&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Anthropic signed a $35 billion cloud-computing deal with Nvidia-backed Lambda for a Texas data center developed by Hut 8, covering about 350 megawatts of capacity.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Anthropic secures massive compute for Claude, easing GPU bottlenecks for its AI services.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.channelnewsasia.com/business/anthropic-signs-35-billion-cloud-deal-nvidia-backed-lambda-source-says-6353306&quot;&gt;Source: CNA&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;DeepSeek releases V4-Flash-Vision open weights&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; DeepSeek published open weights for DeepSeek-V4-Flash-Vision-Exp, a 305B-parameter multimodal model, under an MIT license. The 168GB repository includes reference inference code.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers can run and fine-tune a frontier vision model locally, enabling multimodal agents without API dependency.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://runtimewire.com/article/deepseek-publishes-v4-flash-vision-open-weights&quot;&gt;Source: RuntimeWire&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Runway unveils Solaris, an interface world model&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Runway introduced Solaris, the first model in a new family called Interface World Models. It generates apps and websites in real time, frame by frame, without an intermediate code representation.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Could eliminate the design-to-code step, letting developers build interactive interfaces directly from models.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://runwayml.com/news/research/introducing-solaris&quot;&gt;Source: Runway&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Perceptron AI launches Isaac 0.5 robotics model&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Perceptron AI released Isaac 0.5, a 36B-parameter open-weight embodied foundation model that combines video understanding, reasoning, and robot control. It outperforms open rivals like π0.5 and GR00T N1.7.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Robotics teams get a frontier open model to control robots or integrate into planning systems, reducing need for proprietary solutions.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ittech-pulse.com/news/perceptron-ai-launches-isaac-0-5-a-frontier-open-weight-robotics-model/&quot;&gt;Source: IT Tech Pulse&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Nvidia unveils Nemotron 3.5 Lightning and NeMo Switchyard&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Nvidia released Nemotron 3.5 Lightning, a 30B-parameter MoE model for agentic workloads, and NeMo Switchyard, an open-source library for routing requests to the best model in multi-agent systems.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers can build efficient multi-agent systems with smart routing across open and proprietary models without rewriting apps.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://futuretechmarkets.com/markets/nvidia-nemotron-3-5-lightning-and-nemo-switchyard-deliver-faster-smarter-more-efficient-agentic-ai-25/&quot;&gt;Source: Future Tech Markets&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;EU designates ChatGPT as a search engine under DSA&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; The European Commission classified ChatGPT as a Very Large Online Search Engine under the DSA, alongside Reddit and Roblox as VLOPs. ChatGPT has 159M monthly EU users.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; ChatGPT now faces strict EU compliance: systemic risk assessments, transparency reports, and data access for researchers, affecting how it operates in Europe.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://thenextweb.com/news/hatgpt-reddit-roblox-dsa-designation&quot;&gt;Source: TNW&lt;/a&gt;&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ai.dosa.dev/blog/today-in-ai-2026-09-01/&quot;&gt;Read the full brief on ai.dosa.dev&lt;/a&gt;&lt;/p&gt;</content:encoded><category>nvidia</category><category>anthropic</category><category>deepseek</category><category>runway</category><category>robotics</category><category>regulation</category></item><item><title>Stripe Buys OpenRouter for $7B; Nvidia Circles Hugging Face</title><link>https://ai.dosa.dev/blog/today-in-ai-2026-08-31/</link><guid isPermaLink="true">https://ai.dosa.dev/blog/today-in-ai-2026-08-31/</guid><description>Stripe acquires OpenRouter at 5x valuation, Nvidia in talks for Hugging Face, SoftBank offers OpenAI warrants, and more.</description><pubDate>Mon, 31 Aug 2026 00:00:00 GMT</pubDate><content:encoded>&lt;p&gt;This brief covers AI news from 2026-08-31 UTC.&lt;/p&gt;&lt;p&gt;Today&amp;#39;s brief covers major acquisitions, funding rounds, and new model releases.&lt;/p&gt;&lt;h2&gt;Stripe acquires OpenRouter for $7B, 5x May valuation&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Stripe has acquired OpenRouter for $7 billion, five times its valuation from May, according to a report from NBTC News.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; OpenRouter is a key gateway for developers to access multiple AI models; Stripe&amp;#39;s ownership could change pricing and access for AI builders.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://news.nbtc.finance/stripe-acquires-openrouter-for-7b-five-times-its-may-valuation/&quot;&gt;Source: NBTC News&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Nvidia in talks to buy Hugging Face for $12.9B&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; The Information reported Nvidia agreed to acquire Hugging Face for $12.9 billion, but neither company has confirmed; Business Insider says talks haven&amp;#39;t produced a signed deal.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Hugging Face is the largest open model hub; Nvidia ownership could reshape model distribution and inference routing for developers.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ai2.work/blog/nvidia-s-12-9b-hugging-face-deal-puts-open-ai-under-chip-control&quot;&gt;Source: AI2Work&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;SoftBank offers OpenAI $5.5B in warrants for data center commitment&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; SB Energy, a SoftBank subsidiary, has offered OpenAI stock warrants valued at roughly $5.5 billion to secure its long-term commitment as a tenant in its data center campuses.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; The deal ties OpenAI&amp;#39;s compute infrastructure to SoftBank&amp;#39;s buildout, potentially affecting OpenAI&amp;#39;s costs and capacity for developers.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://cryptobriefing.com/softbank-openai-data-center-warrants/&quot;&gt;Source: Crypto Briefing&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Tencent&amp;#39;s Hy4 preview: 770B params, Apache 2.0, $0.83/M input&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Tencent released Hy4 preview, a 770B-parameter MoE model with 49B active, 1M-token context, under Apache 2.0, priced at $0.83 per million input tokens on OpenRouter.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Open-weight, low-cost coding model with permissive license gives developers a viable alternative to proprietary APIs for agentic coding.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://byteiota.com/tencent-hy4-open-weights-model-benchmarks-and-access/&quot;&gt;Source: byteiota&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;AWS open-sources Kiro Crew for async coding agents&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Amazon announced Kiro Crew, an open-source system for running multiple Kiro coding agents asynchronously, with shared memory, scheduled jobs, and MCP integration.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers can now run coding agents without babysitting, enabling parallel tasks and unattended workflows.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.infoq.com/news/2026/08/kiro-crew-coding-agents/&quot;&gt;Source: InfoQ&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Huawei Cloud launches CodeArts Agent in Singapore&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Huawei Cloud commercially launched CodeArts Agent in Singapore, an AI coding tool with 16 specialized agents covering the full software development lifecycle.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Expands AI coding tool availability in Asia, offering an alternative for enterprises in the region.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://itbrief.asia/story/huawei-cloud-launches-the-codearts-agent-in-singapore&quot;&gt;Source: IT Brief Asia&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Perplexity seeks $30B valuation with Nvidia investment&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Perplexity AI is reportedly in talks to raise a new round at a $30 billion valuation, with Nvidia negotiating to participate.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Nvidia&amp;#39;s potential investment could secure compute for Perplexity&amp;#39;s inference-heavy search, affecting AI search competition.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://i10x.ai/news/perplexity-ai-30-billion-valuation-nvidia&quot;&gt;Source: i10x.ai&lt;/a&gt;&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ai.dosa.dev/blog/today-in-ai-2026-08-31/&quot;&gt;Read the full brief on ai.dosa.dev&lt;/a&gt;&lt;/p&gt;</content:encoded><category>acquisitions</category><category>funding</category><category>open-source</category><category>models</category><category>coding-agents</category><category>chips</category></item><item><title>OpenAI Cuts Cursor Access; Sony, Warner Sue Anthropic</title><link>https://ai.dosa.dev/blog/today-in-ai-2026-08-30/</link><guid isPermaLink="true">https://ai.dosa.dev/blog/today-in-ai-2026-08-30/</guid><description>OpenAI ends Cursor model access, music publishers sue Anthropic, and new AI chips and models debut.</description><pubDate>Sun, 30 Aug 2026 00:00:00 GMT</pubDate><content:encoded>&lt;p&gt;This brief covers AI news from 2026-08-30 UTC.&lt;/p&gt;&lt;p&gt;Today&amp;#39;s brief covers OpenAI&amp;#39;s Cursor cutoff, a major copyright lawsuit, and new silicon and model releases.&lt;/p&gt;&lt;h2&gt;OpenAI to Cut Off Cursor Model Access by Nov. 12&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; OpenAI plans to wind down its contract supplying models to Cursor, with a proposed Nov. 12 cutoff, citing inability to trust SpaceX&amp;#39;s compliance with terms. Cursor CEO says OpenAI models serve about 5% of user traffic.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers using Cursor may lose access to OpenAI models, forcing a switch to alternatives or self-hosted options.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.cnbc.com/2026/08/29/openai-cursor-spacex-model-access.html&quot;&gt;Source: CNBC&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Sony, Warner Chappell Sue Anthropic Over Training Data&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Sony Music Publishing and Warner Chappell filed a lawsuit against Anthropic and co-founders, alleging illegal torrenting and scraping of copyrighted works for Claude training. They seek up to $150, 000 per work.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; This adds legal risk for AI training data sourcing, pushing developers to ensure provenance and licensing.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://techcrunch.com/2026/08/29/sony-music-warner-sue-anthropic-alleging-a-brazen-campaign-of-intellectual-property-theft/&quot;&gt;Source: TechCrunch&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;OpenAI&amp;#39;s Jalapeño Chip Beats Nvidia on Power and Latency&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; OpenAI published benchmarks for its Jalapeño inference chip, showing 1.5x-1.9x better work per watt and 1.7x-3.6x lower latency than Nvidia GB300, at half the power draw.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Custom silicon could lower inference costs and improve agent performance for OpenAI API users.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://byteiota.com/openai-jalapeno-chip-benchmarks-developer-guide/&quot;&gt;Source: byteiota&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Z.ai&amp;#39;s GLM-5.3-Flash Debuts on Cloudflare Workers AI&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Cloudflare&amp;#39;s AI Search now supports GLM-5.3-Flash, a model with a 1, 048, 576-token context window, running on Workers AI.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers can now use this efficient model on Cloudflare&amp;#39;s edge platform for long-context tasks.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://developers.cloudflare.com/changelog/post/2026-08-30-glm-5.3-flash/&quot;&gt;Source: Cloudflare&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Salesforce Commits $600M to Make Claude Default&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Salesforce and Anthropic announced &amp;#39;Claudeforce, &amp;#39; making Claude the default reasoning engine across Salesforce products, with $300M in token procurement and a $300M equity stake.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Enterprises on Salesforce will get Claude integrated by default, potentially simplifying AI adoption but reducing model choice.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.winzheng.com/en/article/salesforce-anthropic-claudeforce-enterprise-ai-saas-binding-&quot;&gt;Source: Winzheng&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Anthropic Abandons $7B MatX Acquisition&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Anthropic reportedly walked away from acquiring AI chip startup MatX for ~$7B, shifting toward a potential partnership instead.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Anthropic&amp;#39;s chip strategy remains uncertain, potentially affecting its ability to control inference costs.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://digitrendz.blog/tech-news/243392/anthropic-abandons-7b-matx-acquisition/&quot;&gt;Source: DigitrendZ&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Agentrys Raises $24.5M for Agentic Chip Design&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Agentrys raised $24.5M to build an agentic design automation platform for chip development, claiming over 90% accuracy on NVIDIA&amp;#39;s CVDP benchmark.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; AI agents that design chips could accelerate hardware development and reduce engineering costs.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://pulse2.com/agentrys-raises-24-5-million-as-agentic-chip-design-platform-tops-90-on-nvidia-benchmark/&quot;&gt;Source: Pulse2&lt;/a&gt;&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ai.dosa.dev/blog/today-in-ai-2026-08-30/&quot;&gt;Read the full brief on ai.dosa.dev&lt;/a&gt;&lt;/p&gt;</content:encoded><category>openai</category><category>anthropic</category><category>chips</category><category>models</category><category>funding</category><category>legal</category></item><item><title>Tencent open-sources Hy4 preview, 770B-param model</title><link>https://ai.dosa.dev/blog/today-in-ai-2026-08-29/</link><guid isPermaLink="true">https://ai.dosa.dev/blog/today-in-ai-2026-08-29/</guid><description>Tencent&apos;s Hy4 preview, Z.ai&apos;s GLM-5.3 weights, a16z&apos;s $1.1B fund, Lambda&apos;s $1B debt, and more.</description><pubDate>Sat, 29 Aug 2026 00:00:00 GMT</pubDate><content:encoded>&lt;p&gt;This brief covers AI news from 2026-08-29 UTC.&lt;/p&gt;&lt;p&gt;Open-weight models, big money for AI infrastructure, and a landmark FTC ruling shape today&amp;#39;s news.&lt;/p&gt;&lt;h2&gt;Tencent open-sources Hy4 preview with 770B parameters&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Tencent released and open-sourced Hy4 preview, a 770B-parameter MoE model with 49B active parameters and a 1M-token context window. It&amp;#39;s available via WorkBuddy, CodeBuddy, Yuanbao, and API through Tencent Cloud TokenHub and OpenRouter.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers get a top-tier open model for coding and productivity, with free access for two weeks on WorkBuddy and CodeBuddy.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.tencent.com/tencent-releases-and-open-sources-tencent-hy4-preview/&quot;&gt;Source: Tencent&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Z.ai releases GLM-5.3 weights for coding and security&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Z.ai released GLM-5.3 weights, a 756GB model with 256 routed experts and 1M-token context, under a custom license. It scored 28.3 on Terminal-Bench 3.0, up from 4.6 for GLM-5.2.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers can self-host a model built for long-running coding and vulnerability hunting, with support for Transformers, vLLM, and SGLang.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://runtimewire.com/article/zai-releases-glm-5-3-open-weights-coding-cyber-defense&quot;&gt;Source: RuntimeWire&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;a16z raises $1.1B Machine Age fund for AI hardware&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Andreessen Horowitz launched a $1.1B fund focused on AI infrastructure: chips, memory, data centers, and robots. The firm says it&amp;#39;s time to &amp;#39;accelerate the physical buildout of AI.&amp;#39;&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; More capital for hardware startups could ease supply constraints that bottleneck AI development.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://techcrunch.com/2026/08/28/a16z-creates-a-1-1b-machine-age-fund-to-accelerate-the-physical-buildout-of-ai/&quot;&gt;Source: TechCrunch&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Lambda raises $1B in debt to buy Nvidia chips&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Lambda secured $1B in private short-dated debt, arranged by JP Morgan, to buy Nvidia AI chips that it will lease to Microsoft. This follows a $926M loan for GB300 GPUs.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Neoclouds are scaling GPU supply fast, giving developers more rental options for high-end chips.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://techcrunch.com/2026/08/28/neocloud-lambda-secures-1b-in-debt-to-buy-more-chips/&quot;&gt;Source: TechCrunch&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;FTC narrows AI tool provider liability in Rytr vacatur&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; The FTC set aside its 2024 consent order against Rytr, establishing a three-part test for the means-and-instrumentalities theory. Providers are liable only if they supply deceptive materials, offer inherently deceptive products, or know of misuse.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers of dual-use AI tools face less legal exposure for how users deploy their products.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://forkast.news/ftc-narrows-ai-tool-provider-liability-in-landmark-rytr-vacatur/&quot;&gt;Source: Forkast&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;AMD ROCm 10 delivers 3.3x inference uplift&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; AMD released ROCm 10 with ROCm.AI, a new developer experience. Systems with ROCm.AI show an average 3.3x increase in inference performance and 2.4x in training versus ROCm 7.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; AMD is closing the software gap, making its GPUs a more viable option for running AI workloads.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://wccftech.com/amd-rocm-10-big-ai-updates-performance-gains/&quot;&gt;Source: Wccftech&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Huawei Cloud CodeArts Agent goes GA in Asia Pacific&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Huawei Cloud CodeArts Agent is now commercially available in Asia Pacific, with Basic and Professional editions. It includes Agent Team with 16 specialized agents and 30+ Huawei engineering skills.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Enterprise developers in the region get an agentic coding tool that handles complex, long-duration tasks.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.prnewswire.com/apac/news-releases/huawei-cloud-codearts-agent-now-available-across-asia-pacific-bringing-agentic-ai-to-software-development-302862642.html&quot;&gt;Source: PR Newswire&lt;/a&gt;&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ai.dosa.dev/blog/today-in-ai-2026-08-29/&quot;&gt;Read the full brief on ai.dosa.dev&lt;/a&gt;&lt;/p&gt;</content:encoded><category>tencent</category><category>open-source</category><category>funding</category><category>amd</category><category>ftc</category><category>huawei</category></item><item><title>Anthropic Wins Court Ruling, Nvidia Pauses Cloud Deals</title><link>https://ai.dosa.dev/blog/today-in-ai-2026-08-28/</link><guid isPermaLink="true">https://ai.dosa.dev/blog/today-in-ai-2026-08-28/</guid><description>Judge rules Pentagon blacklist unlawful; Nvidia pauses revenue-share deals; new models and tools launch.</description><pubDate>Fri, 28 Aug 2026 00:00:00 GMT</pubDate><content:encoded>&lt;p&gt;This brief covers AI news from 2026-08-28 UTC.&lt;/p&gt;&lt;p&gt;A court win for Anthropic, a strategic pause from Nvidia, and a wave of model and tool releases.&lt;/p&gt;&lt;h2&gt;Judge: Pentagon&amp;#39;s Anthropic blacklist was unlawful retaliation&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; A federal judge ruled that the Pentagon&amp;#39;s blacklisting of Anthropic was unconstitutional, calling it &amp;#39;unlawful retaliation in violation of the First Amendment&amp;#39; and arbitrary and capricious.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; The ruling could reshape how the US government can restrict AI companies, giving labs more freedom to set ethical red lines without fear of retribution.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.theverge.com/ai-artificial-intelligence/985947/anthropic-supply-chain-risk-lawsuit-judge-ruling&quot;&gt;Source: The Verge&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Nvidia pauses AI cloud revenue-share deals&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Nvidia has paused some deals under its AI Compute Partnership financing program, which offered credit support to AI cloud companies in exchange for revenue share, amid internal antitrust concerns.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; The pause could affect how smaller AI cloud firms finance compute, and signals Nvidia is being cautious about antitrust exposure.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.reuters.com/business/nvidia-pauses-revenue-sharing-deals-with-ai-cloud-companies-wsj-reports-2026-08-27/&quot;&gt;Source: Reuters&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Google DeepMind ships Gemini Omni 1.1 Flash with 4K video&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Google DeepMind released Gemini Omni 1.1 Flash, adding 4K upscaling, scene extension up to 40 seconds, and first-and-last-frame interpolation, live in the Gemini API and Google AI Studio.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers building video workflows get production-ready control features and lower draft costs, with Adobe, Figma, and GMI Cloud already integrating it.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.aichatdaily.com/ai-models/google-deepmind-ships-gemini-omni-1-1-flash&quot;&gt;Source: AI Chat Daily&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Alibaba&amp;#39;s Qwen3.8-Flash matches DeepSeek V4 Pro at quarter price&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Alibaba released Qwen3.8-Flash, a 125B-parameter MoE model with 6B active, beating DeepSeek V4 Pro on SWE-bench Pro (62.5 vs 55.4) at roughly a quarter of the cost.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; The open-weight model offers near-frontier coding performance on commodity hardware, making advanced AI more accessible and affordable.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.intelligentliving.co/qwen38-flash-matches-deepseek-v4-pro/&quot;&gt;Source: Intelligent Living&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Anthropic launches Model Hardware Standard for physical machines&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Anthropic unveiled the Model Hardware Standard (MHS), a spec letting AI agents operate programmable hardware like microscopes and robotic arms, starting as a research preview.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; MHS could cut integration time for lab and manufacturing equipment from weeks to hours, and is model-agnostic, so any agent can use it.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://techstartups.com/2026/08/27/anthropic-launches-model-hardware-standard-to-let-ai-agents-control-physical-machines/&quot;&gt;Source: Tech Startups&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;AWS Agent Toolkit gives coding agents cloud skills&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; AWS released an Agent Toolkit that installs curated AWS skills into coding agents like Kiro, Claude Code, Codex, and Cursor, via the AWS CLI.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Coding agents get up-to-date AWS knowledge, reducing errors from stale training data and improving cloud development workflows.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://adtmag.com/articles/2026/08/27/aws-toolkit-gives-coding-agents-specialized-cloud-development-skills.aspx&quot;&gt;Source: ADTmag&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;OpenAI models on Bedrock for in-country inferencing in India&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Amazon Bedrock now supports OpenAI GPT-5.6 models Terra and Luna in India with cross-Region inference, keeping data processing within the country.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Indian developers in regulated sectors can use OpenAI models while meeting data residency requirements, without managing capacity.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://aws.amazon.com/blogs/machine-learning/introducing-openai-models-on-amazon-bedrock-for-in-country-inferencing-in-india/&quot;&gt;Source: AWS&lt;/a&gt;&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ai.dosa.dev/blog/today-in-ai-2026-08-28/&quot;&gt;Read the full brief on ai.dosa.dev&lt;/a&gt;&lt;/p&gt;</content:encoded><category>anthropic</category><category>nvidia</category><category>google</category><category>alibaba</category><category>aws</category><category>agents</category></item><item><title>Nvidia to Buy Hugging Face for $12.9B; Qwen, Z.ai Drop Open Models</title><link>https://ai.dosa.dev/blog/today-in-ai-2026-08-27/</link><guid isPermaLink="true">https://ai.dosa.dev/blog/today-in-ai-2026-08-27/</guid><description>Nvidia acquires Hugging Face, Qwen and Z.ai release open-weight models, plus Nvidia earnings and more.</description><pubDate>Thu, 27 Aug 2026 00:00:00 GMT</pubDate><content:encoded>&lt;p&gt;This brief covers AI news from 2026-08-27 UTC.&lt;/p&gt;&lt;p&gt;A blockbuster acquisition, two open-weight releases, and Nvidia&amp;#39;s record quarter headline today&amp;#39;s AI news.&lt;/p&gt;&lt;h2&gt;Nvidia Agrees to Buy Hugging Face for $12.9B&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Nvidia has agreed to acquire Hugging Face, the open-source model repository, for $12.9 billion, according to The Information. Reuters later confirmed the report. Neither company has commented.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; If completed, Nvidia would own the default hub for model distribution, extending its dominance from chips to the developer ecosystem.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.theinformation.com/articles/nvidia-agrees-buy-open-source-model-repository-hugging-face-12-9-billion&quot;&gt;Source: The Information&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Qwen3.8-Flash-Next Open-Sourced as Qwen4 Preview&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Alibaba&amp;#39;s Qwen released Qwen3.8-Flash-Next, a 125B-parameter multimodal MoE with 6B active, using Gated DeltaNet and Qwen Sparse Attention. It&amp;#39;s an early preview of the Qwen4 architecture, with API pricing at $0.16/1M input and $0.47/1M output tokens.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers get a frontier-adjacent open model with long context (262K, up to 1M) at low cost, and day-0 support in SGLang.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://gigazine.net/gsc_news/en/20260827-qwen3-8-flash-next/&quot;&gt;Source: GIGAZINE&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Z.ai Confirms Ox Alpha Is Its GLM Model&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Z.ai confirmed that Ox Alpha, the anonymous model that topped leaderboards on OpenRouter, is the newest iteration of its GLM series. Weights will be released Wednesday. Z.ai also released GLM-5.3-Flash, a 320B-parameter multimodal model served on Chinese chips.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Open-weight models from China continue to challenge expensive frontier models, giving developers cheaper alternatives for coding and agentic workloads.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://techcrunch.com/2026/08/26/surprise-z-ai-is-the-ai-lab-behind-the-mysterious-ox-alpha-model/&quot;&gt;Source: TechCrunch&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Nvidia Posts $96.2B Quarter, Data Center Up 117%&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Nvidia reported Q2 FY2027 revenue of $96.2B, up 106% year over year, with data center revenue at $89B. It guided Q3 to $108B, excluding China data center compute revenue.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; The numbers confirm sustained demand for AI compute, but the China exclusion signals geopolitical risk that could affect global supply chains.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.unite.ai/nvidia-posts-96-2b-quarter-as-data-center-revenue-hits-89b/&quot;&gt;Source: Unite.AI&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Gemini 3.5 Transcribe Launches for Developers&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Google introduced Gemini 3.5 Transcribe, a speech-to-text model with real-time streaming via the Live API and pre-recorded processing via the Interactions API, available in Google AI Studio and Gemini Enterprise Agent Platform.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers can build voice agents and captioning tools with sub-second latency and speaker attribution, directly in the Gemini API.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-5-transcribe/&quot;&gt;Source: Google&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Intel Unveils Three-Pronged Agentic AI Strategy&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; At Hot Chips 2026, Intel detailed Xeon Scalable 7 (Diamond Rapids), Crescent Island GPU for inference, and Wildcat Lake for Core Series 3, positioning agentic AI as a system-level workload.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Intel is betting on CPU-GPU co-design to handle agent orchestration, offering developers an alternative to GPU-only stacks.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.networkworld.com/article/4214452/intel-unwraps-three-pronged-architecture-strategy-to-go-after-agentic-ai.html&quot;&gt;Source: Network World&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Twitch Class Action Alleges AI Training Without Consent&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; A class action filed Aug. 20 against Twitch and Amazon claims they enrolled all channels in AI training by default and used streams, clips, and chat logs without permission. The suit seeks to represent millions of creators.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers using user-generated content for training face legal exposure; this case could set precedent on consent and opt-out defaults.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.claimdepot.com/cases/twitch-class-action-alleges-amazon-used-creators-content-to-train-ai-without-consent&quot;&gt;Source: ClaimDepot&lt;/a&gt;&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ai.dosa.dev/blog/today-in-ai-2026-08-27/&quot;&gt;Read the full brief on ai.dosa.dev&lt;/a&gt;&lt;/p&gt;</content:encoded><category>nvidia</category><category>hugging-face</category><category>qwen</category><category>z-ai</category><category>open-models</category><category>earnings</category></item><item><title>Kimi K3 Tops Coding Benchmarks, Undercuts Opus 4.8</title><link>https://ai.dosa.dev/blog/today-in-ai-2026-08-26/</link><guid isPermaLink="true">https://ai.dosa.dev/blog/today-in-ai-2026-08-26/</guid><description>Moonshot&apos;s open-weight Kimi K3 beats proprietary models on agent coding, plus Stability funding, XPeng robotics, and more.</description><pubDate>Wed, 26 Aug 2026 00:00:00 GMT</pubDate><content:encoded>&lt;p&gt;This brief covers AI news from 2026-08-26 UTC.&lt;/p&gt;&lt;p&gt;Open-weight models made a splash today, with Kimi K3 and Qwen-UI-Agent challenging proprietary leaders.&lt;/p&gt;&lt;h2&gt;Kimi K3 Open-Weight Model Beats Claude and GPT on Coding&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Moonshot AI released Kimi K3, a 2.8 trillion parameter open-weight model. It scores 42.0 on SWE Marathon, beating Claude Fable 5&amp;#39;s 35.0 and GPT-5.6 Sol, and costs about $0.94 per task versus Opus 4.8&amp;#39;s $1.80.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers get frontier-level agentic coding at half the cost, with an OpenAI-compatible API for quick integration.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://byteiota.com/kimi-k3-open-weight-frontier-model-2/&quot;&gt;Source: byteiota&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Stability AI Raises $76M Backed by Record Labels and AMD&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Stability AI raised $76 million in Series B funding, bringing total to $232 million. Investors include Universal Music Group, Sony Music Group, Warner Music Group, EA, AMD Ventures, and Pacific Alliance Ventures.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; The funding fuels Stability&amp;#39;s creative production suite and professional services, signaling entertainment industry buy-in for AI music and video tools.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://techcrunch.com/2026/08/25/stability-ai-maker-of-image-generator-stable-diffusion-raises-76-million-in-fresh-funding/&quot;&gt;Source: TechCrunch&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;XPeng Robotics Raises Over $900M at $6.3B Valuation&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; XPeng&amp;#39;s robotics unit raised more than $900 million at a post-money valuation above $6.3 billion. The round was led by IDG Capital, with Tencent and Alibaba participating. Funds will support humanoid robot production and Physical AI research.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Big capital inflow accelerates humanoid robot development, with mass production targeted by end of 2026, impacting AI hardware and robotics developers.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://technode.com/2026/08/25/xpengs-robotics-unit-raises-over-900m-at-a-6-3bn-valuation/&quot;&gt;Source: TechNode&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;IBM Releases Granite 4.2 with Native Reasoning&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; IBM released Granite 4.2 language models in 3B, 8B, and 30B sizes, featuring native step-by-step reasoning and tool-calling for enterprise agents. They are Apache 2.0 licensed and deployable across cloud, on-premises, and edge.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Teams can build and fine-tune reasoning-capable agents without licensing fees, with flexible sizes for different workloads.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://research.ibm.com/blog/introducing-granite-4-2&quot;&gt;Source: IBM Research&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Perplexity Launches Portable Computer with Nvidia&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Perplexity launched Portable Computer, a fully local AI agent that runs on Nvidia DGX Spark and RTX-equipped Linux machines. Work stays on-device by default, with zero token costs and permission prompts before cloud offload.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers can run agentic workloads locally, cutting cloud costs and keeping data on-premises, with Nvidia hardware as the enabler.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://venturebeat.com/infrastructure/perplexity-partners-with-nvidia-to-launch-portable-computer-a-fully-local-ai-agent-with-zero-token-costs&quot;&gt;Source: VentureBeat&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;OpenAI&amp;#39;s Jalapeño Chip Shows Strong Inference Performance&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; OpenAI shared first measured results for Jalapeño, its custom inference chip. On InferenceX with GPT-OSS 120B, it delivered higher peak throughput per kilowatt and lower token latency than commercial systems, also performing well on DeepSeek R1 and Kimi K2.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; First-party silicon gives OpenAI more control over serving economics, potentially lowering costs for developers using its models.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://openai.com/index/the-full-stack-behind-abundant-intelligence/&quot;&gt;Source: OpenAI&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Delhi High Court Rules AI Training as Fair Dealing&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; The Delhi High Court denied ANI Media&amp;#39;s interim injunction against OpenAI, ruling that using copyrighted news content for AI training qualifies as &amp;#39;fair dealing&amp;#39; under Section 52(1)(a)(i) of India&amp;#39;s Copyright Act.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Provides temporary legal clarity for AI companies training on public data in India, reducing immediate legal blockers.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.whalesbook.com/news/English/lawcourt/Delhi-HC-Rules-AI-Training-Fair-Dealing-in-ANI-OpenAI-Case/6a8d111d84d2dd5c12e42601&quot;&gt;Source: Whalesbook&lt;/a&gt;&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ai.dosa.dev/blog/today-in-ai-2026-08-26/&quot;&gt;Read the full brief on ai.dosa.dev&lt;/a&gt;&lt;/p&gt;</content:encoded><category>open-weights</category><category>funding</category><category>robotics</category><category>enterprise</category><category>chips</category><category>policy</category></item><item><title>Nvidia Eyes Perplexity at $30B+ as OpenAI Slashes GPT-5.6 Sol Prices</title><link>https://ai.dosa.dev/blog/today-in-ai-2026-08-25/</link><guid isPermaLink="true">https://ai.dosa.dev/blog/today-in-ai-2026-08-25/</guid><description>Nvidia in talks to invest in Perplexity, OpenAI cuts GPT-5.6 Sol API prices, Meta launches Muse Code, and more.</description><pubDate>Tue, 25 Aug 2026 00:00:00 GMT</pubDate><content:encoded>&lt;p&gt;This brief covers AI news from 2026-08-25 UTC.&lt;/p&gt;&lt;p&gt;Today&amp;#39;s AI news: Nvidia&amp;#39;s investment talks, OpenAI&amp;#39;s price cuts, and new model releases.&lt;/p&gt;&lt;h2&gt;Nvidia in Talks to Invest in Perplexity at $30B+ Valuation&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Nvidia is discussing an investment in Perplexity as part of an equity round valuing the AI search startup at more than $30 billion, per The Information. Perplexity&amp;#39;s annualized revenue has risen above $750 million, driven partly by Perplexity Computer.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; A Nvidia investment would deepen ties between the chip giant and a fast-growing agent startup, potentially shaping AI infrastructure and agent adoption.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://techstartups.com/2026/08/24/nvidia-in-talks-to-invest-in-perplexity-at-30-billion-valuation-as-revenue-tops-750-million/&quot;&gt;Source: Tech Startups&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;OpenAI Cuts GPT-5.6 Sol API Prices by Over 20%&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; OpenAI cut developer pricing for its frontier GPT-5.6 Sol model by more than 20% for three months, effective on API and eligible plans for ChatGPT Work and Codex. New prices: $4 per 1M input tokens and $20 per 1M output tokens, down from $5 and $30.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Cheaper frontier tokens lower the cost of building agentic applications, intensifying price competition with Anthropic and Chinese models.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.thehindu.com/sci-tech/technology/openai-cuts-developer-pricing-for-frontier-gpt-56-sol-model-by-more-than-20/article71383057.ece&quot;&gt;Source: The Hindu&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Meta Launches Muse Code Terminal Agent with Muse Spark 1.2&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Meta released Muse Code, a terminal-based coding agent in beta for macOS and Linux, powered by the new Muse Spark 1.2 model. It keeps background agents active across long sessions and logs every action for crash recovery.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Long-horizon coding agents that retain context and survive crashes could handle multi-hour tasks, reducing developer oversight.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://theaisoftwarereport.com/meta-releases-muse-code-terminal-agent-powered-by-muse-spark-1-2/&quot;&gt;Source: The AI &amp;amp;amp; Software Report&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Alibaba Launches Wan3.0 Video Model with 30-Second Generation&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Alibaba Cloud launched Wan3.0, a video-generation model that creates clips up to 30 seconds and accepts DOC, XLS, PPT, PDF, and Markdown as inputs. It had been in public beta since early August.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Document-to-video generation opens new workflows for developers building automated content pipelines.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://technode.com/2026/08/24/alibaba-launches-wan3-0-video-model-with-30-second-generation-and-document-input/&quot;&gt;Source: TechNode&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Thomson Reuters Launches Proprietary LLM &amp;#39;Thomson&amp;#39;&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Thomson Reuters announced Thomson, its first proprietary large language model, trained in-house for $40 million from an open-source foundation. The company says it runs at a fraction of the cost of comparable frontier models.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; A domain-focused model built on proprietary data at low cost could challenge general-purpose frontier models in legal and professional services.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.thomsonreuters.com/en/press-releases/2026/august/thomson-reuters-leverages-its-world-class-data-assets-to-launch-its-own-frontier-model&quot;&gt;Source: Thomson Reuters&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Nvidia Unveils Vera Rubin LPX and CPX Inference Platforms&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Nvidia announced the Vera Rubin rack-scale system with Groq 3 LPX in full production, delivering 3, 400 output tokens per second on Gemma 4 31B in an Artificial Analysis benchmark, 4x faster than the nearest alternative. CoreWeave and Nebius are adopting the platforms.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Faster, cheaper inference for long-context agentic workloads could lower token costs and enable more complex AI applications.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://blogs.nvidia.com/blog/vera-rubin-lpx-spectrum-x-nvlink-fusion/&quot;&gt;Source: NVIDIA&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Japan Drafts IP Code for Generative AI Training Data&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Japan&amp;#39;s Cabinet Office proposed a revised Principle-Code for generative AI businesses, covering IP protection, avoiding pirate sites, respecting paywalls, and increasing transparency. It uses a &amp;#39;comply or explain&amp;#39; approach and applies to foreign businesses serving Japan.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Builders targeting Japan may need to document training data practices and IP risk management, even if non-binding.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.medianama.com/2026/08/223-japan-ip-generative-ai-training/&quot;&gt;Source: MediaNama&lt;/a&gt;&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ai.dosa.dev/blog/today-in-ai-2026-08-25/&quot;&gt;Read the full brief on ai.dosa.dev&lt;/a&gt;&lt;/p&gt;</content:encoded><category>nvidia</category><category>openai</category><category>perplexity</category><category>meta</category><category>alibaba</category><category>japan</category></item><item><title>Mystery Model Ox Alpha Tops Coding Test, Traced to Zhipu</title><link>https://ai.dosa.dev/blog/today-in-ai-2026-08-24/</link><guid isPermaLink="true">https://ai.dosa.dev/blog/today-in-ai-2026-08-24/</guid><description>Anonymous coding model stuns benchmarks, Nvidia&apos;s Poolside deal details, OpenAI agent users hit 20M, Japan builds AI factory.</description><pubDate>Mon, 24 Aug 2026 00:00:00 GMT</pubDate><content:encoded>&lt;p&gt;This brief covers AI news from 2026-08-24 UTC.&lt;/p&gt;&lt;p&gt;An anonymous model is winning coding tests, Nvidia&amp;#39;s Poolside deal gets granular, and OpenAI&amp;#39;s agent count climbs.&lt;/p&gt;&lt;h2&gt;Ox Alpha Mystery Model Traced to Zhipu, Beats GPT-5.6 Sol&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; An anonymous model, Ox Alpha, appeared on OpenRouter on Aug 20 with a 1M-token context and free access. Developer Ben Davis reported it passing 8 of 10 DeepSWE tasks, beating Claude Fable 5 (65%) and GPT-5.6 Sol (52%). Server forensics point to Zhipu AI&amp;#39;s GLM-5.3.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; If confirmed, it shows Zhipu can ship a frontier coding model anonymously, and developers get a free million-token context for large codebases.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://startupfortune.com/ox-alpha-topped-coding-benchmarks-and-forensics-now-point-to-zhipu/&quot;&gt;Source: Startup Fortune&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Nvidia&amp;#39;s $6B Poolside Deal: License, $1B Invest, 109 Engineers&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Nvidia agreed to pay $6B for a non-exclusive license to Poolside&amp;#39;s Model Factory platform, invest $1B at a $12B pre-money valuation, and offer jobs to 109 Poolside engineers to work on Nemotron. Co-founders stay with Poolside.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Nvidia gets a ready-made model-building operation to challenge DeepSeek and OpenAI, while Poolside keeps its tech and independence.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://runtimewire.com/article/nvidia-pays-6b-to-license-poolside-s-factory-for-u-s-open-weight-models&quot;&gt;Source: RuntimeWire&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;OpenAI Agent Users Hit 20M Weekly, Enterprise Revenue Up 50%&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; OpenAI disclosed at an all-hands that its coding and agent features reached 20M weekly active users as of Aug 19, about 2% of its 1B ChatGPT users. Overall revenue run rate climbed 35% quarter-to-date, enterprise revenue over 50%.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Agent adoption is still early, but enterprise demand is surging, signaling where OpenAI&amp;#39;s growth is heading.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://cryptobriefing.com/openai-agent-users-20m-revenue-growth/&quot;&gt;Source: CryptoBriefing&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Japan Launches FRONTia Project with Nvidia for Physical AI&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Japan&amp;#39;s METI is backing the FRONTia Project to build national physical AI infrastructure. Nvidia and Noetra Corp. will build an AI factory with 13, 750 Vera CPUs and 27, 500 Rubin GPUs in a 140-megawatt data center.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; This gives Japanese robotics and industrial AI developers a dedicated national compute platform for physical AI models.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://time.news/japan-launches-frontia-project-with-nvidia-for-physical-ai-infrastructure/&quot;&gt;Source: Time News&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Claude Managed Agents Add Session Budgets and More Controls&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Anthropic shipped four new production controls for Claude Managed Agents, including session budgets that pause sessions when a cost cap is reached, advisor models, geo controls, and skills. The update targets governance issues like runaway costs.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers can now set hard spending caps on agent sessions, preventing surprise bills and enabling safer production deployments.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://byteiota.com/claude-managed-agents-session-budgets-advisor-geo-github-skills/&quot;&gt;Source: Byteiota&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;FreeToken Runs 753B GLM-5.2 on a Single Workstation GPU&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Researchers from UC Berkeley and UT Austin released FreeToken, an edge-native MoE serving engine that runs large models on consumer hardware. It runs 753B GLM-5.2 on a single workstation GPU, 284B on a gaming desktop, and 35B on an 8GB laptop GPU.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers can serve frontier open-weight models locally, cutting token costs and keeping data on their own machines.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.marktechpost.com/2026/08/23/meet-freetoken-an-edge-native-moe-serving-engine-that-runs-753b-glm-5-2-on-a-single-workstation-gpu/&quot;&gt;Source: MarkTechPost&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Cursor Launches Origin, Agent-Native Code Hosting in Beta&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Cursor launched Origin, a code-hosting platform with repositories, pull requests, and GitHub sync, in early beta on Aug 17 for paid users. It aims to put AI coding agents inside the same workspace as code.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers get an alternative to GitHub that integrates AI agents directly into the repository workflow, potentially reducing tool switching.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://memeburn.com/cursor-launches-origin-to-challenge-github-with-agent-native-code-hosting/&quot;&gt;Source: Memeburn&lt;/a&gt;&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ai.dosa.dev/blog/today-in-ai-2026-08-24/&quot;&gt;Read the full brief on ai.dosa.dev&lt;/a&gt;&lt;/p&gt;</content:encoded><category>zhipu</category><category>nvidia</category><category>openai</category><category>agents</category><category>japan</category><category>anthropic</category></item><item><title>Grok 4.6 Hits Vertex AI, Anthropic Hires TPU Legend</title><link>https://ai.dosa.dev/blog/today-in-ai-2026-08-23/</link><guid isPermaLink="true">https://ai.dosa.dev/blog/today-in-ai-2026-08-23/</guid><description>Grok 4.6 lands on Google Cloud, Anthropic hires TPU creator, OpenAI buys InstantDB, and more.</description><pubDate>Sun, 23 Aug 2026 00:00:00 GMT</pubDate><content:encoded>&lt;p&gt;This brief covers AI news from 2026-08-23 UTC.&lt;/p&gt;&lt;p&gt;Today&amp;#39;s brief covers model launches, a key hire, an acquisition, and funding news.&lt;/p&gt;&lt;h2&gt;Grok 4.6 Now on Google Cloud Vertex AI&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; xAI&amp;#39;s Grok 4.6 is now available on Google Cloud&amp;#39;s Vertex AI, with a 500, 000-token context window and four reasoning levels. It&amp;#39;s priced for enterprise adoption and accessible via Vertex Model Garden.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers can now use Grok 4.6 for long-running agents and coding tasks directly within Google Cloud, expanding deployment options.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://bitcoinethereumnews.com/tech/xais-grok-4-6-launches-on-google-cloud-vertex-ai/&quot;&gt;Source: bitcoinethereumnews.com&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Anthropic Hires TPU Creator Amir Salek&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Anthropic confirmed on August 21 that Amir Salek, founder and head of Google&amp;#39;s TPU program from 2013 to 2022, is joining its compute team. Anthropic is Google&amp;#39;s largest external TPU tenant, with access to up to one million TPUs.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Hiring the architect behind the TPU signals Anthropic&amp;#39;s push to shape its own silicon strategy, potentially reducing reliance on rented chips.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.ainvest.com/news/anthropic-hires-man-built-tpu-million-chip-renter-draws-silicon-2608/&quot;&gt;Source: ainvest.com&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;OpenAI Acquires InstantDB Team&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; OpenAI acquired the team behind InstantDB, a real-time database platform serving over 17, 000 users and 400, 000 apps. Instant Cloud is being phased out, with support until August 2027.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; OpenAI gains backend infrastructure expertise to support persistent state management for AI agents, a key requirement for agentic workflows.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://cryptobriefing.com/openai-acquires-instantdb-team/&quot;&gt;Source: cryptobriefing.com&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Anthropic Launches $1.5B Ode AI Consultancy&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Anthropic, with Blackstone, Hellman &amp;amp;amp; Friedman, and Goldman Sachs, launched Ode, a $1.5 billion AI services operation. It embeds Claude engineers in midsized private-equity portfolio companies, with first customers in Q4 2026.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; This creates a new revenue stream for Anthropic beyond API consumption and offers a turnkey way for enterprises to integrate Claude.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.saasrise.com/news/anthropic-launches-15b-ode-ai-consultancy-for-privateequity-portfolios-ab05c9c7-4add-4fee-b3c9-ff7efa789130&quot;&gt;Source: SaasRise&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Starcloud Raises $250M at $2.3B Valuation&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Starcloud, an orbital data center startup, raised $250 million at a $2.3 billion valuation. Manhattan West led, with Nvidia and Cisco Investments joining. It&amp;#39;s working with Nvidia on the Space-1 Vera Rubin Module.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Orbital data centers could provide AI compute capacity beyond Earth, a novel approach to scaling infrastructure.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://metapress.net/business/2026/08/22/orbital-data-center-startup-starcloud-valued-at-2-3-billion-in-latest-funding/&quot;&gt;Source: Metapress&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Blacksmith Raises $45M for AI CI Infrastructure&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Blacksmith, a San Francisco startup building CI infrastructure for GitHub Actions, closed a $45 million Series B led by Peak XV Partners at a $550 million valuation. It also launched Codesmith, a cloud coding agent that auto-fixes CI failures.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; As AI-generated code surges, faster CI and automated fixes address the new bottleneck in software delivery.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://byteiota.com/blacksmith-ai-code-validation/&quot;&gt;Source: byteiota&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Colorado AI Rules Face FTC Preemption Claim&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Colorado is writing rules for the Automated Decision-Making Technology Act, with comments open until October 26. The FTC published a policy statement on July 1 asserting federal law may preempt state AI output mandates.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers building AI systems for housing, employment, and insurance face uncertain compliance requirements as state and federal rules clash.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://forkast.news/colorado-is-writing-ai-transparency-rules-the-ftc-says-federal-law-may-override-them/&quot;&gt;Source: Forkast&lt;/a&gt;&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ai.dosa.dev/blog/today-in-ai-2026-08-23/&quot;&gt;Read the full brief on ai.dosa.dev&lt;/a&gt;&lt;/p&gt;</content:encoded><category>xai</category><category>anthropic</category><category>openai</category><category>funding</category><category>infrastructure</category><category>regulation</category></item><item><title>Gemini 3.7 Flash Gains on Coding, Price Doubles in 2027</title><link>https://ai.dosa.dev/blog/today-in-ai-2026-08-22/</link><guid isPermaLink="true">https://ai.dosa.dev/blog/today-in-ai-2026-08-22/</guid><description>Gemini 3.7 Flash&apos;s pricing clock, Liquid AI&apos;s DSpark draft models, Rillet&apos;s $100M round, and more in today&apos;s AI brief.</description><pubDate>Sat, 22 Aug 2026 00:00:00 GMT</pubDate><content:encoded>&lt;p&gt;This brief covers AI news from 2026-08-22 UTC.&lt;/p&gt;&lt;p&gt;Google&amp;#39;s pricing clock starts ticking, Liquid AI speeds up decoding, and Rillet raises $100M in two days.&lt;/p&gt;&lt;h2&gt;Gemini 3.7 Flash: Big Benchmarks, Price Doubles in 2027&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Google&amp;#39;s Gemini 3.7 Flash, released August 13, shows major coding gains: DeepSWE up to 65.3%, AutomationBench to 30.4%. Intro pricing of $0.75/$3.75 per million tokens doubles on January 1, 2027.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; A 62% input cost cut versus Claude Sonnet 5 until year-end, but teams must plan for the rate hike or lock in savings now.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://byteiota.com/gemini-3-7-flash-benchmark-jumps-a-pricing-clock-and-one-real-problem/&quot;&gt;Source: byteiota&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Nvidia Pays $6B for Poolside AI Tech and Talent&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Nvidia struck a $6 billion licensing agreement with Poolside AI, committed $1 billion in fresh capital, and extended job offers to 109 engineers. The deal values Poolside at $12 billion pre-money.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Gives Nvidia broad access to Poolside&amp;#39;s model-development software without absorbing the company, a new template for chipmakers to secure AI software leverage.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.webpronews.com/nvidia-pays-6-billion-for-poolside-ai-tech-and-talent-in-latest-power-play/&quot;&gt;Source: WebProNews&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;DeepSeek Launches V4-Flash Vision Model with Free Files API&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; DeepSeek released DeepSeek-V4-Flash-Vision-Exp on August 21, adding multimodal support. A new Files API lets developers reuse uploaded images for free via file_id references.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Cuts repeated image upload costs for multimodal agent workflows, and DeepSeek claims agent performance near Opus-4.8 on visual tasks.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://meyka.com/blog/deepseek-launches-v4-flash-vision-model-on-api-with-free-image-reuse-via-files-api-2608/&quot;&gt;Source: Meyka&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Liquid AI&amp;#39;s DSpark Draft Models Speed Decoding Up to 3.18x&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Liquid AI released DSpark draft checkpoints for three LFM2.5 models, adding speculative decoding paths. Up to 3.18x faster on H100, 2.87x on M4 Max, with identical outputs.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Self-hosted developers get faster inference without accuracy loss, with day-one llama.cpp and SGLang support.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://thenews92.com/liquid-ai-releases-lfm2-5-dspark-draft-models-that-deliver-up-to-3-18x-faster-decoding-without-changing-model-outputs/&quot;&gt;Source: The News 92&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;NVIDIA AVO Hits 100% on ARC-AGI-3 Benchmark&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; NVIDIA&amp;#39;s Agentic Variation Operators (AVO) architecture achieved a 100.00 RHAE score on ARC-AGI-3, completing all 183 levels. It also beat FlashAttention-4 by up to 10.5% in GPU-kernel optimization.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Shows system design, not just model capability, can unlock frontier-level long-horizon agent performance, a blueprint for building robust agent harnesses.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://developer.nvidia.com/blog/nvidia-avo-reaches-100-on-arc-agi-3-demonstrating-a-frontier-level-general-purpose-architecture-for-long-horizon-autonomous-agents/&quot;&gt;Source: NVIDIA Technical Blog&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Rillet Raises $100M at $1B Valuation in 48 Hours&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; AI accounting startup Rillet raised $100 million at a $1 billion valuation, becoming a unicorn in 48 hours. It has 600 customers and doubled annualized revenue in the last quarter.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Signals strong investor appetite for AI-native vertical SaaS that displaces legacy ERP systems like Oracle and NetSuite.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://techcrunch.com/2026/08/21/how-ai-accounting-startup-rillet-raised-100m-and-became-a-unicorn-in-48-hours/&quot;&gt;Source: TechCrunch&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Civil Society Groups Urge FTC Probe into AI Book Destruction&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Civil society organizations sent a letter to the FTC asking for an antitrust investigation into practices like Anthropic&amp;#39;s effort to &amp;#39;destructively scan all the books in the world&amp;#39;, calling it a potential unfair method of competition under Section 5.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Could reshape how AI companies acquire training data, potentially raising costs for startups and affecting data sourcing strategies.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.mlex.com/mlex/articles/2516403/ai-industry-book-destruction-merits-us-ftc-antitrust-probe-civil-society-groups-say&quot;&gt;Source: MLex&lt;/a&gt;&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ai.dosa.dev/blog/today-in-ai-2026-08-22/&quot;&gt;Read the full brief on ai.dosa.dev&lt;/a&gt;&lt;/p&gt;</content:encoded><category>nvidia</category><category>funding</category><category>models</category><category>agents</category><category>regulation</category><category>benchmarks</category></item><item><title>Nvidia Pays $6B for Poolside AI Tech and Talent</title><link>https://ai.dosa.dev/blog/today-in-ai-2026-08-21/</link><guid isPermaLink="true">https://ai.dosa.dev/blog/today-in-ai-2026-08-21/</guid><description>Nvidia&apos;s $6B Poolside deal, DeepSeek vision model, Qwen3.8-27B, Grok 4.6 in Copilot, and more AI news.</description><pubDate>Fri, 21 Aug 2026 00:00:00 GMT</pubDate><content:encoded>&lt;p&gt;This brief covers AI news from 2026-08-21 UTC.&lt;/p&gt;&lt;p&gt;Today&amp;#39;s brief covers Nvidia&amp;#39;s $6B Poolside deal, new model releases, and key funding and security moves.&lt;/p&gt;&lt;h2&gt;Nvidia Pays $6B for Poolside AI Tech and Talent&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Nvidia struck a $6 billion licensing agreement with Poolside AI, committed $1 billion in fresh capital, and extended job offers to 109 engineers. The deal values Poolside at $12 billion pre-money. Founders stay and the company remains independent.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Gives Nvidia broad access to Poolside&amp;#39;s AI coding software without a full acquisition, potentially reshaping the AI development tool landscape.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.webpronews.com/nvidia-pays-6-billion-for-poolside-ai-tech-and-talent-in-latest-power-play/&quot;&gt;Source: WebProNews&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;DeepSeek Launches V4-Flash Vision Model with Free Files API&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; DeepSeek released DeepSeek-V4-Flash-Vision-Exp, adding multimodal support to V4-Flash. It matches text performance and nears Opus-4.8 on multimodal agent benchmarks. The Files API allows free image reuse via file_id.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers building multimodal agents get a cheaper vision model with reduced upload costs, potentially lowering barriers for visual AI applications.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://meyka.com/blog/deepseek-launches-v4-flash-vision-model-on-api-with-free-image-reuse-via-files-api-2608/&quot;&gt;Source: Meyka&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Alibaba&amp;#39;s Qwen3.8-27B Matches GPT-5.6 Luna&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Alibaba released Qwen3.8-27B, an open-weights model with 27B parameters, 256k context, and multimodal support. It scores 52.02 on Artificial Analysis Intelligence Index, close to GPT-5.6 Luna&amp;#39;s 52.31.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Open-weights models now rival top proprietary models, giving developers a high-performance, locally deployable option without API costs or privacy trade-offs.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://www.thenextgentechinsider.com/pulse/alibaba-releases-qwen38-27b-matching-gpt-56-luna-intelligence&quot;&gt;Source: TheNextGenTechInsider&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Grok 4.6 Lands in GitHub Copilot&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Grok 4.6 is now selectable in GitHub Copilot across eight surfaces including VS Code, JetBrains, and Xcode. It&amp;#39;s available on all major Copilot plans, with gradual rollout and admin enablement for Business/Enterprise.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Developers get a new agentic coding model with 500K context and lower pricing, potentially improving long-running coding tasks.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://byteiota.com/grok-4-6-is-now-in-github-copilot-what-devs-must-know/&quot;&gt;Source: byteiota&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Callosum Raises $100M for AI Workload Optimization&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; London-based Callosum raised $100M in seed funding led by Atomico. Its Tailored Inference service routes tasks to optimal models and chips, claiming 3.7x faster inference than GPT-5.6 Luna with better quality.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Could cut inference costs and latency for developers by intelligently distributing workloads across models and hardware.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://siliconangle.com/2026/08/20/ai-workload-optimization-startup-callosum-raises-100m/&quot;&gt;Source: SiliconANGLE&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;Fortinet Acquires Virtue AI for AI Security&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; Fortinet acquired AI security company Virtue AI, adding validation and runtime protection tools. Financial terms were not disclosed. Virtue AI offers automated red-teaming for agents and continuous AI validation.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Strengthens AI security for enterprises, addressing risks like prompt injection and malicious tool calls in agentic systems.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://securitybrief.co.nz/story/fortinet-buys-virtue-ai-to-boost-ai-security-tools&quot;&gt;Source: SecurityBrief NZ&lt;/a&gt;&lt;/p&gt;&lt;h2&gt;NVIDIA AVO Hits 100% on ARC-AGI-3&lt;/h2&gt;&lt;p&gt;&lt;strong&gt;What happened:&lt;/strong&gt; NVIDIA&amp;#39;s AVO agent architecture achieved a 100.00 RHAE score on ARC-AGI-3, completing all 183 levels. It elevates Claude Opus 5 from 30% to 100%, showing system design can unlock frontier performance.&lt;/p&gt;&lt;p&gt;&lt;strong&gt;Why it matters:&lt;/strong&gt; Demonstrates that agent harnesses, not just models, are key to long-horizon autonomy, guiding developers on building more capable agent systems.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://developer.nvidia.com/blog/nvidia-avo-reaches-100-on-arc-agi-3-demonstrating-a-frontier-level-general-purpose-architecture-for-long-horizon-autonomous-agents/&quot;&gt;Source: NVIDIA&lt;/a&gt;&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ai.dosa.dev/blog/today-in-ai-2026-08-21/&quot;&gt;Read the full brief on ai.dosa.dev&lt;/a&gt;&lt;/p&gt;</content:encoded><category>nvidia</category><category>funding</category><category>models</category><category>agents</category><category>security</category></item></channel></rss>