logo Nvidia to Buy Hugging Face for $12.9B; Qwen, Z.ai Drop Open Models

Nvidia to Buy Hugging Face for $12.9B; Qwen, Z.ai Drop Open Models

Dosa AI Tools 3 min read nvidia hugging-face qwen z-ai
In this brief · 8 sections

This brief covers AI news from 2026-08-27 UTC.

A blockbuster acquisition, two open-weight releases, and Nvidia’s record quarter headline today’s AI news.

Nvidia Agrees to Buy Hugging Face for $12.9B

What happened: Nvidia has agreed to acquire Hugging Face, the open-source model repository, for $12.9 billion, according to The Information. Reuters later confirmed the report. Neither company has commented.

Why it matters: If completed, Nvidia would own the default hub for model distribution, extending its dominance from chips to the developer ecosystem.

Source: The Information

Qwen3.8-Flash-Next Open-Sourced as Qwen4 Preview

What happened: Alibaba’s Qwen released Qwen3.8-Flash-Next, a 125B-parameter multimodal MoE with 6B active, using Gated DeltaNet and Qwen Sparse Attention. It’s an early preview of the Qwen4 architecture, with API pricing at $0.16/1M input and $0.47/1M output tokens.

Why it matters: Developers get a frontier-adjacent open model with long context (262K, up to 1M) at low cost, and day-0 support in SGLang.

Source: GIGAZINE

Z.ai Confirms Ox Alpha Is Its GLM Model

What happened: Z.ai confirmed that Ox Alpha, the anonymous model that topped leaderboards on OpenRouter, is the newest iteration of its GLM series. Weights will be released Wednesday. Z.ai also released GLM-5.3-Flash, a 320B-parameter multimodal model served on Chinese chips.

Why it matters: Open-weight models from China continue to challenge expensive frontier models, giving developers cheaper alternatives for coding and agentic workloads.

Source: TechCrunch

Nvidia Posts $96.2B Quarter, Data Center Up 117%

What happened: Nvidia reported Q2 FY2027 revenue of $96.2B, up 106% year over year, with data center revenue at $89B. It guided Q3 to $108B, excluding China data center compute revenue.

Why it matters: The numbers confirm sustained demand for AI compute, but the China exclusion signals geopolitical risk that could affect global supply chains.

Source: Unite.AI

Gemini 3.5 Transcribe Launches for Developers

What happened: Google introduced Gemini 3.5 Transcribe, a speech-to-text model with real-time streaming via the Live API and pre-recorded processing via the Interactions API, available in Google AI Studio and Gemini Enterprise Agent Platform.

Why it matters: Developers can build voice agents and captioning tools with sub-second latency and speaker attribution, directly in the Gemini API.

Source: Google

Intel Unveils Three-Pronged Agentic AI Strategy

What happened: At Hot Chips 2026, Intel detailed Xeon Scalable 7 (Diamond Rapids), Crescent Island GPU for inference, and Wildcat Lake for Core Series 3, positioning agentic AI as a system-level workload.

Why it matters: Intel is betting on CPU-GPU co-design to handle agent orchestration, offering developers an alternative to GPU-only stacks.

Source: Network World

What happened: A class action filed Aug. 20 against Twitch and Amazon claims they enrolled all channels in AI training by default and used streams, clips, and chat logs without permission. The suit seeks to represent millions of creators.

Why it matters: Developers using user-generated content for training face legal exposure; this case could set precedent on consent and opt-out defaults.

Source: ClaimDepot

Sources

Share

Discuss with AI

© 2026 dosa.dev