← AM·PM Brief

Qwen 3.8-Flash Runs on Phone CPU, Altman Says ChatGPT Queries Use Water Like Almonds - AI Daily Brief (Sep 5)

· Morning brief · 8 news · 8:31

Audio in Mandarin Chinese · English transcript below

Qwen 3.8-Flash-Next runs on mobile CPU; GPT-6, Gemini & 8 new models drop; Ling-3.0 boosts vision Agent—edge deployment & multimodal inference are now AI's main battleground.

The LLM race has shifted from parameter bragging to real-world scenario penetration. What's striking is the parallel momentum: as Qwen achieves on-device operation on mobile CPUs, Ling-3.0-flash-VL simultaneously pushes visual understanding and Agent capabilities to new heights. Taken together, these developments reveal AI simultaneously sinking to the compute limits of edge devices while penetrating hardware domains like circuit board design. The industry has officially entered a pragmatic phase where lightweight deployment and multimodal capability advance in parallel, marking the end of pure scale worship and the beginning of functional integration across the full stack.

Qwen 3.8-Flash Runs on Phone CPU, Altman Says ChatGPT Queries Use Water Like Almonds - AI Daily Brief (Sep 5)

Today's Top 3 Headlines

  1. AI Industry News

    🤖 Qwen 3.8-Flash-Next Achieves Efficient Phone CPU Inference

    Alibaba's Tongyi team revealed Qwen3.8-Flash-Next LLM now runs efficiently on mobile CPUs, enabling on-device inference without cloud dependency. For mobile developers and edge AI, this means high-performance LLMs deployable on ordinary phones, dramatically lowering barriers for edge intelligence applications.

    Source
  2. AI Industry News

    🤖 38,000 ChatGPT queries = water for one California almond; OpenAI CEO says data center water use comparable to office buildings

    OpenAI CEO Sam Altman revealed 38,000 ChatGPT queries use water equivalent to producing one California almond, noting data center consumption matches office buildings. For AI sustainability watchers, this suggests public water anxiety may be overstated—yet transparency remains critical for the industry.

    Source
  3. AI Industry News

    🤖 AI Race: 10 New Models Drop, Led by Gemini, GPT-6, Claude, Qwen, DeepSeek

    Google, OpenAI, Anthropic, Microsoft, and Meta launched 10 new LLMs this week, including Gemini 3.8 Flash, GPT-6 Astra, and Claude 5.1. For developers and investors, this suddenly expands model choices and signals intensifying multimodal competition.

    Source

+5 more headlines

  • 🤖 B.AI launches full-stack infrastructure, free DeepSeek-V4-Flash to power Agent era
  • 🤖 Ling-3.0-flash-VL drops with stronger vision understanding & visual Agent capabilities
  • 🤖 Can Claude Opus Design PCBs? AI Auto-Layout Sparks Debate
  • 🤖 Tesla Cybercab Without Steering Wheel Under Federal Safety Probe
  • 🤖 Zuck to Trump: Hassabis Vision to Dominate US AI Policy
Unlock all 8 headlines + deep analysis →Free 7-day trial · cancel anytime
Browse all past briefings →