← AM·PM Brief

Qwen 3.8-Flash Runs on Phone CPU, Altman Says ChatGPT Queries Use Water Like Almonds - AI Daily Brief (Sep 5)

· Morning brief · 8 news · 8:31

Audio in Mandarin Chinese · English transcript below

Qwen 3.8-Flash-Next runs on mobile CPU; GPT-6, Gemini & 8 new models drop; Ling-3.0 boosts vision Agent—edge deployment & multimodal inference are now AI's main battleground.

The LLM race has shifted from parameter bragging to real-world scenario penetration. What's striking is the parallel momentum: as Qwen achieves on-device operation on mobile CPUs, Ling-3.0-flash-VL simultaneously pushes visual understanding and Agent capabilities to new heights. Taken together, these developments reveal AI simultaneously sinking to the compute limits of edge devices while penetrating hardware domains like circuit board design. The industry has officially entered a pragmatic phase where lightweight deployment and multimodal capability advance in parallel, marking the end of pure scale worship and the beginning of functional integration across the full stack.

Today's Top 3 Headlines

  1. AI Industry News

    🤖 Qwen 3.8-Flash-Next Achieves Efficient Phone CPU Inference

    Alibaba's Tongyi team revealed Qwen3.8-Flash-Next LLM now runs efficiently on mobile CPUs, enabling on-device inference without cloud dependency. For mobile developers and edge AI, this means high-performance LLMs deployable on ordinary phones, dramatically lowering barriers for edge intelligence applications.

    Source
  2. AI Industry News

    🤖 Altman: 38,000 ChatGPT queries use water equivalent to one almond

    OpenAI CEO Sam Altman said each 38,000 ChatGPT queries uses water equivalent to one California almond, with total data center consumption comparable to office buildings. For the AI industry, this suggests water usage pressure may be far lower than public perception, and related environmental concerns could be overestimated.

    Source
  3. AI Industry News

    🤖 AI Race: 10 New Models Drop, Led by Gemini, GPT-6, Claude, Qwen, DeepSeek

    Google, OpenAI, Anthropic, Microsoft & Meta dropped 10 new LLMs this week—Gemini 3.8 Flash, GPT-6 Astra & Claude 5.1 among them. For devs and investors, this sudden expansion of options means multimodal competition is now white-hot, with inference costs likely to plunge further.

    Source

+5 more headlines

  • 🤖 B.AI launches full-stack infrastructure, free DeepSeek-V4-Flash to power Agent era
  • 🤖 Ling-3.0-flash-VL Launched: Built on Ling-3.0-flash, Enhanced Vision & Agent Capabilities
  • 🤖 Can Claude Opus Design PCBs? AI Auto-Layout Sparks Debate
  • 🤖 Tesla Cybercab Without Steering Wheel Under Federal Safety Probe
  • 🤖 Zuck to Trump: Hassabis Vision to Dominate US AI Policy
Unlock all 8 headlines + deep analysis →Free 3-day trial · cancel anytime
Browse all past briefings →