The Technology

An 80B Qwen Model Runs in 4.3 GB of RAM on a Mac

via github.com·yesterday

A new quantization and streaming approach lets an 80-billion-parameter Qwen model run in roughly 4.3 GB of RAM on a Mac, with a 35B variant running on an iPhone. The demonstration continues a steady collapse in the hardware floor for large models, pushing capable inference from data centers toward devices people already own.

Read Full Story at github.com
AITechnology

Related Stories

SpaceX Is Absorbing Surging AI Costs as Insiders Prepare to Sell Shares

MSNBC·12h ago

Texas Halts New Data Center Connections to Its Power Grid

Ars Technica·13h ago

Mistral Releases Shieldstral, a 3B Open-Weights Model for Multimodal Moderation

mistral.ai·17h ago

Apple Says More Former Employees May Have Taken Confidential Data to OpenAI

TechCrunch·18h ago
DiscussSoon
← Front Page