The Technology
Ongoing Story — 95 related articles

An 80B Qwen Model Runs in 4.3 GB of RAM on a Mac

via github.com·Aug 3

A new quantization and streaming approach lets an 80-billion-parameter Qwen model run in roughly 4.3 GB of RAM on a Mac, with a 35B variant running on an iPhone. The demonstration continues a steady collapse in the hardware floor for large models, pushing capable inference from data centers toward devices people already own.

Read Full Story at github.com
AITechnology

Related Stories

Suspecting the Court Used AI, a Man Injected Prompts Into His Filings

Ars Technica·Aug 14

Google Says Homomorphic Encryption Can Make Private AI Practical

blog.google·Aug 14

Teens Are Turning to AI Chatbots for Emotional Support

Phys.org·Aug 13

Inside the Safety Reckoning at OpenAI After Its Rogue Agent Hack

Wired·Aug 13
DiscussSoon
← Front Page