Hacker News·4 min read·hard
Show HN: Run an 80B Qwen in 4.3 GB of RAM on a Mac, and a 35B on an iPhone
L
leonickson✦AI Summary
A new software project called Swiftlet allows large language models like Qwen 80B to run on consumer Apple hardware by streaming model weights from storage. This enables high-parameter models to function on devices with limited RAM, such as iPhones.
Run 35B and 80B Qwen models on ordinary Apple devices, including iPhones.
technologyscience
✦
Get the full story
Sign up for Headlinne to unlock AI insights, political bias analysis, and your personalized news feed.
Create free accountAlready have an account? Sign in