Observed arrival · 2026-09-21
Quadpoint AI Studio Puts Oversized Models on a Small Mac
A native macOS inference engine that pages open-weight model weights from an SSD, letting Apple silicon machines run models larger than their available RAM.
○Open
⊠Login
$Paid
†Ads
✦Pretty
●Pro
◎Niche
◉Human
⚑Risk
ƒJS
- For
- Mac developers running local open-weight models
- Worth noticing
- The page reports a 144 GB model bundle running on a 16 GB Mac through SSD-backed demand paging.
Field notes
The engine organizes weights, tokenizer, configuration, and provenance into one stamped .qpx bundle, then uses RAM as a cache while reading other weights from SSD storage. The homepage reports 16 and 19 tokens-per-second results for two models and describes a 290B model as a 144 GB bundle running on a 16 GB Mac. A localhost server and terminal Companion extend the same engine into Xcode and coding workflows.
Observed signals
Read the marks
Editorial observations of this landing page, not a rating.
○OpenPublic substance visible
✦PrettyNotable craft visible
●ProPolished or operationally mature
◎NicheUnusually specific use
One card from the complete issue