19 August 2026
System enables massive AI models to run on personal computers
First reported
TLDR AI ran this on .
- FreeToken, a new system, allows Mixture of Experts models (AI models split into specialized components) to run on individual laptops and workstations by dynamically adjusting how much data moves between the device and the cloud.
- The system works with over 20 different large models, ranging from 35 billion parameters (a measure of model size) on laptops with 8GB of GPU memory to 753 billion parameter models on single workstation GPUs.
- Rather than requiring specialized data center hardware, FreeToken remaps which model components load into available memory based on what the specific device can handle in real time.
How it was covered
TLDR AITLDR editorial team
FreeToken continuously remaps experts and model state to available bandwidth and memory on personal machines. The system supports over 20 MoE models, from 35B models on 8GB laptop GPUs to 753B models on single workstation GPUs.