August 1, 2026 Running a 2.78T Flagship AI on 29GB RAM: 0.5 Tokens Per Second aikimilocal-inferencelocal-ai