# How to Run Qwen3.6-35B on Your Mac at 77 tok/s

Level: intermediate Estimated time: 20-40 minutes (most of it is the model download) Minimum requirements: Mac with Apple Silicon (M1/M2/M3/M4) and 48 GB of unified RAM What are we setting up? A local server compatible with the OpenAI API that runs the Qwen3.6-35B-A3B model (quantized to 4 bits) using MLX, Apple's Machine Learning framework for Silicon. When you're done, you'll have an endpoint at http://127.0.0.1:7979 that you can point any OpenAI-compatible client to (OpenCode,...

Original Source

Read the full article at Dev →

KhanList aggregates and links to publicly available news content. We do not host full articles from third-party sources. Always verify important information with original sources.