FreeToken brings datacenter-scale model serving to your desktop. Run massive models locally, fast and efficiently.
Tags
Similar apps
More in this categoryNative MLX runtime for Laya typed decision models — 7–14 ms short decisions on M3 Max. No text generation, PyTorch, or cloud API.
6.9kPythonApache-2.0Native LLM inference server for Apple Silicon. OpenAI + Anthropic API compatible. No Python.
1.8kZigmacOSDesktopmacOS menubar app for fast local DeepSeek V4.1, with 1M context.
916SwiftMITmacOSDesktop