Personal fork of Rapid-MLX (Apple Silicon local LLM engine) adding user-configurable model aliases, path overrides, and a resumable rapid-mlx pull download manager with stall/Xet-fallback handling.
mlx inference-engine huggingface openai-api hugging-face llm huggingface-hub antrophic claude-code rapid-mlx mlx-inference
-
Updated
Aug 18, 2026 - Python