edge-lm
Run compressed large language models on iPhones and Apple Silicon using MLX with 7x smaller Gemma checkpoints optimized for on-device performance.
Run compressed large language models on iPhones and Apple Silicon using MLX with 7x smaller Gemma checkpoints optimized for on-device performance.