srvsngh99/Krill
Fast local LLM inference CLI for Apple Silicon. 1.57x faster than Ollama, 58% less memory.
GitHub repository with 6 stars and 0 forks.
Language: Swift
Topics: apple-silicon, gemma, llm, llm-inference, local-llm, macos, mlx, ollama, on-device-ai, swift