The fastest LLM inference runtime for Apple Silicon. Pull a model from HuggingFace, chat with it, or serve an OpenAI-compatible API
$ curl -LsSf https://basecompute.co/install.sh | sh