Base Compute · Documentation

BaseRT

The fastest LLM inference runtime for Apple Silicon. Pull a model from HuggingFace, chat with it, or serve an OpenAI-compatible API

$ curl -LsSf https://basecompute.co/install.sh | sh