6.4x faster than llama.cpp, 3.9x faster than MLX
BaseRT is the fastest LLM runtime on Apple Silicon. Install it with one command and run local models on your own device.