What is BaseRT?
Running AI locally means no per token cost + full privacy. BaseRT is the fastest LLM runtime on Apple Silicon. Install it with one command and run local models on your own device.
6.4x faster than llama.cpp, 3.9x faster than MLX
Running AI locally means no per token cost + full privacy. BaseRT is the fastest LLM runtime on Apple Silicon. Install it with one command and run local models on your own device.
Ask the makers a question, share feedback or tell others how you use BaseRT.
Sign in to comment, ask the makers a question or share useful feedback.
No comments yet. Be the first to share your thoughts on BaseRT.