InFeeo
Language

6.4x faster than llama.cpp, 3.9x faster than MLX(github.com)

×
Link preview GitHub - basecompute/baseRT: Fastest LLM inference runtime for Apple Silicon Fastest LLM inference runtime for Apple Silicon. Contribute to basecompute/baseRT development by creating an account on GitHub. GitHub · github.com
BaseRT is the fastest LLM runtime on Apple Silicon. Install it with one command and run local models on your own device.

Comments

Log in Log in to comment.

No comments yet.