Back to Tools Directory

omlx

An LLM inference server designed for Apple Silicon that enhances performance through continuous batching and SSD caching, easily managed from the macOS menu bar. It streamlines the process of running large language models, making it more efficient for users on Mac systems.

Categories & Topics

LLMmacos

Repository Stats

37
GitHub Stars
Repository
jundot/omlx

Related Tools