# oMLX — LLM Inference Server with Continuous Batching for Apple Silicon > A high-performance LLM inference server optimized for Apple Silicon that provides continuous batching and SSD caching, manageable via a macOS menu bar app. ## Install Copy the content below into your project: --- Source: https://tokrepo.com/en/workflows/asset-39d63ff2 Author: