Github Top Repositories
13.9K subscribers
2.36K photos
59 videos
10 files
2.66K links
Top GitHub repositories in one place 🚀
Explore the best projects in programming, AI, data science, and more.
Download Telegram
Github Top Repositories
Photo
🔥 jundot/omlx is trending — and it deserves your attention.

🔗 https://github.com/jundot/omlx
📝 LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar
──────────────────────────────

oMLX is an optimized solution for running Large Language Models (LLMs) on Macs, providing a convenient and controlled experience. The key features include continuous batching and tiered KV caching, which enable efficient model inference and caching. omlx can be installed via Homebrew, from source, or by downloading the macOS app.

To get started, simply run omlx serve --model-dir ~/models to discover and serve LLMs, VLMs, embedding models, and rerankers from subdirectories. The server includes a built-in chat UI and supports OpenAI-compatible clients.

Some technical highlights of oMLX include its ability to handle concurrent requests, support for vision-language models, and a web-based admin dashboard for real-time monitoring and model management. The dashboard also features a model downloader, allowing users to search and download MLX models directly.

oMLX is designed for developers, researchers, and anyone looking to run LLMs on their Macs. With its user-friendly interface and robust feature set, it's an excellent choice for those seeking a seamless and efficient LLM experience.

In a nutshell, oMLX makes running Large Language Models on your Mac a breeze - giving you the perfect blend of convenience, control, and performance.

──────────────────────────────
🧠 Channel: https://t.iss.one/GithubRe
1