01
jundot/omlx
LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar
- Stars
- 19.1K
- Forks
- 1.6K
- Pushed
- Today
Current GitHub adoption and maintenance signals.
LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar
Lossless DFlash speculative decoding for MLX on Apple Silicon
MLX: An array framework for Apple silicon
Docs of the Hugging Face Hub
Integrate Magisk root and Google Apps (OpenGApps) into WSA (Windows Subsystem for Android)