jundot/omlx
steadyLLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar
Python
View on GitHub
Stars
22,026
Forks
1,902
Open issues
968
24h
+41
+0.2%
7d
+242
+1.1%
Refresh
1h
Star history (7 days)
Last checked
59m ago
Last pushed
1h ago
Next check
just now