jundot/omlx
jundot/omlxOtherPython
20.0k3回登場
LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar
apple-siliconinference-serverllmmacosmlxopenai-api
このトピックのトレンドリポジトリ(2件)
LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar
Open-source inference server and production cluster for all the models your agent needs.