jundot/omlx
jundot/omlxOtherPython
19.2k
LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar
apple-siliconinference-serverllmmacosmlxopenai-api
このトピックのトレンドリポジトリ(2件)
LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar
Deepseek to API: A lightweight, high-performance full-stack middleware converting client protocols to universal APIs. Supports multi-account rotation, compiled binaries, Vercel Serverless, and Docker. Compatible with Google, Claude, and OpenAI API formats.