Home / 🧱 AI Foundation Stack / ⚑ Inference / Serving

jundot/omlx

LLM inference server with continuous batching & SSD caching for Apple Silicon β€” managed from the macOS menu bar

🧱 AI Foundation Stack βœ“ Commercial OK β˜… 16,432Python

Commercial license

βœ“ Commercial OKApache-2.0

ε―ε•†η”¨οΌŒι€šεΈΈεͺιœ€δΏη•™θ‘—δ½œζ¬Šθ²ζ˜Ž/授權撝款

Topics

apple-siliconinference-serverllmmacosmlxopenai-api
Ad