Home / 🧱 AI Foundation Stack / ⚑ Inference / Serving

brontoguana/krasis

Krasis is a Hybrid LLM runtime which focuses on efficient running of larger models on consumer grade VRAM limited hardware

🧱 AI Foundation Stack License unclear β˜… 469C++

Commercial license

License unclearNOASSERTION

ζœͺζ¨™η€ΊζŽˆζ¬Š β€” 商用前務必璺θͺ(ι θ¨­θ¦–η‚ΊδΏη•™ζ‰€ζœ‰ζ¬Šεˆ©)

Topics

cpu-inferencegguf-model-supportgpu-inferencehigh-performance-inferencehybrid-inferenceinference-engineinference-optimizationlarge-language-modelsllama-cpp-alternativellm-inferencemixture-of-expertstransformer
Ad