Home / π§± AI Foundation Stack / π€ Agent Frameworks / Orchestration
huawei-csl/KVarN
KVarN is a native vLLM KV-cache quantization backend for your agents: 3-5x more context, throughput above FP16, and FP16-level accuracy. Calibration-free, one flag.
Commercial license
β Commercial OKApache-2.0
ε―εη¨οΌιεΈΈεͺιδΏηθδ½ζ¬θ²ζ/ζζ¬ζ’ζ¬Ύ
Topics
agentic-aikv-cachellmllm-inferencelong-contextquantizationvllm
Ad