Home / π§± AI Foundation Stack / β‘ Inference / Serving
tinyBigGAMES/VindexLLM
VindexLLM is a pure Delphi, GPU-powered LLM inference engine that uses Vulkan compute shaders to run GGUF models entirely on the GPU. It performs full transformer inference without relying on Python, CUDA, or other external runtimes, requiring only vulkan-1.dll, which is typically included with modern GPU drivers.
Commercial license
License unclearNOASSERTION
ζͺζ¨η€Ίζζ¬ β εη¨εεεΏ η’Ίθͺ(ι θ¨θ¦ηΊδΏηζζζ¬ε©)
Topics
delphillama-cppllm-inferenceobject-pascalollamawin64windows-10windows-11
Ad