llama.cpp
C++
★ 7
updated 3mo ago
⑂ fork
LLM inference in C/C++ (fork of PrismML fork that enables CPU (incl AVX2 and AVX512) and ROCm for AMD GPUs
No plain-English explanation yet — one is being written right now. Check back in a minute.