What it does
A local AI runtime for running GGUF LLMs from a desktop app or Python on everyday computers.
77,399C++
License: MITRepository last updated: May 27, 2025
View details OSS TAGS
llama.cpp related
Popular starting points
Newest
What it does
A local AI runtime for running GGUF LLMs from a desktop app or Python on everyday computers.
What it does
A C/C++ LLM inference runtime spanning GGUF, quantization, CPU/GPU backends, CLI, and an OpenAI-compatible server.