Back to library
Build the GameLibraryOpen sourceFree

llama.cpp

llama.cpp is an open C/C++ inference project with GGUF, many quantizations, CPU and GPU backends, a server API, and broad hardware support.

Visit official site
Content updated Aug 7, 2026Automated check reached the official site · Checked Aug 24, 2026

Why use it

Game tools and experimental NPCs can run text models offline or privately while controlling deployment cost and data flow.

Where it fits

Find engines, frameworks, code libraries, version control, and production automation.

aiengine-integration

What to check

Repository licensing does not cover downloaded models; govern model rights, prompt safety, latency, memory, and unreliable output separately.

Search the site

Search resources and field guides

    Privacy settings

    Your language choice, open home-page sections, favorites, comparisons, and recent views stay in this browser. Nothing is uploaded.