Open Source
Llamma.cpp
If you run LLMs locally on a laptop, phone, or Raspberry Pi, you're probably already running llama.cpp — you just didn't know it. ggml-org/llama.cpp is LLM inference in C/C++ — no Python stack, no cloud required. Created in March 2023 and now with 127K+