llama.cpp
llama.cpp is LLM inference in C/C++. Contribute to ggml-org/llama.cpp development by creating an account on GitHub.. It is ranked #281 on the AI Radar, in Open & local models, first seen 8 days ago and shared in 3 posts (50.1k views).
What people said about llama.cpp on X
If you want to be in the 1% who run AI locally with $0, start with these 10 GitHub repos. 1. llama.cpp (128k stars) The lean C/C++ core that makes models run fast on both CPUs and GPUs. https://github.com/ggml-org/llama.cpp 2. Jan (44.4k stars) A free ChatGPT swap that keeps everything running locally on your…
— @hasantoxr, 8 days ago · 215 likes · see the post
llama.cpp v0.4.1 https://github.com/ggml-org/llama.cpp/releases/tag/v0.4.1
— @ggml_org, 5 days ago · 210 likes · see the post
DeepSeek-V4.1-Flash GGUF is LIVE. 🔥 I converted the new 748B model into a full GGUF ladder on ONE DGX Spark, built the llama.cpp conversion support it needed, and submitted it upstream as PR #28696. There was no DeepSeek-V4.1 conversion path in llama.cpp when I started. Now there is. 𝗗𝗘𝗘𝗣𝗦𝗘𝗘𝗞-𝗩𝟰.𝟭 𝗜𝗦…
— @ViC305, 8 days ago · 142 likes · see the post
Alternatives to llama.cpp
- Pirate Face — The un-bannable, checksum-verified mirror for open models. Claim your handle before a squatter does.
- atomic.chat — Run AI models locally ->
- deepseek-v4.1-flash-exl3-2x-dgx-sparks — DeepSeek v4.1 Flash EXL3 2.9 bpw for 2x DGX Sparks - MiaAI-Lab/DeepSeek-v4.1-Flash-EXL3-2x-DGX-Sparks
- bespoke-nimble-9b — We’re on a journey to advance and democratize artificial intelligence through open source and open science.
- Cactus Compute — It runs on mobiles, wearables, smart home devices, small robots and microcontrollers, with prebuilt engines for macOS,…
- orcabonsai-27b-uncensored — Open source
llama.cpp in numbers
- Rank on the AI Radar: #281 of 1192
- Shared in 3 posts by 3 accounts: @hasantoxr, @ggml_org, @ViC305
- 50.1k views on those posts
- First seen 8 days ago, last shared 5 days ago
- Pricing seen by Jev: open source
- Market: Open & local models
FAQ
What is llama.cpp?
LLM inference in C/C++. Contribute to ggml-org/llama.cpp development by creating an account on GitHub. It was first shared on X 8 days ago and is ranked #281 on the AI Radar.
Is llama.cpp free?
It is open source.
Who shared llama.cpp?
3 accounts on X, including @hasantoxr, @ggml_org, @ViC305, in 3 posts totalling 50.1k views.
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 27.2k posts from 4.8k X accounts over the last 14 days, 1.2k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-19 18:40 UTC. Full method.