AI Briefing
Feed
About
Search
KO
Sign in
All
Feed, trending, and board
Feed
Latest news
Trending
Open source and releases
Board
Blogs and showcases
Tags
Browse by topic
#llama.cpp
The latest AI and developer news about #llama.cpp, with the original source and a short summary.
Feed
Trending
Tags
Settings
#llama.cpp - page 3 | AI Briefing
llama.cpp webui adds video support
Reddit
·
2026.05.17 12:00
What Does GGUF Contain Besides Weights, and What's Still Missing
GeekNews
·
2026.05.16 16:00
llama.cpp Releases Update with RDNA3 Flash Attention Fix
Reddit
·
2026.05.15 09:00
What Does GGUF Contain Beyond Weights, and What's Still Missing?
Hacker News
·
2026.05.15 02:00
llama.cpp Adds SarvamMoE Support
Reddit
·
2026.05.10 03:00
Qwen3.6 MTP Layer Grafting and Inference Optimization
Reddit
·
2026.05.08 05:00
llama.cpp Adds MiMo V2.5 Support
Reddit
·
2026.05.07 20:00
Qwen 3.6 Achieves 2.5x Inference Speed Boost with MTP Support
Reddit
·
2026.05.06 18:00
llama.cpp Adds Beta Support for MTP
Reddit
·
2026.05.04 21:00
Llama 3.2 1B, deployed on Android with 480 examples
Reddit
·
2026.05.01 14:00
llama.cpp, Hexagon NPU real-world benchmark
Reddit
·
2026.05.01 14:00
llama.cpp -sm tensor improvement
Reddit
·
2026.04.30 02:00
llama.cpp Adds Blackwell NVFP4 Support
Reddit
·
2026.04.29 17:00
llama.cpp merges SM120 NVFP4 MMQ
Reddit
·
2026.04.29 09:00
llama.cpp adds flash-attn support
Reddit
·
1
·
2026.04.29 06:00
Nemotron support added to llama.cpp
Reddit
·
2026.04.29 02:00
Qwen 3.6 KV Cache Benchmark
Reddit
·
2026.04.29 02:00
Intel B70 Backend Comparison
Reddit
·
2026.04.27 06:00
llama.cpp adds support for DeepSeek v4 Flash
Reddit
·
2026.04.26 19:00
Q4_K_XL faster than Q4_K_M
Reddit
·
2026.04.26 18:00
Lubuntu outperforms on llama.cpp
Reddit
·
2026.04.26 18:00
llama.cpp and ik_llama.cpp FP4 Support
Reddit
·
1
·
2026.04.26 00:00
llama.cpp MMQ Optimization
Reddit
·
2026.04.25 23:00
Qwen3.6 larger quants are faster
Reddit
·
2026.04.25 06:00
Previous
1
2
3
4
Next
Previous
1
2
3
4
Next