AI Briefing
Feed
About
Search
KO
Sign in
All
Feed, trending, and board
Feed
Latest news
Trending
Open source and releases
Board
Blogs and showcases
Tags
Browse by topic
#llama-cpp
The latest AI and developer news about #llama-cpp, with the original source and a short summary.
Feed
Trending
Tags
Settings
#llama-cpp - page 3 | AI Briefing
llama.cpp Improves Prompt Processing Speed for MTP
Reddit
·
2026.05.18 00:00
llama.cpp b9180 Adds MTP Support
Reddit
·
2026.05.17 02:00
llama.cpp Adds MTP Support
Reddit
·
2026.05.16 21:00
llama.cpp adds MTP support
Reddit
·
2026.05.16 21:00
Ollama Adds Direct llama.cpp Support
Reddit
·
1
·
2026.05.14 23:00
Hugging Face releases ml-intern
Reddit
·
2026.05.14 19:00
TextGen Desktop App Released
Reddit
·
2026.05.13 22:00
Hermes, a Self-Improving AI Agent Built on NVIDIA RTX PCs and DGX Spark
NVIDIA Blog
·
2026.05.13 22:00
llama.cpp Adds 'Continue Generation' Support for Reasoning Models
Reddit
·
2026.05.13 19:00
llama.cpp adds local model evaluation feature
Reddit
·
2026.05.12 21:00
MiniCPM-V 4.6 Released
Product Hunt
·
2026.05.12 18:00
llama.cpp Prepares to Support Integration of MTP and Multimodal Features
Reddit
·
2026.05.12 07:00
llama.cpp Adds Blackwell Tensor Parallelism Support
Reddit
·
2026.05.10 22:00
BeeLlama.cpp: llama.cpp Optimization
Reddit
·
2026.05.10 01:00
RTX6k Qwen3.5 Benchmark Updated
Reddit
·
2026.05.03 08:00
Qwen3.5 DFlash Running on RTX 2080 SUPER
Reddit
·
2026.05.01 20:00
IK_LLAMA Adds Qwen MTP Support
Reddit
·
2026.04.29 23:00
llama.cpp NVFP4 Comparison
Reddit
·
2026.04.29 21:00
MiMo-V2.5 GGUF Released
Reddit
·
2026.04.29 14:00
Qwen3.6-27B IQ4_XS Regression
Reddit
·
2026.04.28 21:00
Per-iPhone Gemma 4 E2B Configuration Report
Reddit
·
2026.04.28 21:00
6900 XT ROCm vs Vulkan Comparison
Reddit
·
2026.04.28 18:00
16GB M4, Qwen 35B SSD bottleneck
Reddit
·
2026.04.28 18:00
R9700 Ngram-Mod Speed Test
Reddit
·
1
·
2026.04.27 13:00
Previous
1
2
3
4
5
Next
Previous
1
2
3
4
5
Next