AI Briefing
Feed
About
Search
KO
Sign in
All
Feed, trending, and board
Feed
Latest news
Trending
Open source and releases
Board
Blogs and showcases
Tags
Browse by topic
#llama-cpp
The latest AI and developer news about #llama-cpp, with the original source and a short summary.
Feed
Trending
Tags
Settings
#llama-cpp - page 2 | AI Briefing
pi 0.81.0 adds llama.cpp support
Reddit
·
2026.07.22 00:00
llama.cpp Improves AMD ROCm Performance and Fixes Bugs
Reddit
·
2026.07.21 15:00
BeeLlama.cpp v0.4.0 Released
Reddit
·
2026.07.20 03:00
Show HN: Reame – a CPU inference server that gets faster the more it runs
Hacker News
·
2026.07.12 01:00
EAGLE3 integration into llama.cpp completed
Reddit
·
2026.06.12 16:00
OSCAR: 2-bit KV Cache Quantization Technique Released
Reddit
·
2026.06.10 04:00
llama.cpp Adds Support for Granite4 Vision
Reddit
·
2026.06.06 01:00
llama.cpp Optimizes MTP Inference for Qwen
Reddit
·
2026.06.04 02:00
Mellum/Granite Embedding Support
Reddit
·
2026.06.03 14:00
mistral.rs Significantly Improves CUDA Inference Performance
Reddit
·
2026.06.01 23:00
llama.cpp Adds Support for EXAONE 4.5
Reddit
·
2026.06.01 18:00
GGUF Release with Integrated MTP Head
Reddit
·
1
·
2026.05.31 14:00
llama.cpp Unified Binary and Website
Reddit
·
2026.05.30 01:00
llama.cpp AMD/ROCm Update
Reddit
·
1
·
2026.05.29 10:00
Harbor Adds Support for Running Local LLM Agents
Reddit
·
2026.05.26 23:00
Qwen3.6 Model Performance Degradation Bug
Reddit
·
1
·
2026.05.24 13:00
llama.cpp Native Tool Support
Reddit
·
1
·
2026.05.24 07:00
llama.cpp Adds NVFP4/MTP Support
Reddit
·
2026.05.24 03:00
Llama.cpp adds PDL support
Reddit
·
2026.05.23 06:00
llama.cpp Fixes VRAM Leak in MTP Models
Reddit
·
2026.05.22 07:00
LM Studio Adds MTP Support
Reddit
·
2026.05.20 12:00
llama.cpp Reflects MTP Improvements
Reddit
·
2026.05.19 21:00
llama.cpp Adds MTP Support
Reddit
·
2026.05.19 04:00
llama.cpp b9200 Release
Reddit
·
2026.05.18 08:00
Previous
1
2
3
4
5
Next
Previous
1
2
3
4
5
Next