AI Briefing
Feed
About
Search
KO
Sign in
All
Feed, trending, and board
Feed
Latest news
Trending
Open source and releases
Board
Blogs and showcases
Tags
Browse by topic
#gguf
The latest AI and developer news about #gguf, with the original source and a short summary.
Feed
Trending
Tags
Settings
#gguf - page 2 | AI Briefing
GGUF Release with Integrated MTP Head
Reddit
·
1
·
2026.05.31 14:00
Qwen 3.6 35B GGUF Released
Reddit
·
2026.05.21 00:00
llama.cpp adds MTP support
Reddit
·
2026.05.16 21:00
What Does GGUF Contain Besides Weights, and What's Still Missing
GeekNews
·
2026.05.16 16:00
What Does GGUF Contain Beyond Weights, and What's Still Missing?
Hacker News
·
2026.05.15 02:00
Ollama Adds Direct llama.cpp Support
Reddit
·
1
·
2026.05.14 23:00
Unsloth releases Qwen models with MTP applied
Reddit
·
2026.05.11 23:00
GGUF Model Uploads Surge 2x
Reddit
·
2026.05.11 19:00
BeeLlama.cpp: llama.cpp Optimization
Reddit
·
2026.05.10 01:00
Qwen 3.6 Achieves 2.5x Inference Speed Boost with MTP Support
Reddit
·
2026.05.06 18:00
Unsloth Fixes Mistral Medium 3.5 GGUF
Reddit
·
2026.05.02 16:00
IBM Granite 4.1 30B GGUF Quantization Released
Reddit
·
2026.04.30 07:00
Mistral Medium 3.5 Released
Reddit
·
2026.04.30 00:00
IK_LLAMA Adds Qwen MTP Support
Reddit
·
2026.04.29 23:00
MiMo-V2.5 GGUF Released
Reddit
·
2026.04.29 14:00
llama.cpp merges SM120 NVFP4 MMQ
Reddit
·
2026.04.29 09:00
Qwen 3.6 27B GGUF Comparison
Reddit
·
2026.04.28 21:00
Qwen3.6-27B throughput doubled
Reddit
·
2026.04.28 01:00
llama.cpp adds support for DeepSeek v4 Flash
Reddit
·
2026.04.26 19:00
Darwin-36B-Opus, GPQA 88.4%
Reddit
·
2026.04.26 04:00
RTX 5070 Ti 44 t/s
Reddit
·
2026.04.24 20:00
Nemotron-3 GGUF Released
Reddit
·
2026.04.24 17:00
Qwen 3.6 27B local 20 TPS
Reddit
·
2026.04.24 16:00
Qwen3.6 64K real-world test
Reddit
·
2026.04.23 14:00
Previous
1
2
3
Next
Previous
1
2
3
Next