AI Briefing
Feed
About
Search
KO
Sign in
All
Feed, trending, and board
Feed
Latest news
Trending
Open source and releases
Board
Blogs and showcases
Tags
Browse by topic
#quantization
The latest AI and developer news about #quantization, with the original source and a short summary.
Feed
Trending
Tags
Settings
#quantization - page 6 | AI Briefing
Mistral 3.5 MLX 4-bit Released
Reddit
·
2026.05.01 06:00
JSON Vector Search Package Released
Reddit
·
2026.05.01 02:00
IBM Granite 4.1 30B GGUF Quantization Released
Reddit
·
2026.04.30 07:00
oQ vs Q vs MXFP vs UD KLD Comparison
Reddit
·
2026.04.30 05:00
Valkyr with TurboQuant applied
Reddit
·
2026.04.30 04:00
llama.cpp NVFP4 Comparison
Reddit
·
2026.04.29 21:00
llama.cpp Adds Blackwell NVFP4 Support
Reddit
·
2026.04.29 17:00
Qwen 3.6 KV Cache Benchmark
Reddit
·
2026.04.29 02:00
Qwen 3.6 27B GGUF Comparison
Reddit
·
2026.04.28 21:00
2x 5060 Ti Qwen3.6 Benchmark
Reddit
·
1
·
2026.04.28 04:00
TurboQuant: An Explanation from First Principles
TLDR AI
·
2026.04.27 14:00
Qwen3.6-27B MLX 3-bit Released
Reddit
·
2026.04.27 12:00
New Inference Engine Hipfire Released for AMD GPUs
Reddit
·
2026.04.27 10:00
Qwen3.6 35B Heretic released
Reddit
·
2026.04.26 20:00
Q4_K_XL faster than Q4_K_M
Reddit
·
2026.04.26 18:00
Qwen3.6 larger quants are faster
Reddit
·
2026.04.25 06:00
MLX Q vs oQ Comparison
Reddit
·
2026.04.25 03:00
Nemotron-3 GGUF Released
Reddit
·
2026.04.24 17:00
22 tok/s on RTX 5060 Ti 16GB
Reddit
·
2026.04.24 09:00
Qwen3.6 64K real-world test
Reddit
·
2026.04.23 14:00
27B won
Reddit
·
2026.04.23 12:00
27B Optimization for the 3090
Reddit
·
2026.04.23 09:00
Qwen3.6-27B: Flagship-Level Coding Performance Achieved with a 27B Dense Model
TLDR AI
·
2026.04.23 09:00
Qwen3.6-27B Uncensored Release Published
Reddit
·
2026.04.23 03:00
Previous
4
5
6
7
8
Next
Previous
1
2
3
4
5
6
7
8
9
10
Next