AI Briefing
Feed
About
Search
KO
Sign in
All
Feed, trending, and board
Feed
Latest news
Trending
Open source and releases
Board
Blogs and showcases
Tags
Browse by topic
#quantization
The latest AI and developer news about #quantization, with the original source and a short summary.
Feed
Trending
Tags
Settings
#quantization - page 2 | AI Briefing
4-bit LLMs Surpass Original Model Performance
HuggingFace Blog
·
1
·
2026.08.25 20:00
Meta Releases MoE LLM for Mobile
Reddit
·
2
·
2026.08.25 03:00
llama-cpp-turboquant Adds ConvRot Quantization Support
Reddit
·
2026.08.24 10:00
Qwen3.8-27B 2-bit Quantization Released
Reddit
·
2026.08.21 02:00
Unsloth Dynamic 3.0 GGUF Released
TLDR AI
·
2026.08.20 09:00
Liquid AI Releases LFM2.5 Q4_0
HuggingFace Blog
·
2026.08.19 22:00
Pick
Uncensored Qwen3.8-27B MLX Model Released
HuggingFace Blog
·
3
·
2026.08.19 11:00
DFlash 2 Adds Support for Qwen 3.8 27B
Reddit
·
1
·
2026.08.19 06:00
Qwen3.8 vs Qwen3.6 vs Gemma 4 on 24GB GPUs (10-minute read)
TLDR AI
·
2026.08.18 09:00
Qwen3.8-27B-Uncensored-MLX (4 min read)
TLDR AI
·
2026.08.18 09:00
WarpQuant: 3-bit Quantization Based on Hadamard Rotation
Reddit
·
2026.08.16 10:00
bitsandbytes Creator Previews New Quantization Method
Reddit
·
2026.08.14 22:00
WARP Runs K3 on 64GB MacBook
PyTorchKR
·
1
·
2026.08.12 18:00
Pick
Needle 2 - A 14MB Agentic LLM for Smartphones, Wearables, Smart Home, and Robots
GeekNews
·
1
·
2026.08.11 13:00
NVFP4 Distillation Preserves Internal Geometry
Reddit
·
2026.08.10 05:00
Kimi K3 GGUF Released
Reddit
·
2026.08.07 17:00
Core Technologies for LLM Inference Optimization: Quantization, KV Cache, and Inference Chips
KT Cloud
·
2026.08.06 14:00
LittleBit-2: Maximizing Spectral Energy Gain in Sub-1-Bit LLMs via Latent Geometric Alignment
Samsung Research
·
1
·
2026.08.04 08:00
Smaller, Faster, Safer: Running Kimi and GLM at Scale
Cloudflare Blog
·
2026.08.03 22:00
MSLK Kernel Reference (Website)
TLDR AI
·
2026.08.03 09:00
NanoQuant: Efficient Sub-1-Bit Quantization for Large Language Models
Samsung Research
·
2026.08.03 08:00
EdgeRazor, 1.88-bit LLM
Reddit
·
2026.08.03 04:00
Escha-W2 (Hugging Face repository)
TLDR AI
·
2026.07.30 09:00
BeeLlama.cpp v0.4.1 Released: Improved KV Cache Quantization Performance
Reddit
·
2026.07.27 01:00
Previous
1
2
3
4
5
Next
Previous
1
2
3
4
5
6
7
8
9
10
Next