AI Briefing
Feed
About
Search
KO
Sign in
All
Feed, trending, and board
Feed
Latest news
Trending
Open source and releases
Board
Blogs and showcases
Tags
Browse by topic
#quantization
The latest AI and developer news about #quantization, with the original source and a short summary.
Feed
Trending
Tags
Settings
TextCLF Releases Calibration-Free TQ Quantization Method and Open-Source Quant Factory
Reddit
·
2026.09.25 07:00
UkisAI Releases Swift Series: Efficient Qwen-Based LLMs with Reduced Thinking Tokens
Reddit
·
2026.09.25 01:00
Hugging Face Introduces Native GGUF Support in Transformers
Reddit
·
2026.09.23 14:00
K2-Horizon GGUF Quantizations Released
Reddit
·
2026.09.23 05:00
TextCLF Releases Calibration-Free 4-bit Quantization Tool
Reddit
·
2026.09.23 03:00
Gewell Inference Engine Optimized for Gemma 4 31B Released
Reddit
·
2026.09.21 18:00
PrismML Releases 'Ternary Bonsai 2 27B', a Ternary Quantized Model Maintaining 98.2% of FP16 Performance
PyTorchKR
·
2026.09.19 11:00
Single-File C Engine for Gemma Multimodal Inference Released
Reddit
·
2026.09.16 19:00
2026 AI Inference Hardware Innovations: New Architectures for Prefill/Decode Disaggregation and Solving Memory Bottlenecks
GeekNews
·
2
·
2026.09.15 22:00
Voodoo Dynamic Quant Released Under MIT License
Reddit
·
2026.09.15 15:00
Pinterest Reduces Costs by Up to 30% with SSD Serving and Quantization in Manas Embedding Search Platform
Karrot
·
2026.09.15 09:00
Edge0 Runs 35B MoE Model on 3GiB Memory via SSD Streaming
PyTorchKR
·
2026.09.15 08:00
Evolution of Pinterest's Embedding Search Platform
Pinterest Engineering
·
2026.09.12 00:00
Bartowski Applies New Tensor Layout Maps for GGUF Models
Reddit
·
2026.09.11 04:00
NVIDIA Releases INT4 Quantized Version of Cosmos3 64B
Reddit
·
2026.09.09 23:00
NVIDIA Releases Model Lightweighting Tool
PyTorchKR
·
1
·
2026.09.02 09:00
The Efficient Frontier of LLM Inference (6-minute read)
TLDR AI
·
2
·
2026.09.02 09:00
NVIDIA Releases NVFP4 Quantized Model of DeepSeek-V4-Pro-0813
TLDR AI
·
2026.08.31 09:00
Qwen3-Coder 4bit TQ Released
Reddit
·
2026.08.31 05:00
Qwen3.8-27B SOTA GGUF Released
Reddit
·
1
·
2026.08.29 06:00
Mismatch between GGUF filenames and actual bit counts
Reddit
·
2026.08.29 05:00
The Pitfalls of Ollama KV Cache Configuration
Reddit
·
2026.08.28 06:00
New Feature Update for ik_llama.cpp
Reddit
·
2026.08.27 16:00
Qwen3.8-27B NVFP4 Quantization Released
Reddit
·
2026.08.26 10:00
Previous
1
2
3
4
5
Next
Previous
1
2
3
4
5
6
7
8
9
10
Next
#quantization | AI Briefing