AI Briefing
Feed
About
Search
KO
Sign in
All
Feed, trending, and board
Feed
Latest news
Trending
Open source and releases
Board
Blogs and showcases
Tags
Browse by topic
#quantization
The latest AI and developer news about #quantization, with the original source and a short summary.
Feed
Trending
Tags
Settings
#quantization - page 9 | AI Briefing
Standard Format Elevated
HuggingFace Blog
·
2026.04.08 09:00
Fujitsu One Compression (3 min read)
TLDR AI
·
2026.04.02 09:00
AI21 Unveils Jamba 1.5 Model Family
AI21 Labs
·
2026.03.25 20:00
[Tech Series] kt cloud AI Retrieval-Augmented Generation (RAG) #4: Embedding and Vector Indexing Technology
KT Cloud
·
2026.03.23 16:00
Beyond 'Strong and Powerful Morning': An Experiment Log on TranslateGemma as a Replacement for GPT-4o-mini
Musinsa
·
2026.03.16 07:00
How Low-Bit Inference Makes AI Efficient
Dropbox Tech
·
2026.02.13 03:00
On-device AI Face Verification Pipeline Optimization
Hyperconnect
·
2026.01.23 09:00
Holistic Optimization of AI Inference Systems
FuriosaAI
·
2025.12.08 09:00
VLM Optimization and Execution Guide for Intel CPUs
HuggingFace Blog
·
2025.10.15 09:00
SD3.5-Flash: Distribution-Guided Distillation for Generative Flow
Stability AI Research
·
2025.09.27 05:00
transformers, gpt-oss Optimization Update
HuggingFace Blog
·
2025.09.11 09:00
FLUX.1-dev, QLoRA Training on Low-Spec GPUs
HuggingFace Blog
·
2025.06.19 09:00
Comparing and Using Diffusers Quantization Backends
HuggingFace Blog
·
2025.05.21 09:00
Intel Unveils AutoRound, a High-Performance Quantization Tool
HuggingFace Blog
·
2025.04.29 09:00
Optimizing and Deploying LLMs for Intel Hardware
HuggingFace Blog
·
2024.09.20 09:00
Llama 3 1.58-bit Fine-tuning Technique Released
HuggingFace Blog
·
2024.09.18 09:00
Getting Started with ggml, the Foundation of Local LLMs
HuggingFace Blog
·
2024.08.13 09:00
Diffusion Memory Optimization Based on Quanto
HuggingFace Blog
·
2024.07.30 09:00
Hugging Face Unveils KV Cache Quantization Feature
HuggingFace Blog
·
2024.05.16 09:00
SetFit Inference Accelerated 7.8x on Intel CPUs
HuggingFace Blog
·
2024.04.03 09:00
Hugging Face Unveils Quantization Backend 'Quanto'
HuggingFace Blog
·
2024.03.18 09:00
StarCoder Inference Accelerated 7x on Intel Xeon
HuggingFace Blog
·
2024.01.30 09:00
Guide to Optimizing LLM Production Deployment
HuggingFace Blog
·
2023.09.15 09:00
🤗 Transformers Supported Quantization Methods Guide
HuggingFace Blog
·
2023.09.12 09:00
Previous
6
7
8
9
10
Next
Previous
1
2
3
4
5
6
7
8
9
10
Next