AI Briefing
Feed
About
Search
KO
Sign in
All
Feed, trending, and board
Feed
Latest news
Trending
Open source and releases
Board
Blogs and showcases
Tags
Browse by topic
#vllm
The latest AI and developer news about #vllm, with the original source and a short summary.
Feed
Trending
Tags
Settings
#vllm - page 4 | AI Briefing
vLLM Adds Support for Mistral 3.5
Reddit
·
2026.04.29 05:00
Power Limits and TG/s on 2x3090
Reddit
·
2026.04.28 13:00
Moat or Commons
TLDR AI
·
2026.04.28 09:00
Qwen3.6-27B vLLM Docker, 118 TPS
Reddit
·
2026.04.27 22:00
Gemma 4 E2B leads on H100
Reddit
·
2026.04.26 18:00
H100 Qwen·Gemma Performance Comparison
Reddit
·
2026.04.25 19:00
Cohere MoE added to vLLM
Reddit
·
2026.04.25 01:00
vLLM Fixes Qwen Tool Calling Bug
Reddit
·
2026.04.24 18:00
vLLM Recipes Revamp - One Click for Model+Hardware Combo Configs
GeekNews
·
2026.04.23 13:00
Building a Chat UI with Qwen3.6-27b
Reddit
·
2026.04.23 11:00
27B GGUF Released
Reddit
·
2026.04.22 23:00
15B Supernet Released
Reddit
·
2026.04.22 23:00
Cautions When Fine-Tuning and Deploying Gemma-4
Reddit
·
2026.04.19 07:00
Super-accelerating with 2x 3090
Reddit
·
2026.04.19 04:00
Qwen Comparative Analysis
Reddit
·
2026.04.18 10:00
The winner flips depending on the model
Reddit
·
2026.04.18 06:00
MIG cuts p95 in half
Reddit
·
2026.04.18 04:00
Qwen 3.6 and Agentic Coding (GitHub Repository)
TLDR AI
·
2026.04.17 09:00
64 tok/s peak on 8x MI50
Reddit
·
2026.04.17 02:00
Strengthening a Summarization Model with GRPO
Reddit
·
2026.04.15 18:00
Arc B70 Struggles
Reddit
·
2026.04.15 07:00
vLLM Bottleneck Tracing
Reddit
·
2026.04.15 05:00
2026 vLLM Korea Meetup
Rebellions
·
2026.04.14 11:00
Breaking Down the Nondeterminism of LLM Inference
TLDR AI
·
2026.04.14 09:00
Previous
1
2
3
4
5
Next
Previous
1
2
3
4
5
Next