AI Briefing
Feed
About
Search
KO
Sign in
All
Feed, trending, and board
Feed
Latest news
Trending
Open source and releases
Board
Blogs and showcases
Tags
Browse by topic
#llama-cpp
The latest AI and developer news about #llama-cpp, with the original source and a short summary.
Feed
Trending
Tags
Settings
#llama-cpp - page 4 | AI Briefing
Unified Offline AI App 'Box' for Android Released
Reddit
·
2026.04.24 18:00
Qwen3.6 35B on a 780M iGPU
Reddit
·
2026.04.24 17:00
Qwen 3.6 27B local 20 TPS
Reddit
·
2026.04.24 16:00
Reka Edge 2603 support added
Reddit
·
2026.04.23 23:00
Local LLM VSCode
Reddit
·
2026.04.23 20:00
Qwen3.6 64K real-world test
Reddit
·
2026.04.23 14:00
Qwen3.6-27B Uncensored Release Published
Reddit
·
2026.04.23 03:00
Gemma 4 VLA Running on Jetson
HuggingFace Blog
·
2026.04.23 00:00
Qwen3.6-35B tops the charts
Reddit
·
2026.04.22 20:00
TurboQuant fixes Qwen3.6
Reddit
·
2026.04.22 08:00
q1_0 speed explosion
Reddit
·
2026.04.21 20:00
Mistral Small 4 update
Reddit
·
2026.04.20 02:00
ARM Local Inference Speed
Reddit
·
2026.04.19 05:00
MoE 54% speedup
Reddit
·
2026.04.18 16:00
35B Benchmark on RTX 5060 Ti
Reddit
·
2026.04.18 09:00
100tps with 5070 Ti + RX 9070
Reddit
·
2026.04.18 07:00
MoE flipped the script
Reddit
·
2026.04.18 02:00
Unsloth Tops the Benchmark
Reddit
·
2026.04.18 01:00
Qwen 3.6 Benchmark
Reddit
·
2026.04.17 21:00
Qwen3.6 fully unlocked
Reddit
·
2026.04.17 09:00
ROCm Turnabout?
Reddit
·
2026.04.17 07:00
Qwen 3.6 1L Benchmark
Reddit
·
2026.04.17 06:00
40 tok/s on a 3080
Reddit
·
2026.04.17 05:00
Removing CUDA checks with graph_reused
Reddit
·
2026.04.16 21:00
Previous
1
2
3
4
5
Next
Previous
1
2
3
4
5
Next