GLM 5.2 Fast via Wafer, Now Available on AI Gateway
Key point
The high-performance Wafer-based GLM 5.2 Fast model is now available on Vercel AI Gateway.
Details
The GLM 5.2 Fast model, provided via Wafer, is now available on Vercel's AI Gateway. According to their own benchmark results, Wafer recorded 2x higher throughput than other GLM-5.2 providers in serverless environments.
Based on test results, the performance of GLM 5.2 Fast is as follows:
- Small context: 170+ tok/s
- Large context: 200+ tok/s
AI Gateway provides a unified API for model calls, and ensures high uptime through usage and cost tracking, retry, failover, and performance optimization features. It also includes features such as custom reporting, Zero Data Retention support, and per-API-key budget settings.
No additional margin or platform fees are charged on any inference, and provider pricing policies are reflected as-is for BYOK (Bring Your Own Key) requests as well.
This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.
Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.