GLM 5.2 Fast via Wafer AI is now available on Vercel AI Gateway. Wafer's inference stack delivers 2x higher throughput than other serverless GLM-5.2 providers, with benchmarks showing 170+ tok/s for small context and 200+ tok/s for large context. Developers can access it by setting the model to `zai/glm-5.2-fast` in the AI SDK. Vercel AI Gateway offers a unified API with usage tracking, retries, failover, Zero Data Retention, and no markup on provider pricing.

2m read timeFrom vercel.com
Post cover image
11 Impressions