Where to Access GLM-5.2: Provider Roundup
Melvin Vivas · X post · 2026-06-23 · Open on X
Topics: LLM Fundamentals, LLMOps, Deployment & Monitoring · Level: intermediate
Summary
Lists the providers that offered GLM-5.2 when the post was written: Z.ai's coding plan, Together AI, Baseten, Fireworks, Vercel AI Gateway and OpenRouter. It is a handy reference for picking an inference provider or gateway for an open-weights model.
Key points
- Z.ai (the maker of GLM) offers GLM-5.2 through its coding plan subscription.
- Inference providers hosting it include Together AI, Baseten and Fireworks.
- Gateways or routers such as Vercel AI Gateway and OpenRouter also give access.
- Open-weights models are often available from many providers, so you can compare price and latency.
Resources mentioned
- Z.ai (@Zai_org) on X · person · x.com · free
Official X account of Z.ai, maker of the GLM models and the GLM coding plan.
Also in: Open models now dominate token volume on Vercel AI Gateway (Melvin Vivas on X · notes), GLM 5.3 Open Weights Release Delayed for Framework Support (Melvin Vivas on X · notes), AI News Roundup: Claude Fable 5, Scientist AI, ZCode, NVIDIA RL (Melvin Vivas on X · notes), Trying ZCode by Z.ai with GLM 5.2 on a Mac (Melvin Vivas on X · notes) and 9 more - Together AI · tool · x.com · paid
Inference platform for hosting and calling open-weights models.
Also in: Fine-tune and Deploy Qwen3.8 27B on Together AI (Melvin Vivas on X · notes), Where to Access GLM 5.2: Inference Providers and Gateways (Melvin Vivas on X · notes), Free GLM-5.2 via Hugging Face Inference Providers in coding agents (Melvin Vivas on X · notes) - Baseten · tool · x.com · paid
Model inference platform offering dedicated GPU deployments for serving open models.
Also in: Serving Qwen3.8 27B FP8 on an H100 with Baseten dedicated inference (Melvin Vivas on X · notes), Coworker Desktop App Works with Any OpenAI-Compatible Endpoint (Melvin Vivas on X · notes), Running GLM 5.3 Flash on Baseten with the Pi Coding Agent (Melvin Vivas on X · notes), GLM 5.3 Flash Now Available on Baseten (Melvin Vivas on X · notes) and 6 more - Fireworks AI · tool · x.com · free
Fireworks AI's official X account, which posts news about inference and serving open models.
Also in: FireRouter: cost-aware model routing between open models and Claude Opus (Melvin Vivas on X · notes), GLM 5.2 Hits 446 tok/s on Fireworks AI (Melvin Vivas on X · notes), GLM 5.2 Speed vs Opus 4.8 and GPT 5.5 (Melvin Vivas on X · notes), GLM 5.2 on Fireworks Makes Opus 4.8 and GPT 5.5 Feel Slow (Melvin Vivas on X · notes) and 4 more - Vercel · tool · x.com · free · recommended by both Bashiri Smith & Melvin Vivas
Deploy web apps and frontends.
Also in: OpenAI DevDay 2026 Recap: Dots, Agents API, Codex Cloud & Marketplace (Melvin Vivas on X · notes), 7 Habits to Become an AI Engineer: Books, Tooling, Research & Shipping (Bashiri Smith on Facebook · notes), Jev model added to the AIBackends API via Vercel AI Gateway (Melvin Vivas on X · notes), Open models now dominate token volume on Vercel AI Gateway (Melvin Vivas on X · notes) and 7 more - OpenRouter · tool · openrouter.ai · free
A single OpenAI-compatible API that routes requests to hundreds of models from many providers, with one bill and model fallbacks.
Also in: Customizing Your Coding Setup with Pi Coding Agent Extensions (Melvin Vivas on X · notes), The Jev model is now on OpenRouter (Melvin Vivas on X · notes), Kev-4B model, an alternative to Jev, now available on OpenRouter (Melvin Vivas on X · notes), Space Bunny Alpha: stealth 1M-context flash model on OpenRouter (Melvin Vivas on X · notes) and 42 more
More in LLM Fundamentals
- Sakana Fugu model now available on OpenRouter
- How GPT Works: From Token Embeddings to Multi-Head Attention
- GLM-5.2 on Fireworks: Top Open-Weights Model on GDPval-AA
- Google Interactions API GA: one endpoint for inference and agents
- Google Interactions API is GA: recommended for Gemini
- GLM 5.2 open model now deployable on Google Cloud