Twitter/X

OpenRouter reports that GLM-5.2 providers in the Zai_org ecosystem are rolling…

Brief

OpenRouter says Zai_org's GLM-5.2 provider pool is increasing inference speed with new fast endpoints from Wafer AI and FireworksAI. By using the model tag "z-ai/glm-5.2:nitro," users are automatically routed to the fastest provider based on live traffic, which OpenRouter frames as delivering the best price and performance from the deepest inference market.

Why it matters

OpenRouter reports that GLM-5.2 providers in the Zai_org ecosystem are rolling out faster inference variants as of 2026-06-26, adding new endpoints from @wafer_ai and @FireworksAI_HQ.

Key details

  • Setting the model specifier to "z-ai/glm-5.2:nitro" will route requests to the currently fastest provider using live traffic data, aiming to optimize price and performance automatically.
Source evidence

Building on the deepest inference market = automatically getting the best price and performance!

OpenRouter (@OpenRouter)

TIP 💡@Zaiorg GLM-5.2 providers are working on faster and faster inference! Today's new endpoints include @waferai and @FireworksAI_HQ fast variants.

Set your model to "z-ai/glm-5.2:nitro" to continuously get the fastest provider based on live traffic data.

— https://nitter.net/OpenRouter/status/2070310815476097314#m