LINUX DO user cooor shared a screenshot-led post stating that GLM-5.3-FlashX had been released, prompting a forum discussion on model availability, costs, and plan eligibility.
In the repliesCommenters noted the model feels faster but drains allowances more rapidly, while one user reported being unable to access it via their Coding Plan and cited official documentation calling it unsupported. Commenter yyf123 linked to official Z.AI pricing, which lists GLM-5.3-FlashX at $0.37 per 1M input tokens, $0.075 per 1M cached input tokens, and $1.25 per 1M output tokens—roughly 2.5 times the rates for GLM-5.3-Flash ($0.15 input, $0.03 cached input, $0.50 output), with cached storage listed as limited-time free for both.
Official documentation corroborates the per-token pricing structure, but community observations regarding speed, allowance burn, and Coding Plan restrictions remain anecdotal. The thread lacks formal benchmark scores, architecture specifications, or a comprehensive plan-eligibility matrix.
LINUX DO thread (https://linux.do/t/topic/2917803) and official Z.AI pricing documentation (https://docs.z.ai/guides/overview/pricing).