Tired of ads? Enjoy an ad-free experience by signing up.
👩‍🍳 How we use AI at Tech in Asia, thoughtfully and responsibly.
🧔‍♂️ A friendly human may check it before it goes live. More news here

Zhipu AI caps GLM coding plan sign-ups after demand surge

Zhipu AI, a Beijing-based AI firm, is limiting new sign-ups for its GLM Coding Plan due to performance issues caused by a surge in demand, the company announced on January 21.

The restrictions cap new user sign-ups at 20% of existing levels to prioritize current users amid a computing capacity crunch.

Zhipu’s global operations head said the limits are to free up resources and denied targeted rate limiting.

The company cited the strain to a fivefold increase in traffic after launching its latest model, GLM 4.7, last month.

Industry analysts note that Chinese AI firms face infrastructure constraints, especially in serving inference requests at scale, which hampers their ability to expand globally.

🔗 Source: South China Morning Post

🧠 Food for thought

Implications, context, and why it matters.

The severity of Zhipu’s compute crunch remains unclear without infrastructure details

  • To judge whether the pause on new sign-ups for the GLM Coding Plan is a short-lived disruption or a longer limit, the scale of Zhipu AI’s inference compute (the chips and servers used to run its AI model for user requests) needs to be spelled out.
  • Its hardware sourcing also needs to be pinned down, including how much it depends on sanctioned Nvidia chips, domestic substitutes, or black-market gear, since some Chinese AI companies are considering buying Nvidia’s high-performance H200 chips from the black market.
  • More clarity should be provided on the split between cloud partners and on-premise data centers, plus how the recent Hong Kong IPO proceeds will be directed toward computing infrastructure, to judge how the bottleneck will be addressed.

Chinese AI’s global push could create opportunities for APAC compute providers

  • International GPU (graphics processing unit) cloud providers can pick up spillover demand from Chinese AI firms with global users when domestic infrastructure runs hot.
  • APAC operators with ready Nvidia H100 or AMD MI300X capacity in hubs such as Singapore or Hong Kong can move fastest, since some providers publish on-demand GPU instance pricing 1.
  • Inference-optimization software vendors can also benefit by helping Chinese AI companies cut compute cost per user, which can ease strain on existing hardware while improving uptime.

Recent Zhipu AI developments

Stay ahead in Asia’s tech landscape

You've reached your 2 free content limit for the month. Sign up for free to read the full story.

🏄 For casual readers / 👶 Free

Basic

US$0

Free forever

Get instant access to this article and more every month

0 premium content

Unlimited news briefs

5

5 articles

Ad-free reading experience

Just US$0 per day

⌛Sign up in 20s. No payment details needed.

📖 For learners / 👍 Starter

Lite

US$4.92/month

Billed annually at US$59/year

Get instant access to this article and more every month

4

4 premium content

Unlimited news briefs & articles

Ad-free reading experience

Just US$0.17 per day

Cancel anytime

Our subscriber community includes professionals from these companies:

Stay updated on the go with our mobile app.

Get latest insights with smoother, more personalized experience through TIA mobile app.