Tired of ads? Enjoy an ad-free experience by signing up.
👩‍🍳 How we use AI at Tech in Asia, thoughtfully and responsibly.
🧔‍♂️ A friendly human may check it before it goes live. More news here

Homegrown AI app HKChat leads HK App Store chart

HKChat, an AI chatbot developed by the Hong Kong Generative AI Research and Development Centre (HKGAI), attracted about 90,000 users in its first week after launching in Hong Kong.

The app topped the free app ranking on Apple’s Hong Kong App Store, supports Cantonese, Mandarin, and English, and offers local information such as bus arrivals, weather, and legal regulations.

HKGAI, funded by Hong Kong’s InnoHK program, was established in 2023 by the Hong Kong University of Science and Technology and other local institutions.

The chatbot uses HKGAI V1, a large language model built on China’s DeepSeek models, and was trained on Nvidia’s H800 AI chips.

🔗 Source: South China Morning Post

🧠 Food for thought

Implications, context, and why it matters.

HKGAI funding and compute will decide if HKChat stays free

  • HKGAI runs on $235 million for Research and Development (R&D) plus operations 1, a HK$200 million donation from the Ng Teng Fong Charitable Foundation (a local philanthropic foundation) 2, with no model cost detail 1.
  • A first‑phase government Large Model Smart Computing Center (a public AI training and inference facility) using NVIDIA DGX H800 systems was near HK$600 million 3.
  • Heavy first‑week traffic raised inference (running the model to produce answers) costs, so slow replies reveal capacity strain that needs more compute or efficiency steps like quantization (reducing precision to shrink then speed up models) to keep access free.
  • HKGAI has not shared a runway (how long current funding lasts) or monetization plans, so without near‑term revenue staying free depends on InnoHK funds plus donations covering compute, with lower per‑query costs via technical gains.

GPU clouds can ease Hong Kong’s low‑latency gap for Large Language Models (LLMs)

  • HKChat latency exposes demand for nearby inference that HKUST’s H800 SuperPOD (a clustered GPU system) 4 plus Cyberport’s Artificial Intelligence Supercomputing Centre (AISC) subscriptions 5 serve only for research and startups.
  • Cyberport offers H800 access from HK$1,800/month 5, but air‑gapped conditions (no internet connectivity) plus 20GB transfer caps 5 block real‑time chatbots.
  • TraxComm runs Tier III+ (high‑availability) sites with carrier‑neutral links plus direct GPU provider connections 6, so firms use colocation (renting space to host your own servers) for hybrid production AI.
  • MLOps vendors can help with inference tuning. Techniques include quantization, caching (reusing prior results) and dynamic batching (grouping requests to maximize GPU utilization) which cut H800 cost per query while improving odds access stays free.

Stay ahead in Asia’s tech landscape

You've reached your 2 free content limit for the month. Sign up for free to read the full story.

🏄 For casual readers / 👶 Free

Basic

US$0

Free forever

Get instant access to this article and more every month

0 premium content

Unlimited news briefs

5

5 articles

Ad-free reading experience

Just US$0 per day

⌛Sign up in 20s. No payment details needed.

📖 For learners / 👍 Starter

Lite

US$4.92/month

Billed annually at US$59/year

Get instant access to this article and more every month

4

4 premium content

Unlimited news briefs & articles

Ad-free reading experience

Just US$0.17 per day

Cancel anytime

Our subscriber community includes professionals from these companies:

Stay updated on the go with our mobile app.

Get latest insights with smoother, more personalized experience through TIA mobile app.