Tired of ads? Enjoy an ad-free experience by signing up.
👩‍🍳 How we use AI at Tech in Asia, thoughtfully and responsibly.
🧔‍♂️ A friendly human may check it before it goes live. More news here

S Korean AI chip firm Rebellions launches AI chip to rival Nvidia

Rebellions, a South Korean AI chip startup, is aiming to compete globally with its inference-focused chips, said CEO Park Sung-hyun.

The company raised 350 billion won (US$237 million) earlier this year from investors including Samsung Electronics and Arm.

The company recently launched the Rebel-Quad chip, which combines four AI inference chips to support large-scale data processing.

Rebellions said the new chip targets governments and companies looking for alternatives to Nvidia’s products.

🔗 Source: Yonhap

🧠 Food for thought

Implications, context, and why it matters.

Rebellions inference claims need outside checks beyond 2023 MLPerf

  • In 2023 MLPerf inference for language and vision (MLPerf is an industry-standard AI benchmarking suite), Rebellions’ ATOM beat Qualcomm and Nvidia 1. Since then, Nvidia topped MLPerf Training v5.1 with Blackwell Ultra (its latest GPU platform), delivering over 4x Llama 3.1 405B pretraining and nearly 5x Llama 2 70B Low-Rank Adaptation (LoRA) fine-tuning versus Hopper 2.
  • REBEL-Quad (a module that combines four Rebellions inference chips) lacks independent benchmarks 3, which makes competitiveness against current Nvidia GPUs hard to judge given Nvidia’s one-year innovation cycle 2.
  • Pricing, total cost of ownership data, and availability timelines for REBEL-Quad remain unpublished. Enterprise IT buyers weighing alternatives to Nvidia are unsure whether switching costs, performance trade-offs justify moving off existing GPU infrastructure.

Software integrators can win work porting to Rebellions chips

  • Rebellions Software Development Kit (SDK) supports PyTorch, TensorFlow, and hundreds of models 4. The focus stays on inference-only acceleration (speeding up model execution rather than training or fine-tuning, no fine-tuning support) 4, which opens room for system integrators (IT consulting and implementation firms) and cloud service providers to migrate LLM inference workloads from Nvidia to Rebellions hardware for cost-sensitive customers in government or enterprise.
  • Support for vLLM (an open-source high-throughput inference engine for large language models) and Triton (Nvidia’s software for model serving plus inference) 3 ships with Rebellions chips, enabling consultancies plus managed service providers to build multi-vendor practices that hedge supply chain risks, plus cut costs.

Stay ahead in Asia’s tech landscape

You've reached your 2 free content limit for the month. Sign up for free to read the full story.

🏄 For casual readers / 👶 Free

Basic

US$0

Free forever

Get instant access to this article and more every month

0 premium content

Unlimited news briefs

5

5 articles

Ad-free reading experience

Just US$0 per day

⌛Sign up in 20s. No payment details needed.

📖 For learners / 👍 Starter

Lite

US$4.92/month

Billed annually at US$59/year

Get instant access to this article and more every month

4

4 premium content

Unlimited news briefs & articles

Ad-free reading experience

Just US$0.17 per day

Cancel anytime

Our subscriber community includes professionals from these companies:

Stay updated on the go with our mobile app.

Get latest insights with smoother, more personalized experience through TIA mobile app.