Tired of ads? Enjoy an ad-free experience by signing up.
👩‍🍳 How we use AI at Tech in Asia, thoughtfully and responsibly.
🧔‍♂️ A friendly human may check it before it goes live. More news here

DeepSeek V4.1-Flash targets lower AI agent costs

DeepSeek, a Hangzhou-based Chinese AI company founded by Liang Wenfeng, said its V4.1-Flash model uses less key-value (KV) cache than the previous generation to lower cache-related costs in AI agent workloads.

DeepSeek said its official partners WorkBuddy, including CodeBuddy, and OpenCode now support V4.1-Flash, and users can select the model through the deepseek-flash setting.

CodeBuddy is described as an AI code editor.

The company also said it plans to work with the open-source community on inference support for V4.1-Flash and explore more deployment options, including potential large-scale setups with 2,000 graphics processing units (GPUs) and a storage cluster.

DeepSeek said the update would help it serve more users at lower cost and that it would pass those savings on to users.

🔗 Source: DeepSeek

Recent DeepSeek developments

Stay ahead in Asia’s tech landscape

You've reached your 2 free content limit for the month. Sign up for free to read the full story.

🏄 For casual readers / 👶 Free

Basic

US$0

Free forever

Get instant access to this article and more every month

0 premium content

Unlimited news briefs

5

5 articles

Ad-free reading experience

Just US$0 per day

⌛Sign up in 20s. No payment details needed.

📖 For learners / 👍 Starter

Lite

US$4.92/month

Billed annually at US$59/year

Get instant access to this article and more every month

4

4 premium content

Unlimited news briefs & articles

Ad-free reading experience

Just US$0.17 per day

Cancel anytime

Our subscriber community includes professionals from these companies:

Stay updated on the go with our mobile app.

Get latest insights with smoother, more personalized experience through TIA mobile app.