Tired of ads? Enjoy an ad-free experience by signing up.
👩‍🍳 How we use AI at Tech in Asia, thoughtfully and responsibly.
🧔‍♂️ A friendly human may check it before it goes live. More news here

Alibaba open-sources WebSailor, a web AI agent

Alibaba’s AI division, Tongyi, has released its web AI agent, WebSailor, as an open-source project.

The agent achieved the highest score on the BrowseComp benchmark, surpassing models such as DeepSeek R1 and Grok-3. BrowseComp is designed to test the reasoning and retrieval capabilities of web AI agents in complex scenarios.

The project is available on GitHub, allowing developers and researchers to explore and build upon the technology.

This move reflects a growing trend of tech companies open-sourcing AI tools to promote collaboration and innovation in the field.

🔗 Source: 36Kr


🧠 Food for thought

1️⃣ Strategic open-sourcing reflects shifting AI business models beyond proprietary advantage

Alibaba’s WebSailor release follows a clear industry pattern where companies increasingly view open-sourcing as strategically beneficial rather than a competitive disadvantage.

A comprehensive 2023 survey found that 76% of technology leaders expect to increase their use of open-source AI technologies in the coming years, reflecting the growing recognition of its business value 1.

This trend is particularly pronounced in the technology sector, where McKinsey research shows 72% of companies are already utilizing open-source AI models 2.

The shift occurs as companies recognize that open-sourcing creates multiplier effects: transparency builds trust, community involvement improves models, and broader adoption expands potential applications beyond what a single organization might develop.

For Chinese tech giants like Alibaba, open-sourcing also serves as a counterbalance to Western AI dominance, potentially accelerating global adoption of their technological frameworks and standards.

2️⃣ Benchmark competitions reflect the intensifying global AI race

Alibaba’s emphasis on WebSailor outperforming models like DeepSeek R1 and Grok-3 on the BrowseComp benchmark highlights how performance metrics have become crucial competitive battlegrounds in AI development.

Comprehensive benchmarking data shows that model performance correlates strongly with estimated training compute resources, with significant performance jumps occurring past certain computational thresholds 3.

These benchmarks have become proxy measures for technological advancement, with US models currently outperforming non-US models on key metrics, though the performance gap is narrowing with each new release 3.

Recent Alibaba developments

Stay ahead in Asia’s tech landscape

You've reached your 2 free content limit for the month. Sign up for free to read the full story.

🏄 For casual readers / 👶 Free

Basic

US$0

Free forever

Get instant access to this article and more every month

0 premium content

Unlimited news briefs

5

5 articles

Ad-free reading experience

Just US$0 per day

⌛Sign up in 20s. No payment details needed.

📖 For learners / 👍 Starter

Lite

US$4.92/month

Billed annually at US$59/year

Get instant access to this article and more every month

4

4 premium content

Unlimited news briefs & articles

Ad-free reading experience

Just US$0.17 per day

Cancel anytime

Our subscriber community includes professionals from these companies:

Stay updated on the go with our mobile app.

Get latest insights with smoother, more personalized experience through TIA mobile app.