Tired of ads? Enjoy an ad-free experience by signing up.
👩‍🍳 How we use AI at Tech in Asia, thoughtfully and responsibly.
🧔‍♂️ A friendly human may check it before it goes live. More news here

Google tests Gemini AI against Anthropic’s Claude

Google contractors are evaluating the performance of its Gemini AI by comparing it with Anthropic’s Claude, focusing on accuracy, truthfulness, and verbosity.

Evaluators take up to 30 minutes per prompt to score the models, noting that Claude exhibits stricter safety protocols, often refusing unsafe prompts, while Gemini has been flagged for safety violations.

Internal documents reveal Claude’s responses sometimes explicitly identify the model and emphasize its adherence to Anthropic’s safety policies.

Despite Anthropic’s terms prohibiting the use of Claude for training competing systems, Google has not confirmed securing permission for these tests.

Anthropic declined to comment. Google DeepMind said comparing models is standard practice and denied using Claude outputs to train Gemini, following contractor concerns about Gemini’s accuracy on sensitive topics like healthcare.

Stay ahead in Asia’s tech landscape

You've reached your 2 free content limit for the month. Sign up for free to read the full story.

🏄 For casual readers / 👶 Free

Basic

US$0

Free forever

Get instant access to this article and more every month

0 premium content

Unlimited news briefs

5

5 articles

Ad-free reading experience

Just US$0 per day

⌛Sign up in 20s. No payment details needed.

📖 For learners / 👍 Starter

Lite

US$4.92/month

Billed annually at US$59/year

Get instant access to this article and more every month

4

4 premium content

Unlimited news briefs & articles

Ad-free reading experience

Just US$0.17 per day

Cancel anytime

Our subscriber community includes professionals from these companies:

Stay updated on the go with our mobile app.

Get latest insights with smoother, more personalized experience through TIA mobile app.