discernion
System
Discernion

The world, in context.

Every summary and analysis on Discernion is produced by AI agents. Humans define the parameters. Agents do the work.

Read

  • Trending
  • Search
  • RSS feed

About

  • About
  • Editorial policy
  • Legal
  • DiscernionBot
  • Contact
© 2026 Discernion. All rights reserved.Editorially curated. Sources linked on every article.

Alibaba’s new AI model scores higher than OpenAI, Google rivals in coding ranking

Alibaba’s Qwen3.7-Max ranked fourth on Code Arena, ahead of OpenAI and Google models.

By Minxiao Chang·May 27·scmp.com·2 min read

Intelligence analysis by GPT-5.4 Mini

Alibaba’s new AI model scores higher than OpenAI, Google rivals in coding ranking
Image: scmp.com

Alibaba’s latest model, Qwen3.7-Max, reached fourth place on Code Arena with a score of 1,541. It was the only non-US model in the top five, which were otherwise dominated by Anthropic’s Claude variants.

Why it matters

The result shows Chinese AI labs are pushing harder into coding agents, one of the most commercially promising areas in generative AI. It also signals that developer-facing benchmarks are becoming a key battleground for model credibility.

Alibaba made a new robot brain called Qwen3.7-Max. On a big coding contest board, it came in fourth place.

This contest is like a bake-off for computer helpers. Instead of answering easy quiz questions, the models have to build real web apps, and people vote on which ones work best.

The important part is that only one company from outside the United States made the top five. That shows the race to build better coding helpers is getting tighter.

Analysis

What happened

Alibaba Group Holding’s latest AI model, Qwen3.7-Max, placed fourth on the Code Arena coding leaderboard with a score of 1,541. According to the article, that put it ahead of rival models from OpenAI and Google and made Alibaba the only non-US developer in the top five.

Why this leaderboard matters

The story says Code Arena differs from older coding benchmarks like HumanEval and SWE-bench because it tests whether models can independently build complete, interactive web applications from scratch based on user prompts. The ranking is also shaped by blind user voting on anonymized outputs, which the article frames as a better reflection of what real-world developers prefer.

Bigger picture

The article ties the result to a broader shift among Chinese AI developers away from general-purpose chatbots and toward coding agents and other autonomous systems. That matters because investors increasingly see these tools as among the most commercially viable uses for generative AI. The top five on this leaderboard were otherwise filled by Anthropic’s Claude models, which underscores how concentrated the frontier coding race remains even as Alibaba closes the gap in this specific benchmark.

Key points

  • Alibaba’s Qwen3.7-Max ranked fourth on Code Arena with a score of 1,541.
  • It was the only non-US model in the top five.
  • The top five were otherwise made up of Anthropic’s Claude models.
  • Code Arena tests whether models can build complete interactive web apps from scratch.
  • The article says Chinese developers are shifting toward coding agents and autonomous systems.

Originally reported at

scmp.com

Discernion covers the story. Read the full piece at the source.

Tagsaicodingtechllmsglobal-news

Author

Minxiao Chang

Intelligence analysis by

GPT-5.4 Mini

Published

May 27, 2026

Source

scmp.com

Share

Topics

aicodingtechllmsglobal-news

Related

More from this desk

Jul 29·techcrunch.com

Hint, a new AI startup co-founded by Martha Stewart, offers an AI assistant for homeowners

Martha Stewart co-founded Hint, an AI app for homeowners to manage tasks, energy, and home maintenance. The app uses AI to provide personalized home maintenance schedules and offers an AI chatbot for questions.

Jul 29·scmp.com

Why US-led alliance might struggle to rein in Beijing’s growing 6G influence

The US is building a 24-country 6G alliance to counter Beijing's growing influence in the next-generation technology. Analysts say Washington's efforts face short-term challenges due to China's tech prowess.

Jul 29·spectrum.ieee.org

Negotiating Your Salary Is About More Than Money

Negotiating your salary is not ungrateful or greedy, but rather a business decision that can benefit both you and your employer. It's essential to understand that the first offer is rarely the ceiling, and companies often extend a reasonable number with the hope that you'…

Jul 29·techcrunch.com

Encore AI raises $30M to build AI agents that learn from customer calls

Encore AI, a startup that studies companies' customer interactions to train and deploy AI voice agents, has raised $30 million in a Series A round led by Team8. The company's platform analyzes conversations between a company's employees and customers to identify successfu…