ON
← Back to feed
Chinese AI agent outperforms Anthropic’s Claude Code in autonomous research
HK💻 Technology8 hr. ago

Chinese AI agent outperforms Anthropic’s Claude Code in autonomous research

A Chinese AI system called Qiushi Engine, developed by a team at Zhejiang University, has surpassed Anthropic’s Claude Code in an international benchmark for autonomous scientific research. As of Tuesday, Qiushi Engine ranked first on the ResearchClawBench leaderboard, followed by Open Science Desktop and Claude Code. The benchmark evaluates AI agents' ability to independently conduct research and compare their findings to human-authored papers. The test was created by researchers at the Shanghai Artificial Intelligence Laboratory to determine whether AI systems can truly perform the complex tasks they claim to handle. Qiushi Engine, launched recently, is described as a large language model-based agent capable of end-to-end autonomous scientific discovery in real-world settings.

How each side covered it

The same event, grouped by the political lean of the outlets covering it.

How each side covered it

Support independent, bias-aware news and unlock the social pulse, community voting, and your personalized For You feed.

Become a Supporter

Covered around the world

The same event as reported in other countries.

Covered around the world

Support independent, bias-aware news and unlock the social pulse, community voting, and your personalized For You feed.

Become a Supporter

Claims check

Key factual claims, and how many sources assert vs dispute each.

Claims check

Support independent, bias-aware news and unlock the social pulse, community voting, and your personalized For You feed.

Become a Supporter

1 reports

South China Morning Post logoSouth China Morning PostIndependentCenter8 hr. ago
Chinese AI agent outperforms Anthropic’s Claude Code in autonomous research

A Chinese AI system called Qiushi Engine, developed by a team at Zhejiang University, has surpassed Anthropic’s Claude Code in an international benchmark for autonomous scientific research. As of Tuesday, Qiushi Engine ranked first on the ResearchClawBench leaderboard, followed by Open Science Desktop and Claude Code. The benchmark evaluates AI agents' ability to independently conduct research and compare their findings to human-authored papers. The test was created by researchers at the Shanghai Artificial Intelligence Laboratory to determine whether AI systems can truly perform the complex tasks they claim to handle. Qiushi Engine, launched recently, is described as a large language model-based agent capable of end-to-end autonomous scientific discovery in real-world settings.

Bias read (Center): The article reports on advancements in AI technology without taking a stance on political issues. It focuses on technical achievements and does not involve political figures, policies, or contentious topics.

Keep the news honest.

ObjectiveNews is reader-funded and ad-free — we show you the bias instead of hiding it. Support independent journalism for €5/month.

Become a Supporter

Related stories