Alibaba’s Qwen3 unseats DeepSeek’s R1, now leads open-source AI models
By: cryptosheadlines|2025/05/07 01:30:02
0
Share
Airdrop Is Live CaryptosHeadlines Media Has Launched Its Native Token CHT. Airdrop Is Live For Everyone, Claim Instant 5000 CHT Tokens Worth Of $50 USDT. Join the Airdrop at the official website, CryptosHeadlinesToken.com Alibaba’s new Qwen3 family of AI models has surpassed DeepSeek’s R1 to become the world’s best open-source model. According to reports, Qwen3 did better than R1 in tests that measure open-source AI models’ abilities in areas like language instruction, math, coding, and data analysis. The Qwen3 family was launched last week by Alibaba’s cloud computing unit. It has eight improved models with between 600 million and 235 billion parameters. In machine learning, parameters are the variables in an AI system while it is being trained.According to LiveBench platform, an independent platform that tests large language models, before these new tests, DeepSeek’s R1 had been the best open-source AI model in the world since it came out in January. But not anymore.Both US and Chinese companies rush to adopt Qwen 3The rise of Qwen3 in the LiveBench rankings shows how quickly AI is developing in China. The Chinese tech industry has grown a lot thanks to open-source tools. The Alibaba open-source method code has allowed other third-party software developers to share the design, fix broken links, or make the program more powerful. However, the overall LiveBench results showed that Qwen3 was not as good as OpenAI’s o3, Google’s Gemini Pro 2.5, and Anthropic’s Claude 3.7, which are the best closed-source AI models in the world. LiveBench says that the o3-mini high, OpenAI’s most popular AI model, was the best in the world overall. Microsoft backs OpenAI.For every 1 million tokens, it takes $10 to run o3. On the other hand, Qwen3 is cheaper to use because it only costs $0.55 per 1 million tokens to run. Because Qwen3 is cheaper and works better, many businesses said they would back Alibaba’s newest AI model as soon as it came out.Huawei Technologies, Moore Threads, Cambricon Technologies, and Hygon Information Technology are all chip companies that have said they will support Qwen3.Cambricon said last Tuesday that it had successfully optimized Qwen3 to run quickly on its graphics processing units. This was done because AI developers in the Philippines wanted chips made in China.Qwen3 is also being used on the cloud computing services of Hyperbolic and Fireworks.ai, two AI infrastructure companies. The American chipmakers Nvidia and Intel have begun to support Qwen3.Many big data centers in China, like those in Beijing, Shanghai, Hangzhou, and the provinces of Hubei, Jilin, and Northwest Shaanxi, have also said they will use Alibaba’s third-generation Qwen AI models. The Supercomputing Network in China has also adopted Qwen3. This network links over 20 data centers in 20 towns across 14 provinces.Anthropic CEO says that DeepSeek was “a bit overblown”At a business event, a co-founder of Anthropic, the company that made the Claude AI models, said that DeepSeek is still “six to eight months behind where US frontier companies are.” He also said that the recent buzz around the Chinese start-up was “perhaps a bit overblown.”DeepSeek got attention around the world in late December 2024 and early January 2025 by sharing two advanced open-source AI models, V3 and R1. These models were made for a small fraction of the cost and computing power that big tech companies usually need for LLM projects.It’s unclear when DeepSeek will release the next generation of its models. The Hangzhou-based company quietly released its 671-billion-parameter Prover-V2 in late April. This was an update to its specialized model for handling math proofs. However, it hasn’t said anything about the progress of its long-awaited R2 reasoning model.KEY Difference Wire helps crypto brands break through and dominate headlines fastSource link
You may also like

BitsLab Deep Production: Nanobot User Security Practice Guide
BitsLab releases AI Agent Security Guidelines: Through a three-pronged strategy of "User Review + Agent Awareness + Script Hard Interception," a zero-trust security defense line is established to prevent prompt injection and sensitive data leakage risks.

What are the common traits of people who founded a $5 Billion+ company before the age of 23?
Trauma, Neurodiversity, Cross-Domain Skills. These characteristics, which may appear as "flaws" on a traditional resume, could instead be the most important signals

Why Hasn't $160 Billion Stripe Gone Public?
The Rise of Private Placements, with Companies like Stripe Rewriting Fundraising Logic.

All the AI News You Need to Know is Here, Lyrical Officially Launches AI News Feed
Users can access key information in real time without switching pages

Bitwise: Why Bitcoin Is Destined to Impact a Million Dollars?
When people talk about Bitcoin, they often overlook one key thing.

Amid Geopolitical Turmoil, Tokenized Gold Emerges Alongside Round-the-Clock On-Chain Markets
When the stock market is closed, the on-chain becomes the sole trading and pricing outlet.

Who Longs War on Polymarket?
The Rug Pull War rages on, with the potential to earn up to 4x gains on your bet

4 AI Trading Strategy Lessons from WEEX Hackathon Finalist
Finalist Bambi shares how AI tools helped turn real trading experience into an automated strategy, why survival-first risk control shaped the system’s design, and how the approach will evolve ahead of WEEX AI Trading Hackathon Season 2.

Hong Kong Crypto Ecosystem 2.0: Stablecoins, RWA, and the New Battleground for Financial Institutions
Hong Kong is no longer just a bystander in the cryptocurrency industry, but may become the core hub of the compliant cryptocurrency market in the Chinese-speaking world and even the entire Asia-Pacific region.

Polymarket Arbitrage Bible: The Real Gap is in the Mathematical Infrastructure
While retail investors are still engaged in simple probability addition, top quantitative teams are systematically harvesting millions of dollars in arbitrage profits on Polymarket using hardcore mathematical infrastructure such as integer programming and Bregman projections.

Crypto Barbarians Jupiter Series: Still Owes the Market an Answer
This entrepreneurial team from Singapore and Malaysia has indeed demonstrated its product execution capabilities to the market over the past three years, but they have also fully arbitraged every regulatory gray area with their business logic.

Bank Card Payment vs. Stablecoin Payment: Which is More Suitable for AI Agents?
Using bank cards to serve humanity and relying on stablecoins for high-frequency micro-trading with machines: Setting aside camp biases, a mixed payment architecture is the ultimate goal of AI entities in business.

Zuck is really out of touch! He actually acquired a dated Lobster-based social platform?
The asset pool Meta can now touch is not on the same level as it was in 2012

Key Market Information Discrepancy on March 11th - A Must-See! | Alpha Morning Report
1. Top News: Iran Reportedly Plants Mines in the Strait of Hormuz, Trump Warns of "Unprecedented" Military Strike
2. Token Unlock: $IO

How to Deal with Trump? Accept this "Art of the Deal Playbook"
The U.S. macro research firm The Kobeissi Letter deconstructs its "10-Step Conflict Pattern": Verbal Pressure, Friday Night Raid, Market Triple Bottom Exploration, Conditional Downgrade... concluding with a single "trade" paper.

AI Computing Power Arms Race Intensifies: This Startup Aims to Mine Bitcoin in Space
The next battleground for AI computing power is extending into space, gradually becoming a new frontier in commercial storytelling.

Claude Code launches the /btw feature, Musk X Money set to launch soon, what's the English community talking about today?
What have foreigners been most interested in over the past 24 hours?

Polymarket Arbitrage Bible: The Real Edge is in the Math Infrastructure
Predictive Market-Making Quantitative Arbitrage Logic.
BitsLab Deep Production: Nanobot User Security Practice Guide
BitsLab releases AI Agent Security Guidelines: Through a three-pronged strategy of "User Review + Agent Awareness + Script Hard Interception," a zero-trust security defense line is established to prevent prompt injection and sensitive data leakage risks.
What are the common traits of people who founded a $5 Billion+ company before the age of 23?
Trauma, Neurodiversity, Cross-Domain Skills. These characteristics, which may appear as "flaws" on a traditional resume, could instead be the most important signals
Why Hasn't $160 Billion Stripe Gone Public?
The Rise of Private Placements, with Companies like Stripe Rewriting Fundraising Logic.
All the AI News You Need to Know is Here, Lyrical Officially Launches AI News Feed
Users can access key information in real time without switching pages
Bitwise: Why Bitcoin Is Destined to Impact a Million Dollars?
When people talk about Bitcoin, they often overlook one key thing.
Amid Geopolitical Turmoil, Tokenized Gold Emerges Alongside Round-the-Clock On-Chain Markets
When the stock market is closed, the on-chain becomes the sole trading and pricing outlet.