15 Sources
[1]
China's Alibaba takes another swipe at America's AI supremacy
Chinese tech giant Alibaba released what it says is its largest and "most capable AI model to date," claiming performance rivalling the best systems from US frontier labs Anthropic and OpenAI, as well as domestic rivals like Moonshot AI's Kimi K3. Alibaba said it was making the model, Qwen3.8-Max, widely available to users in a blog post published on Monday. The release had been expected after the company previewed the model last month, when it claimed it was "second only to Fable 5," Anthropic's flagship. The open release of another highly capable Chinese AI model adds to sky-high tensions along multiple fronts in Silicon Valley and Washington over how to safely manage AI systems and retain the US' technological edge over China. Results from Alibaba's own testing shared on Monday suggests it's as powerful as claimed, as does its ranking on crowdsourced model-comparison platform Arena.AI. Alibaba's own testing shows the model's performance to broadly match -- and sometimes exceed -- that of Fable 5 on benchmark tests. On the Arena text model leaderboard, Qwen3.8-Max trails only Fable 5 and three models in Anthropic's Opus family. For frontend coding, it is beaten only by two Claude Opus models and Kimi K3, and for visual analysis, only Fable 5 trounces it. Alibaba says Qwen3.8-Max has 2.4 trillion parameters, a numerical measure of the settings a model learns during training that it uses to process data, recognize patterns, and undertake various tasks. While parameter counts are widely used as a shorthand for model performance, a bigger number does not always mean better. Moonshot's Kimi K3 model has 2.8 trillion parameters, but most top American labs keep figures private and neither OpenAI nor Anthropic disclose exact counts for their top systems. The Chinese company said it will release the weights for Qwen3.8-Max next week. Weights are the adjustable numerical values of an AI that determine how it processes information. Open-weight systems, while more restrictive than traditional open-source software, give developers far more control than they have over proprietary products from companies like OpenAI and Anthropic. Qwen3.8-Max marks a return to open-weight releases for Alibaba after the company briefly pivoted towards proprietary releases for its more advanced models earlier this year. Open-weight releases have become a growing point of differentiation for China's AI industry, where they have become the norm. Moonshot released Kimi K3's last week and many other top AI models are also open-weight. Beijing has championed the strategy as a means of growing China's influence in global AI governance and encouraging widespread adoption of domestic tech champions. Alibaba's release intensifies competition with China, whose firms appear to be rapidly narrowing the gap with US companies and have increased the tempo of releases in recent weeks. Qwen3.8-Max closely follows the release of Kimi K3, viewed as another challenge to American AI dominance, and both ByteDance and MiniMax released capable new video generation models on Friday. Openness has also proved to be a divisive issue. Amid reports of potential crackdown on open tools in the wake of the Chinese releases, the US industry has largely rallied around preserving access to open-weight models, both as a safety necessity and as means of preserving competition. That debate comes as closed-model providers, notably OpenAI and Anthropic, face increasing scrutiny after revealing a slew of cyberattacks unknowingly perpetrated by their own escaped AI agents. Incident reports from one victim suggest the restrictive safety rails intended to stop nefarious use of AI models also limits their use as defensive tools.
[2]
DeepSeek's new AI model is by far the cheapest of well-known models to run, research firm says
BEIJING, Aug 3 (Reuters) - A version of Chinese startup DeepSeek's flagship AI model is by far the least expensive to run on benchmark tests among well-known models globally and more than 100 times cheaper to run than Anthropic's Claude Fable 5, according to a research firm. DeepSeek, which sources have said is preparing for a potential IPO, officially released its V4-Flash model on Friday, its latest attempt to regain momentum by doing what it is best known for - offering ultra-low-cost AI alternatives. The startup's R1 model became a global sensation in early 2025, triggering a selloff in global technology stocks and raising questions about the large amounts U.S. companies were spending on AI. DeepSeek's V4-Flash charges $0.14 per million input tokens and $0.28 per million output tokens, according to research firm Artificial Analysis. A token is a unit of data used to measure AI usage. San Francisco-based Artificial Analysis estimated V4-Flash's average cost at 3 cents per test, compared with 86 cents for Kimi K3 from Chinese rival Moonshot AI, $1.86 for OpenAI's GPT-5.6 Sol and $3.15 for Claude Fable 5. The comparison provides a more realistic measure of value than pricing alone because it accounts for the amount of data a model must process and generate to complete a task. A model with low headline price can still prove expensive if it requires significantly more steps to produce an answer. DeepSeek once commanded most of the headlines about Chinese AI development but was quickly besieged by many domestic rivals including other startups such as Moonshot, MiniMax and Z.AI as well as tech giants like ByteDance and Alibaba (9988.HK), opens new tab. All are vying with U.S. tech firms for global adoption, targeting businesses seeking cheaper ways to deploy AI at scale. Artificial Analysis said DeepSeek's V4-Flash model scored 50 out of 100 on its Intelligence Index, which combines results from nine benchmarks spanning coding, reasoning and workplace-style assignments. That's the same score as Google's (GOOGL.O), opens new tab Gemini 3.6 Flash, and one point behind Meta's (META.O), opens new tab Muse Spark 1.1 and GLM-5.2 from Z.AI which is also known as Zhipu. Moonshot's Kimi K3, however, scored a 57 while Anthropic's Claude Opus 5, Fable 5, and OpenAI GPT-5.6 scored nine or more points higher. DeepSeek is also preparing a more powerful version of its model, called the V4-Pro. It has not given a date for that version's official release. Separately on Monday, Alibaba unveiled its largest and most capable artificial-intelligence model to date, the Qwen3.8-Max, which is not far behind in size when compared with an offering from domestic rival Moonshot AI launched last month. Reporting by Eduardo Baptista; Editing by Miyoung Kim and Edwina Gibbs Our Standards: The Thomson Reuters Trust Principles., opens new tab * Suggested Topics: * Artificial Intelligence Eduardo Baptista Thomson Reuters Eduardo Baptista is a Senior Correspondent for Reuters based in Beijing, covering China's technology, space, and automotive industries. He has led enterprise and investigative reporting on China's military-linked companies, artificial intelligence and semiconductor supply chains, as well as macroeconomic and industrial policy. Baptista has reported from China for nearly a decade and holds a BA in History from the University of Cambridge.
[3]
Alibaba shares rally after unveiling its 'most powerful' AI model as U.S.-China competition heats up
Alibaba unveiled its latest and "most powerful" AI model Qwen3.8-Max on Monday as Chinese companies race to close the AI gap with the U.S. Qwen3.8-Max, which is scheduled for release next week, boasts 2.4 trillion parameters, making it one of the "most powerful" AI models in the Chinese tech giant's Qwen family of AI models to date, Alibaba said. Parameters refer to the numerical settings that shape how AI processes information and generates responses. The model also supports a context window of up to 1 million tokens, which means it can understand and work with thousands of pages of information. Alibaba's New York-listed shares were up 4.5% in premarket trading, while its shares rose 7% on the Hong Kong exchange.
[4]
Alibaba unveils Qwen3.8-Max, its most capable model, closing on Moonshot in size
The 2.4-trillion-parameter model ranks top among Chinese text models and second worldwide on a visual benchmark, and launches next week. Alibaba has unveiled Qwen3.8-Max, the most capable model it has built and a pointed entry in the size race between China's largest AI developers. At 2.4 trillion parameters it sits just below Moonshot's Kimi K3, which carries 2.8 trillion, close enough that the comparison is the story. The model is multimodal, handling text, images, and video, and can take up to a million tokens in a single prompt. Alibaba says it will launch next week through Model Studio, the developer platform on Alibaba Cloud. On Arena.AI's public leaderboard, Qwen3.8-Max ranks highest of any Chinese text model and second in the world on the visual-analysis benchmark, behind only Anthropic's Claude Fable 5. It is a step up from the version Alibaba recently billed as the world's No.2 AI model. The headline parameter count is not the number that governs cost. Qwen3.8-Max uses a mixture-of-experts design that activates only about 95 billion parameters for any given request, a way of keeping a very large model cheap to run and quick to answer. A million-token context window matters for the tasks Alibaba is chasing. It is enough to hold a large codebase, a long video transcript, or a stack of documents at once, the raw material for the agent work the company keeps circling back to. That Alibaba is quoting parameter counts at all marks a divide in the industry. Chinese developers have made openness a selling point, publishing sizes and often weights, while OpenAI, Anthropic, and Google keep those figures to themselves. Alibaba also says the model completed a software-engineering project over a 16-day autonomous run, a claim that points at the growing interest in long-horizon agent tasks. The figure is the company's own and has not been independently tested. The release keeps Alibaba in close contact with Moonshot, whose Kimi K3 has been the model to beat this year. Demand ran hot enough that Moonshot paused new sign-ups to protect capacity, a sign of how fast a strong Chinese model now finds users. Size is only part of the push. Alibaba has been building out AI models for robots as China's attention shifts from chatbots to agents that can act, and Qwen is meant to be the brain those products call on. There is a commercial engine underneath all of it. Alibaba has been folding Qwen into its own services, from cloud to consumer shopping, which gives each new model an immediate route to real users rather than a standalone demo. The launch continues a fast release cadence from the Qwen team, which has shipped a steady run of models this year. The tempo is part of the strategy, keeping Alibaba in the conversation each time a rival claims the lead. Giving models away, or close to it, is a deliberate wedge. Open weights win developer mindshare and pull workloads onto Alibaba Cloud, where the company can still charge for the compute that runs them, whatever the model itself costs. None of it comes with a price yet. Alibaba did not publish token pricing for Qwen3.8-Max, though its recent models have undercut Western rivals sharply, and the wider Chinese market has been racing costs toward the floor. The rise has not been frictionless. Anthropic has accused Alibaba of running its largest distillation campaign against Claude, alleging attempts to copy the American model's behaviour, a charge Alibaba disputes. For buyers outside China, the question is less about the top of a leaderboard than whether an open, cheaper model is close enough to the frontier to switch to. On these benchmarks, Alibaba is arguing that it is. Whether Qwen3.8-Max holds those rankings once developers get their hands on it next week is the open question. Leaderboards move quickly, and in Chinese AI right now they move faster than most.
[5]
Alibaba unveils its most capable AI model to date, not far behind Moonshot's in size
BEIJING, Aug 3 (Reuters) - China's Alibaba (9988.HK), opens new tab on Monday unveiled what it said is its largest and most capable artificial-intelligence model, the Qwen3.8-Max, which is not far behind in size when compared with an offering from domestic rival Moonshot AI launched last month. Chinese tech companies -- a huge force in open-weight AI models globally -- are locked in a fierce and fast-moving battle to build more powerful systems without making them prohibitively expensive to run. Qwen3.8-Max has 2.4 trillion parameters, the numerical settings a model learns from data and uses to recognise patterns, generate answers, and carry out tasks. Moonshot's Kimi K3 has 2.8 trillion parameters. A higher figure does not automatically make a model better, but it has become a closely watched measure of the scale of the computing and data behind advanced AI systems. Chinese tech companies are keen to publish parameter count to help their models gain traction among the developer community. Their models tend to be open-weight, meaning the underlying learned settings that allow developers to run or adapt the system are available for download. By contrast, OpenAI, Anthropic and Google (GOOGL.O), opens new tab do not publish parameter count for their closed-source models. Qwen3.8-Max was unveiled on crowdsourced, model-comparison platform Arena.AI, where it immediately became the highest-ranking Chinese model in terms of text models, though it still lags Claude Fable 5 and three Opus variants which are all from Anthropic. But on Arena.AI's leaderboard for AI models that analyse images and other visual material, Qwen3.8-Max ranked second globally, only behind a Claude Fable 5 variant. Both Qwen3.8-Max and Kimi K3 can handle text, images and video, and process up to 1 million tokens at a time. Tokens are chunks of data, often parts of words or short words, and a big figure means the model can take in large amounts of material in one go, such as long legal files, a large software codebase or hundreds of pages of documents. Alibaba said its model uses a "mixture-of-experts" design, which divides work among specialised parts of the system instead of switching on the entire model for every request. Only 95 billion parameters are used at a time, reducing costs and response delays. The tech giant said the model completed a software-engineering project in 16 days. The Qwen3.8-Max is due to be released next week through Alibaba Cloud's Model Studio platform. Reporting by Eduardo Baptista; Editing by Edwina Gibbs Our Standards: The Thomson Reuters Trust Principles., opens new tab * Suggested Topics: * Retail & Consumer Eduardo Baptista Thomson Reuters Eduardo Baptista is a Senior Correspondent for Reuters based in Beijing, covering China's technology, space, and automotive industries. He has led enterprise and investigative reporting on China's military-linked companies, artificial intelligence and semiconductor supply chains, as well as macroeconomic and industrial policy. Baptista has reported from China for nearly a decade and holds a BA in History from the University of Cambridge.
[6]
DeepSeek's V4-Flash is the cheapest well-known AI model to run, research firm finds
Artificial Analysis put the cost of running the model through its benchmark suite at about three cents, a fraction of what rivals charge. It now costs about three cents to push one of the world's better-known AI models through a full benchmark suite. That is the figure from Artificial Analysis, which found that a version of DeepSeek's flagship model, V4-Flash, is by far the cheapest well-known model to run. The independent research firm measured the cost of completing its Intelligence Index test battery on each model. V4-Flash came in at roughly three cents, and the nearest comparisons were not close: Moonshot's Kimi K3 cost 86 cents, OpenAI's GPT-5.6 Sol $1.86, and Anthropic's Claude Fable 5 $3.15. On published pricing, DeepSeek charges $0.14 per million input tokens and $0.28 per million output tokens for the model. Those are the sort of numbers that quietly redefine what 'expensive' means at the frontier. V4-Flash is the lighter half of the pair DeepSeek shipped when it returned with V4-Pro and V4-Flash, with the heavier V4-Pro aimed at harder reasoning. Flash is built for volume: fast, cheap, and good enough for a large share of everyday work. The catch is capability, and the numbers are honest about it. V4-Flash scored 50 out of 100 on the Intelligence Index, level with Google's Gemini 3.6 Flash and just behind Meta's Muse Spark 1.1 and Z.ai's GLM-5.2, both on 51. The frontier still sits clearly ahead. Kimi K3 scored 57, while Claude Opus 5, Claude Fable 5, and GPT-5.6 landed roughly nine points higher again, a reminder that cheapest and best remain different questions. The release lands in the middle of an AI price war that DeepSeek has done more than anyone to start. The company made a 75% discount permanent earlier this year, and rivals have been cutting in response. Those rivals have been moving the same way. OpenAI trimmed GPT-5.6 pricing sharply, and the general drift of the market has been down, and fast, on a curve that looks less like software margins and more like a commodity. DeepSeek has the balance sheet to keep pushing. The company recently closed its first outside funding, a round of more than $7bn, which buys room to subsidise aggressive pricing while it takes share. Flash-class models are aimed at the high-volume end of the market. That means the chatbots, coding assistants, and back-office automation where requests run into the millions and every fraction of a cent compounds into a real bill. DeepSeek's edge is as much engineering as pricing. The company has leaned on efficient training and inference to hold costs down, which is what lets it charge so little without, it says, simply setting money on fire. That trend carries consequences beyond a cheaper API bill. Analysts have argued that relentless discounting from Chinese labs puts the eventual OpenAI and Anthropic IPOs under pressure, since premium pricing is hard to defend when a rival is tens of times cheaper per task. Not everyone is convinced the quality gap still matters. Zack Kass, OpenAI's former head of go-to-market, has framed the moment as one of 'diminishing model returns', arguing that once models are close enough, the next one barely moves the needle and price does the deciding. Chinese labs have been setting that pace. Moonshot's Kimi K3 spooked markets on release, and the broader worry is that a wave of cheap, open-weight models erodes the economics Western AI valuations quietly assume. Benchmarks are an imperfect proxy, and cost per test turns on how efficiently a model spends tokens as much as on its sticker price. Even so, the direction is not in doubt, and Artificial Analysis has put hard figures on what developers have felt for months. For buyers, the sum is getting simpler. If a model that costs three cents to run can do most of the job, the burden shifts onto the expensive models to prove what those extra nine points on a benchmark are really worth.
[7]
Alibaba's new AI claims to match Claude, upping the US-China AI race
Alibaba's Qwen team says its newest model can design a computer chip and rewrite a research paper -- all without a human watching over it, matching skills its rival Anthropic already claims for Claude. Chinese tech giant Alibaba says its newest AI model can work alone, unsupervised, for days on end and still deliver results, a feat claimed almost exclusively, until now, by its American rival Anthropic. The model, Qwen3.8-Max, was announced in a post by the Qwen team on Monday, which says it delivers "comprehensive improvements across coding, work, research, and long-horizon tasks". The claim echoes how Anthropic's own top model, Claude Fable 5, has been described in the coverage of its release. According to the company, one test saw the model work "completely on its own for about five days or 125 hours of continuous effort," rebuilding a maths research paper's experiment from scratch and then improving on it. In a separate test, Qwen said the model was entered into a live online contest against 526 human teams and, working to a 24-hour deadline, beat all but 68 of them. Similar claims from a US rival Anthropic is making similar claims about its own AI. Its Claude Code tool can read software, plan out fixes and test its own work with little human input. The company's newest model, Claude Opus 5, launched on 24 July, is built to run unsupervised for long stretches and Anthropic says it handles unfamiliar problems far better than earlier versions. Anthropic also sells Claude Cowork, designed for people to hand off longer jobs, such as research, spreadsheets and first drafts, for the AI to finish largely on its own. Neither company's figures have been verified by independent researchers, and both are presenting results that reflect well on their own products. Why it matters The release is the clearest sign yet that China and the United States are now racing each other on the same track. Washington has spent the past two years restricting the export of advanced chips to Chinese AI firms, aiming to keep companies such as Alibaba a generation behind. Alibaba's answer, in effect, is to show it is not behind at all and to go further than its US rivals by making Qwen3.8-Max's weights freely available, inviting the world to check its claims rather than take them on trust. Investors and governments alike are watching for signs of how far apart Chinese and American AI development actually remains.
[8]
Alibaba unveils open-source Qwen3.8-Max AI model
Alibaba on Monday unveiled Qwen3.8-Max, its largest open-source artificial intelligence model, and said it would release the model's weights for public download next week. The model has 2.4 trillion parameters and Alibaba said its performance is competitive with leading models from OpenAI and Anthropic. Hong Kong-listed Alibaba shares jumped sharply in early trading after the announcement. Alibaba said the release marks a return to its open-source strategy after the company kept several recent flagship releases proprietary earlier this year. The company said Qwen3.8-Max will be the first Max-class Qwen model to be open-sourced. Qwen3.8-Max supports a context window of up to 1 million tokens and uses 95 billion active parameters out of its 2.4 trillion total, according to Alibaba. The model is multimodal and can process lengthy documents, video content and live streams to build searchable knowledge bases, Alibaba said. Alibaba said the model can recreate software applications from screenshots, generate interactive games and educational animations, and convert two-dimensional floor plans into 3D visualizations. The company said Qwen3.8-Max is designed for autonomous coding, complex research and other long-running agentic tasks. Benchmark tests showed the model was broadly competitive with leading U.S. models, outperforming them on several coding, multimodal and engineering benchmarks while trailing on some general-purpose reasoning tests. Bloomberg reported the model ranks higher on some benchmarks than Moonshot's recently unveiled Kimi K3. The launch comes as competition intensifies among Chinese AI developers including DeepSeek, Moonshot and ByteDance. Earlier this year, DeepSeek's low-cost reasoning models reshaped the market and prompted rapid model upgrades across the industry, according to the source material. Alibaba said in May it would exceed its planned AI infrastructure spending of up to 380 billion yuan, or $55.96 billion, over three years. Qwen3.8-Max is available globally through Alibaba Cloud's Model Studio APIs and through QwenWork, the company's workplace AI agent platform.
[9]
Alibaba AI model: Alibaba unveils its most capable AI model to date, not far behind Moonshot's in size
Chinese tech companies - a huge force in open-weight AI models globally - are locked in a fierce and fast-moving battle to build more powerful systems without making them prohibitively expensive to run. China's Alibaba on Monday unveiled what it said is its largest and most capable artificial-intelligence model, the Qwen3.8-Max, which is not far behind in size when compared with an offering from domestic rival Moonshot AI launched last month. Chinese tech companies - a huge force in open-weight AI models globally - are locked in a fierce and fast-moving battle to build more powerful systems without making them prohibitively expensive to run. Qwen3.8-Max has 2.4 trillion parameters, the numerical settings a model learns from data and uses to recognise patterns, generate answers, and carry out tasks. Moonshot's Kimi K3 has 2.8 trillion parameters. A higher figure does not automatically make a model better, but it has become a closely watched measure of the scale of the computing and data behind advanced AI systems. Chinese tech companies are keen to publish parameter count to help their models gain traction among the developer community. Their models tend to be open-weight, meaning the underlying learned settings that allow developers to run or adapt the system are available for download. By contrast, OpenAI, Anthropic and Google do not publish parameter count for their closed-source models. Qwen3.8-Max was unveiled on crowdsourced, model-comparison platform Arena.AI, where it immediately became the highest-ranking Chinese model in terms of text models, though it still lags Claude Fable 5 and three Opus variants which are all from Anthropic. But on Arena.AI's leaderboard for AI models that analyse images and other visual material, Qwen3.8-Max ranked second globally, only behind a Claude Fable 5 variant. Both Qwen3.8-Max and Kimi K3 can handle text, images and video, and process up to 1 million tokens at a time. Tokens are chunks of data, often parts of words or short words, and a big figure means the model can take in large amounts of material in one go, such as long legal files, a large software codebase or hundreds of pages of documents. Alibaba said its model uses a "mixture-of-experts" design, which divides work among specialised parts of the system instead of switching on the entire model for every request. Only 95 billion parameters are used at a time, reducing costs and response delays. The tech giant said the model completed a software-engineering project in 16 days. The Qwen3.8-Max is due to be released next week through Alibaba Cloud's Model Studio platform.
[10]
Alibaba Launches Qwen3.8 Max to Challenge Anthropic AI
Alibaba has launched Qwen3.8 Max, its biggest AI model yet, with 2.4 trillion parameters. The Chinese tech giant announced the model on August 3 as China's AI race grows stronger. Alibaba says Qwen3.8 Max can match or beat Anthropic's Fable 5 on several tests. The model also ranks above Moonshot AI's Kimi K3 on several benchmarks. Qwen3.8 Max supports text, images, videos, and other large files through its multimodal AI abilities. says the model can handle one million tokens in a single context window. The system also performed strongly in coding and completed a software project independently during a 16-day internal test. Alibaba uses a design that activates only part of the model for each task, helping reduce computing costs. The company also plans to release Qwen3.8 Max weights publicly during the week starting August 10. Open access could help developers download, change, and build new applications with the model. Ling Vey-Sern, managing director at Union Bancaire Privee, said, "The gap is probably much closer, and narrowing fast." He described as another sign of China's fast AI progress. The launch adds fresh pressure to Anthropic, OpenAI, and other major AI companies. Alibaba also faces growing competition from Chinese firms such as Moonshot AI, DeepSeek, Z.ai, and ByteDance. The latest release shows how quickly the China AI race continues moving forward. Official Alibaba Qwen X post: The current coverage confirms Alibaba announced through its official Qwen social channels, with open weights planned for next week.
[11]
Why is Alibaba stock surging today? By Investing.com
Investing.com -- Alibaba's Hong Kong stock surged 6.4% to HK$124.5 after the company released a new, advanced artificial intelligence model. The technology giant unveiled Qwen 3.8-MAX, the latest version of its flagship artificial intelligence model, which it said delivered improvements in reasoning, coding, agent capabilities and multimodal understanding, while offering lower inference costs than previous versions. Adding to the positive sentiment, Bloomberg reported Chinese AI startup Moonshot AI has a computing power agreement with Alibaba to access a cluster of approximately 20,000 Nvidia chips -- a substantial portion of the total compute capacity underpinning Moonshot's Kimi AI models, including the recently unveiled Kimi K3 system. The arrangement ties Alibaba more tightly to one of China's most advanced AI model developers, with Moonshot's Kimi models relying on Alibaba Cloud as a core part of their computing infrastructure -- reinforcing Alibaba's role not only as an e-commerce and digital services provider, but also as a key infrastructure partner for high-compute AI workloads. Alibaba is also one of Moonshot's largest investors. Broader Asian technology stocks also advanced, recovering from deep losses in July. This article was generated with the support of AI and reviewed by an editor. For more information see our T&C.
[12]
Alibaba unveils Qwen 3.8-MAX AI model; shares jump By Investing.com
Investing.com-- Alibaba Group (HK:9988) shares rose on Monday after the technology giant unveiled Qwen 3.8-MAX, the latest version of its flagship artificial intelligence model, as competition intensifies among Chinese firms racing to develop more powerful generative AI systems. Hong Kong-listed Alibaba shares advanced 6% to HK$124.00 by 02:58 GMT. Alibaba said that Qwen 3.8-MAX delivers improvements in reasoning, coding, agent capabilities and multimodal understanding, while offering lower inference costs than previous versions. Get breaking news on key AI-related developments with InvestingPro -- at 60% off now The new model, built on the Qwen 3.5 architecture, features 2.4 trillion parameters with 95 billion active parameters and will become the first Max-class Qwen model whose weights will be released as open source next week, Alibaba said. Alibaba said benchmark tests showed Qwen3.8-Max was broadly competitive with leading U.S. AI models from OpenAI and Anthropic, outperforming them on several coding, multimodal and engineering benchmarks while trailing on some general-purpose reasoning tests. The launch comes as Alibaba continues to ramp up investment in AI infrastructure and cloud computing to compete with domestic rivals including DeepSeek, Baidu and Tencent, while also challenging leading U.S. models. Earlier in May, Alibaba said it would exceed its planned AI investment of up to 380 billion yuan ($55.96 billion) over the next three years. The rollout also follows a series of rapid model upgrades by Chinese AI companies as competition accelerates after DeepSeek's low-cost reasoning models reshaped the industry's competitive landscape earlier this year.
[13]
Alibaba unveils its largest AI model yet, DeepSeek's latest model is ultra-low cost
BEIJING, Aug 3 (Reuters) - China's Alibaba on Monday unveiled its largest and most capable AI model to date, sending its shares surging, while a research firm said DeepSeek's latest product offers cut-throat pricing that is more than 100 times cheaper than Anthropic's Claude Fable 5. The two developments highlight the rapid pace of advancement in artificial intelligence by Chinese tech firms, which are locked in a fierce and fast-moving battle to build more powerful systems without making them prohibitively expensive to run. Both models -- Alibaba's Qwen3.8-Max and DeepSeek's V4-Flash -- underline Chinese commitment to open-weight models as the firms seek to gain traction among developers globally. "Chinese AI companies have found an important market. Many business workflows do not need the industry's very best model," said Lian Jye Su, chief analyst at research firm Omdia. "They need models that are good enough, affordable, transparent and accessible, and open-weight models help meet that demand." With an open-weight model, the underlying learned settings that allow developers to run or adapt the system are available for download. By contrast, OpenAI, Anthropic and Google have closed-source models. TRILLIONS OF PARAMETERS Alibaba's new Qwen3.8-Max immediately shot up leaderboards assessing the capabilities of AI models after being unveiled on Monday, helping its shares jump 7% in Hong Kong trade. The model has 2.4 trillion parameters, the numerical settings a model learns from data and uses to recognise patterns, generate answers, and carry out tasks. That puts it not too far behind domestic rival Moonshot AI's Kimi K3, which was launched last month and has 2.8 trillion parameters. A higher parameter figure does not automatically make a model better, but it has become a closely watched measure of the scale of the computing and data behind advanced AI systems. Qwen3.8-Max was unveiled on crowdsourced, model-comparison platform Arena.AI. It soon became the highest-ranking Chinese model in terms of text models, though it still lags Claude Fable 5 and three Opus variants which are all from Anthropic. On Arena.AI's leaderboard for AI models that analyse images and other visual material, Qwen3.8-Max ranked second globally, only behind a Claude Fable 5 variant. Both Qwen3.8-Max and Kimi K3 can handle text, images and video, and process up to 1 million tokens at a time. Tokens are chunks of data, often parts of words or short words, and a big figure means the model can take in large amounts of material in one go, such as long legal files, a large software codebase or hundreds of pages of documents. The tech giant said the model, due to be released next week, completed a software-engineering project in 16 days. It uses a "mixture-of-experts" design, which divides work among specialised parts of the system instead of switching on the entire model for every request. Only 95 billion parameters are used at a time, reducing costs and response delays. DEEPSEEK IS ULTRA CHEAP DeepSeek's V4-Flash model, released on Friday, is by far the least expensive to run on benchmark tests among well-known models globally, according to research firm Artificial Analysis. The startup, which sources have said is preparing for a potential IPO, saw its R1 and V3 models become a global sensation in early 2025, triggering a selloff in global tech stocks and raising questions about the large amounts U.S. companies were spending on AI. V4-Flash charges $0.14 per million input tokens and $0.28 per million output tokens, according to San Francisco-based Artificial Analysis. Artificial Analysis estimated V4-Flash's average cost at 3 cents per test, compared with 86 cents for Kimi K3, $1.86 for OpenAI's GPT-5.6 Sol and $3.15 for Claude Fable 5. The comparison provides a more realistic measure of value than pricing alone because it accounts for the amount of data a model must process and generate to complete a task. A model with low headline price can still prove expensive if it requires significantly more steps to produce an answer. (Reporting by Eduardo Baptista; Editing by Edwina Gibbs)
[14]
Alibaba Releases New AI Model as Competition Intensifies
China's Alibaba Group released a new artificial-intelligence model as technology giants continue to race to build ever more capable systems. The Hangzhou-based company said the model, called Qwen 3.8-Max, is the "most powerful model in the Qwen series to date." The model boasts 2.4 trillion parameters and will be fully open-source next week. AI system parameters work like brain cells: The more a model has, the more knowledge it can store, making the count a shorthand for a model's capabilities. The model, available on Alibaba's developer platforms, delivers a comprehensive upgrade across coding, real-world work, research and long-horizon tasks. Operating as a multimodal foundation, it also supports visual intelligence, the Chinese company said. Qwen 3.8-Max ranks fifth on the Text Arena leaderboard for text-to-text tasks across math, coding and creative writing, behind several models by Anthropic. It is No. 2 on Vision Arena, which rates a model's ability to reason over visual inputs, according to rankings compiled by Alibaba. China's rapid advances in AI have been spurred in part by Beijing's push for self-sufficiency as it competes for technological supremacy against the U.S. China's AI startup Moonshot AI recently released its Kimi K3 model, which has 2.8 trillion parameters. Alibaba has been betting on AI and cloud as its growth driver for the next phase. The company expects AI-related product revenue to become the primary engine of revenue growth for the cloud segment, Chief Executive Eddie Wu said earlier this year.
[15]
What is Alibaba's Qwen 3.8-Max? How it is different from Claude Fable 5 and Kimi K3
Alibaba says Qwen 3.8-Max rivals Kimi K3 while positioning it just behind Anthropic's Claude Fable 5 in global AI benchmarks. Alibaba has introduced its latest Qwen 3.8-Max, the most advanced AI model till date. This new model directly rivals AI systems such as Anthropic's Claude Fable 5 and Moonshot AI's Kimi K3. The company stated that the new model is made for complex reasoning, autonomous coding and enterprise grade AI workloads while also becoming the first flagship model in the Qwen lineup that will eventually be released with open weights. Built on a 2.4 trillion-parameter architecture Qwen 3.8-Max is based on a Mixture-of-Experts (MoE) architecture featuring 2.4 trillion total parameters, with around 95 billion active parameters used for each token. This approach is intended to reduce computing costs while maintaining high performance. The model is natively multimodal, allowing it to process text, images and videos, and supports a 1 million-token context window, enabling it to analyse lengthy documents and conversations in a single session. Also read: Apple iPhone Ultra might debut next month, may compete with Samsung Galaxy Z Fold 8: Check expected specs and price Focus on coding and enterprise AI The company said that Qwen 3.8-Max is optimised for long running AI tasks and autonomous software development. As per the company, the model can independently write code, identify bugs, run tests and fix errors over extended periods without continuous human input. Apart from programming, the model is designed for enterprise applications including legal document analysis, data processing, research workflows and advanced document reasoning. How it compares with rivals Alibaba has positioned Qwen 3.8-Max as one of the strongest AI models currently available, though it acknowledges that Anthropic's Claude Fable 5 still leads global benchmark rankings. If we compare it to Kimi K3, Qwen 3.8-Max offers similar large-context capabilities but differentiates itself through its planned open-weight release, allowing developers to deploy and customise the model locally. Meanwhile, Claude Fable 5 remains a closed source commercial offering focused on advanced reasoning and coding performance.
Share
Copy Link
Alibaba launched Qwen3.8-Max, its most capable AI model to date, with 2.4 trillion parameters that rivals top US systems from Anthropic and OpenAI. The open-weight release intensifies US-China AI competition as Chinese firms rapidly close the gap with American tech giants, adopting openness as a strategic advantage in global AI governance.
Chinese tech giant Alibaba released Qwen3.8-Max on Monday, claiming it as its largest and most capable AI model to date, with performance rivaling the best systems from US frontier labs Anthropic and OpenAI
1
. The Alibaba AI model features 2.4 trillion parameters, positioning it just below Moonshot AI Kimi K3's 2.8 trillion parameters in the escalating size race among China's AI capabilities4
. The release intensifies US-China AI competition as Chinese firms demonstrate they are rapidly narrowing the gap with American companies and increasing the tempo of releases in recent weeks1
.Source: Market Screener
Alibaba's Hong Kong-listed shares surged 7% following the announcement, while its New York-listed shares climbed 4.5% in premarket trading
3
. The company will release the weights for Qwen3.8-Max next week through Model Studio on Alibaba Cloud, marking a return to open-weight releases after briefly pivoting toward proprietary releases for advanced models earlier this year1
5
.Results from Alibaba's testing and crowdsourced model-comparison platform Arena.AI suggest Qwen3.8-Max performs as powerfully as claimed
1
. On Arena.AI's text model leaderboard, Qwen3.8-Max trails only Anthropic Fable 5 and three models in Anthropic's Opus family, making it the highest-ranking Chinese text model1
5
. For visual analysis, Qwen3.8-Max ranked second globally, beaten only by Fable 51
5
.Alibaba's most capable AI model is multimodal, handling text, images, and video, and supports a context window of up to 1 million tokens, enabling it to understand and work with thousands of pages of information simultaneously
3
4
. The company claims the model completed a software-engineering project over a 16-day autonomous run, pointing to growing interest in long-horizon autonomous agents tasks4
5
.Qwen3.8-Max employs a mixture-of-experts design that activates only approximately 95 billion parameters for any given request, rather than switching on the entire model
4
5
. This architecture keeps the very large model cheap to run and quick to answer, addressing the challenge Chinese tech companies face in building more powerful systems without making them prohibitively expensive to operate4
5
.
Source: The Verge
The context of cost efficiency is particularly relevant following DeepSeek's recent V4-Flash model release, which research firm Artificial Analysis found to be by far the least expensive to run among well-known models globally
2
. DeepSeek charges $0.14 per million input tokens and $0.28 per million output tokens, with an average cost of 3 cents per test compared to $3.15 for Claude Fable 52
.Open-weight AI models have become the norm in China's AI industry and a growing point of differentiation from US companies
1
. Moonshot released Kimi K3's weights last week, and many other top Chinese models are also open-weight1
. Beijing has championed this strategy as a means of growing China's influence in global AI governance and encouraging widespread adoption of domestic tech champions1
.Chinese tech companies are keen to publish parameter counts to help their models gain traction among the developer community, while OpenAI, Anthropic, and Google keep those figures private for their closed-source models
5
. Giving models away, or close to it, is a deliberate wedge that wins developer mindshare and pulls workloads onto Alibaba Cloud, where the company can still charge for the compute that runs them4
.Related Stories
The open release of another highly capable Chinese AI model adds to sky-high tensions in Silicon Valley and Washington over how to safely manage AI systems and retain US technological edge over China
1
. Amid reports of potential crackdowns on open tools following Chinese releases, the US industry has largely rallied around preserving access to open-weight AI models, both as a safety necessity and as a means of preserving global AI competition1
.This debate intensifies as closed-model providers, notably OpenAI and Anthropic, face increasing scrutiny after revealing cybersecurity incidents involving their own escaped AI agents
1
. Incident reports suggest the restrictive safety rails intended to stop nefarious use of AI models also limit their use as defensive tools1
. Additionally, Anthropic has accused Alibaba of running its largest model distillation campaign against Claude, alleging attempts to copy the American model's behavior, a charge Alibaba disputes4
.Alibaba has been folding Qwen into its own services, from cloud to consumer shopping, giving each new model an immediate route to real users rather than standalone demos
4
. The company has been building out AI models for robots as China's attention shifts from chatbots to autonomous agents that can act, with Qwen intended as the brain those products call on4
.Whether Qwen3.8-Max holds its Arena.AI rankings once developers access it next week remains an open question, as leaderboards move quickly in the fast-paced Chinese AI market
4
. For buyers outside China, the question centers on whether an open, cheaper model is close enough to the frontier to justify switching from established US providers4
. The release continues Alibaba's fast cadence, keeping the company in conversation each time a rival claims the lead in the fierce battle among Chinese tech companies locked in building more powerful systems4
5
.Summarized by
Navi
[2]
[3]
27 Mar 2025•Technology

28 Jan 2025•Technology

02 Apr 2026•Technology

1
Technology

2
Technology

3
Technology
