tag

Monday, August 17, 2026 | Daily Newspaper published by GPPC Doha, Qatar.

Tag Results for "DeepSeek" (3 articles)

Alibaba's Qwen3.8-Max and DeepSeek's V4-Flash underline Chinese commitment to open-weight models as the firms seek to gain traction among developers globally.
Business

Alibaba unveils its largest AI model yet; DeepSeek's latest model is ultra-low cost

China's Alibaba on Monday unveiled its largest and most capable AI model to date, sending its shares surging, while a research firm said DeepSeek's latest product offers cut-throat pricing that is more than 100 times cheaper than Anthropic's Claude Fable 5.The two developments highlight the rapid pace of advancement in artificial intelligence by Chinese tech firms, which are locked in a fierce and fast-moving battle to build more powerful systems without making them prohibitively expensive to run.Both models — Alibaba's Qwen3.8-Max and DeepSeek's V4-Flash — underline Chinese commitment to open-weight models as the firms seek to gain traction among developers globally."Chinese AI companies have found an important market. Many business workflows do not need the industry's very best model," said Lian Jye Su, chief analyst at research firm Omdia. "They need models that are good enough, affordable, transparent and accessible, and open-weight models help meet that demand."With an open-weight model, the underlying learned settings that allow developers to run or adapt the system are available for download. By contrast, OpenAI, Anthropic and Google have closed-source models.Alibaba's new Qwen3.8-Max immediately shot up leaderboards assessing the capabilities of AI models after being unveiled on Monday, helping its shares jump 7% in Hong Kong trade.The model has 2.4tn parameters, the numerical settings a model learns from data and uses to recognise patterns, generate answers, and carry out tasks. That puts it not too far behind domestic rival Moonshot AI's Kimi K3, which was launched last month and has 2.8 tn parameters.A higher parameter figure does not automatically make a model better, but it has become a closely watched measure of the scale of the computing and data behind advanced AI systems.Qwen3.8-Max was unveiled on crowdsourced, model-comparison platform Arena.AI. It soon became the highest-ranking Chinese model in terms of text models, though it still lags Claude Fable 5 and three Opus variants which are all from Anthropic.On Arena.AI's leaderboard for AI models that analyse images and other visual material, Qwen3.8-Max ranked second globally, only behind a Claude Fable 5 variant.Both Qwen3.8-Max and Kimi K3 can handle text, images and video, and process up to 1 mn tokens at a time.Tokens are chunks of data, often parts of words or short words, and a big figure means the model can take in large amounts of material in one go, such as long legal files, a large software codebase or hundreds of pages of documents.The tech giant said the model, due to be released next week, completed a software-engineering project in 16 days. It uses a "mixture-of-experts" design, which divides work among specialised parts of the system instead of switching on the entire model for every request. Only 95 bn parameters are used at a time, reducing costs and response delays.DeepSeek's V4-Flash model, released on Friday, is by far the least expensive to run on benchmark tests among well-known models globally, according to research firm Artificial Analysis.The startup, which sources have said is preparing for a potential IPO, saw its R1 and V3 models become a global sensation in early 2025, triggering a selloff in global tech stocks and raising questions about the large amounts US companies were spending on AI.V4-Flash charges $0.14 per million input tokens and $0.28 per mn output tokens, according to San Francisco-based Artificial Analysis.Artificial Analysis estimated V4-Flash's average cost at 3 cents per test, compared with 86 cents for Kimi K3, $1.86 for OpenAI's GPT-5.6 Sol and $3.15 for Claude Fable 5.The comparison provides a more realistic measure of value than pricing alone because it accounts for the amount of data a model must process and generate to complete a task. A model with low headline price can still prove expensive if it requires significantly more steps to produce an answer.

A DeepSeek AI sign is seen at a building where its office is located in Beijing. Chinese AI startup DeepSeek is planning to launch a fresh fundraising round ‌at a valuation of about 500bn yuan ($74bn) ahead of a potential mainland ​initial public offering, two people ‌with knowledge of the matter said.
Business

China's DeepSeek seen to raise fresh capital at $74bn valuation ahead of onshore IPO

DeepSeek eyes listing on Shanghai's STAR market, sources saidCompany plans fresh fundraising after $7.4bn raised in JuneFresh fundraising signals growing cost of frontier AIChinese AI startup DeepSeek is planning to launch a fresh fundraising round ‌at a valuation of about 500bn yuan ($74bn) ahead of a potential mainland ​initial public offering, two people ‌with knowledge of the matter said.The plan comes just weeks after the Hangzhou-based ‌company, which drew global ⁠attention with its ‌low-cost AI models in 2025, raised about $7.4bn ‌in June a post-money valuation of about 450bn yuan, the people said.The back-to-back fundraising ⁠plans underscore strong investor appetite for one of China's most closely watched AI companies, but also point to the rising costs of competing in AI, which requires large amounts of computing power, data-centre capacity and engineering talent.DeepSeek is looking to raise as much as 50bn yuan in the new funding round, according to a third person briefed on the matter.It has also started early deliberations on a potential IPO on Shanghai's Nasdaq-style STAR Market, the ​three sources and two other people with knowledge of the plan said.The company has set an internal target to complete an IPO filing this year, one of them said.All the people declined to be ‌identified because the information is not public. ⁠The fundraising and ​IPO plans are at early stages and terms and timetable may change, they ​said.DeepSeek did not immediately respond to a request for comment.Bloomberg News first reported on Tuesday that DeepSeek was preparing for a possible IPO filing, while the Financial Times reported that the company was weighing a fresh fundraising round at a valuation of at least 480bn yuan.DeepSeek shook global technology markets last year after releasing models that appeared to rival leading US systems at lower training and operating costs.Soon after its maiden fundraising round in June, DeepSeek said it planned to double staff across departments, including in areas such as data centres and AI agents, systems capable of performing tasks with limited prompting.Some ‌of those initiatives will require ‌significant capital expenditure. Reuters reported earlier this ⁠month that DeepSeek was looking to develop its own AI inference chip and had discreetly increased hiring ⁠of chip-design engineers for the project.DeepSeek had ⁠long stood out in China's AI sector for rejecting outside funding. Founder Liang Wenfeng had largely bankrolled the company using his quantitative hedge fund High-Flyer before its recent external financing, sources previously told Reuters.But the cost of staying at the frontier of AI has risen sharply, forcing a change in strategy.DeepSeek has in the past year faced stiff competition at home from tech giants including ​ByteDance and Alibaba, as well as well-funded AI startups such as Z.ai, Moonshot, and MiniMax. 

The makers of short-video ​platform TikTok released the Doubao 2.0 chatbot, the company said on Saturday. The new model is capable of deep thinking and long, multi-step task execution that matches the capabilities of US rival OpenAI's GPT 5.2 and Google's Gemini 3 Pro, Bytedance said.
Business

Chinese AI models festoon spring festival a year after DeepSeek shock

As China prepares for Lunar New Year holidays starting on Sunday, rivals to DeepSeek are scrambling to release artificial-intelligence models a year after it burst onto the scene with its game-changing R1 and V3 models.With DeepSeek set to launch its next-generation V4 ‌model soon, according to tech news site The Information, many other Chinese AI firms have released ​or are preparing to launch their own ‌models in the hopes of stealing the spotlight — or at least avoid being off-guard again ‌during this year's ⁠Spring Festival.Below are the companies ‌and models looking to make a Spring Festival ‌splash. DEEPSEEKThe Hangzhou-based startup's V4 would replace last year's V3 model, which powered the AI assistant app that overtook ChatGPT to ⁠become the top-rated free application available on Apple's App Store in the US Investors and industry insiders are also on the lookout for R2, successor to the R1 model.This week DeepSeek fuelled anticipation when its web and mobile chatbot upgraded its "context window" — the amount of information it can remember and handle in a single task, from 128,000 to1 mn tokens, the unit of data processed by the AI model.This means the chatbot can now process book-length passages of text to answer a single user command. BYTEDANCEThe makers of short-video ​platform TikTok released the Doubao 2.0 chatbot, the company said on Saturday. The new model is capable of deep thinking and long, multi-step task execution that matches the capabilities of US rival OpenAI's GPT 5.2 and Google's Gemini 3 Pro, Bytedance said.Doubao ‌is China's most popular AI chatbot app, ⁠in terms of weekly active ​users, according to data published by Quest Mobile in December.Thursday's release of video-generation AI model ​Seedance 2.0 has generated comparisons to DeepSeek's global rise, going viral on Chinese social media and drawing widespread praise on X, including from the platform's owner, Elon Musk. Seedance 2.0 can produce high-quality cinematic videos based on a few prompts, or even one.The tech giant released picture-generation model Seedream 5.0 Lite on Friday. ALIBABAAlibaba, the first Chinese firm to respond to DeepSeek's viral ascent last year, with Qwen 2.5-Max, is preparing to launch Qwen 3.5.The e-commerce giant's Qwen app is riding a wave of growing domestic usage after it spent 3bn yuan ($400mn) last week on a coupon giveaway campaign to promote "agentic commerce", where AI handles consumers' online shopping.This drove more than 120mn consumer orders in the six days through Wednesday, the company said. ZHIPUZhipu ‌AI released its open-source GLM-5 model on ‌Wednesday, with enhanced coding capabilities and the ability ⁠to perform long-running agent tasks.Zhipu is considered one of China's "AI tigers" — promising startups vying with the US to win ⁠the AI race. Zhipu went public on the ⁠Hong Kong Stock Exchange last month, alongside rival MiniMax, another AI tiger.Both stocks have rallied strongly as investors bet on the companies benefiting from China's AI boom. Zhipu plans a secondary listing in Shanghai, a regulatory filing on Friday showed. MINIMAXMiniMax released its M2.5 open-source model on its overseas agent website on Wednesday. The company's Hong Kong listing raised HK$4.8bn ($620mn), higher than Zhipu's $558mn.Shanghai-based MiniMax has developed popular apps like Hailuo AI, a video generation tool, and Talkie, a ​character interaction app that enables users to engage with AI-powered virtual personas. TENCENTTencent's Hunyuan team on Tuesday released a low-storage, compressed AI model, HY-1.8B-2Bit, designed to be used on consumer hardware including mobile phones. IFLYTEKiFlytek on Wednesday released Spark X2, trained entirely on Chinese-made chips. The company said the upgrade focuses on practical deployment in sectors including education, healthcare, automotive and agent-based applications. NETEASE YOUDAONetEase Youdao on Wednesday launched LobsterAI, a desktop-level personal assistant agent that can perform tasks such as information retrieval, scheduling and data analysis by executing workflows locally on a user's computer after authorisation.The product supports mobile and PC connections and allows remote interaction via enterprise apps popular among Chinese companies such as DingTalk and Feishu. DEXMALEmbodied-intelligence startup ‌Dexmal on Tuesday unveiled DM0, ​an AI model designed for robot-related scenarios. DM0 integrates multimodal internet data with driving, navigation and robotic operation data, and was trained across multiple robot platforms.