AI资讯大全AIPPXP.CN搜索 ↗
AI 快讯 · 自动更新

DeepSeek V4.1 Flash takes full advantage, and the gap between the LiveBench scores of top models in China and the United States narrows to 3%DeepSeek V4.1 Flash 发力,中美顶尖模型 LiveBench 跑分差距缩至 3%

Beijing time on October 5, according to Bloomberg Industry Research, in the past few months, as the models of Chinese AI laboratories such as DeepSeek have continued to improve, the performance lead of U.S. AI companies relative to their Chinese counterparts has shrunk sharply to an all-time low, putting the U.S.’s technological dominance under threat. …

北京时间 10 月 5 日,据彭博行业研究称,过去几个月,随着 DeepSeek 等中国 AI 实验室的模型不断进步,美国 AI 公司相对中国同行的性能领先优势急剧缩小,降至历史最低水平,这让美国的科技主导地位面临威胁。…

2026年10月5日 发布 · 2 分钟阅读 · 免费

Key Points要点速览

Beijing time on October 5, according to Bloomberg Industry Research, in the past few months, as the models of Chinese AI laboratories such as DeepSeek have continued to improve, the performance lead of U.S. AI companies relative to their Chinese counterparts has shrunk sharply to an all-time low, putting the U.S.’s technological dominance under threat. After DeepSeek released V4.1 Flash in September, top Chinese models trailed U.S. rivals by just 3% in benchmark scores, Robert Lea, senior analyst at DeepSeek Bloomberg Intelligence, wrote in a report on Monday. The gap narrowed from about 9% in May and 15% earlier this year. This continued improvement in performance means Chinese competitors will further expand their market share, he said. Bloomberg pointed out that the rise of China’s model technology is due to the continuous deepening of its AI professional capabilities and the ability of researchers to optimize models for domestic hardware. These developments also raise questions about the effectiveness of U.S. restrictions on technology exports. Leah said China's progress "further calls into question the long-term sustainability of U.S. technological dominance in AI." DeepSeek's V4.1 Flash ranked sixth in the LiveBench global rankings last month, becoming the highest-ranked Chinese model since the startup's breakthrough in 2025 with its inference model R1. LiveBench scores an AI model's responses and analysis of a question, puzzle, or task, a process similar to measuring human IQ. DeepSeek’s most recent score on LiveBench was 81.1, lower than Anthropic’s highest score of 83.4. Leah said this means the DeepSeek model's performance is already "comparable to leading AI systems from Anthropic and OpenAI." However, although the score gap between the two sides has narrowed to just 3%, only three of the top 15 models selected by LiveBench are from China. Leah cautions that such rankings are constantly changing, and monetization can be difficult regardless of a model's position on the leaderboard.

中文

北京时间 10 月 5 日,据彭博行业研究称,过去几个月,随着 DeepSeek 等中国 AI 实验室的模型不断进步,美国 AI 公司相对中国同行的性能领先优势急剧缩小,降至历史最低水平,这让美国的科技主导地位面临威胁。 DeepSeek 彭博行业研究高级分析师罗伯特 · 利亚 (Robert Lea) 在周一的一份报告中写道,在 DeepSeek 9 月发布 V4.1 Flash 之后,中国顶尖模型在基准测试得分上仅落后美国竞争对手 3%。这一差距较 5 月份的约 9% 和今年早些时候的 15% 有所收窄。 他表示,这种性能的持续提升意味着中国竞争者将进一步扩大市场份额。 彭博指出,中国模型技术的崛起得益于其 AI 专业能力的不断深化,以及研究人员针对国产硬件优化模型的能力。这些进展也让人们开始质疑美国限制技术出口的做法是否有效。 利亚称,中国取得的进展“进一步让人质疑美国在 AI 领域的科技主导地位能否长期维持”。 DeepSeek 的 V4.1 Flash 上月在 LiveBench 全球排名中位列第六,成为自这家创业公司 2025 年凭借其推理模型 R1 取得突破以来排名最高的中国模型。LiveBench 根据 AI 模型对问题、谜题或任务的回答与分析进行评分,这一过程类似于衡量人类智商。 DeepSeek 最近一次在 LiveBench 上的得分为 81.1 分,低于 Anthropic 最高的 83.4 分。利亚表示,这意味着 DeepSeek 模型的表现已经“可以与 Anthropic、OpenAI 的领先 AI 系统相媲美”。不过,尽管双方的得分差距已经缩小至仅 3%,LiveBench 评选出的前 15 个模型中只有 3 个来自中国。 利亚提醒称,这类排名会不断变化,而且无论模型在排行榜上的排名如何,变现都可能很困难。

读完接着看 · Keep reading

想马上用起来?去「AI工具」栏挑一个直接下载,或在「AI教程」里跟着图文步骤做一遍。

去 AI工具 → 看 AI教程 →
0阅读0 条评论

阅读与点赞数据保存在你的浏览器本地,欢迎留下你的想法。

评论 文明发言,让讨论更有价值

正能量公益广告今日正能量学一点,用一点;今天种下的种子,会长成明天的能力。去免费下载专区 →广告