5 min read

The AI Model Wars Heat Up: DeepSeek V4 Pro, Grok 4.6 and Qwen3.8

The AI model race is accelerating. This week, DeepSeek, xAI and Alibaba's Qwen team have all made significant moves, giving developers a fresh selection of increasingly capable models.

The AI model race is accelerating.

This week, DeepSeek, xAI and Alibaba's Qwen team have all made significant moves, giving developers a fresh selection of increasingly capable models.

DeepSeek V4 Pro 0813, Grok 4.6 and Qwen3.8-2.4T-A95B represent three different approaches to the same challenge: building more capable AI while competing on performance, cost, scale and accessibility.

DeepSeek V4 Pro: Leaving Preview Behind

DeepSeek has moved DeepSeek V4 Pro out of preview, with the production V4 Pro 0813 release becoming available on August 12.

The model features a 1-million-token context window and is designed for demanding workloads including coding, reasoning and agentic tasks. It is also available through platforms such as OpenRouter.

One of DeepSeek's biggest advantages has been pricing. The model launched with particularly low API rates compared with many competing frontier models, although that advantage is about to change.

DeepSeek has announced a new peak and off-peak pricing structure for its V4 models, taking effect on August 16 at 16:00 UTC. Reuters reports that the changes represent increases of between 50% and 1,100%, depending on the model, token type and usage period.

That makes the V4 Pro release particularly interesting: developers get a more mature production model, but its long-term cost advantage will need to be reassessed once the new pricing takes effect.

Grok 4.6: xAI Focuses on AI Agents

xAI also announced Grok 4.6 on August 12.

The company says the new model builds on Grok 4.5 with a particular focus on long-running agents, coding, knowledge work and more ambitious interactive and visual tasks.

This reflects a wider change across the AI industry.

The competition is increasingly moving beyond traditional chatbots. Companies are building models that can work through longer tasks, interact with tools and assist with complex workflows.

Independent testing has placed Grok 4.6 at 61 on Artificial Analysis' Intelligence Index, putting it among the current frontier models. Artificial Analysis

Qwen3.8: Alibaba Opens a 2.4 Trillion-Parameter Model

Alibaba's Qwen team has made one of the most striking moves in terms of raw model scale.

Qwen3.8-2.4T-A95B is an open-weight mixture-of-experts model with 2.4 trillion total parameters and 95 billion active parameters. The model weights were released publicly this week, giving developers access to a model at an enormous scale. Hugging Face

The model uses a sparse mixture-of-experts architecture, meaning that although the model contains 2.4 trillion parameters overall, only a portion is activated for each token.

That distinction is important. A 2.4-trillion-parameter model does not require all 2.4 trillion parameters to be processed for every piece of text.

Qwen3.8 is particularly notable because it brings a model of this scale into the open-weight ecosystem, giving researchers and developers another option beyond closed commercial AI systems.

Why This Matters

These releases highlight how competitive the AI market has become.

The race is no longer simply about building the largest model. Companies are competing across several areas:

  • Performance: Better reasoning, coding and knowledge capabilities
  • Cost: More competitive pricing for developers and businesses
  • Context: The ability to process increasingly large amounts of information
  • Agents: Models capable of handling longer, multi-step tasks
  • Open weights: Greater access for developers and researchers
  • Infrastructure: More efficient ways to deploy increasingly large models

For developers, this competition creates more choice.

A project that previously required an expensive proprietary model may now have several alternatives. Open-weight models such as Qwen3.8 also give developers greater flexibility for experimentation and deployment.

The Developer Perspective

There is no single model that is automatically the best choice for every application.

Developers increasingly need to consider:

What does the model cost?

How large is its context window?

How well does it perform on the specific task?

Can it use tools or operate as an agent?

Can it be deployed independently?

How reliable is it in real-world use?

Those factors can matter more than a model's position on a general benchmark.

What Comes Next?

The pace of development shows little sign of slowing.

DeepSeek is pushing further into the frontier while maintaining a strong focus on efficiency. xAI is developing Grok around increasingly capable agents and complex knowledge work. xAI Blog

Alibaba is expanding the open-weight ecosystem with a model operating at an enormous scale.

And the competition extends far beyond these three companies. OpenAI, Anthropic, Google, Meta, Mistral and other AI developers are continuing to improve their own systems. OpenRouter

For developers and businesses, the result is a rapidly expanding selection of AI models.

The question is becoming less about "Which company has the best AI?"

Instead, it is:

"Which model is best for what I need to build?"

As the AI frontier continues moving, that answer could change much faster than it did even a year ago.

Discover more technology, AI, software, and digital content at Digital Infohub:

Digital Infohub
digitalinfohub.net

読了時間 5 分

AI モデル戦争が激化:DeepSeek V4 Pro、Grok 4.6、Qwen3.8

AI モデル競争が加速しています。今週、DeepSeek、xAI、アリババの Qwen チームがすべて重要な動きを見せ、開発者にますます高性能なモデルの新たな選択肢を提供しました。

AI モデル競争は加速しています。

今週、DeepSeek、xAI、アリババの Qwen チームがすべて重要な動きを見せ、開発者にますます高性能なモデルの新たな選択肢を提供しました。

DeepSeek V4 Pro 0813、Grok 4.6、Qwen3.8-2.4T-A95Bは、パフォーマンス、コスト、スケール、アクセシビリティで競争しながら、より有能な AI を構築するという同じ課題に対する 3 つの異なるアプローチを表しています。

DeepSeek V4 Pro:プレビューを卒業

DeepSeek はDeepSeek V4 Proをプレビューから移行し、本番用のV4 Pro 0813リリースが 8 月 12 日に利用可能になりました。

このモデルは100 万トークンのコンテキストウィンドウを備え、コーディング、推論、エージェントタスクなどの要求の厳しいワークロード向けに設計されています。OpenRouter などのプラットフォームでも利用可能です。

DeepSeek の最大の利点の 1 つは価格でした。このモデルは多くの競合するフロンティアモデルと比較して特に低い API レートで起動しましたが、その利点はまもなく変わろうとしています。

DeepSeek は V4 モデル向けの新しいピークおよびオフピーク価格構造を発表し、8 月 16 日 16:00 UTCに発効します。ロイターによると、変更はモデル、トークンタイプ、使用期間に応じて50% から 1,100%の増加を表しています。

これにより、V4 Pro リリースは特に興味深いものになります:開発者はより成熟した本番モデルを取得しますが、新しい価格が適用されると、長期的なコスト優位性を再評価する必要があります。

Grok 4.6:xAI は AI エージェントに注力

xAI も8 月 12 日に Grok 4.6を発表しました。

同社によると、新しいモデルは Grok 4.5 をベースにし、長時間実行されるエージェント、コーディング、知識作業、より野心的なインタラクティブおよび視覚タスクに特に焦点を当てています。

これは AI 業界全体のより広い変化を反映しています。

競争は従来のチャットボットを超えて進化しています。企業は、より長いタスクを処理し、ツールと対話し、複雑なワークフローを支援できるモデルを構築しています。

独立したテストでは、Grok 4.6 はArtificial Analysis のインテリジェンスインデックスで 61にランクされ、現在のフロンティアモデルの中に位置づけられています。Artificial Analysis

Qwen3.8:アリババが 2.4 兆パラメータモデルを公開

アリババの Qwen チームは、生モデルスケールの面で最も注目すべき動きの 1 つを行いました。

Qwen3.8-2.4T-A95Bは、総計 2.4 兆パラメータ、アクティブパラメータ 950 億を持つオープンウェイトのエキスパート混合モデルです。モデル重みは今週公開され、開発者が膨大なスケールのモデルにアクセスできるようになりました。Hugging Face

このモデルはスパースエキスパート混合アーキテクチャを使用しており、モデル全体で 2.4 兆のパラメータを含んでいますが、各トークンに対して一部のみが活性化されます。

その区別は重要です。2.4 兆パラメータモデルは、すべてのテキストに対して 2.4 兆すべてのパラメータを処理する必要はありません。

Qwen3.8 が特に注目すべきなのは、このスケールのモデルをオープンウェイトエコシステムにもたらし、研究者と開発者に閉鎖された商業 AI システム以外の別のオプションを提供することです。

これが重要な理由

これらのリリースは、AI 市場がどのように競争的になっているかを浮き彫りにしています。

競争は単に最大のモデルを構築することだけではなくなりました。企業はいくつかの領域で競争しています:

  • パフォーマンス:より優れた推論、コーディング、知識機能
  • コスト:開発者と企業のためのより競争力のある価格
  • コンテキスト:ますます大量の情報を処理する能力
  • エージェント:より長く、マルチステップのタスクを処理できるモデル
  • オープンウェイト:開発者と研究者へのより大きなアクセス
  • インフラストラクチャ:ますます大規模なモデルをデプロイするためのより効率的な方法

開発者にとって、この競争はより多くの選択肢を生み出します。

以前は高価な専有モデルが必要だったプロジェクトには、 теперьいくつかの代替手段があります。Qwen3.8 などのオープンウェイトモデルは、実験とデプロイメントのためのより大きな柔軟性も開発者に提供します。

開発者の視点

すべてのアプリケーションにとって自動的に最適な選択である単一のモデルは存在しません。

開発者はますます以下を考慮する必要があります:

モデルのコストは?

コンテキストウィンドウはどのくらい大きいか?

特定のタスクでどのくらいうまく機能するか?

ツールを使用したり、エージェントとして動作したりできるか?

独立してデプロイできるか?

実世界での使用でどのくらい信頼できるか?

これらの要因は、一般的なベンチマークでのモデルの位置よりも重要になる可能性があります。

次に何が来るか?

開発のペースは鈍化する兆しをほとんど示していません。

DeepSeek は効率性に強い焦点を維持しながら、さらにフロンティアを押し進めています。xAI は、ますます有能なエージェントと複雑な知識作業を中心に Grok を開発しています。xAI ブログ

アリババは、膨大なスケールで動作するモデルでオープンウェイトエコシステムを拡大しています。

そして、競争はこの 3 社をはるかに超えて広がっています。OpenAI、Anthropic、Google、Meta、Mistral およびその他の AI 開発者は、独自のシステムの改善を続けています。OpenRouter

開発者と企業にとって、結果は AI モデルの急速に拡大する選択肢です。

質問は「どの会社が最高の AI を持っているか?」ではなくなりつつあります。

代わりに:

「私が構築する必要があるものに最適なモデルはどれか?」

AI フロンティアが動き続けるにつれて、その答えは 1 年前でさえあったよりもはるかに速く変わる可能性があります。

Digital Infohub で、テクノロジー、AI、ソフトウェア、デジタルコンテンツをもっと発見:

Digital Infohub
digitalinfohub.net

阅读时间 5 分钟

AI 模型战争升温:DeepSeek V4 Pro、Grok 4.6 和 Qwen3.8

AI 模型竞赛正在加速。本周,DeepSeek、xAI 和阿里巴巴的 Qwen 团队都取得了重大进展,为开发者提供了越来越多功能强大模型的新选择。

AI 模型竞赛正在加速。

本周,DeepSeek、xAI 和阿里巴巴的 Qwen 团队都取得了重大进展,为开发者提供了越来越多功能强大模型的新选择。

DeepSeek V4 Pro 0813、Grok 4.6 和 Qwen3.8-2.4T-A95B代表了三种不同的方法,应对相同的挑战:在性能、成本、规模和可访问性方面竞争的同时构建更强大的 AI。

DeepSeek V4 Pro:告别预览版

DeepSeek 已将DeepSeek V4 Pro移出预览阶段,生产版V4 Pro 0813于 8 月 12 日发布。

该模型具有100 万 token 上下文窗口,专为包括编码、推理和代理任务在内的高要求工作负载而设计。它还可通过 OpenRouter 等平台使用。

DeepSeek 最大的优势之一是价格。与许多竞争的前沿模型相比,该模型推出时的 API 费率特别低,尽管这一优势即将改变。

DeepSeek 已宣布为其 V4 模型采用新的峰值和非峰值定价结构,将于8 月 16 日 16:00 UTC生效。路透社报道,这些变化代表着根据模型、token 类型和使用时期的不同,增幅在50% 到 1,100%之间。

这使得 V4 Pro 发布特别有趣:开发者获得了更成熟的生产模型,但一旦新定价生效,其长期成本优势将需要重新评估。

Grok 4.6:xAI 专注于 AI 代理

xAI 也于8 月 12 日宣布了 Grok 4.6。

该公司表示,新模型建立在 Grok 4.5 的基础上,特别关注长时间运行的代理、编码、知识工作和更具雄心的交互式及视觉任务。

这反映了 AI 行业更广泛的变化。

竞争正日益超越传统聊天机器人。公司正在构建能够处理更长任务、与工具交互并协助复杂工作流的模型。

独立测试将 Grok 4.6 置于Artificial Analysis 智能指数第 61 位,使其跻身当前前沿模型之列。Artificial Analysis

Qwen3.8:阿里巴巴开放 2.4 万亿参数模型

阿里巴巴的 Qwen 团队在原始模型规模方面做出了最引人注目的举动之一。

Qwen3.8-2.4T-A95B是一个开源权重的专家混合模型,具有2.4 万亿总参数和 950 亿活跃参数。模型权重本周公开发布,使开发者能够访问规模巨大的模型。Hugging Face

该模型使用稀疏专家混合架构,这意味着虽然模型总共包含 2.4 万亿参数,但每个 token 只激活一部分。

这种区别很重要。2.4 万亿参数模型不需要为每段文本处理所有 2.4 万亿参数。

Qwen3.8 特别引人注目,因为它将如此规模的模型带入了开源权重生态系统,为研究人员和开发者提供了封闭商业 AI 系统之外的另一个选择。

为什么这很重要

这些发布凸显了 AI 市场的竞争有多么激烈。

竞赛不再仅仅是构建最大的模型。公司在几个领域展开竞争:

  • 性能:更好的推理、编码和知识能力
  • 成本:对开发者和企业更具竞争力的价格
  • 上下文:处理越来越多信息的能力
  • 代理:能够处理更长、多步骤任务的模型
  • 开放权重:为开发者和研究人员提供更大访问权限
  • 基础设施:部署越来越大型模型的更高效方法

对于开发者来说,这种竞争创造了更多选择。

以前需要昂贵专有模型的项目现在可能有几种替代方案。Qwen3.8 等开放权重模型也为开发者提供了更大的实验和部署灵活性。

开发者视角

没有哪个单一模型自动成为每个应用的最佳选择。

开发者越来越需要考虑:

模型成本是多少?

它的上下文窗口有多大?

它在特定任务上表现如何?

它能使用工具或作为代理运行吗?

它能独立部署吗?

它在实际使用中有多可靠?

这些因素可能比模型在一般基准测试中的位置更重要。

接下来会发生什么?

发展步伐几乎没有放缓的迹象。

DeepSeek 在保持对效率的强烈关注的同时,进一步推向前沿。xAI 正围绕越来越强大的代理和复杂知识工作开发 Grok。xAI 博客

阿里巴巴正在以超大规模运行的模型扩展开放权重生态系统。

而且,竞争远远超出了这三家公司。OpenAI、Anthropic、Google、Meta、Mistral 和其他 AI 开发者继续改进自己的系统。OpenRouter

对于开发者和企业来说,结果是 AI 模型的快速扩展选择。

问题正变得不再是"哪家公司拥有最好的 AI?"

而是:

"哪个模型最适合我需要构建的东西?"

随着 AI 前沿继续前进,这个答案的变化速度可能比一年前快得多。

在 Digital Infohub 发现更多技术、AI、软件和数字内容:

Digital Infohub
digitalinfohub.net