AlibabaはQwen3.8-27Bのオープンウェイト版を8月14日にリリースする予定している。

AlibabaのQwenチームは、27Bの密結合型ビジョン・ランゲージモデルを8月14日に予定したが、提供された資料はダウンロード可能な重みがいつ公開されるかを明らかにしていない。

By · Published

Primary source: Qwen / Alibaba

Why it matters

The planned 27B model gives developers a much smaller Qwen3.8 target than Alibaba's 2.4T checkpoint, provided downloadable weights become available.

Alibaba schedules Qwen3.8-27B open-weight release for August 14

Alibaba's Qwen project scheduled Qwen3.8-27B, a dense vision-language checkpoint, for release on August 14, 2026. Developers could test the smaller model locally if downloadable files become available. Alibaba is led by co-founder and CEO Eddie Wu.

本レポートで確認した資料は、重み(weights)が実際にダウンロード可能になったか、また正確にいつダウンロード可能になったかを確定していません。

The official Qwen account (@Alibaba_Qwen) said in a countdown post that the scheduled release was less than two hours away.

Qwen / Alibaba on X

Preliminary repository material collected for this report described Qwen3.8-27B as a dense, 27-billion-parameter vision-language model with a 262,144-token native context window that can be extended to roughly 1 million tokens. Alibaba's promotional image calls Qwen3.8-27B a "renewal of the beloved Qwen model" delivering "intelligence density," a company description that independent testing has not established.

予備的に収集したリポジトリ資料は、Qwen3.8-27B を、ネイティブのコンテキストウィンドウが262,144トークンで、概ね100万トークンに拡張可能な、27億(27 billion)のパラメータを持つディープなビジョン・ランゲージモデルとして説明しています。Alibaba のプロモーション画像は Qwen3.8-27B を「renewal of the beloved Qwen model」(愛される Qwen モデルの刷新)であり「intelligence density」(インテリジェンス密度)を提供すると称していますが、独立したテストでその評価は確定していません。

A smaller checkpoint for Qwen3.8

Alibaba has already published weights for Qwen3.8-2.4T-A95B, a mixture-of-experts model with 2.4 trillion total parameters and 95 billion activated for each token. Qwen3.8-27B is listed as a separate model; Alibaba has described Qwen3.8-Max as a 2.4-trillion-parameter model.

Alibaba は既に、合計で2.4兆パラメータ、各トークンにつき950億(95 billion)を活性化する Mixture-of-Experts モデルである Qwen3.8-2.4T-A95B の重みを公開しています。Qwen3.8-27B は別個のモデルとして掲載されており、Alibaba は Qwen3.8-Max を 2.4 兆パラメータのモデルとして説明しています

If released, Qwen3.8-27B would give the Qwen3.8 line a smaller dense checkpoint alongside the 2.4-trillion-parameter model. The supplied materials do not establish that the 27B model was distilled from, or otherwise technically derived from, the larger model. A 27B dense checkpoint remains computationally demanding, particularly at long context lengths, although its parameter count makes local evaluation and serving more plausible than with the 2.4T model. The supplied material does not specify a hardware configuration or serving cost for Qwen3.8-27B.

もし公開されれば、Qwen3.8-27B は Qwen3.8 系列において、2.4 兆パラメータモデルと並ぶより小さなデンス(dense)チェックポイントを提供することになります。提供された資料は、27B モデルが大きなモデルから蒸留(distilled)されたか、あるいは技術的に派生したかを示していません。27B のデンスチェックポイントは、特に長いコンテキスト長では計算的負荷が高いままですが、そのパラメータ数は 2.4T モデルに比べてローカルでの評価やサービングをより現実的にします。提供資料は Qwen3.8-27B のハードウェア構成やサービングコストを明示していません。

The preliminary description lists image and video input alongside text, with intended uses spanning coding, research, professional work and long-running agent tasks. It also lists compatibility with Transformers, vLLM, SGLang and TokenSpeed, giving teams several established inference frameworks to test once the artifacts become available.

予備的な説明では、テキストに加えて画像および動画入力を列挙しており、コーディング、研究、専門的業務、長時間稼働するエージェントタスクなどへの利用を想定しています。また Transformers、vLLM、SGLang、TokenSpeed との互換性も記載されており、アーティファクトが利用可能になった際に複数の既存推論フレームワークでテストできるようになっています。

Alibaba's material says thinking mode is enabled by default and can be adjusted through reasoning_effort settings labeled xhigh, medium and low. Those controls are intended to help operators balance response quality, latency and inference expense. Their practical cost will depend on the final weights, serving implementation and the amount of reasoning the model generates.

Alibaba の資料によれば、thinking mode はデフォルトで有効になっており、reasoning_effort 設定(xhighmediumlow とラベル付け)で調整できるとしています。これらのコントロールは、応答品質、レイテンシ、推論コストのバランスを運用者がとるのを助けることを目的としています。実際のコストは最終的な重み、サービング実装、およびモデルが生成する推論量に依存します。

Open weights and managed inference

Wu joined Alibaba as technology director in 1999, the year its 18 co-founders established the group. He later served as chief technology officer of Alipay and Taobao, ran Alibaba's search, advertising and mobile operations, and founded the technology-focused investment firm Vision Plus Capital in 2015. Eddie Wu became Alibaba Group's chief executive officer on September 10, 2023.

Wu は、グループを設立した18人の共同設立者が集まった1999年にテクノロジーディレクターとして Alibaba に参加しました。その後、Alipay と Taobao の最高技術責任者(CTO)を務め、Alibaba の検索、広告、モバイル事業を統括し、2015年にテクノロジーに特化した投資会社 Vision Plus Capital を設立しました。Eddie Wu は 2023年9月10日に Alibaba Group の最高経営責任者(CEO)に就任しました

In a May 2024 shareholder letter signed by Wu and Chairman Joe Tsai, the executives wrote that training large language models and using them for development or inference require computing resources. They also said open-sourcing Qwen created "additional demand" for Alibaba's proprietary model and related computing resources. The letter does not establish that Qwen3.8-27B will produce paid Alibaba Cloud usage or that self-hosted users will buy agent services.

Wu と会長 Joe Tsai の署名入りの2024年5月の株主向け書簡で、役員らは大規模言語モデルのトレーニングやそれを開発・推論に使用するには計算資源が必要であると述べました。また、Qwen をオープンソース化したことで Alibaba の独自モデルおよび関連する計算資源に「追加需要(additional demand)」が生じたとも記しています。この書簡は、Qwen3.8-27B が有料の Alibaba Cloud 利用を生むこと、あるいはセルフホストユーザーがエージェントサービスを購入することを確定するものではありません。

Alibaba distributes Qwen models through Hugging Face and operates Qwen Cloud, which provides hosted model access and APIs. The supplied materials do not establish the August 14 checkpoint's pricing or cloud revenue impact.

Alibaba は Hugging Face を通じて Qwen モデルを配布しており、ホストされたモデルアクセスと API を提供する Qwen Cloud を運営しています。提供資料は、8月14日のチェックポイントの価格設定やクラウド収益への影響を示していません。

Alibaba reported in May 2026 that Cloud Intelligence Group external revenue grew 40% year over year during the March quarter. Alibaba did not attribute that growth to Qwen3.8-27B.

Alibaba は2026年5月に報告しており、Cloud Intelligence Group の外部収益が3月四半期で前年同期比40%成長したとしています。Alibaba はその成長を Qwen3.8-27B に帰属させていません。

From flagship access to local testing

Qwen3.8-27B is a separate development from RuntimeWire's recent article, "Alibaba's Qwen Cloud recap lays out an agent-first developer platform", which focused on Alibaba's hosted developer service. The August 14 announcement concerns a smaller checkpoint that developers may be able to inspect and run outside that platform.

Qwen3.8-27B は、Alibaba のホスト型開発者サービスに焦点を当てた RuntimeWire の最近の記事 "Alibaba's Qwen Cloud recap lays out an agent-first developer platform" とは別個の開発です。8月14日の発表は、開発者が当該プラットフォーム外で検査・実行できる可能性のあるより小さなチェックポイントに関するものです。

RuntimeWire reported earlier in August that Alibaba planned model weights for its 2.4-trillion-parameter flagship while license terms remained unresolved. A later look at its cloud preview found that developers still lacked a stable target for evaluating the hosted version.

RuntimeWire は8月初めに 報じた とおり、Alibaba はライセンス条件が未解決の間に 2.4 兆パラメータのフラッグシップのモデル重みを計画していました。後の クラウドプレビューの検証 では、開発者はホスト版を評価するための安定したターゲットを依然として欠いていることがわかりました。

Alibaba also publishes multimodal plugins that connect Qwen capabilities to agent environments including Claude Code, Codex and Gemini CLI. Qwen Cloud offers agents that can call tools and carry work across multiple steps. These are documented product offerings, although the supplied materials do not connect them commercially to Qwen3.8-27B.

Alibaba はまた、Claude Code、Codex、Gemini CLI を含むエージェント環境に Qwen の機能を接続する multimodal plugins を公開しています。Qwen Cloud はツールを呼び出し、複数のステップにわたって作業を行えるエージェントを提供します。これらは文書化された製品提供ですが、提供資料はそれらを商業的に Qwen3.8-27B と結び付けてはいません。

If the checkpoint becomes downloadable, developers can inspect its behavior, measure memory use and throughput, and compare local inference with hosted alternatives. Reproducible evaluations could then test its coding, vision and agent performance across the supported inference frameworks. Quantized versions, if released, would provide another route for evaluating the model on smaller hardware configurations.

チェックポイントがダウンロード可能になれば、開発者はその挙動を検査し、メモリ使用量やスループットを測定し、ローカル推論とホスト提供の代替手段を比較できます。再現可能な評価により、サポートされる推論フレームワーク全体でコーディング、ビジョン、エージェント性能をテストできるようになります。もし量子化(quantized)バージョンが公開されれば、より小さなハードウェア構成でモデルを評価する別の手段を提供します。

For now, Alibaba's official Qwen promotion set August 14 as the release date. The supplied materials do not establish downloadable availability or its exact timing.

Reader comments

Conversation for this story loads after sign-in.