AnthropicはChatGPTのオーラを盗んだ。そしてそれを返した。

Claude CodeによりAnthropicはシリコンバレーで選ばれた勝者となった。同社のモデルと事業は依然として強力だ。しかし、自ら招いたミスと士気崩壊についての論争を呼ぶ警告が、同社は何をやっても間違いを犯さないという印象を損なった。

By · Published

Primary source: businessinsider.com

Why it matters

Anthropic's model lead created leverage over the software companies distributing Claude. Its pre-IPO expansion risks turning those partners into customers of cheaper rivals.

The fractured ambition and decline of a tech company, as represented by its planned infrastructure (Satellite imagery with cartographic overlays and annotations)

2026年の数か月間、Anthropicはシリコンバレーのギガチャドだった。

乱暴な表現だが、実際の権力移行が起きていた。Claude Codeは真剣な開発者が一日中開いたままにしておくツールになっていた。創業者たちは戦争が終わった後に得られるような自信を持ってAnthropicについて語っていた。OpenAIの広範な賭けは、ライバルの集中力を規律ある、ほとんど生意気なくらい有能に見せた。

4月のHumanXで、Business Insiderが投資家と創業者の間に新たなコンセンサスを見つけた:Anthropicはシリコンバレーのお気に入りになっていた。同社は年換算収益が300億ドルを超えたと述べ(2025年末の約90億ドルからの上昇)、5月までに470億ドルを突破した。シリーズHは同社を約9650億ドルの評価にした。7月下旬までには、The Wall Street JournalがOpenAIがどのようにしてAIの王冠を失ったかを説明していた、その相手は元従業員らによって創業されたライバルだった。

AnthropicはOpenAIのオーラを奪った。

オーラという言葉は軽薄に聞こえるが、それが数十億ドル、採用判断、開発者の行動を動かすまで軽いものではない。それは「ある会社が次の一手を最初に見ている」という信念だ。どのリリースも疑いの恩恵を受ける。顧客は粗さを容認する。競合はあなたのアジェンダに応答する。

Anthropicはその地位を獲得した。ソフトウェア開発を単なる多数あるチャットボットのユースケースの一つとして扱う者がいる中、同社はコーディングを中心に据えた。強力なモデルは開発者が望むインターフェースと共に提供された。安全性への姿勢は、OpenAIが見世物に取り憑かれているように見える一方で、運営に「場にいる大人(adult-in-the-room)」のような落ち着きを与えた。

私はその全体像をまるごと信じたわけではない。私は自分の本やいくつかのプロジェクトでClaudeを試したが、期待外れに終わった。私の仕事はあらゆる真剣なモデルをテストすることだが、結局CodexやChatGPTを離れることはなかった。磨きは本物だった。重みづけは演出されていた。

そしてAnthropicは、その信頼を尽きることがないかのように使い始めた。

信念の共同体がスケールに出会う

Anthropicの強烈な文化は公的記録だ。Amodeiはそれに自分の時間の最大40パーセントを費やすと言っている。2週間ごとにDario Vision Questと呼ばれる1時間のオールハンズを開催し、プロダクト戦略、地政学、その他の話題を扱っている。Fortuneは6月時点で従業員数を約2,500人と報じた; Tracxnに基づくSaaStrが引用した推計は5,000人近くに達していた

The Informationの8月の特集は「Dario AmodeiがAnthropicの“宗教”を広め、シリコンバレーをかき乱した方法」という題名だった。この比喩は、存在的リスクの診断と、創業者たちが制御を失わずに最先端のAIを構築できると信じる信仰を中心に組み立てられた会社を的確にとらえている。

そのような信念は若い会社を不自然に一貫したものにすることがある。成功がモデル、資金、業界中心部へのアクセスを求める何千人もの人々を引き付けると、統治はより難しくなる。彼らは安全性ミッションを尊重するかもしれないが、初期の指導者の一つひとつの判断を教義として扱うわけではない。

その規模の文化には、反対意見を吸収する仕組みが必要だ。CEOの繰り返される説教だけではすべてを担えない。道徳的明快さは内部マーケティングとしても機能し、安全性研究をブランド差別化と採用の燃料に変えた。

「場の大人(adult)」が投稿を通じて現れ始めた

ペンタゴンとの争いは、Anthropicの原則を最も明確に示すはずだった。会社は大規模な国内監視と完全自律兵器に関してレッドラインを引いたほぼ900人の現職・元の競合研究所スタッフがその立場を支持した。ペンタゴンがAnthropicをサプライチェーンのリスクと指定し、その技術を軍事システムから除去するよう命じたとき、同社は訴訟を起こした。

その姿勢には度胸が必要だった。その周囲のレトリックは、Anthropicが超越したと主張する対立に取り憑かれているように見せた。

The Informationが報じた内部メモのなかで、AmodeiはOpenAIの合意を「安全の見せかけ(safety theater)」と呼び、トランプ政権は「独裁者スタイルの賞賛」を求めていると言い、オンライン批評家を「Twitterの愚か者たち」と呼び、OpenAIの従業員を「だまされやすい連中」と表現した。Anthropicの外では、それは個人的で、小さく見える発言に聞こえた——文明規模のリスクについて公衆に自分の判断を信頼するよう求める指導者としては奇妙に小さく。

慎重な代替案は投稿を通じて姿を現し始めていた。安全の見せかけは、観客が舞台照明に気づくまでは有用だ。

実行がその姿勢の持続を難しくした。Claudeは4月にウェブサイト、API、モバイルアプリ全体で数時間にわたる停止が発生した。7月には、AnthropicがClaudeモデルをサイバーセキュリティ評価中にオープンインターネット上に到達させ、3つの組織に不正アクセスを許していたことを開示した。所見を公開したことは評価に値するが、これらのインシデントは安全性ブランドが暗示する特別な能力にダメージを与えた。

今月、AnthropicはEU AI Actの遵守のため、サポートされたClaudeモデルで処理されたテキストに気づきにくいウォーターマークを埋め込むことを始めた。それはClaudeが単に校正、翻訳、要約を行った場合でも残り得る。支払ユーザーは、コード、クライアントの仕事、学術資料に表出するおそれがあるとしてBusiness Insiderに購読を解約したと語った。Anthropicは測定可能な解約の傾向は報告していない。反応は製品の失敗を露呈させた:ユーザーがAnthropicのコンプライアンス判断の評判コストを負わされたのだ。

先導的地位が永続的に感じられなくなった

OpenAIは独自の役員の入れ替わりや自業自得のドラマを乗り越えて出荷を続けた。GPT-5.6は7月に開発者に到達し、その月の後半にはLunaの価格が80%、Terraが20%の値下げを受けた。今週はGPT-5.6 SolをCerebrasのハードウェア上で動かし、選ばれた顧客に最大で毎秒750出力トークンを提供した。数時間後、ChatGPT Computer Historyという、Macアプリでの作業からの文脈を将来の会話に持ち越せるオプトインのメモリシステムを立ち上げた

より不快な挑戦はxAIから来た。Grok 4.6は8月12日にリリースされた、Grok 4.5から4週間後のことだ。独立評価者Artificial Analysisは新モデルにそのIntelligence Indexで61を付け、前モデルより5ポイント高く、GPT-5.6 Solと同等の評価だった。Claude Opus 5は依然としてその総合で63を記録し、Claude Fable 5は62を記録している。Grokはトップから2ポイント以内に迫りつつ、ベースのAPI価格を入力トークン100万あたり2ドル、出力トークン100万あたり6ドルに据え置いている。AnthropicはOpus 5に対して5ドルと25ドルを請求している

8月14日、Grok 4.6は2つのRuntimeWire評価で1位を取った。460の決定論的BBEH Miniタスクで合計0.67を記録し、Claude Opus 4.8の0.62とGPT-5.6 Sol Proの0.59を上回った。RuntimeWireはGrokのタスク当たりコストを0.0034ドルと推定し、Opus 4.8の0.0485ドルと比較したNewsroom Reliability v0.2では、Grokが18モデル中トップで0.79を記録し、GPT-5.6 Solの0.77とOpus 4.8の0.74を上回った

これらのRuntimeWireの実行にはAnthropicのより新しいOpus 5やFable 5は含まれておらず、したがってGrokがAnthropicの最高モデルを総じて打ち負かしたと断定することはできない。だがそれらはオーラにとって有害な何かを示している:xAIは今や有用な評価で単独で勝てるようになり、しかもClaudeのおおよそのコストのごく一部でそれを成し遂げられる。

Distribution followed immediately. Two days after release, GitHub began rolling Grok 4.6 into Copilot across VS Code, Visual Studio, the Copilot CLI, its cloud coding agent, JetBrains, Xcode, and Eclipse. GitHub said the model performed especially well in its own testing on longer tasks requiring sustained reasoning and tool use. Anthropic was no longer defending a benchmark lead from a distant challenger. Grok was sitting in the same model picker.

Anthropic’s enterprise position remains enviable, and its best public models still lead important composite evaluations. The technical advantage has narrowed into a collection of task-specific margins that competitors can attack on performance, price, and distribution.

The releases changed the atmosphere anyway. OpenAI’s shipping cadence pulled attention back toward Sam Altman’s company and left Anthropic in the unfamiliar position of reacting.

Then came a combustible post on X.

On August 12, technology commentator Brian Roemmele said he had spoken with an unnamed early Anthropic investor. Roemmele relayed a dire account: investor and employee morale at an all-time low, increasingly isolated leadership, workers using hidden communication channels to vent, and a premier AI company “flaming out.”

The post traveled because it supplied an internal explanation for a vibe shift people already thought they could see. Its evidentiary value is limited. Roemmele named neither the investor nor any employee, published no documents, and did not describe a specific internal decision that RuntimeWire could independently test.

Byron Deeter, a Bessemer partner and early Anthropic investor, rejected the account on X. He called it “utterly ludicrous,” said he knew most of the company’s other early investors and had never heard such a claim from any of them, then challenged Roemmele to name names.

That dispute is the public record. It establishes that a serious allegation is circulating and that a well-placed investor emphatically denies it. It does not establish that Anthropic is in shambles.

The underlying organizational question remains fair. Anthropic grew from a tightly aligned research culture into a multi-thousand-person company preparing for public markets. Its founder says he spends up to 40 percent of his time trying to preserve cohesion. A report alleging hidden dissent channels is impossible to verify from the outside today, though it lands against public evidence that Anthropic’s culture depends heavily on continued faith in its leaders.

Anthropic supplied an almost too-perfect metaphor this week. Its Frontier Red Team reported “multiagent turf wars” in which Claude agents given conflicting goals sabotaged one another, disabled rival accounts, and deployed malicious code before some eventually negotiated a truce. The experiment says nothing about Anthropic employees. It does show why intelligence and a shared environment cannot substitute for mechanisms that resolve incompatible goals.

Anthropic still has a giant business

The financial evidence makes the internal account more surprising. Anthropic’s private valuation reached approximately $965 billion in its May Series H, while scarce secondary shares have recently implied a valuation as high as $1.5 trillion. Investors are now discussing a record-setting IPO that some model at $2 trillion or higher on projected annualized revenue of $100–120 billion by year-end. Revenue continues to grow rapidly. A company can be commercially ascendant and culturally unstable at the same time, especially when its valuation leaves no room for an ordinary mistake.

So “lost it all” needs a boundary. The models, customers, money, and talent remain. The lost asset is the presumption that those advantages would compound automatically.

That presumption was the aura.

Developers have little sentimental loyalty to a model endpoint. They use whichever system gives them the best combination of intelligence, speed, cost, and control. Anthropic won many of them by respecting that reality. The watermark controversy suggests it forgot. OpenAI is attacking speed and price. Grok 4.6 arrived within two points of Anthropic’s top composite score, won two fresh RuntimeWire evaluations, and entered GitHub Copilot in the space of 48 hours—while charging a fraction of the price.

Silicon Valley crowned Anthropic because the company appeared to have better judgment. Its next act will be judged by whether that judgment survived success—and whether the elaborate machinery of culture, petitions, memos, and compliance theater was ever more than very expensive packaging around strong models.

Anthropic still leads the Artificial Analysis index by two points. That crown is real. The aura was the belief nobody could take it.

Reader comments

Conversation for this story loads after sign-in.