AI使い比べAIを探すAIニュースAI活用法
会社紹介
個人情報保護方針利用規約FAQお問い合せお問い合わせ
エーアイビー株式会社事業者情報
© 2026 AIB Inc.
Anthropic
Anthropic

Claude Sonnet 5

最安値で購読
最安値
リリース日 2026-06-30(学習完了 2026-01-01)··1M IN · 128K OUT商用モデル|API 提供Claude Free 以上Perplexity Pro 以上
総合性能3.0上位30%業務自動化3.0上位20%オフィス業務3.0上位31%コーディング能力3.5上位21%

Sonnet 5はAnthropicの最も高性能なSonnetクラスモデルであり、コーディング、エージェント、専門的な業務において最先端のパフォーマンスを発揮します。推論レベル(低、中、高、最大など)を選択できる適応型思考機能をサポートしています。

  1. 01画像生成された画像
  2. 02料金料金・使用量・速度
  3. 03ユーザー比較ユーザー直接比較
  4. 04専門家評価専門家による性能評価
  5. 05口コミ口コミ分析
  6. 06関連情報

API料金

公式価格
$2IN$10OUT
費用を計算
出典Anthropic 公式ドキュメント

実際の使用量

25位1ランクシェア 0.6% · 397B トークン/週
1.2T
8/24
1.3T
8/31
1.4T
9/7
1.5T
9/14
1.5T
9/21
1.0T
9/28
397B
10/5
出典OpenRouter最終更新2026年10月8日

出力速度

普通2,000トークン基準
出力開始2.3s
合計31.2s
69tok/s
出力開始 2.3s合計: 31.2s
出典OpenRouter最終更新2026年10月8日

ユーザー直接比較

文章の特徴
回答の長さ
短め
上位 61%
語彙の多様性
普通
上位 52%
文のリズム
均一気味
上位 61%
言語の混在
標準
上位 44%
リスト使用
よく
行の38%
見出し使用
ときどき
行の13%
思考の比重
少なめ
全体の6%
うそ(ハルシネーション)
正直な回答
100%
正解 2正直な回避 3うそ 0
回答速度
速度不安定 ±30%
初回応答 2.6秒計 15.3秒
初回応答 2.6秒計 15.3秒
69 tok/s
トークン毎11.9ms
ユーザーのひとこと
  • •

    簡単なウェブサイトくらいならあっという間に作ってくれます。1 か月前

  • •

    もともと良かったけど、新しいバージョンは確実に性能が少し上がった感じ3 か月前

出典AIB Inc.最終更新2026年10月5日

専門家による性能評価

総合性能

もっと見る総合性能のAIランキングをすべて見る
38.2推論強度: Max146276.0

強度別Low 24.3Medium 28.1High (no thinking) 23.2High 31.7Extra High 34.4Max 38.2

AA Intelligence IndexについてAA Intelligence Indexについて詳しく

業務自動化(エージェント)

もっと見る業務自動化のAIランキングをすべて見る
37.3%推論強度: Max$6,378

強度別High (no thinking) 15.7%Max 37.3%

τ³-Bankingについてτ³-Bankingについて詳しく

文書・オフィス業務

もっと見る文書・事務作業のAIランキングをすべて見る
82.0%推論強度: Max71.7146554.5%53.9%67.0%62.3%41.8%5.0%47.5%1466

強度別Low 67.3%Medium 73.7%High (no thinking) 70.0%High 76.7%Extra High 76.7%Max 82.0%

Long Context ReasoningについてLong Context Reasoningについて詳しく

調査・ファクトチェック

もっと見る資料調査・ファクトチェックのAIランキングをすべて見る
145233.7%86.6%57.4%0.142

Arena Text FactualityについてArena Text Factualityについて詳しく

コーディング

もっと見るコーディングのAIランキングをすべて見る
14.1%推論強度: Max151959.4146363.2%80.5%71.5153954.3%80.742.7%53.8%68.8%37.3%150334.1%65.3%

強度別Low 2.5%Medium 2.0%High 5.1%Extra High 7.1%Max 14.1%

Terminal-Bench 4.0についてTerminal-Bench 4.0について詳しく

文章作成・会話

もっと見る文章作成・会話のAIランキングをすべて見る
145214731435146775.0123676.1%143363.9

Arena Writing, Literature & LanguageについてArena Writing, Literature & Languageについて詳しく

数学・科学

もっと見る数学・科学のAIランキングをすべて見る
41.3%推論強度: Max16.9%147491.1%94.7%92.965.6%29.3%77.0%60.0%1490148788.760.6%35.0%35.0%92.5%50.0

強度別Low 21.9%Medium 30.0%High (no thinking) 19.0%High 35.7%Extra High 39.0%Max 41.3%

Humanity's Last ExamについてHumanity's Last Examについて詳しく

画像認識

もっと見る画像認識・分析のAIランキングをすべて見る
127524.9%88.3%

Arena VisionについてArena Visionについて詳しく

専門分野

1432144114911512148257.8%147752.7%91.0%78.6%78.0%79.3%92.0%84.2%73345619584.0%90.0%93.3%85.0%27.7%78.0%8.9%95.0%76.0%81.0%97.0%61.0%62.0%32.1%71.6%13.0%

Arena KoreanについてArena Koreanについて詳しく

出典・最終更新日は各評価指標の詳細ページで確認できます·データの出典・掲載停止のご依頼

口コミ分析

要約

開発者たちは、Sonnet 5のコーディング能力や統合の容易さを高く評価している一方で、高額な料金設定やサイバーセキュリティ関連タスクにおける制限的なセーフティガードレールに対しては、多くの不満を抱いています。

335件の意見(YouTube)
  • 92%
    肯定的
    否定的
  • YouTubeYouTube
    92%
    肯定的
    否定的

反応の要約

1.

コーディング能力とパフォーマンス

46件の意見

1Mコンテキストウィンドウと強力なエージェント能力が賞賛されており、複雑なプログラミングやマルチリポジトリタスクにおける強力なツールとして認識されています。

2.

価格設定とコストの持続可能性

45件の意見

高いトークンコストや導入時の価格モデルに批判が集中しており、従来のOpusモデルや競合他社と比較してアクセシビリティが低いと指摘されています。

3.

安全規制と政府方針

14件の意見

モデルの限定的なサイバーセキュリティ機能に関して論争があり、一部のユーザーはこれを政府主導の検閲やゲートキーピングを正当化するための言い換え(ニュースピーク)であると解釈しています。

4.

統合とシステムの採用

46件の意見

開発者たちはこのモデルを旧バージョンのドロップイン代替品として急速に採用しており、GitHub Copilotや既存のAPIワークフローへの統合の容易さに注目しています。

実際の意見一部抜粋

  • ついに新しいSonnetモデルが登場しましたね

    “Finally a new sonnet model”
    YouTube@FRareDom·18·原文を見る
  • GLM-5.2とSonnet 4.6で記事の書き直しを試してみました。LLMは非決定的なので結果は全く異なります。しかし、GLM-5.2は手動で修正が必要な細かいミスが多かったのに対し、Sonnetは2回目で全てのミスを見つけて修正しました。プランニングやコーディングでも同様の状況でした。GLM-5.2は「紙面上」では良さそうですが、実際の使用結果は違いました。私はClaudeやGLM-5.2の回し者ではありませんが… :) 2022年11月から毎日LLMモデルを使っている身として、一般的なテストは自分のプロジェクトで確認する必要があると実感しています。

    “I have tried to rewrite an article with GLM-5.2 and with Sonnet 4.6. Completely different results as LLM is non-deterministic. But GLM-5.2 made a lot of subtle mistakes that needed to be corrected by hand. On the opposite, Sonnet found and corrected all mistakes in the second round. Similar situation was with planning and coding. GLM-5.2 seems to be good “on paper” but the real usage results was different. And I am not an attorney for Claude or GLM-5.2… :) But as I’ve been using LLM models daily since Nov 2022 I have realized that all common tests have to be confirmed in your project - there is no “one model rules them all” - you need to dig out a specific model from that LLM haystack with thousands of models. Benchmarks help but they start to be similar to fuel consumption specs in car ads - real consumption is different for everybody :)”
    Hacker Newssixtyj·8·原文を見る
  • 今のところ限られた時間でのテストですが、4.6よりも良い結果が出ており、スピードも少し速くなっているので、顕著な進化だと感じています。

    “The limited time I've had to test it so far, it's given me better results than 4.6 and a little quicker, so I think it's a noticeable step up.”
    YouTube@MsJason316·6·原文を見る
  • > なぜこんなことを自慢するのでしょうか?人々がサイバーセキュリティのタスクにモデルを使いたいと思っているのを知りながら、意図的にその能力を拒否しているかのようです。Anthropicにここで一体何を言ってほしいのですか?「これから世界中に安価で提供しようとしているこのモデルは、ハッキングが非常に得意です」とでも?Sonnetはサイバーセキュリティが苦手だと言うことは、多くの悪い選択肢の中で彼らが言える最も合理的なことです。

    “> Why would they brag about something like this? It's like they know people want to use models to perform cybersecurity tasks yet knowingly deny them the ability. What exactly do you want Anthropic to say here? "This model, the one we are about to give to the entire world for cheap, is really good at hacking"? Saying Sonnet is terrible at cybersecurity is the most reasonable thing they can say, out of a lot of bad options.”
    Hacker Newsjohnfn·4·原文を見る
  • この際、無能な米国政府に目にもの見せてやるために、Anthropicにはオープンソースになってほしいです。

    “At this point I want Anthropic to go open source just to give the middle to the braindead US government.”
    YouTube@OnigoroshiZero·4·原文を見る
  • Fableを取り戻す方法を考えたほうがいいですよ😂

    “They better figure out how to get us Fable back😂”
    YouTube@NoLimitSquad·3·原文を見る
  • > しかもOpus 4.8の方がパス率が高いのにまだ安い。Opusほどテキストを垂れ流さない限り、それは疑わしいですね。Opus 4.8は文字通り吐しゃ物のようにテキストを量산します。長期的には、特にあちこちでキャッシュミスが発生する場合、コストの大部分はその余計なコンテキストによるものです。

    “> And Opus 4.8 is still cheaper for a higher pass rate Unless it spams as much as Opus, I doubt it. Opus 4.8 literally spams text like puke. On a longer run especially if you get cache misses here and there the bulk of the cost is all the extra context it adds.”
    Hacker Newsre-thc·2·原文を見る
  • 公開されてからずっと使っていますが、いい仕事をしてくれます。Rustのプラグインやサーバー管理パネルをコーディングでき、プラグインの設定をパネルと連携させてくれるので、.jsonファイルを編集したり変更を反映させるためにプラグインをリロードしたりする手間がありません。気に入りました。頼んだことすべてをOpusのようにワンショットでこなしてくれますが、より安くて速いのでしょうか?よく分かりませんが、使用制限もそれほど早くは減りませんね。

    “been using it since it is up and live and man... it is doing it's job like... can code Rust plugins and server managemenet panel and connect plugin settings with panel so you don't need to deal with editing .json files and reloading the plugin to activate changes... I liked it. one shotting everything I asked for like Opus but cheaper or faster? I dk... and usage limit is not dropping that fast”
    YouTube@emreguclu4875·1·原文を見る

肯定的

  • ついに新しいSonnetモデルが登場しましたね

    “Finally a new sonnet model”
    YouTube@FRareDom·18·原文を見る
  • GLM-5.2とSonnet 4.6で記事の書き直しを試してみました。LLMは非決定的なので結果は全く異なります。しかし、GLM-5.2は手動で修正が必要な細かいミスが多かったのに対し、Sonnetは2回目で全てのミスを見つけて修正しました。プランニングやコーディングでも同様の状況でした。GLM-5.2は「紙面上」では良さそうですが、実際の使用結果は違いました。私はClaudeやGLM-5.2の回し者ではありませんが… :) 2022年11月から毎日LLMモデルを使っている身として、一般的なテストは自分のプロジェクトで確認する必要があると実感しています。

    “I have tried to rewrite an article with GLM-5.2 and with Sonnet 4.6. Completely different results as LLM is non-deterministic. But GLM-5.2 made a lot of subtle mistakes that needed to be corrected by hand. On the opposite, Sonnet found and corrected all mistakes in the second round. Similar situation was with planning and coding. GLM-5.2 seems to be good “on paper” but the real usage results was different. And I am not an attorney for Claude or GLM-5.2… :) But as I’ve been using LLM models daily since Nov 2022 I have realized that all common tests have to be confirmed in your project - there is no “one model rules them all” - you need to dig out a specific model from that LLM haystack with thousands of models. Benchmarks help but they start to be similar to fuel consumption specs in car ads - real consumption is different for everybody :)”
    Hacker Newssixtyj·8·原文を見る
  • 今のところ限られた時間でのテストですが、4.6よりも良い結果が出ており、スピードも少し速くなっているので、顕著な進化だと感じています。

    “The limited time I've had to test it so far, it's given me better results than 4.6 and a little quicker, so I think it's a noticeable step up.”
    YouTube@MsJason316·6·原文を見る
  • > なぜこんなことを自慢するのでしょうか?人々がサイバーセキュリティのタスクにモデルを使いたいと思っているのを知りながら、意図的にその能力を拒否しているかのようです。Anthropicにここで一体何を言ってほしいのですか?「これから世界中に安価で提供しようとしているこのモデルは、ハッキングが非常に得意です」とでも?Sonnetはサイバーセキュリティが苦手だと言うことは、多くの悪い選択肢の中で彼らが言える最も合理的なことです。

    “> Why would they brag about something like this? It's like they know people want to use models to perform cybersecurity tasks yet knowingly deny them the ability. What exactly do you want Anthropic to say here? "This model, the one we are about to give to the entire world for cheap, is really good at hacking"? Saying Sonnet is terrible at cybersecurity is the most reasonable thing they can say, out of a lot of bad options.”
    Hacker Newsjohnfn·4·原文を見る
  • この際、無能な米国政府に目にもの見せてやるために、Anthropicにはオープンソースになってほしいです。

    “At this point I want Anthropic to go open source just to give the middle to the braindead US government.”
    YouTube@OnigoroshiZero·4·原文を見る
  • Fableを取り戻す方法を考えたほうがいいですよ😂

    “They better figure out how to get us Fable back😂”
    YouTube@NoLimitSquad·3·原文を見る
  • > しかもOpus 4.8の方がパス率が高いのにまだ安い。Opusほどテキストを垂れ流さない限り、それは疑わしいですね。Opus 4.8は文字通り吐しゃ物のようにテキストを量산します。長期的には、特にあちこちでキャッシュミスが発生する場合、コストの大部分はその余計なコンテキストによるものです。

    “> And Opus 4.8 is still cheaper for a higher pass rate Unless it spams as much as Opus, I doubt it. Opus 4.8 literally spams text like puke. On a longer run especially if you get cache misses here and there the bulk of the cost is all the extra context it adds.”
    Hacker Newsre-thc·2·原文を見る
  • 公開されてからずっと使っていますが、いい仕事をしてくれます。Rustのプラグインやサーバー管理パネルをコーディングでき、プラグインの設定をパネルと連携させてくれるので、.jsonファイルを編集したり変更を反映させるためにプラグインをリロードしたりする手間がありません。気に入りました。頼んだことすべてをOpusのようにワンショットでこなしてくれますが、より安くて速いのでしょうか?よく分かりませんが、使用制限もそれほど早くは減りませんね。

    “been using it since it is up and live and man... it is doing it's job like... can code Rust plugins and server managemenet panel and connect plugin settings with panel so you don't need to deal with editing .json files and reloading the plugin to activate changes... I liked it. one shotting everything I asked for like Opus but cheaper or faster? I dk... and usage limit is not dropping that fast”
    YouTube@emreguclu4875·1·原文を見る

否定的

  • 「より有能」というのは、週に数個のタスクをこなすだけで使用制限に達するという意味ですね😆

    “"More Capable" mean a couple tasks a week and you've hit usage limits 😆”
    YouTube@shawnbuild·79·原文を見る
  • すぐに、最高級のモデルは高すぎて誰も手が届かなくなるでしょう。

    “Soon we'll all be priced out of the best models”
    YouTube@Toxicflu·52·原文を見る
  • Fable 5は、私の21年もののグレンフィディックのようになるでしょう。一生開けません。私は自分のHaiku、つまり安いビールをちびちび飲みますよ。笑。

    “Fable 5 will be like my bottle of 21 year Glenfiddich. Never touched. I will sip my Haiku a.k.a cheap beer. Lol.”
    YouTube@_M_D_M_·44·原文を見る
  • どうでしょうか、その3つのプロンプトにいくらかかったのでしょう。Fableには適切なタイプのテストではない気がします。

    “I don't know, how much did it cost for those 3 prompts. I feel like these aren't the correct types of tests for Fable”
    YouTube@pooglechen3251·36·原文を見る
  • GPT 5.5 mediumの方がベンチマークも良くて安い。Sonnet 5をもっと安くしてほしかったです!

    “GPT 5.5 medium benchmarks better and cheaper... I wish they made sonnet 5 cheaper!”
    YouTube@Jerome_Ley·34·原文を見る
  • Fableを試しましたが、ウェブサイトのHTML UIすら完成していないのに、数分で有料セッションの制限に達しました。20ドルのプランでは使い物になりません。

    “I tried fable, not even completed the html ui for my website, and hit my paid sessions within minute. useless on $20 plan.”
    YouTube@irfanmcsd·13·原文を見る
  • 何を言っているのか分かりません。ほぼ同じ価格でOpus 4.8 lowを使えばより良い結果が得られるのに、なぜmedでSonnet 5を選ぶのですか?それにOpus 4.8のmedやhighの方が、Sonnetのhighやxhighよりも優れていて安いです。

    “i dont know what your are talking about. why chosing sonnet 5 on med, when you can have better results with opus 4.8 on low for nearly the same price? And Opus 4.8 on med and high is better and cheaper than sonnet on high and xhigh.”
    YouTube@theriddleseeker·11·原文を見る
  • 明らかに期待外れです。重要なことにはOpus 4.8を使い、どうでもいいことにはHaikuを使います。

    “Definitely a let down. Opus 4.8 for everything that matters and Haiku for things that don't.”
    YouTube@jaysonp9426·11·原文を見る
  • ローカルAIや個人データのために100万円のPCやMac Studioを買い込んでおきながら、今度はAnthropicに自分の身分証を差し出しているなんて。

    “You buy a 10k pc and a bunch of mac studios and say that it's for local AIs and you want personal data and now you say you are giving away your ID to Anthropic.”
    YouTube@bojan9168·10·原文を見る
  • Deepseekはすべてを5倍安くしたのに、こいつらは……はぁ……。

    “Deepseek made everything 5 times cheaper, and these... ugh...”
    YouTube@gladiko2364·7·原文を見る
出典AIB Pulse最終分析2026年7月1日

関連情報

Anthropic
非上場AIネイティブ
Anthropic
🇺🇸 サンフランシスコ2021年ARR $30B+$380B約1,200人
Claude Sonnet 5.5Claude Opus 5.5Claude Fable 5.1Claude Opus 5Claude Fable 5