xAIプロプライエタリ

Grok 4.5

このモデルを比較

xAIのfoundationモデル。

シェア:XはてブLINE

パラメータ

非公開

コンテキスト長

ライセンス

プロプライエタリ

リリース日

2026-07-08

日本語性能

高品質日本語

多言語対応モデルのうち、日本語処理に優れた性能を持つモデル。

API料金

入力料金(1Mトークンあたり)

$2

出力料金(1Mトークンあたり)

$

課金モード: standard

強み

    弱み

      活用例

        深度分析

        Artificial Analysis Intelligence Index

        54

        #4 of 168 models; behind Fable 5 (60), Opus 4.8 (56), GPT-5.5 (55)

        Agentic Tool Use (τ³-Banking)

        33%

        #1 overall, ahead of GPT-5.5 (31%) and Claude Sonnet 4.6 (31%)

        Coding Agent Index (Grok Build)

        76

        Tied with GPT-5.5 (Codex), 1 point behind Fable 5 (Claude Code)

        Input / Output Price

        $2.00 / $6.00 per 1M

        60-75% cheaper than Opus 4.8 and GPT-5.5

        Intelligence Index Task Cost

        $0.31 per task

        5x cheaper than Claude Sonnet 5 (max) with higher score

        Token Efficiency

        1.9M tokens/task

        vs 6.2M (GPT-5.5 Codex) and 7.2M (Fable 5 Claude Code)

        強み

        • Best-in-class agentic tool use (#1 on τ³-Banking) with exceptional token efficiency across agent workflows
        • Dramatically cheaper than frontier peers: $2/$6 per 1M tokens vs $5-$10 input and $25-$50 output for competitors
        • Fast output (~80-86 tokens/sec) and highly concise, completing tasks with 4x fewer tokens than comparable models

        弱み

        • Not the smartest model available — #4 on Intelligence Index; trails Fable 5 and Opus 4.8 on raw reasoning and harder coding benchmarks
        • Hallucination rate rose to 54% (up from 25% on Grok 4.3) as accuracy improved — more confident when wrong
        • Context window reduced to 500K tokens (down from Grok 4.3's 1M); no batch discount at launch; EU availability delayed

        競合比較

        ModelPrice
        Claude Fable 5 (max)$10/$50
        Claude Opus 4.8 (max)$5/$25
        GPT-5.5 (xhigh)$5/$30

        Grok 4.5, released July 8, 2026 by SpaceXAI (the merged SpaceX/xAI entity), represents a significant leap forward for xAI's model lineup, jumping 16 points on the Artificial Analysis Intelligence Index from Grok 4.3's score of 38 to 54, placing it fourth overall behind only Claude Fable 5, Claude Opus 4.8, and GPT-5.5. The model was jointly trained with Cursor on trillions of tokens of real developer-agent interaction data, following SpaceX's $60 billion acquisition of Anysphere (Cursor's maker) in June 2026. It is a 1.5 trillion parameter mixture-of-experts model (per Musk's disclosure, not officially confirmed by SpaceXAI) trained on tens of thousands of NVIDIA GB300 GPUs.

        Grok 4.5's positioning is deliberate: it does not claim to be the absolute smartest model, but rather the best intelligence-per-dollar option in the near-frontier tier. Its standout achievement is the #1 ranking on agentic tool use (τ³-Banking at 33%), and its token efficiency is remarkable — averaging just 1.9 million tokens per coding agent task compared to 6.2-7.2 million for competitors. At $2/$6 per 1M input/output tokens, it costs 60-75% less than Claude Opus 4.8 or GPT-5.5, while completing tasks at roughly $2.49 per Coding Agent Index task versus $5.07 for GPT-5.5 and $11.80 for Fable 5.

        The model ships as the default in Grok Build (xAI's coding agent CLI), is available across all Cursor plans, and is integrated into Microsoft Office add-ins for Word, PowerPoint, and Excel. Access is also available through OpenRouter, Vercel, Cloudflare, Snowflake, and Databricks Mosaic. The trade-offs are real: it trails the top models on raw coding accuracy (64.7% on SWE-Bench Pro vs. Fable 5's 80.4%), its hallucination rate has increased, and it reduced the context window from 1M to 500K tokens. But for high-volume agentic workloads where cost and speed matter as much as peak accuracy, Grok 4.5 has established itself as the most compelling value proposition on the market.

        分析生成日: 2026-07-17