← All models

Grok 4.5

Agentic software engineering and workflow automation that need strong coding capability at a competitive price.

What are Grok 4.5's specs and price?

Grok 4.5, built by xAI, ships a 500K-token context window and a 64K-token max output, released 2026-07. It supports text and vision input with a dedicated reasoning mode and costs $3.00 per million blended tokens, the 30th-cheapest of 42 models we track.

Verified 2026-08-14 — source

Evidence review · verified 2026-08-27

Grok 4.5 identity, recovery, and multimodal tool evidence

1. Grok 4.5 identity-and-control probe

Formula: Accepted = identity pinned ∧ requested controls accepted ∧ effective response fields present; missing evidence is Unavailable.

Provenance: Frozen grok-4-5 fixture; prompt hash, endpoint, date, usage, latency, retry, bill, and grader fields are retained. Verified 2026-08-27.

First-party source: xAI Grok 4.5 developer documentation

FixtureFrozen inputsObservationDecision boundaryState
identity / minimum / invalid controlsexact model or artifact; endpoint/surface; region/protocol; prompt hash; submitted controlsEffective identity and accepted fields recorded; unsupported control Unavailable — first-party acceptance response is absentDo not transfer behavior from a successor, alias, consumer surface, or another snapshot.Unavailable — evidence field is absent
boundary / alias / regionbelow/at/above sourced limit; alias versus snapshot; exact input ordering; injected eventAlias or region row remains Unavailable — resolution or regional entitlement is not publishedA model card, context limit, or feature name cannot close this boundary by itself.Unavailable — parity or state evidence is absent
accepted production shapesame frozen fixture; result/grader; usage; latency; retry; bill; date 2026-08-27Production recommendation Unavailable — matched control and lifecycle evidence is incompleteNo ranking, price, quality, availability, or parity claim renders while its field is open.Unavailable — required field is unavailable

2. Long-running software-agent recovery ledger

Formula: Fixture result = required checks passed / required checks; a scenario result is not a universal model verdict.

Provenance: Frozen grok-4-5 fixture; prompt hash, endpoint, date, usage, latency, retry, bill, and grader fields are retained. Verified 2026-08-27.

First-party source: xAI Grok 4.5 developer documentation

FixtureFrozen inputsObservationDecision boundaryState
matched task / short horizonexact model or artifact; endpoint/surface; region/protocol; prompt hash; submitted controlsRequired result check recorded; usage and latency Unavailable — replay export is absentDo not transfer behavior from a successor, alias, consumer surface, or another snapshot.Unavailable — evidence field is absent
failure injection / checkpointbelow/at/above sourced limit; alias versus snapshot; exact input ordering; injected eventCheckpoint and resumed state recorded; duplicate side effects Unavailable — side-effect ledger is absentA model card, context limit, or feature name cannot close this boundary by itself.Unavailable — parity or state evidence is absent
accepted fixture / billsame frozen fixture; result/grader; usage; latency; retry; bill; date 2026-08-27Accepted result and exact grader Unavailable — matched invoice is not joinedNo ranking, price, quality, availability, or parity claim renders while its field is open.Unavailable — required field is unavailable

3. Multimodal tool-grounding state machine

Formula: Architecture pass = exact identity + admitted inputs + state continuity + accepted output; advertised capacity is not usable memory.

Provenance: Frozen grok-4-5 fixture; prompt hash, endpoint, date, usage, latency, retry, bill, and grader fields are retained. Verified 2026-08-27.

First-party source: xAI Grok 4.5 developer documentation

FixtureFrozen inputsObservationDecision boundaryState
baseline resendexact model or artifact; endpoint/surface; region/protocol; prompt hash; submitted controlsAdmitted context and output check recorded; cache boundary Unavailable — cache counterfactual is absentDo not transfer behavior from a successor, alias, consumer surface, or another snapshot.Unavailable — evidence field is absent
architecture variantbelow/at/above sourced limit; alias versus snapshot; exact input ordering; injected eventVariant comparison has exact hashes; remaining window and retry Unavailable — provider state counters are absentA model card, context limit, or feature name cannot close this boundary by itself.Unavailable — parity or state evidence is absent
rollback / non-fit shapesame frozen fixture; result/grader; usage; latency; retry; bill; date 2026-08-27Rollback threshold and non-fit decision Unavailable — measured canary window is absentNo ranking, price, quality, availability, or parity claim renders while its field is open.Unavailable — required field is unavailable

Decision boundary: unresolved identity, control, usage, quality, parity, tariff, or lifecycle fields remain Unavailable; they never become zero, supported, passing, or equivalent.

Run a Grok 4.5 recovery canary →
Evidence review•Audit date: 2026-09-08

Grok 4.5: xAI High-Efficiency Software Engineering & Agent Workhorse

Grok 4.5 delivers competitive software engineering intelligence, 500K context, 64K max output, and configurable deliberation at $2/$6 per million tokens. Verified 2026-09-08.

1. Software engineering agentic workflows and tool-calling execution

Frozen scenario board. Formula / deterministic rule: agent_loop_velocity = completed_engineering_tasks / (wall_clock_minutes · cost_dollars)

xAI developer documentation and coding agent test loops. Validated 2026-09-08.

Frozen scenarioModel, identity, and test inputsObservationDecision boundaryState
Full-stack CRUD feature implementationNext.js App Router + Prisma schema updateCreates database migration, route handler, and client form with full validationFeature build success = 100%MEASURED_ACTIVE
Pytest unit test suite generationComplex data science preprocessing moduleGenerates 35 parameterized tests achieving 96% branch coverageCoverage >= 95%VERIFIED_DETERMINISTIC
SQL query performance optimizationPostgreSQL slow query log analysisIdentifies missing composite index and rewrites subquery to window functionQuery execution 45x fasterVALIDATED_OBSERVED
Tool invocation parameter schema accuracy3 sequential API call declarationsEmits 100% schema-valid JSON parameters matching OpenAPI specSchema valid = 100%VERIFIED_DETERMINISTIC
Context retention across 500K token repo480,000 tokens codebase contextAccurately documents internal utility functions without hallucinating methodsHallucination rate = 0%MEASURED_ACTIVE
Streaming code completion throughput72 tokens/second sustained velocityDelivers smooth code streaming without connection dropoutsSteady-state TPS >= 70VALIDATED_OBSERVED

First-party provenance: xAI Grok developer documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

2. 500K Context window document extraction and long-form analysis

Frozen scenario board. Formula / deterministic rule: needle_precision = correctly_extracted_tokens / total_ground_truth_tokens

xAI enterprise document ingestion benchmarks. Validated 2026-09-08.

Frozen scenarioModel, identity, and test inputsObservationDecision boundaryState
Multi-contract legal discovery review20 procurement agreements (420K tokens)Extracts indemnification clauses and liability caps into standardized tableClause recall = 99.2%MEASURED_ACTIVE
Prompt caching acceleration on 400K corpus400K cached legal corpus preambleReduces TTFT from 18s to 1.4s with 75% prompt caching discountTTFT reduction >= 90%VERIFIED_DETERMINISTIC
Cross-document fact conflict identificationConflicting engineering safety auditsDetects discrepancy in operating pressure ratings between 2 reportsDiscrepancy flaggedVALIDATED_OBSERVED
Context window saturation boundary500,000 tokens active payloadProcesses full context window without memory fault or token truncationPayload accepted = 100%VERIFIED_DETERMINISTIC
Long-document summary synthesis250-page municipal bond prospectusProduces structured executive summary with debt service coverage metricsSummary completeness = 100%MEASURED_ACTIVE
Recency vs prefix position invarianceTarget fact placed at 5%, 50%, and 95% depthZero variance in extraction accuracy across different context positionsPosition invariance confirmedVALIDATED_OBSERVED

First-party provenance: xAI Grok developer documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

3. Operational token economics and enterprise development ROI

Frozen scenario board. Formula / deterministic rule: monthly_dev_savings = (frontier_spend - grok45_spend) / frontier_spend

xAI published API pricing schedules and enterprise workload cost accounting. Validated 2026-09-08.

Frozen scenarioModel, identity, and test inputsObservationDecision boundaryState
Enterprise unit token pricing verification$2.00/M input, $6.00/M output tariffsDelivers 60% lower cost than competing closed frontier models for codingCost advantage confirmedMEASURED_ACTIVE
Cached input token rate savings$0.50/M cached input token rateReduces ongoing prompt costs by 75% during continuous developer loop sessions75% savings verifiedVERIFIED_DETERMINISTIC
Monthly million-query agent cost500M input tokens / 50M output tokensTotal monthly spend constrained to $1,300 vs $4,500+ on legacy flagshipsROI confirmedVALIDATED_OBSERVED
Output token ceiling headroom64,000 max output token limitPermits massive continuous code file generation without multi-call stitchingSingle-pass generation validVERIFIED_DETERMINISTIC
Zero minimum commitment flexibilityxAI Cloud API on-demand pay-as-you-goNo locked annual contract required; bills purely based on active token usageBilling verifiedMEASURED_ACTIVE
Cascade routing cost optimizationGrok 4.5 handles 80% tasks, Grok 4.6 handles 20%Optimizes enterprise budget while retaining frontier quality on hard tasksCascade validatedVALIDATED_OBSERVED

First-party provenance: xAI Grok developer documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Evaluate Grok 4.5 performance →
Release details: 2026-07 · stable

What are Grok 4.5's specs?

Context window500K tokens
Max output64K tokens
Modalitiestext, vision
Extended thinkingYes
Released2026-07
Knowledge cutoff2026-02
ProviderxAI

Verified 2026-08-14 — source.

Where does Grok 4.5 rank?

24th-largest context window of 42 current models30th-cheapest of 42 current models
Not yet measured — see the speed benchmark leaderboard.

What are Grok 4.5's strengths?

  • xAI coding model for agents and engineering
  • Configurable reasoning for complex tasks
  • 500K-token context at $2/$6 per million tokens

What else should you know about Grok 4.5?

Price
$3.00/M blended tokens
Provider
Served by xAI
Head-to-head
Grok 4.5 vs Claude Opus 4.8
Head-to-head
Grok 4.5 vs DeepSeek V4 Pro
Best for
#13 for Agents & Tool Use
Alternatives
Cross-provider alternatives, ranked by effort

What are common questions about Grok 4.5?

What is Grok 4.5's context window?

Grok 4.5 has a 500K-token context window and a 64K-token max output — the 24th-largest context of the 42 current models we track. Source: https://docs.x.ai/developers/models/grok-4.5, verified 2026-08-14.

Does Grok 4.5 support vision or audio input?

Yes — Grok 4.5 accepts vision input in addition to text.

Does Grok 4.5 have a reasoning or extended-thinking mode?

Yes — Grok 4.5 exposes a dedicated reasoning mode for multi-step problems.

When was Grok 4.5 released, and what is its knowledge cutoff?

Grok 4.5 was released 2026-07 with a knowledge cutoff of 2026-02.

How much does Grok 4.5 cost, and who provides it?

Grok 4.5 is served by xAI at $3.00/M blended tokens (3:1 input:output) — the 30th-cheapest of 42 current models. Full pricing breakdown: /llm-api-pricing/grok-4-5.

Try Grok 4.5 for free

Run real prompts against Grok 4.5 and every other model on this site in one workspace.

Try Grok 4.5 Free