AIREITER
GrokText Chat

Grok 4.5

Try Grok 4.5 online for coding, agent workflows, structured analysis, streaming output, prompt caching, and Chat Completions API access.

InputOfficial $2.00 per 1M tokensAIReiter $0.40 per 1M tokensOutputOfficial $6.00 per 1M tokensAIReiter $1.20 per 1M tokensCache readOfficial $0.30 per 1M tokensAIReiter $0.06 per 1M tokens
Run with API

INPUT

OUTPUT

Example
Generated in
42.7 seconds
Input tokens
134
Output tokens
2354
Tokens per second
55.13 tokens / second
Time to first token
-

Model details

Use the same model key in Playground, API requests, and internal workflows.

Model ID
grok-4.5
Provider
Grok
Protocol
OpenAI Chat Completions
Context window
500,000 tokens
Max output
8,192 tokens
Input tokens
40 credits / 1M tokens
Output tokens
120 credits / 1M tokens
Cache read
6 credits / 1M tokens
Cache write
-

What You Can Do with Grok 4.5

A coding and agentic workflow model with configurable generation and prompt caching.

Coding and debugging

Review implementation plans, trace defects, explain unfamiliar code, and propose focused fixes.

Agentic workflows

Plan multi-step tasks and evaluate the outputs of tool-driven software and engineering flows.

Structured analysis

Return decisions, risks, and action lists in a format that downstream systems can consume.

Cache-aware API traffic

Reuse stable prompt prefixes and verify savings through prompt_tokens_details.cached_tokens.

Grok 4.5 Use Cases

Best suited to engineering work that benefits from reasoning and iterative refinement.
01

Pull request review

Find risky assumptions, missing tests, and compatibility problems before a change ships.

02

Incident investigation

Combine logs, symptoms, and code context into ranked causes and a verification plan.

03

Developer assistants

Support implementation, documentation, test design, and technical explanations in one workflow.

04

API automation

Run repeated analysis and transformation tasks with streaming output and visible token usage.

Should You Use Grok 4.5 or Grok 4.6?

Keep this page useful for existing Grok 4.5 integrations while making the upgrade path explicit.

Keep Grok 4.5 when

Your evaluation set already passes, the integration is stable, and changing models would add regression risk.

Evaluate Grok 4.6 when

You are starting a new coding workflow or want to retest difficult agent tasks against the newer route.

Do not migrate on a headline

Compare both models on your own code, latency target, cache behavior, and accepted answer rate.

Use routing when useful

Keep predictable traffic on 4.5 and send only harder or newly evaluated work to 4.6.

How to Use Grok 4.5

Test the same workload you plan to send through the API.

01

Run a representative prompt

Use real code or workflow context instead of a generic benchmark question.

02

Check output and usage

Inspect completion quality, latency, prompt tokens, output tokens, and cached input tokens.

03

Connect Chat Completions

Call POST https://aireiter.com/api/v1/chat/completions with model "grok-4.5" and stream=true when needed.

Grok 4.5 API Questions

Practical questions about compatibility, caching, and upgrading.

/ 01

Is Grok 4.5 still worth using?

Yes for workloads already validated on it. New integrations should compare it directly with Grok 4.6 before choosing a default.

/ 02

Which API protocol does this route use?

Use the OpenAI-compatible Chat Completions endpoint with model "grok-4.5".

/ 03

What context window is listed?

The page lists a 500K-token context window. Requests at or above 200K may have different upstream economics, so inspect the current AIReiter rate.

/ 04

How do I confirm prompt caching?

Check usage.prompt_tokens_details.cached_tokens. A repeated request without a positive cached token count is not a confirmed hit.

/ 05

How should I compare Grok 4.5 and 4.6?

Use the same code task, response limit, and acceptance rubric, then compare pass rate, latency, and actual token usage.