Model icon
grok-4.6
Chat
Model:
Grok 4.6 is xAI's 500K-context frontier model for long-running agents, coding, knowledge work, and ambitious interactive projects.
Chat

This Playground is for demo purposes. Data is only valid in the current window and will be cleared on refresh.

Type a message...

Configuration
1.0
1.0
Pricing details

Transparent pricing with no hidden fees. Pay as you go.

xAIgrok-4.6
Input
PoYo price
$1.60/1M tokens
320 credits
Official price
$2.00/1M tokens
Official
You save
20%
xAIgrok-4.6
Output (including reasoning tokens)
PoYo price
$4.80/1M tokens
960 credits
Official price
$6.00/1M tokens
Official
You save
20%

* Actual fees are based on the final output.

Available on PoYo

Complete guide to using Grok 4.6 API for Long-Running Agents

Grok 4.6 API for Long-Running Agents

Grok 4.6 is xAI's frontier model for coding, engineering, knowledge work, and ambitious interactive or visual projects. It can stay with complex tasks across many steps, research unfamiliar domains, work through a codebase, and self-check results before moving forward.

Use grok-4.6 through /v1/chat/completions or /v1/responses. PoYo charges 320 input and 960 output credits per million tokens—$1.60 and $4.80, exactly 20% below xAI's $2 and $6 standard short-context rates.

Grok 4.6 API Features

The official xAI launch focuses on sustained agent work and stronger first passes for interactive and visual projects.

01

Build, Test, and Refine over Longer Trajectories

Grok 4.6 was trained for knowledge work, general coding, kernel optimization, web development, computer-aided design, and other agentic environments. xAI reports more self-testing and verification on longer trajectories.

  • 500,000-token context window
  • Text and image input with text output
  • Low, medium, high, and xhigh reasoning
  • Responses API context compaction for long agent loops
Official Grok 4.6 launch artwork from xAI

02

Turn Bigger Ideas into Interactive, Visual First Passes

xAI designed Grok 4.6 to take on more ambitious interactive and visual work. It can carry a broad brief through planning and implementation, then produce a more substantial first version of an application or polished work artifact.

  • Stronger first-pass implementations
  • Interactive web and UI projects
  • Visual applications and work artifacts
  • Continued refinement across multiple steps

03

Use One Model across Research, Code, and Delivery

Connect Grok 4.6 through either PoYo endpoint and keep the same public model ID from quick chat tests to long-running agent workflows. Combine streaming, tool calls, reasoning control, and context compaction without maintaining a second integration.

  • OpenAI-compatible Chat Completions
  • Responses API for stateful agent work
  • Function calling and external tools
  • Streaming and non-streaming responses
PoYo Grok 4.6 workflow cover

What Can You Build with Grok 4.6?

01

Long-Running Agents

Sustain research, planning, implementation, testing, and refinement across many steps without losing the objective.

02

Software Engineering

Navigate repositories, diagnose defects, implement changes, run tools, and verify results in agentic coding workflows.

03

Interactive and Visual Apps

Turn broad product ideas into substantial first versions with coherent structure, interactions, and visual language.

04

Knowledge Work

Analyze reports and complex information, synthesize evidence, and create polished work artifacts within a 500K context.

05

Tool-Using Systems

Combine function calling with web search, X search, code execution, structured outputs, and external application tools.

06

Adjustable Reasoning

Choose low, medium, high, or xhigh reasoning effort to match latency and depth to each task.

Grok 4.6 Benchmark Comparison

xAI's official high-reasoning evaluations compare Grok 4.6 with Grok 4.5, GPT-5.6 Sol Max, and Fable 5 Max across composite intelligence, coding agents, terminal work, and professional knowledge tasks.

BenchmarkGrok 4.6 HighGrok 4.5 HighGPT-5.6 Sol MaxFable 5 Max
AA Intelligence Index61566162
GDPVal-AA v21753152617281741
CursorBench v3.269.9%66.7%67.2%70.5%
DeepSWE v1.165.9%54%73%70%
FrontierCode v1.1 Extended61.3%56.6%60.6%63.6%
APEX-Agents57.5%47.1%56.7%59.2%
Terminal-Bench v3.026%15.7%34.6%34.1%
APEX-SWE56.4%53.6%—58.8%
AA-Briefcase1577131315021574
Harvey LAB (Vals)15.8%12.9%2.5%11.3%

Official xAI launch evaluations, published August 12, 2026. Third-party model scores are the best self-reported or publicly available results used by xAI.

How to Use Grok 4.6 on PoYo

1. Create an API key. Sign in and create a PoYo key.

2. Choose an endpoint. Use POST /v1/chat/completions with a messages array, or POST /v1/responses with input. Both support streaming.

3. Set the model. Send grok-4.6 as the model ID. For reliable prompt-cache affinity, keep a stable conversation identifier when your client supports it. Read the PoYo API documentation.

Grok 4.6 API Frequently Asked Questions

Is Grok 4.6 available on PoYo?

Yes. Use the grok-4.6 model ID with either supported endpoint.

Which endpoints does Grok 4.6 support?

PoYo supports POST /v1/chat/completions and POST /v1/responses, including streaming and non-streaming requests.

How much does Grok 4.6 cost on PoYo?

Input costs 320 credits ($1.60) per million tokens and output, including reasoning tokens, costs 960 credits ($4.80) per million tokens. Both rates are 20% below xAI's $2/$6 short-context rates.

Does the 20% comparison apply above 200K input tokens?

No. The displayed comparison is against xAI's standard short-context rate. xAI applies higher long-context rates once a prompt reaches 200,000 tokens; confirm workload-specific billing before very long requests.

What is the Grok 4.6 context window?

xAI lists a 500,000-token context window and no separate text output limit.

Can Grok 4.6 understand images?

Yes. The model accepts text and image input and produces text output.

Which reasoning levels are supported?

Grok 4.6 supports low, medium, high (the default), and xhigh reasoning effort.

What tools can Grok 4.6 use?

xAI lists function calling, web search, X search, and code execution. Tool availability and extra tool charges depend on the endpoint and configuration.

How is the Responses API different from Chat Completions?

Chat Completions uses a messages history. Responses uses input, can continue with prior response state, and supports context compaction for long-running workflows.

What is Grok 4.6's knowledge cutoff?

xAI lists February 1, 2026. Use current search or retrieval tools when the task needs newer information.

How does Grok 4.6 compare with Grok 4.5?

The official evaluations show gains across all ten listed benchmarks, with particular improvements in long-running agents, coding, knowledge work, self-testing, and first-pass visual projects.

Should I use prompt caching?

For repeated conversation prefixes and long agent loops, xAI recommends a stable prompt cache key or conversation ID. PoYo's displayed 320/960 rates do not advertise a separate cached-token tier.

Why Use Grok 4.6 on PoYo

01

Two OpenAI-Style Endpoints

Use Chat Completions or Responses with the same public model ID and PoYo credential.

02

Transparent 20% Savings

See input and output rates in both credits and US-dollar equivalents against official short-context pricing.

03

One PoYo Key

Evaluate Grok alongside other model providers without adding another gateway integration.

04

Playground and Monitoring

Switch endpoints, test streaming, review service status, and inspect token usage before production rollout.