Grok 4.2 is xAI’s latest large language model, built for strong reasoning, multimodal understanding, and enterprise use. It improves instruction following, honesty, and calibration over earlier Grok versions, while supporting both single‑agent and multi‑agent workflows. Designed as a general‑purpose, truth‑seeking assistant, Grok 4.2 is well suited for research, analysis, coding, and complex professional tasks when deployed with appropriate guardrails.
Pricing
- Input Tokens: $2 /M tokens
- Output Tokens: $6 /M tokens
- Cache Read: $0.2 /M tokens
Input Modalities
- Text
- Vision
Output Modalities
- Text
Context length
- 1M tokens
Capabilities
- Thinking
- Streaming
- Tool calling
- Web search
- URL context
- Code interpreter
- Computer use
- File search
- Memory tool
- Structured outputs
- Citations
- Prompt caching
- Background mode
- Server-side sessions
Providers
Azure grok-4-20-reasoning
Pricing$2$6
Cache$0.2
Context2M
Max output2M
Latency2.9S
Throughput42.3TPS
Uptime
100.00% uptime 3 days ago
100.00% uptime 2 days ago
100.00% uptime yesterday
Grok grok-4-20-reasoning
Pricing$2$6
Cache$0.2
Context2M
Max output2M
Latency0.5S
Throughput78.1TPS
Uptime
0.00% uptime 3 days ago
0.00% uptime 2 days ago
0.00% uptime yesterday
Performance for grok-4-20-reasoning
Uptime is the percentage of requests that succeeded over the past 72 hours. AIHubMix continuously monitors every provider and automatically retries with the next-best provider when one returns an error or responds too slowly; Latency is total round-trip time (lower is better); Throughput is how fast the model writes (tokens per second, higher is better).
Uptime
Loading...
Latency
Loading...
Throughput
Loading...
Try this model
Python
