Grok 4.20 (non-reasoning)
- API model IDs
- grok-4.20-0309-non-reasoninggrok-4.20-non-reasoninggrok-4.20-non-reasoning-latest
- Context window
- 1,000,000 tokens
- Max output
- 1,000,000 tokens
- Released
- 2026-03-16
- Lifecycle
- active
- License
- proprietary
Lifecycle verified 2026-09-08 (0 days ago). Source
Data handling
- Trains on API data by default
- No
- Retention
- 30 days
- Zero data retention
- yes
- Residency
- us
- HIPAA BAA
- Available
- SOC 2 / FedRAMP
- SOC 2 yes, FedRAMP Not published
xAI's API security FAQ states that xAI never trains on API inputs or outputs without explicit permission, and that requests and responses are stored encrypted at rest for 30 days for abuse-auditing purposes before automatic deletion. Zero data retention is a self-serve toggle a team admin can enable from the xAI Console, applying automatically to all of that team's API keys. xAI states it is SOC 2 Type 2 compliant and offers a Business Associate Agreement on request via a BAA questionnaire; no FedRAMP authorization is published. The Data Processing Addendum states personal information is processed in the United States; this audit could read that page's existence and title but the live fetch returned a 403 during the audit, so the residency claim here rests on the DPA's indexed description rather than a direct read.
Verified 2026-09-08 (0 days ago). Source 1, Source 2
Pricing snapshot
$1.25 input, $2.5 output, per million tokens. Cached input $0.2 per million tokens.
This is a snapshot verified on 2026-09-08, 0 days ago, not a live price. Confirm on the provider’s pricing page before you budget against it.
Where you can call it
- direct · grok-4.20-0309-non-reasoning First-party xAI API.
xAI does not publish a max output token figure for this model separate from context window on its model-card page; maxOutputTokens is recorded here as equal to the published context window, consistent with the explicit "no text output limit" statement on Grok 4.5 and Grok 4.6's model cards. Prompts at or above 200,000 tokens bill at $2.50 input, $0.40 cached input, and $5 output per million tokens instead of the standard $1.25, $0.20, and $2.50.