GOOGLE VS XAI

Gemini 3.1 Flash-Lite vs Grok 4.3

Gemini 3.1 Flash-Lite is about 2.8x cheaper than Grok 4.3 on a blended 3:1 input:output rate — the same 10K-in/2K-out request costs $0.0055 vs $0.0175. Gemini 3.1 Flash-Lite has the larger context window (1,048,576 vs 1,000,000 tokens).

Pricing verified

API pricing per 1M tokens

ModelInput / 1M tokensOutput / 1M tokensBlended (3:1 mix)
Gemini 3.1 Flash-Lite$0.25$1.50$0.563
Grok 4.3$1.25$2.50$1.563

Rates from official Google pricing and official xAI pricing, checked 2026-08-11.

What a real workload costs

Example: one request with 10,000 input tokens and 2,000 output tokens, then that request at 100,000 runs per month.

ModelPer request (10K in / 2K out)Per month (100K requests)Monthly delta
Gemini 3.1 Flash-Lite$0.0055$550—
Grok 4.3$0.0175$1,750+$1,200

Model your own prompt sizes and volumes in the token cost analyzer, or check caching savings with the prompt caching calculator.

Specs side by side

SpecGemini 3.1 Flash-LiteGrok 4.3
ProviderGooglexAI
Context window1,048,576 tokens1,000,000 tokens
Model classGeneral textReasoning
TokenizerProprietary (counts estimated)Proprietary (counts estimated)

Count tokens for each model on its own page: Gemini 3.1 Flash-Lite token counter and Grok 4.3 token counter. All Google models are on the Gemini Token Counter. All xAI models are on the Grok Token Counter.

FAQ

Gemini 3.1 Flash-Lite vs Grok 4.3 FAQ

Which is cheaper, Gemini 3.1 Flash-Lite or Grok 4.3?

Gemini 3.1 Flash-Lite is cheaper. At rates checked 2026-08-11, Gemini 3.1 Flash-Lite costs $0.25 input / $1.50 output per 1M tokens, while Grok 4.3 costs $1.25 input / $2.50 output — about a 2.8x difference on a typical 3:1 input-heavy workload.

How much does a real request cost on Gemini 3.1 Flash-Lite vs Grok 4.3?

A request with 10,000 input tokens and 2,000 output tokens costs about $0.0055 on Gemini 3.1 Flash-Lite and $0.0175 on Grok 4.3. At 100,000 such requests per month that is $550 vs $1,750 — a difference of $1,200 per month.

Which has the bigger context window, Gemini 3.1 Flash-Lite or Grok 4.3?

Gemini 3.1 Flash-Lite. It supports 1,048,576 tokens of context versus 1,000,000 tokens — roughly 786,000 English words in one request.

Do Gemini 3.1 Flash-Lite and Grok 4.3 count tokens the same way?

Not exactly. Gemini 3.1 Flash-Lite uses a proprietary tokenizer (counts are estimates) and Grok 4.3 uses a proprietary tokenizer (counts are estimates). The same prompt will tokenize differently, which shifts effective cost per word.