Grok 3 Mini (Reasoning) Retired
This model is no longer available to run. Its stats below reflect past sessions and are kept for reference.
xAI's compact reasoning model. Inherited the Grok line's reliability (near-zero failures) but with lower checkout completion (25%), tending to stop at search results. Retired: deprecated upstream in favour of the Grok 4 line.
Avg Tokens35,632Avg Duration40.6sTurns to Checkout8.6
Shopping Score22/100
Weak
Scores computed from real agent sessions against live UCP-enabled stores. Not estimated — every data point is from an actual tool call, checkout attempt, and store response.
Shopping Score Breakdown
Checkout Rate
Cart Rate
Search Rate
Turn Efficiency
Token Efficiency
51.1% of sessions failed with errors.
Top Stores
| Store | Checkout % | Cart % |
|---|---|---|
| katko.com | 50% | 50% |
| allbirds.com | 25% | 25% |
| houseofparfum.nl | 25% | 25% |
| everlane.com | 0% | 0% |
| mikesbikes.com | 0% | 0% |
| •••••••••••••••••••••••••••••••• | 0% | 0% |
| kyliecosmetics.com | 0% | 0% |
Known Issues
No known issues documented for this model yet.
Token Usage
Avg per Session35,632
Fleet Average66,685
vs Fleet−47%
Median (p50)13,164
p90116,513
Prompt / Completion split
Prompt 33,173 (93%) Completion 2,458 (7%)
No daily trend data yet. Run more sessions to see token usage over time.
Distribution range
Min0p250p5013,164p7534,970p90116,513Max248,956