xAI

Grok 4.6 (medium effort)

Ranked 6 of 15 on the Tripstitch AI Travel Index

Travel Score

85.0

#6 of 15 ± 1.9 across 5 runs

Grounding

88.3%

#5 of 15 field median 84.5%

Route excess

22.7%

#10 of 15 field median 10.5%

Constraints satisfied

100.0%

#1 of 15 field median 98.9%

Grok 4.6 scores 85.0 on the Travel Score, in the middle of the field. It recommends a real place 88.3% of the time from memory, inventing fewer places than most. Its strongest family is geographic knowledge at 91 of 100; its weakest is spatial reasoning at 59, against a field median of 59. At $0.0048 per scored case it is expensive relative to the field, with the slowest calls of any model here.

Model id
x-ai/grok-4.6
Reasoning effort
medium
Tested
12 August 2026
Suite version
1.6
Repeats
5
Cases scored
733
Cost per case
$0.0048
Median call
42.7s
Usable output
96.8%
Tokens per call
4,977

Against the field

Family scores for Grok 4.6 (medium effort) next to the median of every model tested.

Accuracy by how known the place is

How far Grok 4.6 (medium effort)'s coordinates land from the real place, for famous landmarks, regional spots and obscure ones.

Travel Score against cost

Grok 4.6 (medium effort) highlighted against the rest of the field.

Every measurement

MeasurementValueSpreadField median
Place grounding 91 30% of score
Geographic knowledge 91 20% of score
Spatial reasoning 59 15% of score
Real-world estimation 80 15% of score
Itinerary construction 90 20% of score

Full definitions of every measurement are on the methodology page. Back to the leaderboard, or see what has changed in the changelog.