OpenAI

GPT 6 Luna (medium effort)

Ranked 5 of 19 on the Tripstitch AI Travel Index

Travel Score

89.3

#5 of 19 ± 0.5 across 5 runs

Grounding

92.5%

#1 of 19 field median 85.6%

Route excess

4.8%

#9 of 19 field median 6.3%

Constraints satisfied

100.0%

#1 of 19 field median 100.0%

GPT 6 Luna scores 89.3 on the Travel Score, among the strongest tested. It recommends a real place 92.5% of the time from memory, inventing fewer places than most. Its strongest family is place grounding at 98 of 100; its weakest is spatial reasoning at 75, against a field median of 67. At $0.0002 per scored case it is cheap to run.

Model id
openai/gpt-6-luna
Reasoning effort
medium
Tested
27 September 2026
Suite version
1.6
Repeats
5
Cases scored
733
Cost per case
$0.000162
Median call
8.9s
Usable output
100.0%
Tokens per call
1,617

Against the field

Family scores for GPT 6 Luna (medium effort) next to the median of every model tested.

Accuracy by how known the place is

How far GPT 6 Luna (medium effort)'s coordinates land from the real place, for famous landmarks, regional spots and obscure ones.

Travel Score against cost

GPT 6 Luna (medium effort) highlighted against the rest of the field.

Every measurement

MeasurementValueSpreadField median
Place grounding 98 30% of score
Geographic knowledge 87 20% of score
Spatial reasoning 75 15% of score
Real-world estimation 77 15% of score
Itinerary construction 98 20% of score

Full definitions of every measurement are on the methodology page. Back to the leaderboard, or see what has changed in the changelog.