OpenAI

GPT 5.6 Luna (medium effort)

Ranked 4 of 13 on the Tripstitch AI Travel Index

Travel Score

86.4

#4 of 13 ± 0.3 across 5 runs

Grounding

89.3%

#1 of 13 field median 84.5%

Route excess

3.5%

#4 of 13 field median 10.5%

Constraints satisfied

98.4%

#8 of 13 field median 98.6%

GPT 5.6 Luna scores 86.4 on the Travel Score, among the strongest tested. It recommends a real place 89.3% of the time from memory, inventing fewer places than most. Its strongest family is place grounding at 98 of 100; its weakest is spatial reasoning at 76, against a field median of 59. At $0.0017 per scored case it is cheap to run.

Model id
openai/gpt-5.6-luna
Reasoning effort
medium
Tested
6 August 2026
Suite version
1.6
Repeats
5
Cases scored
733
Cost per case
$0.0017
Median call
8.2s
Usable output
100.0%
Tokens per call
3,049

Against the field

Family scores for GPT 5.6 Luna (medium effort) next to the median of every model tested.

Accuracy by how known the place is

How far GPT 5.6 Luna (medium effort)'s coordinates land from the real place, for famous landmarks, regional spots and obscure ones.

Travel Score against cost

GPT 5.6 Luna (medium effort) highlighted against the rest of the field.

Every measurement

MeasurementValueSpreadField median
Place grounding 98 30% of score
Geographic knowledge 83 20% of score
Spatial reasoning 76 15% of score
Real-world estimation 79 15% of score
Itinerary construction 86 20% of score

Full definitions of every measurement are on the methodology page. Back to the leaderboard, or see what has changed in the changelog.