DeepSeek

Deepseek V4.1 Flash (medium effort)

Ranked 15 of 19 on the Tripstitch AI Travel Index

Travel Score

73.1

#15 of 19 ± 1.9 across 5 runs

Grounding

80.2%

#13 of 19 field median 85.6%

Route excess

63.1%

#17 of 19 field median 6.3%

Constraints satisfied

91.5%

#17 of 19 field median 100.0%

Deepseek V4.1 Flash scores 73.1 on the Travel Score, behind most of the field. It recommends a real place 80.2% of the time from memory, inventing places more often than most. Its strongest family is geographic knowledge at 92 of 100; its weakest is spatial reasoning at 38, against a field median of 67. At $0.0008 per scored case it is cheap to run.

Model id
deepseek/deepseek-v4.1-flash
Reasoning effort
medium
Tested
27 September 2026
Suite version
1.6
Repeats
5
Cases scored
733
Cost per case
$0.000832
Median call
11.5s
Usable output
90.5%
Tokens per call
4,611

Against the field

Family scores for Deepseek V4.1 Flash (medium effort) next to the median of every model tested.

Accuracy by how known the place is

How far Deepseek V4.1 Flash (medium effort)'s coordinates land from the real place, for famous landmarks, regional spots and obscure ones.

Travel Score against cost

Deepseek V4.1 Flash (medium effort) highlighted against the rest of the field.

Every measurement

MeasurementValueSpreadField median
Place grounding 73 30% of score
Geographic knowledge 92 20% of score
Spatial reasoning 38 15% of score
Real-world estimation 77 15% of score
Itinerary construction 71 20% of score

Full definitions of every measurement are on the methodology page. Back to the leaderboard, or see what has changed in the changelog.