Latency you can feel.
Shorter network paths can help a chat response start sooner. Türkiye-based compute keeps local workloads close; model size, queueing and prompt length still shape first-token time.
Measure first token, P50 and P95 from your application.