We ran benchmarks for 30 days with real client loads in CDMX and Bogotá. Latency to US-East APIs is acceptable for async; for real-time sync, you need edge or aggressive caching.
Cost per 1M tokens varies up to 3x by provider and time of day. We document the real numbers to help size budgets.