I'm curious to hear more about this. My experience has been that inference speeds are the #1 cause of delay by orders of magnitude, and I'd assume those won't go down substantially on edge devices because the cloud will be getting faster at approximately the same rate.
Have people outside the US benchmarked OpenAI's response times and found network lag to be a substantial contributor to slowness?
I'm curious to hear more about this. My experience has been that inference speeds are the #1 cause of delay by orders of magnitude, and I'd assume those won't go down substantially on edge devices because the cloud will be getting faster at approximately the same rate.
Have people outside the US benchmarked OpenAI's response times and found network lag to be a substantial contributor to slowness?