Thanks, I see this thread got flagged internally so @Neil_Dwyer may have additional feedback.
I should have picked up on this with my earlier answer, but since we only serve the gemma model from from US, I’m sure that will be a factor in the latency (and explain why your dashboard doesn’t match the numbers I’m seeing internally). Please see the thread below that clarifies our plans to bring the model to the EU soon: