Gemma Chat vLLM on GKE

Ask me anything

Running on vLLM, served from a GPU node pool in your GKE cluster. Turn on Think to see the model's reasoning before it answers.

Enter to send · Shift+Enter for a new line