vibehacker

GKE Inference

Skill

Deploy and optimize AI inference workloads on GKE with GPUs and TPUs

by GoogleCoding & Dev ToolsAgents & Automation Listed Mar 31, 2026

Clone or download it from github.com, then drop the folder into ~/.claude/skills/ (Claude Code) or your agent's skills directory.

Product badge

Share this product's name and rating in your README or on your website.

GKE Inference: rating on VibeHacker

Updates may be delayed by image caching.

GKE Inference cover

About GKE Inference

Official Google skill for deploying LLM/inference workloads on GKE with GPUs/TPUs, Inference Quickstart patterns, and model-server configuration—not for batch HPC or AlloyDB RAG migrations.

Used GKE Inference?

Log in to write a review.

No reviews yet

Used it? Write the first review.

Similar skills

View all

ASP.NET Core

Skill

Build and review ASP.NET Core apps with current official guidance