I ported the whole Gemma-4 family — E2B, E4B, 12B, 31B, and the 26B-A4B MoE — to run on AWS

Inferentia2. Each has its own write-up in this series; this is the map. What's shared, what's different,

how the recipe evolved from "trace it and pray" to a single-rank-compile pipeline that carries a

30-billion-parameter dense model and a 128-expert MoE — and the one bug that appears, in some costume, in

**every* port.*