An evaluation of Google's new multi-modal Gemma 4 model family, testing its performance across various sizes ranging from compact E2B versions to larger mixture-of-experts (MoE) models. The article explores how these models handle vision, audio, reasoning, and code generation tasks on consumer-grade hardware using tools such as LM Studio.