Gemma 4 models are designed to deliver frontier-level performance at each size. They are well-suited for reasoning, agentic workflows, coding, and multimodal understanding.
"_comment": "Declares the MLX-layout tensors only. NVFP4 tensors use the compressed-tensors layout, which forces its own type and group size; group_size is deliberately omitted here so that default applies."