What it is
The job it was built for, and the job it is not for.
Gemma 3 4B is a small multimodal language model: vision-language in, text out, with support for over 140 languages and improved maths, reasoning and chat behaviour over earlier Gemma releases. It is built for high-volume, well-defined work such as classification, extraction, tagging, summarising and routing. It is not the model to reach for when a task needs deep multi-step reasoning or long-form authored output.