Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...
Specifications
- Provider
- Google DeepMind
- Type
- Open-source / open-weight
- Modality
- text+image->text
- Context window
- 131,072
- Knowledge cutoff
- 2024-08-31
- Released
- March 13, 2025
Capabilities
Input: textInput: imageOutput: text
Strengths
frequency_penaltylogit_biasmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetop_ktop_p
