Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...
| Type | Model |
| Section | Models & platforms / text+image->text |
| Pricing | paid (от $0.05/mo) |
| Platform | Self-hosted |
| Systems | api, python, self-hosted |
| Hosting | cloud |
| Site language | en |
| Vendor | |
| Launched | 2025-03-13 |