google/translategemma-12b-it

VISIONPricing:Input $0.2 / Output $0.6Concurrent Unit Cost:1Model Size:12BQuant:FP8Context Size:32kPublished:Jan 12, 2026License:gemmaArchitecture:Transformer0.3K Gated Featherless Exclusive Cold

TranslateGemma is a family of lightweight, 12 billion parameter open translation models from Google, based on the Gemma 3 architecture. Designed to handle translation tasks across 55 languages, this model supports both text-to-text translation and text extraction and translation from images. Its relatively small size and 32768 token context length make it suitable for deployment in resource-limited environments, democratizing access to advanced translation capabilities.

Loading preview...

TranslateGemma: Lightweight Multilingual Translation

TranslateGemma is a family of 12 billion parameter open translation models developed by Google, built upon the Gemma 3 architecture. These models are specifically designed for efficient and accurate translation across 55 languages, making them highly versatile for global applications. A key differentiator is their ability to perform both direct text translation and text extraction and translation from images (normalized to 896x896 resolution).

Key Capabilities

  • Multilingual Translation: Supports translation for 55 languages, utilizing ISO 639-1 Alpha-2 codes or regionalized variants.
  • Image-to-Text Translation: Can extract text from images and translate it, offering a powerful feature for diverse use cases.
  • Resource-Efficient Deployment: Its lightweight nature allows for deployment on devices with limited resources, such as laptops, desktops, or private cloud infrastructure.
  • Flexible Input Handling: Features a specific chat template that supports both text and image inputs, with a total input context of 2K tokens.

Good For

  • On-device Translation: Ideal for applications requiring translation capabilities in environments with constrained computational resources.
  • Multilingual Content Processing: Excellent for translating text across a wide array of languages.
  • Visual Text Translation: Suitable for scenarios where text needs to be extracted and translated from images, such as signs or documents.
  • Innovation in Translation: Fosters innovation by providing accessible, state-of-the-art translation models for developers and researchers.