Moklas/gemma-3-270m

TEXT GENERATIONConcurrent Unit Cost:1Model Size:0.3BQuant:BF16Context Size:32kPublished:Jun 11, 2026Architecture:Transformer Featherless Exclusive Cold

Moklas/gemma-3-270m is a 0.3 billion parameter language model based on the Gemma architecture. This model is a smaller variant, likely intended for efficient deployment and tasks where computational resources are limited. Its compact size suggests suitability for on-device applications or rapid prototyping.

Loading preview...

Model Overview

Moklas/gemma-3-270m is a compact language model, part of the Gemma family, featuring 0.3 billion parameters. This model is designed for scenarios requiring a smaller footprint and faster inference times compared to larger models.

Key Characteristics

  • Parameter Count: 0.3 billion parameters, making it highly efficient.
  • Context Length: Supports a substantial context window of 32768 tokens, allowing it to process longer inputs despite its small size.
  • Architecture: Based on the Gemma architecture, known for its performance in various language tasks.

Good For

  • Resource-constrained environments: Ideal for deployment on devices with limited memory or processing power.
  • Rapid prototyping: Its small size enables quick experimentation and iteration.
  • Specific, narrow tasks: Potentially suitable for fine-tuning on highly specialized tasks where a larger model might be overkill.