Dohyeon1/ERNIE-HC-SMoE-ngroups32

TEXT GENERATIONPricing:Input $0.32 / Cached $0.016 / Output $1.6Concurrent Unit Cost:1Model Size:21BQuant:FP8Context Size:32kPublished:Sep 19, 2026Architecture:Transformer Featherless Exclusive Cold

Dohyeon1/ERNIE-HC-SMoE-ngroups32 is a 21 billion parameter language model. This model is based on the Mixture-of-Experts (MoE) architecture, specifically using 32 groups, and supports a substantial context length of 32768 tokens. As a Hugging Face Transformers model, it is designed for general language understanding and generation tasks, though specific differentiators and primary use cases are not detailed in its current model card.

Loading preview...

Model Overview

The Dohyeon1/ERNIE-HC-SMoE-ngroups32 is a 21 billion parameter language model available on the Hugging Face Hub. This model leverages a Sparse Mixture-of-Experts (SMoE) architecture, configured with 32 expert groups, which typically allows for efficient scaling and potentially better performance on diverse tasks compared to dense models of similar parameter count. It is designed to handle long sequences with a context window of 32768 tokens, making it suitable for applications requiring extensive contextual understanding.

Key Characteristics

  • Architecture: Sparse Mixture-of-Experts (SMoE) with 32 expert groups.
  • Parameters: 21 billion.
  • Context Length: Supports up to 32768 tokens.
  • Framework: Implemented as a Hugging Face Transformers model.

Current Status and Limitations

The provided model card indicates that many details regarding its development, training data, specific language support, license, and intended use cases are currently marked as "More Information Needed." Therefore, specific performance benchmarks, training methodologies, and recommended applications are not yet available. Users should be aware that without further details, the model's exact capabilities, biases, risks, and optimal use scenarios remain to be fully documented.