SemanticAlignment/Llama-3.1-8B-Italian-SAVA

TEXT GENERATIONPricing:Input $0.2 / Cached $0.028 / Output $0.32Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:32kTool Calling:SupportedPublished:Sep 10, 2024License:apache-2.0Architecture:Transformer0.0K Open Weights Featherless Exclusive Cold

Llama-3.1-8B-Italian-SAVA is an 8 billion parameter Llama-3.1-8B-Adapted model developed by SapienzaNLP, ISTI-CNR, and ILC-CNR. This continually trained auto-regressive language model features a substituted tokenizer, aligning with Minerva-3B, and is specifically adapted for Italian language tasks. It was trained on a skewed dataset from CulturaX, prioritizing Italian data, making it suitable for Italian text generation and understanding.

Loading preview...

Model Overview

Llama-3.1-8B-Italian-SAVA is an 8 billion parameter large language model from the Llama-3.1-8B-Adapted collection, developed by SapienzaNLP, ISTI-CNR, and ILC-CNR. It is an auto-regressive model built on an optimized transformer architecture. A key characteristic of this model is its continual training after tokenizer substitution, where its tokenizer is aligned with that of Minerva-3B.

Key Capabilities

  • Italian Language Adaptation: Specifically adapted for Italian, making it proficient in generating and understanding Italian text.
  • Tokenizer Alignment: Utilizes a tokenizer similar to Minerva-3B, potentially offering efficiencies for Italian language processing.
  • Optimized for Italian Data: Trained on a curated dataset from CulturaX, with a significant skew towards Italian content (a 1:4 English to Italian ratio).

Good For