Zed-fx/llama31-8b-sft

TEXT GENERATIONPricing:Input $0.37 / Cached $0.074 / Output $0.38Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:8kTool Calling:SupportedPublished:Jun 16, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

The Zed-fx/llama31-8b-sft is an 8 billion parameter instruction-tuned causal language model, finetuned from unsloth/meta-llama-3.1-8b-instruct-unsloth-bnb-4bit. Developed by Zed-fx, this model was trained using Unsloth and Huggingface's TRL library, resulting in a 2x faster finetuning process. It is designed for general instruction-following tasks, leveraging the efficiency of Unsloth for rapid deployment.

Loading preview...

Model Overview

The Zed-fx/llama31-8b-sft is an 8 billion parameter instruction-tuned language model developed by Zed-fx. It is finetuned from the unsloth/meta-llama-3.1-8b-instruct-unsloth-bnb-4bit base model, leveraging the Unsloth library for accelerated training.

Key Characteristics

  • Architecture: Llama 3.1 family, 8 billion parameters.
  • Training Efficiency: Finetuned 2x faster using Unsloth and Huggingface's TRL library.
  • Base Model: Built upon a 4-bit quantized version of Meta Llama 3.1 8B Instruct, optimized for efficient deployment.
  • Context Length: Supports an 8192-token context window.

Use Cases

This model is suitable for a variety of instruction-following applications where the efficiency of Unsloth-trained models is beneficial. Its 8B parameter size makes it a good candidate for tasks requiring a balance between performance and computational resources.