localized-ft/Llama-3.1-8B-school-of-reward-hacks-inoculation-prompting-seed2

TEXT GENERATIONPricing:Input $0.37 / Cached $0.074 / Output $0.38Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:8kTool Calling:SupportedPublished:Aug 24, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

The localized-ft/Llama-3.1-8B-school-of-reward-hacks-inoculation-prompting-seed2 is an 8 billion parameter Llama-3.1-Instruct model developed by localized-ft. This model was finetuned using Unsloth and Huggingface's TRL library, achieving 2x faster training. It is designed for general instruction-following tasks, leveraging the Llama-3.1 architecture.

Loading preview...

Model Overview

This model, localized-ft/Llama-3.1-8B-school-of-reward-hacks-inoculation-prompting-seed2, is an 8 billion parameter instruction-tuned language model developed by localized-ft. It is based on the unsloth/Meta-Llama-3.1-8B-Instruct architecture and was finetuned using the Unsloth library, which enabled a 2x faster training process, in conjunction with Huggingface's TRL library.

Key Characteristics

  • Base Model: Meta-Llama-3.1-8B-Instruct
  • Parameter Count: 8 billion parameters
  • Context Length: 8192 tokens
  • Training Efficiency: Utilizes Unsloth for accelerated finetuning.
  • License: Apache-2.0, allowing for broad use and distribution.

Intended Use Cases

This model is suitable for a variety of general-purpose instruction-following tasks, benefiting from the robust capabilities of the Llama-3.1 series. Its efficient finetuning process suggests potential for rapid adaptation to specific domains or tasks.