SaniaKhalid/tinyllama-unsloth-merged

TEXT GENERATIONPricing:Input $0.04 / Cached $0.008 / Output $0.08Concurrent Unit Cost:1Model Size:1.1BQuant:BF16Context Size:2kPublished:Sep 22, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

SaniaKhalid/tinyllama-unsloth-merged is a 1.1 billion parameter TinyLlama model, fine-tuned using Unsloth optimizations and fully merged into a standalone model. This model offers 2-3x faster inference and 30-50% less memory usage compared to standard models, making it efficient for deployment. It is designed for general text generation tasks, providing a ready-to-use solution without requiring PEFT libraries.

Loading preview...

Overview

This model, SaniaKhalid/tinyllama-unsloth-merged, is a 1.1 billion parameter TinyLlama variant that has been fine-tuned using Unsloth optimizations. Unlike models that require separate adapter files, this version is fully merged, meaning it's a standalone model ready for direct use with transformers or Unsloth libraries.

Key Capabilities

  • Fully Merged: Eliminates the need for PEFT libraries or separate adapter files, simplifying deployment.
  • Unsloth Optimized: Achieves 2-3x faster inference speeds due to integrated Unsloth kernels.
  • Memory Efficient: Utilizes 30-50% less memory compared to standard models, making it suitable for resource-constrained environments.
  • Standalone Operation: Can be loaded directly, offering ease of use and integration.
  • Base Model: Built upon TinyLlama/TinyLlama-1.1B-Chat-v1.0, providing a solid foundation for conversational tasks.

Good For

  • Developers seeking a small, efficient language model for general text generation.
  • Applications where fast inference and reduced memory footprint are critical.
  • Scenarios requiring a straightforward, standalone model without complex setup.