Wade5/MyModel2

TEXT GENERATIONPricing:Input $0.04 / Cached $0.008 / Output $0.08Concurrent Unit Cost:1Model Size:1.5BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Feb 13, 2025License:mitArchitecture:Transformer Open Weights Featherless Exclusive Cold

MyModel2 by Wade5 is a 1.5 billion parameter causal language model, fine-tuned from deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B. It features a 32768 token context length and is available in both SafeTensors and GGUF formats, enabling efficient inference with tools like llama.cpp and ctransformers. This model is suitable for various natural language processing tasks, leveraging its fine-tuned architecture for general-purpose applications.

Loading preview...

MyModel2 Overview

MyModel2 is a 1.5 billion parameter causal language model, fine-tuned by Wade5 from the deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B base model. It offers a substantial context length of 32768 tokens, making it capable of processing longer sequences of text. This model is provided in both SafeTensors and GGUF formats, enhancing its versatility for deployment across different environments.

Key Capabilities

  • Efficient Inference: The GGUF format allows for optimized inference using popular tools like llama.cpp and ctransformers, making it accessible for local and edge deployments.
  • General NLP Tasks: Designed for a variety of natural language processing applications, leveraging its fine-tuned base for broad utility.
  • Flexible Deployment: Availability in both SafeTensors and GGUF formats provides options for different hardware and software setups.

Good For

  • Developers seeking a compact yet capable language model for general NLP tasks.
  • Applications requiring efficient local inference, particularly on consumer hardware, due to GGUF support.
  • Projects that benefit from a model with a large context window for processing extensive text inputs.