fiveflow/rq_4b_64

TEXT GENERATIONPricing:Input $0.4 / Cached $0.08 / Output $0.8Concurrent Unit Cost:1Model Size:4BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Aug 30, 2026Architecture:Transformer Featherless Exclusive Cold

The fiveflow/rq_4b_64 is a 4 billion parameter language model with a 32,768 token context length. This model is provided by fiveflow, though specific architectural details and training objectives are not publicly available. It is intended for general language generation tasks where a compact model size and extended context window are beneficial.

Loading preview...

Model Overview

The fiveflow/rq_4b_64 is a 4 billion parameter language model developed by fiveflow. It features a substantial context length of 32,768 tokens, which allows it to process and generate longer sequences of text compared to models with smaller context windows. While specific details regarding its architecture, training data, and fine-tuning objectives are not provided in the available documentation, its parameter count suggests it is designed for efficient deployment.

Key Capabilities

  • Extended Context Window: The 32,768 token context length enables the model to maintain coherence and draw information from extensive input texts, making it suitable for tasks requiring deep contextual understanding.
  • Compact Size: With 4 billion parameters, it offers a balance between performance and computational efficiency, potentially allowing for faster inference and reduced resource consumption compared to larger models.

Good For

  • Applications requiring processing of long documents or conversations.
  • Scenarios where a balance between model capability and operational efficiency is crucial.
  • General text generation and understanding tasks where the extended context can be leveraged.