The casperhansen/llama-3-70b-fp16 model is a 70 billion parameter large language model developed by Meta, part of the Llama 3 family. This auto-regressive transformer model is instruction-tuned for dialogue use cases, outperforming many open-source chat models on industry benchmarks. It features an 8k context length and utilizes Grouped-Query Attention (GQA) for improved inference scalability, making it suitable for commercial and research applications requiring high-performance conversational AI.
No reviews yet. Be the first to review!