casperhansen/llama-3-70b-fp16
TEXT GENERATIONConcurrency Cost:4Model Size:70BQuant:FP8Ctx Length:8kPublished:Apr 18, 2024License:llama3Architecture:Transformer0.0K Warm

The casperhansen/llama-3-70b-fp16 model is a 70 billion parameter large language model developed by Meta, part of the Llama 3 family. This auto-regressive transformer model is instruction-tuned for dialogue use cases, outperforming many open-source chat models on industry benchmarks. It features an 8k context length and utilizes Grouped-Query Attention (GQA) for improved inference scalability, making it suitable for commercial and research applications requiring high-performance conversational AI.

Loading preview...

Popular Sampler Settings

Top 3 parameter combinations used by Featherless users for this model. Click a tab to see each config.

temperature
top_p
top_k
frequency_penalty
presence_penalty
repetition_penalty
min_p