casperhansen/llama-3-70b-fp16
TEXT GENERATIONConcurrency Cost:4Model Size:70BQuant:FP8Ctx Length:8kPublished:Apr 18, 2024License:llama3Architecture:Transformer0.0K Warm
The casperhansen/llama-3-70b-fp16 model is a 70 billion parameter large language model developed by Meta, part of the Llama 3 family. This auto-regressive transformer model is instruction-tuned for dialogue use cases, outperforming many open-source chat models on industry benchmarks. It features an 8k context length and utilizes Grouped-Query Attention (GQA) for improved inference scalability, making it suitable for commercial and research applications requiring high-performance conversational AI.
Loading preview...
Popular Sampler Settings
Top 3 parameter combinations used by Featherless users for this model. Click a tab to see each config.
temperature
–
top_p
–
top_k
–
frequency_penalty
–
presence_penalty
–
repetition_penalty
–
min_p
–