PKU-Baichuan-MLSystemLab/Llama3-PBM-Nova-70B

TEXT GENERATIONPricing:Input $3.5 / Cached $0.7 / Output $8.3Concurrent Unit Cost:4Model Size:70BQuant:FP8Context Size:8kTool Calling:SupportedPublished:Aug 13, 2024License:llama3Architecture:Transformer0.0K Featherless Exclusive Cold

Llama3-PBM-Nova-70B is a 70 billion parameter chat model developed by PKU-Baichuan-MLSysLab, based on the Llama3-70B architecture. It enhances the base model's performance through data deduplication, quality filtering, synthesis, Supervised Fine-Tuning (SFT), and Reinforcement Learning from Human Feedback (RLHF). This model is optimized for chat applications, demonstrating strong performance on benchmarks like Arena-Hard and Alpaca-Eval 2.0, surpassing Llama-3.1-70B-Instruct.

Loading preview...

Overview

PKU-Baichuan-MLSysLab's Llama3-PBM-Nova-70B is a 70 billion parameter chat model built upon the Llama3-70B base. Released in August 2024, this model has undergone a rigorous training process involving data deduplication, quality filtering, and data synthesis to optimize its performance for conversational tasks. It leverages both Supervised Fine-Tuning (SFT) and Reinforcement Learning from Human Feedback (RLHF) to significantly improve its capabilities.

Key Capabilities & Performance

  • Enhanced Chat Performance: Achieves a 74.5% score on Arena-Hard and 56.9% on Alpaca-Eval 2.0, outperforming Llama-3.1-70B-Instruct and Llama-3-70B-Instruct.
  • Robust Training: Benefits from advanced data processing and a combination of SFT and RLHF for improved conversational quality.

When to Use This Model

  • Chat Applications: Ideal for developing chatbots and conversational AI systems that require high-quality responses.
  • Benchmarking: Suitable for use cases where strong performance on general conversational benchmarks is critical.
  • Llama3 Ecosystem: A strong candidate for developers already working within the Llama3 framework who need an instruction-tuned and RLHF-optimized variant.