Fine-Tuning Open-Source LLMs (LLaMA, Mistral, Qwen, etc.) Training Course
Fine-tuning open-source large language models (LLMs) has become a best practice for organizations looking to tailor AI capabilities within secure, cost-effective, and private environments.
This instructor-led, live training (available online or onsite) is designed for intermediate-level machine learning practitioners and AI developers who want to fine-tune and deploy open-weight models such as LLaMA, Mistral, and Qwen for specific business or internal applications.
Upon completion of this training, participants will be able to:
- Grasp the ecosystem and key distinctions among open-source LLMs.
- Prepare datasets and fine-tuning configurations for models like LLaMA, Mistral, and Qwen.
- Execute fine-tuning pipelines using Hugging Face Transformers and PEFT.
- Evaluate, save, and deploy fine-tuned models in secure environments.
Format of the Course
- Interactive lecture and discussion.
- Ample exercises and practice.
- Hands-on implementation in a live-lab environment.
Course Customization Options
- To request a customized training for this course, please contact us to arrange.
Course Outline
Introduction to Open-Source LLMs
- What are open-weight models and why they're important
- Overview of LLaMA, Mistral, Qwen, and other community models
- Use cases for private, on-premise, or secure deployments
Environment Setup and Tools
- Installing and configuring Transformers, Datasets, and PEFT libraries
- Choosing appropriate hardware for fine-tuning
- Loading pre-trained models from Hugging Face or other repositories
Data Preparation and Preprocessing
- Dataset formats (instruction tuning, chat data, text-only)
- Tokenization and sequence management
- Creating custom datasets and data loaders
Fine-Tuning Techniques
- Standard full fine-tuning vs. parameter-efficient methods
- Applying LoRA and QLoRA for efficient fine-tuning
- Using Trainer API for quick experimentation
Model Evaluation and Optimization
- Assessing fine-tuned models with generation and accuracy metrics
- Managing overfitting, generalization, and validation sets
- Performance tuning tips and logging
Deployment and Private Use
- Saving and loading models for inference
- Deploying fine-tuned models in secure enterprise environments
- On-premise vs. cloud deployment strategies
Case Studies and Use Cases
- Examples of enterprise use of LLaMA, Mistral, and Qwen
- Handling multilingual and domain-specific fine-tuning
- Discussion: Trade-offs between open and closed models
Summary and Next Steps
Requirements
- An understanding of large language models (LLMs) and their architecture
- Experience with Python and PyTorch
- Basic familiarity with the Hugging Face ecosystem
Audience
- ML practitioners
- AI developers
Need help picking the right course?
southafrica@nobleprog.co.za or +27 (0)10 005 5793