Catálogo · Deep Learning · Aprendizaje por Refuerzo

LLM Post-Training: Fine-Tuning and Reinforcement Learning Basics

Name: LLM Post-Training: Fine-Tuning and Reinforcement Learning Basics
Price: 23000 CLP
Availability: InStock

Master the essentials of LLM post-training to align, specialize, and improve model safety using supervised fine-tuning and reinforcement learning techniques.

⏱ 1 h 20 min 📚 8 lecciones

Sobre este curso

Pre-trained large language models are powerful, but adapting them to specific tasks and aligning them with human preferences requires post-training. Understanding how to guide these models is essential for building safe, reliable, and specialized AI applications. In this text-based course, you will learn the fundamental concepts and practical workflows behind LLM post-training, moving from raw models to helpful, aligned AI assistants.

What you'll learn:
- Understand the key differences between pre-training, supervised fine-tuning (SFT), and reinforcement learning.
- Apply parameter-efficient fine-tuning (PEFT) methods like LoRA to adapt models with minimal computational resources.
- Explore Reinforcement Learning from Human Feedback (RLHF) and modern alignment alternatives like Direct Preference Optimization (DPO).
- Evaluate model behavior and safety to ensure outputs are helpful, honest, and harmless.
- Analyze code snippets and written walkthroughs to prepare datasets for custom fine-tuning tasks.

The course begins with foundational definitions of post-training paradigms before guiding you through data preparation, fine-tuning configurations, and alignment strategies. You will progress from theoretical concepts to reading and analyzing real-world implementation code.

This course is designed for software developers, data enthusiasts, and AI beginners who want to understand how LLMs are customized. No prior experience with advanced machine learning is required, though basic Python familiarity is helpful.

Start reading today to unlock the power of custom model alignment and post-training.

Lo que obtendrás

📜 Certificado de finalización
Añádelo a tu perfil de LinkedIn
💬 Tutor AI personal
¿Atascado en una lección? Pregúntale a tu tutor integrado lo que quieras, cuando quieras.
♾️ Acceso de por vida
Vuelve cuando quieras, sin caducidad
📱 Teléfono o computadora
Funciona en cualquier dispositivo
💸 Reembolso de 14 días
Sin preguntas
⚡ Breve y enfocado
1 h 20 min de contenido práctico

Reseñas

Aún no hay reseñas — sé el primero en compartir tu experiencia.

Otros también tomaron

💼 Listo para trabajar

Preguntas frecuentes

¿Qué necesito para tomar este curso? +

Solo un teléfono o computadora con internet. Sin instalaciones ni hardware especial.

¿Cómo pago? +

Con tarjeta a través de Stripe. No almacenamos datos de tarjeta — Stripe los gestiona de forma segura.

¿Puedo obtener un reembolso? +

Sí — reembolso completo en 14 días, sin preguntas.

¿Por cuánto tiempo tendré acceso? +

Para siempre. Una vez comprado, el curso es tuyo para revisarlo cuando quieras.

¿Obtendré un certificado? +

Sí. Al finalizar recibirás un certificado que puedes añadir a tu perfil de LinkedIn.

Diseñado para profesionales en

Tecnología Diseño Finanzas Marketing Salud Educación Hostelería Manufactura

LLM Post-Training: Fine-Tuning and Reinforcement Learning Basics

Sobre este curso

Lo que obtendrás

Reseñas

Escribir una reseña

Otros también tomaron

Fundamentos de Aprendizaje por Refuerzo Profundo

Aprendizaje por Refuerzo: Predicción y Control con Aproximación de Funciones

Introducción al Aprendizaje por Refuerzo: De Q-Learning a Deep RL

Aprendizaje profundo por refuerzo en Python: una introducción moderna

Preguntas frecuentes