Reinforcement Learning from Human Feedback (RLHF)

Add to wishlistAdded to wishlistRemoved from wishlist 0

Add to compare+

Duration	40m
level	Beginner
Course Creator	Jerry Kurata
Last Updated	31-Oct-23

Pluralsight

Category: Machine Learning

In this course we explore one corner of the expanding AI universe, and review some of the basic principles found in reinforcement learning from human feedback (RLHF), the technology underlying great AI tools such as ChatGPT, Bard, and more.

Add your review

Description
Reviews (0)

Have you ever wondered how tools like ChatGPT and Bard are able to generate great responses to the questions we pose? How they can respond to a prompt like “Plan a trip to Italy this fall and suggest great things to see,” and produce a response containing a full itinerary with places to see, the best time to visit, and the sites you shouldn’t miss? In this course, Reinforcement Learning from Human Feedback (RLHF), you’ll gain the ability to understand what is going on behind the scenes to create responses to your prompts. First, you’ll explore why having all the information available is not enough to create a great response. Next, you’ll discover how we teach a machine learning model to handle all that data and craft a response that people like. Finally, you’ll learn how none of it is magic, just some really great engineering by some bright people. When you’re finished with this course, you’ll have the skills and knowledge of reinforcement learning with human feedback needed to understand how this great engineering works and produces its amazing results.
Author Name: Jerry Kurata
Author Description:
Jerry has Bachelor of Science degrees in Geology and Physics. His plans to work in the oil exploration industry were sidetracked when he discovered he preferred to work with computers on simulation and data processing, instead of reading mud and core samples in the North Sea. His love of computers and tech resulted in him spending many additional hours working on computers while getting his Master’s degree in Computer Science. His current areas of interests include Machine Learning, Big Data,… more

Course Overview
1min
Understanding Text-generative Applications
6mins
What Is Wrong with the Pre-trained GPT Model?
5mins
Supervised Fine-tuning
4mins
Reward Model Training
11mins
Fine-tuning via Reinforcement Learning
5mins
Challenges and Limitations of RLHF
5mins

User Reviews

0.0 out of 5

★★★★★

Write a review

There are no reviews yet.

Be the first to review “Reinforcement Learning from Human Feedback (RLHF)” Cancel reply

Reinforcement Learning from Human Feedback (RLHF)

Description
Reviews (0)

Start Course

All Categories

Reinforcement Learning from Human Feedback (RLHF)

Table of Contents

User Reviews

Be the first to review “Reinforcement Learning from Human Feedback (RLHF)” Cancel reply

COURSE PROVIDERS

CATEGORIES

Quick Links

Contact Us

Compare items

All Categories

Reinforcement Learning from Human Feedback (RLHF)

Table of Contents

User Reviews

Be the first to review “Reinforcement Learning from Human Feedback (RLHF)” Cancel reply

Related Products

Getting Started with Natural Language Processing

Learn R

Data and Programming Foundations for AI

Machine Learning: Logistic Regression

Apply Natural Language Processing with Python

Build a Machine Learning Model

COURSE PROVIDERS

CATEGORIES

Quick Links

Contact Us

Compare items