ViewTube

ViewTube
Sign inSign upSubscriptions
Filters

Upload date

Type

Duration

Sort by

Features

Reset

32,289 results

IBM Technology
Reinforcement Learning from Human Feedback (RLHF) Explained

Want to play with the technology yourself? Explore our interactive demo → https://ibm.biz/BdKSby Learn more about the ...

11:29
Reinforcement Learning from Human Feedback (RLHF) Explained

99,355 views

2 years ago

StatQuest with Josh Starmer
Reinforcement Learning with Human Feedback (RLHF), Clearly Explained!!!

Generative Large Language Models, like ChatGPT and DeepSeek, are trained on massive text based datasets, like the entire ...

18:02
Reinforcement Learning with Human Feedback (RLHF), Clearly Explained!!!

67,622 views

1 year ago

Sebastian Raschka
Reinforcement Learning with Human Feedback (RLHF) in 4 minutes

Understanding Reinforcement Learning with Human Feedback (RLHF) – A short clip from my talk at the 2023 Optimized AI ...

4:06
Reinforcement Learning with Human Feedback (RLHF) in 4 minutes

17,076 views

1 year ago

Shaw Talebi
Fine-tuning LLMs on Human Feedback (RLHF + DPO)

Your team not maximizing Claude? I run 1:1 and team AI workshops for companies doing $10M+ per year: ...

28:53
Fine-tuning LLMs on Human Feedback (RLHF + DPO)

26,076 views

1 year ago

Zachary Huang
RLHF in 90 min

Don't like the Sound Effect?:* https://youtu.be/6xEXyJAbYns *LLM Training Playlist:* ...

1:30:36
RLHF in 90 min

7,610 views

11 months ago

Luis Serrano Academy
Reinforcement Learning with Human Feedback (RLHF) - How to train and fine-tune Transformer Models

Reinforcement Learning with Human Feedback (RLHF) is a method used for training Large Language Models (LLMs). In the heart ...

15:31
Reinforcement Learning with Human Feedback (RLHF) - How to train and fine-tune Transformer Models

37,894 views

2 years ago

Mark Hennings
RLHF Explained

Learn how Reinforcement Learning from Human Feedback (RLHF) actually works and why Direct Preference Optimization (DPO) ...

19:39
RLHF Explained

20,082 views

2 years ago

CodeEmporium
Reinforcement Learning through Human Feedback - EXPLAINED! | RLHF

We talk about reinforcement learning through human feedback. ChatGPT among other applications makes use of this. ABOUT ME ...

10:17
Reinforcement Learning through Human Feedback - EXPLAINED! | RLHF

30,466 views

2 years ago

freeCodeCamp.org
LLM Fine-Tuning Course – From Supervised FT to RLHF, LoRA, and Multimodal

Learn how to tailor massive models to specific tasks with this comprehensive, deep dive into the modern LLM ecosystem. You will ...

11:56:26
LLM Fine-Tuning Course – From Supervised FT to RLHF, LoRA, and Multimodal

100,439 views

5 months ago

Ashwani Kumar
RLHF from scratch, step-by-step, in code

Reinforcement Learning from Human Feedback (RLHF) has been instrumental in turning a pretrained large language model ...

3:14:37
RLHF from scratch, step-by-step, in code

4,154 views

1 year ago

Hugging Face
Reinforcement Learning from Human Feedback: From Zero to chatGPT

In this talk, we will cover the basics of Reinforcement Learning from Human Feedback (RLHF) and how this technology is being ...

1:00:38
Reinforcement Learning from Human Feedback: From Zero to chatGPT

190,206 views

Streamed 3 years ago

Nathan Lambert
RLHF Foundations, IFT, Reward Modeling, Rejection Sampling | Post-Training Course Lecture 2

These three methods, instruction-finetuning (IFT, also called supervised finetuning, SFT), reward modeling (creating a model that ...

49:49
RLHF Foundations, IFT, Reward Modeling, Rejection Sampling | Post-Training Course Lecture 2

8,584 views

4 months ago

Rubén López Hernández
RLHF (Aprendizaje reforzado con feedback humano)

RLHF significa "Aprendizaje por Reforzamiento a partir de Retroalimentación Humana". Es un tipo de aprendizaje automático ...

2:34
RLHF (Aprendizaje reforzado con feedback humano)

291 views

3 years ago

Unfold Data Science
Reinforcement Learning with Human Feedback (RLHF) | Reinforcement Learning with Human Feedback LLM

Reinforcement Learning with Human Feedback (RLHF) | Reinforcement Learning with Human Feedback LLM #RLHF #LLM #coding ...

25:03
Reinforcement Learning with Human Feedback (RLHF) | Reinforcement Learning with Human Feedback LLM

2,450 views

1 year ago

Lex Clips
Yann LeCun: Why RL is overrated | Lex Fridman Podcast Clips

Lex Fridman Podcast full episode: https://www.youtube.com/watch?v=5t1vTLU7s40 Please support this podcast by checking out ...

5:30
Yann LeCun: Why RL is overrated | Lex Fridman Podcast Clips

39,463 views

2 years ago

machinelearnear
[#49] Curso LLM-RLHF (3/n) - Reinforcement Learning from Human Feedback explicado por Data Scientist

Este va a ser el tercer video de una serie que estoy haciendo sobre modelos de lenguaje gigantes con el objetivo de poder ...

49:57
[#49] Curso LLM-RLHF (3/n) - Reinforcement Learning from Human Feedback explicado por Data Scientist

2,464 views

3 years ago

Graphics in 5 Minutes
Reinforcement Learning:  ChatGPT and RLHF

Reinforcement Learning from human feedback, and how it's used to help train large language models like ChatGPT. Part 3 of RL ...

6:31
Reinforcement Learning: ChatGPT and RLHF

25,986 views

3 years ago

Dwarkesh Patel
John Schulman (OpenAI Cofounder) — Reasoning, RLHF, & plan for 2027 AGI

John Schulman on how posttraining tames the shoggoth, and the nature of the progress to come... EPISODE LINKS ...

1:35:51
John Schulman (OpenAI Cofounder) — Reasoning, RLHF, & plan for 2027 AGI

191,343 views

2 years ago