Using RLHF to train champion-level drone racing agents

https://www.nature.com/articles/s41586-023-06419-4

A cool application of RLHF (Reinforcement Learning w/ Human Feedback - the same approach as what OpenAI used to train ChatGPT).

The authors trained an agent to fly FPV drones at a level surpassing world champions.

3 points · 0 comments · view on lemmy.world

0 Comments

No comments yet.