Robert Miles
YouTuber who explains AI safety ideas like corrigibility in plain language
A science communicator whose YouTube channel is the clearest free introduction to AI safety concepts — corrigibility, instrumental convergence, reward hacking — for people with no research background.
In the library
Nothing of theirs is filed in the hub yet. Browse the full library.
Recommended next
Hand-picked from the hub based on what Robert Miles covers.
What Is RLCD? Reinforcement Learning from Contrast Distillation
A plain-language explainer on RLCD, a way of aligning language models by learning from contrasting outputs rather than human ratings alone.
Why this: Covers AI safety and explainers too
Partnership on AI — resource library
Partnership on AI's open library of guidance, frameworks, and case studies on responsible AI: synthetic media, labor and the economy, AI safety, fairness, and inclusive AI development.
Why this: Also about AI safety
Two Minute Papers (YouTube)
This YouTube channel offers simple video breakdowns of artificial intelligence and machine learning research papers. You can watch these quick guides to stay informed about new AI developments and understand how the technology is evolving.
Why this: Also about explainers
Yoshua Bengio
Deep learning pioneer and Turing Award winner, now focused on AI risk. His site holds papers, talks and written positions; his Google Scholar list has the full publication record, most-cited first.
Why this: Also about AI safety
