Today we conclude our Black in AI series with Sicelukwanda Zwane, a masters student at the University of Witwatersrand and graduate research assistant at the CSIR.
At the workshop, he presented on “Safer Exploration in Deep Reinforcement Learning using Action Priors,” which explores transferring action priors between robotic tasks to reduce the exploration space in reinforcement learning, which in turn reduces sample complexity. In our conversation, we discuss what “safer exploration” means in this sense, the difference between this work and other techniques like imitation learning, and how this fits in with the goal of “lifelong learning.”
The complete show notes for this episode can be found at https://twimlai.com/talk/235. To follow along with the Black in AI series, visit https://twimlai.com/blackinai19.
Supercharging Developer Productivity with ChatGPT and Claude with Simon Willison - #701
Automated Design of Agentic Systems with Shengran Hu - #700
The EU AI Act and Mitigating Bias in Automated Decisioning with Peter van der Putten - #699
The Building Blocks of Agentic Systems with Harrison Chase - #698
Simplifying On-Device AI for Developers with Siddhika Nevrekar - #697
Genie: Generative Interactive Environments with Ashley Edwards - #696
Bridging the Sim2real Gap in Robotics with Marius Memmel - #695
Building Real-World LLM Products with Fine-Tuning and More with Hamel Husain - #694
Mamba, Mamba-2 and Post-Transformer Architectures for Generative AI with Albert Gu - #693
Decoding Animal Behavior to Train Robots with EgoPet with Amir Bar - #692
How Microsoft Scales Testing and Safety for Generative AI with Sarah Bird - #691
Long Context Language Models and their Biological Applications with Eric Nguyen - #690
Accelerating Sustainability with AI with Andres Ravinet - #689
Gen AI at the Edge: Qualcomm AI Research at CVPR 2024 with Fatih Porikli - #688
Energy Star Ratings for AI Models with Sasha Luccioni - #687
Language Understanding and LLMs with Christopher Manning - #686
Chronos: Learning the Language of Time Series with Abdul Fatir Ansari - #685
Powering AI with the World's Largest Computer Chip with Joel Hestness - #684
AI for Power & Energy with Laurent Boinot - #683
Controlling Fusion Reactor Instability with Deep Reinforcement Learning with Aza Jalalvand - #682
Create your
podcast in
minutes
It is Free
20/20
The Dropout
10% Happier with Dan Harris
World News Tonight with David Muir
NEJM This Week