In this episode, we’re joined by Hanbyul Joo, a PhD student in the Robotics Institute at Carnegie Mellon University.
Han, who is on track to complete his thesis at the end of the year, is working on what is called the “Panoptic Studio,” a multi-dimension motion capture studio with over 500 camera sensors that are used to capture human body behavior and body language. While robotic and other artificially intelligent systems can interact with humans, Han’s work focuses on understanding how humans interact and behave so that we can teach AI-based systems to react to humans more naturally. In our conversation, we discuss his CVPR best student paper award winner “Total Capture: A 3D Deformation Model for Tracking Faces, Hands, and Bodies.” Han also shares a complete overview of the Panoptic studio, and we dig into the creation and performance of the models, and much more.
For the complete show notes for this episode, visit https://twimlai.com/talk/180.
Supercharging Developer Productivity with ChatGPT and Claude with Simon Willison - #701
Automated Design of Agentic Systems with Shengran Hu - #700
The EU AI Act and Mitigating Bias in Automated Decisioning with Peter van der Putten - #699
The Building Blocks of Agentic Systems with Harrison Chase - #698
Simplifying On-Device AI for Developers with Siddhika Nevrekar - #697
Genie: Generative Interactive Environments with Ashley Edwards - #696
Bridging the Sim2real Gap in Robotics with Marius Memmel - #695
Building Real-World LLM Products with Fine-Tuning and More with Hamel Husain - #694
Mamba, Mamba-2 and Post-Transformer Architectures for Generative AI with Albert Gu - #693
Decoding Animal Behavior to Train Robots with EgoPet with Amir Bar - #692
How Microsoft Scales Testing and Safety for Generative AI with Sarah Bird - #691
Long Context Language Models and their Biological Applications with Eric Nguyen - #690
Accelerating Sustainability with AI with Andres Ravinet - #689
Gen AI at the Edge: Qualcomm AI Research at CVPR 2024 with Fatih Porikli - #688
Energy Star Ratings for AI Models with Sasha Luccioni - #687
Language Understanding and LLMs with Christopher Manning - #686
Chronos: Learning the Language of Time Series with Abdul Fatir Ansari - #685
Powering AI with the World's Largest Computer Chip with Joel Hestness - #684
AI for Power & Energy with Laurent Boinot - #683
Controlling Fusion Reactor Instability with Deep Reinforcement Learning with Aza Jalalvand - #682
Create your
podcast in
minutes
It is Free
20/20
The Dropout
10% Happier with Dan Harris
World News Tonight with David Muir
NEJM This Week