Today we’re joined by Nataniel Ruiz, a research scientist at Google. In our conversation with Nataniel, we discuss his recent work around personalization for text-to-image AI models. Specifically, we dig into DreamBooth, an algorithm that enables “subject-driven generation,” that is, the creation of personalized generative models using a small set of user-provided images about a subject. The personalized models can then be used to generate the subject in various contexts using a text prompt. Nataniel gives us a dive deep into the fine-tuning approach used in DreamBooth, the potential reasons behind the algorithm’s effectiveness, the challenges of fine-tuning diffusion models in this way, such as language drift, and how the prior preservation loss technique avoids this setback, as well as the evaluation challenges and metrics used in DreamBooth. We also touched base on his other recent papers including SuTI, StyleDrop, HyperDreamBooth, and lastly, Platypus.
The complete show notes for this episode can be found at twimlai.com/go/648.
Advancing Deep Reinforcement Learning with NetHack, w/ Tim Rocktäschel - #527
Building Technical Communities at Stack Overflow with Prashanth Chandrasekar - #526
Deep Learning is Eating 5G. Here’s How, w/ Joseph Soriaga - #525
Modeling Human Cognition with RNNs and Curriculum Learning, w/ Kanaka Rajan - #524
Do You Dare Run Your ML Experiments in Production? with Ville Tuulos - #523
Delivering Neural Speech Services at Scale with Li Jiang - #522
AI’s Legal and Ethical Implications with Sandra Wachter - #521
Compositional ML and the Future of Software Development with Dillon Erb - #520
Generating SQL Database Queries from Natural Language with Yanshuai Cao - #519
Social Commonsense Reasoning with Yejin Choi - #518
Deep Reinforcement Learning for Game Testing at EA with Konrad Tollmar - #517
Exploring AI 2041 with Kai-Fu Lee - #516
Advancing Robotic Brains and Bodies with Daniela Rus - #515
Neural Synthesis of Binaural Speech From Mono Audio with Alexander Richard - #514
Using Brain Imaging to Improve Neural Networks with Alona Fyshe - #513
Adaptivity in Machine Learning with Samory Kpotufe - #512
A Social Scientist’s Perspective on AI with Eric Rice - #511
Applications of Variational Autoencoders and Bayesian Optimization with José Miguel Hernández Lobato - #510
Codex, OpenAI’s Automated Code Generation API with Greg Brockman - #509
Spatiotemporal Data Analysis with Rose Yu - #508
Create your
podcast in
minutes
It is Free
20/20
The Dropout
Ten Percent Happier with Dan Harris
World News Tonight with David Muir
NEJM This Week