Download - The new AI app stack | Podbean

Discover

Podcast Features
Your all-in-one podcasting solution.

Podcast Studio
Easy-to-use audio recorder app.
Livestream
High-performing audio live, without limits.

Podcast App
The best podcast player & podcast app.
Podbean AI
AI-Enhanced Audio Quality and Content Generation.

Ads Marketplace
Join Ads Marketplace to earn money
through sponsorship on your podcast.

PodAds
Manage your ads with dynamic ad insertion capability.
Patron & Paid Content
The seamless way for fans to support you directly
from your podcast.
Apple Podcasts Subscriptions Integration
Effortlessly publish and manage exclusive episodes for your
Apple Podcasts subscribers directly from Podbean.

All Arts Business Comedy Education
Fiction Government Health & Fitness History Kids & Family
Leisure Music News Religion & Spirituality Science
Society & Culture Sports Technology True Crime TV & Film
Live

How to Start a Podcast
How to Start a Live Podcast
How to Monetize a podcast
How to Promote Your Podcast
How to Use Group Recording

Log in
Start your podcast for free

Podcasting
Monetization
Enterprise
Pricing
Discover

Practical AI: Machine Learning, Data Science

Technology

The new AI app stack

2023-08-23

Download Right click and do "save link as"

Recently a16z released a diagram showing the “Emerging Architectures for LLM Applications.” In this episode, we expand on things covered in that diagram to a more general mental model for the new AI app stack. We cover a variety of things from model “middleware” for caching and control to app orchestration.

Leave us a comment

Changelog++ members save 2 minutes on this episode because they made the ads disappear. Join today!

Sponsors:

Fastly – Our bandwidth partner. Fastly powers fast, secure, and scalable digital experiences. Move beyond your content delivery network to their powerful edge cloud platform. Learn more at fastly.com
Fly.io – The home of Changelog.com — Deploy your apps and databases close to your users. In minutes you can run your Ruby, Go, Node, Deno, Python, or Elixir app (and databases!) all over the world. No ops required. Learn more at fly.io/changelog and check out the speedrun in their docs.
Typesense – Lightning fast, globally distributed Search-as-a-Service that runs in memory. You iterlly can’t get any faster!
Changelog News – A podcast+newsletter combo that’s brief, entertaining & always on-point. Subscribe today.

Featuring:

Chris Benson – Twitter, GitHub, LinkedIn, Website
Daniel Whitenack – Twitter, GitHub, Website

Show Notes:

Emerging Architectures for LLM Applications

Something missing or broken? PRs welcome!

Timestamps:

(00:07) - Welcome to Practical AI
(00:43) - Deep dive into LLMs
(02:25) - Emerging LLM app stack
(04:35) - Playgrounds
(08:07) - App Hosting
(10:46) - Stack orchestration
(15:50) - Maintenance breakdown
(19:08) - Sponsor: Changelog News
(20:43) - Vector databases
(22:36) - Embedding models
(24:27) - Benchmarks and measurements
(26:59) - Data & poor architecture
(29:42) - LLM logging
(33:01) - Middleware Caching
(37:32) - Validation
(40:53) - Key takeaways
(42:36) - Closing thoughts
(44:23) - Outro

More Episodes

AI in the U.S. Congress

2024-05-29

First impressions of GPT-4o

2024-05-22

Full-stack approach for effective AI agents

2024-05-15

Autonomous fighter jets?!

2024-05-08

Private, open source chat UIs

2024-04-30

2024-04-24

Udio & the age of multi-modal AI

2024-04-16

RAG continues to rise

2024-04-10

Should kids still learn to code?

2024-04-02

AI vs software devs

2024-03-26

Prompting the future

2024-03-20

Generating the future of art & entertainment

2024-03-12

YOLOv9: Computer vision is alive and well

2024-03-06

Representation Engineering (Activation Hacking)

2024-02-28

Leading the charge on AI in National Security

2024-02-20

Gemini vs OpenAI

2024-02-14

Data synthesis for SOTA LLMs

2024-02-06

Large Action Models (LAMs) & Rabbits 🐇

2024-01-30

Collaboration & evaluation for LLM apps

2024-01-23

Advent of GenAI Hackathon recap

2024-01-17

←
1
2
3
4
5
6
7
8
9
10
→

012345678910111213141516171819

Get this podcast on your
phone, FREE

Download Podbean app on App Store

Download Podbean app on Google Play

Create your
podcast in
minutes

Full-featured podcast site
Unlimited storage and bandwidth
Comprehensive podcast stats
Distribute to Apple Podcasts, Spotify, and more
Make money with your podcast

It is Free

Podcast Services
MONETIZATION & MORE
KNOWLEDGE BASE
Support
Podbean

Privacy Policy
Cookie Policy
Terms of Use
Consent Preferences
Copyright © 2015-2024 Podbean.com