Listen to a sample
Generative Deep Learning
- Author
- David Foster
- Narrator
- Mike Cooper
- Publisher
- Ascent Audio
- Publish Date
- 22 July 2025
- Run Time
- 9 hours 7 minutes
- Format
- Audiobook
What to expect
The book starts with the basics of deep learning and progresses to cutting-edge architectures. Through tips and tricks, you'll understand how to make your models learn more efficiently and become more creative. You will discover how VAEs can change facial expressions in photos; train GANs to generate images based on your own dataset; build diffusion models to produce new varieties of flowers; train your own GPT for text generation; learn how large language models like ChatGPT are trained; explore state-of-the-art architectures such as StyleGAN2 and ViT-VQGAN; compose polyphonic music using Transformers and MuseGAN; understand how generative world models can solve reinforcement learning tasks; and dive into multimodal models such as DALL.E 2, Imagen, and Stable Diffusion.
This book also explores the future of generative AI and how individuals and companies can proactively begin to leverage this remarkable new technology to create competitive advantage.
User Reviews
No reviews yet.
Details
- Author
- David Foster
- Narrator
- Mike Cooper
- Duration
- 9 hours 7 minutes
- Release Date
- 22 July 2025
- ISBN
- 9781663754080
- Format
- Audiobook
- Publisher
- Ascent Audio
- Genre
- Computer vision, Pattern recognition, Enterprise software, Operational research, Artificial intelligence, Data capture and analysis, Data science and analysis: general, Machine learning, Mathematical theory of computation, Machine learning
Synopsis
Generative AI is the hottest topic in tech. This practical book teaches machine learning engineers and data scientists how to use TensorFlow and Keras to create impressive generative deep learning models from scratch, including variational autoencoders (VAEs), generative adversarial networks (GANs), Transformers, normalizing flows, energy-based models, and denoising diffusion models.The book starts with the basics of deep learning and progresses to cutting-edge architectures. Through tips and tricks, you'll understand how to make your models learn more efficiently and become more creative. You will discover how VAEs can change facial expressions in photos; train GANs to generate images based on your own dataset; build diffusion models to produce new varieties of flowers; train your own GPT for text generation; learn how large language models like ChatGPT are trained; explore state-of-the-art architectures such as StyleGAN2 and ViT-VQGAN; compose polyphonic music using Transformers and MuseGAN; understand how generative world models can solve reinforcement learning tasks; and dive into multimodal models such as DALL.E 2, Imagen, and Stable Diffusion.This book also explores the future of generative AI and how individuals and companies can proactively begin to leverage this remarkable new technology to create competitive advantage.
About xigxag
xigxag is an independent, UK-based audiobook platform and the UK's only B Corp certified audiobook company. We offer nearly 150,000 audiobooks, all with professional human narration. No AI-generated voices. xigxag is rated an Ethical Consumer Best Buy for audiobooks and a top-rated ethical bookshop.
There is no subscription. You buy the audiobooks you want, you own them permanently, and prices start at $13.95. The more you listen over the year, the less you pay per book, and there is always a selection in Listen for Less from $7.95. You can also send audiobooks as gifts using our audiobook gift cards or in-app gifting.
Every audiobook on xigxag has honest reviews covering both the book and the narration, and our community features make it easy to share what you are listening to, discover what others recommend, and find your next listen.
Download or stream on iOS, Android, or in your browser. Browse the full catalogue at xigxag.co.uk.