Leanpub Header

Skip to main content
Packt Publishing Ltd

Modern Computer Vision with PyTorch - Second Edition

A practical roadmap from deep learning fundamentals to advanced applications and Generative AI

This book provides a hands-on approach to solving over 30 prominent real-world computer vision problems using PyTorch 2.x on actual datasets. Here you’ll learn to build a neural network from scratch and optimize hyperparameters, perform image classification, multi-object detection, segmentation, and more. You'll also explore facial expression manipulation and combining CV with NLP and RL techniques, build generative AI applications, and take your model to production on AWS. By the end of this book, you'll master modern NN architectures and confidently solve real-world CV problems.

The author is letting you choose the price you pay for this book!

Pick Your Price...
PDF
EPUB
About

About

About the Book

Whether you are a beginner or are looking to progress in your computer vision career, this book guides you through the fundamentals of neural networks (NNs) and PyTorch and how to implement state-of-the-art architectures for real-world tasks.

The second edition of Modern Computer Vision with PyTorch is fully updated to explain and provide practical examples of the latest multimodal models, CLIP, and Stable Diffusion.

You’ll discover best practices for working with images, tweaking hyperparameters, and moving models into production. As you progress, you'll implement various use cases for facial keypoint recognition, multi-object detection, segmentation, and human pose detection. This book provides a solid foundation in image generation as you explore different GAN architectures. You’ll leverage transformer-based architectures like ViT, TrOCR, BLIP2, and LayoutLM to perform various real-world tasks and build a diffusion model from scratch. Additionally, you’ll utilize foundation models' capabilities to perform zero-shot object detection and image segmentation. Finally, you’ll learn best practices for deploying a model to production.

By the end of this deep learning book, you'll confidently leverage modern NN architectures to solve real-world computer vision problems.

Share this book

Categories

Price

Pick Your Price...

Minimum price

$48.99

$48.99

You pay

$48.99
$

All prices are in US $. You can pay in US $ or in your local currency when you check out.

EU customers: prices exclude VAT, which is added during checkout.

...Or Buy With Credits!

Number of credits (Minimum 4)

4
The author will earn $48.00 from your purchase!
You can get credits monthly with a Reader Membership

Author

About the Author

Packt Publishing Ltd

Packt Publishing are an established global technical learning content provider, founded in Birmingham, UK with over twenty years’ experience in delivering premium rich content from ground-breaking authors on a wide range of emerging and popular technologies. Our titles have global relevance our multimedia portfolio includes over 9,000 books, e-books, audiobooks and video courses. www.packtpub.com

Contents

Table of Contents

  1. Artificial Neural Network Fundamentals
  2. PyTorch Fundamentals
  3. Building a Deep Neural Network with PyTorch
  4. Introducing Convolutional Neural Networks
  5. Transfer Learning for Image Classification
  6. Practical Aspects of Image Classification
  7. Basics of Object Detection
  8. Advanced Object Detection
  9. Image Segmentation
  10. Applications of Object Detection and Segmentation
  11. Autoencoders and Image Manipulation
  12. Image Generation Using GANs
  13. Advanced GANs to Manipulate Images
  14. Combining Computer Vision and Reinforcement Learning
  15. Combining Computer Vision and NLP Techniques
  16. Foundation Models in Computer Vision
  17. Applications of Stable Diffusion
  18. Moving a Model to Production
  19. Appendix

About the Publisher

About the Publisher

This book is published on Leanpub by Packt Publishing Ltd

Packt Publishing are an established global technical learning content provider, founded in Birmingham, UK with over twenty years’ experience in delivering premium rich content from ground-breaking authors on a wide range of emerging and popular technologies. Our titles have global relevance our multimedia portfolio includes over 9,000 books, e-books, audiobooks and video courses. www.packtpub.com

The Leanpub 60 Day 100% Happiness Guarantee

Within 60 days of purchase you can get a 100% refund on any Leanpub purchase, in two clicks.

Now, this is technically risky for us, since you'll have the book or course files either way. But we're so confident in our products and services, and in our authors and readers, that we're happy to offer a full money back guarantee for everything we sell.

You can only find out how good something is by trying it, and because of our 100% money back guarantee there's literally no risk to do so!

So, there's no reason not to click the Add to Cart button, is there?

See full terms...

Earn $8 on a $10 Purchase, and $16 on a $20 Purchase

We pay 80% royalties on purchases of $7.99 or more, and 80% royalties minus a 50 cent flat fee on purchases between $0.99 and $7.98. You earn $8 on a $10 sale, and $16 on a $20 sale. So, if we sell 5000 non-refunded copies of your book for $20, you'll earn $80,000.

(Yes, some authors have already earned much more than that on Leanpub.)

In fact, authors have earned over $14 million writing, publishing and selling on Leanpub.

Learn more about writing on Leanpub

Free Updates. DRM Free.

If you buy a Leanpub book, you get free updates for as long as the author updates the book! Many authors use Leanpub to publish their books in-progress, while they are writing them. All readers get free updates, regardless of when they bought the book or how much they paid (including free).

Most Leanpub books are available in PDF (for computers) and EPUB (for phones, tablets and Kindle). The formats that a book includes are shown at the top right corner of this page.

Finally, Leanpub books don't have any DRM copy-protection nonsense, so you can easily read them on any supported device.

Learn more about Leanpub's ebook formats and where to read them

Write and Publish on Leanpub

You can use Leanpub to easily write, publish and sell in-progress and completed ebooks and online courses!

Leanpub is a powerful platform for serious authors, combining a simple, elegant writing and publishing workflow with a store focused on selling in-progress ebooks.

Leanpub is a magical typewriter for authors: just write in plain text, and to publish your ebook, just click a button. (Or, if you are producing your ebook your own way, you can even upload your own PDF and/or EPUB files and then publish with one click!) It really is that easy.

Learn more about writing on Leanpub