{"product_id":"build-a-text-to-image-generator-from-scratch-with-transformers-and-diffusions-isbn-9781633435421","title":"Build a Text-to-Image Generator (from Scratch): With transformers and diffusions","description":"\u003cb\u003eGet a free eBook (PDF or ePub) from Manning as well as access to the online liveBook format (and its AI assistant that will answer your questions in any language) when you purchase the print book.\u003c\/b\u003e\u003cbr\u003e\u003cbr\u003eThis book takes you step-by-step through creating your own AI models that can generate images from text. You’ll explore two methods of image generation—vision transformers and diffusion models—and learn vital AI development techniques as you go.\u003cbr\u003e\u003cbr\u003eDive into the powerful models behind AI image generators. The best way to learn is to build something from scratch, and in this book you’ll build your very own diffusion model and vision transformer. As you work through each stage of development, you’ll develop an understanding of how these models can be customized, applied, and integrated for impressive multimodal AI.\u003cbr\u003e\u003cbr\u003e\u003ci\u003eBuild a Text-to-Image Generator (from Scratch)\u003c\/i\u003e teaches you how to:\u003cul\u003e\n\u003cli\u003eBuild and train models to generate high resolution images based on text descriptions\u003c\/li\u003e\n\u003cli\u003eEdit an existing image based on text prompts\u003c\/li\u003e\n\u003cli\u003eBuild and train a model to add captions to images\u003c\/li\u003e\n\u003cli\u003eBuild and train a vision transformer to classify images\u003c\/li\u003e\n\u003cli\u003eFine-tune LLMs for downstream tasks such as classification, text or image generation\u003c\/li\u003e\n\u003cli\u003eBetter differentiate real images from deepfakes\u003c\/li\u003e\n\u003c\/ul\u003e\u003cb\u003eAbout the technology\u003c\/b\u003e\u003cbr\u003e\u003cbr\u003eAI-generated images appear everywhere from high-end advertising to casual social media feeds. Text-to-image tools like Dall-e, Midjourney, and Flux make it easy to create AI art, but how do they work? In this book, you’ll find out by building your own text-to-image generator!\u003cbr\u003e\u003cbr\u003e\u003cb\u003eAbout the book\u003c\/b\u003e\u003cbr\u003e\u003cbr\u003e\u003ci\u003eBuild a Text-to-Image Generator (from Scratch) \u003c\/i\u003eexplores both transformer-based image generation and diffusion models. You’ll work hands-on to build a pair of simple generation models that can classify images, automatically add captions, reconstruct images, and enhance existing graphics. Author \u003cb\u003eMark Liu\u003c\/b\u003e guides you every step of the way with clear explanations, informative diagrams, and eye-opening examples you can build on your own laptop.\u003cbr\u003e\u003cbr\u003e\u003cb\u003eWhat's inside\u003c\/b\u003e\u003cul\u003e\n\u003cli\u003eBuild a vision transformer to classify images\u003c\/li\u003e\n\u003cli\u003eEdit images using text prompts\u003c\/li\u003e\n\u003cli\u003eFine-tune image models\u003c\/li\u003e\n\u003c\/ul\u003e\u003cb\u003eAbout the reader\u003c\/b\u003e\u003cbr\u003e\u003cbr\u003eRequires basic knowledge of generative AI models and intermediate Python skills.\u003cbr\u003e\u003cbr\u003e\u003cb\u003eAbout the author\u003c\/b\u003e\u003cbr\u003e\u003cbr\u003e\u003cb\u003eMark Liu\u003c\/b\u003e is the founding director of the Master of Science in Finance program at the University of Kentucky. He is also the author of \u003ci\u003eLearn Generative AI with PyTorch\u003c\/i\u003e.\u003cbr\u003e\u003cbr\u003e\u003cb\u003eTable of Contents\u003c\/b\u003e\u003cbr\u003e\u003cbr\u003ePart 1\u003cbr\u003e1 A tale of two models: Transformers and diffusions\u003cbr\u003e2 Build a transformer\u003cbr\u003e3 Classify images with a vision transformer\u003cbr\u003e4 Add captions to images\u003cbr\u003ePart 2\u003cbr\u003e5 Generate images with diffusion models\u003cbr\u003e6 Control what images to generate in diffusion models\u003cbr\u003e7 Generate high-resolution images with diffusion models\u003cbr\u003ePart 3\u003cbr\u003e8 CLIP: A model to measure the similarity between image and text\u003cbr\u003e9 Text-to-image generation with latent diffusion\u003cbr\u003e10 A deep dive into Stable Diffusion\u003cbr\u003ePart 4\u003cbr\u003e11 VQGAN: Convert images into sequences of integers\u003cbr\u003e12 A minimal implementation of DALL-E\u003cbr\u003ePart 5\u003cbr\u003e13 New developments and challenges in text-to-image generation\u003cbr\u003eA Installing PyTorch and enabling GPU training locally and in Colab","brand":"Simon \u0026 Schuster","offers":[{"title":"Default Title","offer_id":48140532482277,"sku":"NP9781633435421","price":59.99,"currency_code":"USD","in_stock":false}],"thumbnail_url":"\/\/cdn.shopify.com\/s\/files\/1\/1842\/7735\/files\/801f7fbc-ff9c-4089-93fe-b35e0027c309.jpg?v=1765323030","url":"https:\/\/k12savings.com\/products\/build-a-text-to-image-generator-from-scratch-with-transformers-and-diffusions-isbn-9781633435421","provider":"K12savings","version":"1.0","type":"link"}