You're reading from Using Stable Diffusion with Python Leverage Python to control and automate high-quality AI image generation using Stable Diffusion

Product type Paperback

Published in Jun 2024

Publisher Packt

ISBN-13 9781835086377

Length 352 pages

Edition 1st Edition

Languages

Python

Concepts

GPT/LLMs

Author (1):

Andrew Zhu (Shudong Zhu)

View More author details

Table of Contents (29) Chapters

Preface

1. Part 1 – A Whirlwind of Stable Diffusion FREE CHAPTER

2. Chapter 1: Introducing Stable Diffusion

3. Chapter 2: Setting Up the Environment for Stable Diffusion

4. Chapter 3: Generating Images Using Stable Diffusion

5. Chapter 4: Understanding the Theory Behind Diffusion Models

6. Chapter 5: Understanding How Stable Diffusion Works

7. Chapter 6: Using Stable Diffusion Models

8. Part 2 – Improving Diffusers with Custom Features

9. Chapter 7: Optimizing Performance and VRAM Usage

10. Chapter 8: Using Community-Shared LoRAs

11. Chapter 9: Using Textual Inversion

12. Chapter 10: Overcoming 77-Token Limitations and Enabling Prompt Weighting

13. Chapter 11: Image Restore and Super-Resolution

14. Chapter 12: Scheduled Prompt Parsing

15. Part 3 – Advanced Topics

16. Chapter 13: Generating Images with ControlNet

17. Chapter 14: Generating Video Using Stable Diffusion

18. Chapter 15: Generating Image Descriptions Using BLIP-2 and LLaVA

19. Chapter 16: Exploring Stable Diffusion XL

20. Chapter 17: Building Optimized Prompts for Stable Diffusion

21. Part 4 – Building Stable Diffusion into an Application

22. Chapter 18: Applications – Object Editing and Style Transferring

23. Chapter 19: Generation Data Persistence

24. Chapter 20: Creating Interactive User Interfaces

25. Chapter 21: Diffusion Model Transfer Learning

26. Chapter 22: Exploring Beyond Stable Diffusion

27. Index

Why subscribe?

28. Other Books You May Enjoy

Understanding the 77-token limitation

The Stable Diffusion (v1.5) text encoder uses the CLIP encoder from OpenAI [2]. The CLIP text encoder has a 77-token limit, and this limitation propagates to the downstream Stable Diffusion. We can reproduce the 77-token limitation with the following steps:

We can take out the encoder from Stable Diffusion and verify it. Let’s say we have the prompt a photo of a cat and dog driving an aircraft and we multiply it by 20 to make the prompt’s token size larger than 77:
```
prompt = "a photo of a cat and a dog driving an aircraft "*20
```
Reuse the pipeline we initialized at the beginning of the chapter and take out tokenizer and text_encoder:
```
tokenizer = pipe.tokenizer
```
```
text_encoder = pipe.text_encoder
```

Use tokenizer to get the token IDs from the prompt:

tokens = tokenizer(

    prompt,

    truncation = False,

    return_tensors = 'pt'

)["input_ids&quot...

The rest of the chapter is locked

You're reading from Using Stable Diffusion with Python Leverage Python to control and automate high-quality AI image generation using Stable Diffusion

Table of Contents (29) Chapters

Understanding the 77-token limitation

Authors (1)

Personalised recommendations for you

You're reading from Using Stable Diffusion with Python Leverage Python to control and automate high-quality AI image generation using Stable Diffusion

Table of Contents (29) Chapters

Understanding the 77-token limitation

Unlock this book and the full library FREE for 7 days

Authors (1)

Personalised recommendations for you