Kandinsky 2.0

Habr post

Demo

pip install "git+https://github.com/ai-forever/Kandinsky-2.0.git"

Model architecture:

It is a latent diffusion model with two multilingual text encoders:

mCLIP-XLMR 560M parameters
mT5-encoder-small 146M parameters

These encoders and multilingual training datasets unveil the real multilingual text-to-image generation experience!

Kandinsky 2.0 was trained on a large 1B multilingual set, including samples that we used to train Kandinsky.

In terms of diffusion architecture Kandinsky 2.0 implements UNet with 1.2B parameters.

Kandinsky 2.0 architecture overview:

How to use:

Check our jupyter notebooks with examples in ./notebooks folder

1. text2img

from kandinsky2 import get_kandinsky2

model = get_kandinsky2('cuda', task_type='text2img')
images = model.generate_text2img('A teddy bear на красной площади', batch_size=4, h=512, w=512, num_steps=75, denoised_type='dynamic_threshold', dynamic_threshold_v=99.5, sampler='ddim_sampler', ddim_eta=0.05, guidance_scale=10)

prompt: "A teddy bear на красной площади"

2. inpainting

from kandinsky2 import get_kandinsky2
from PIL import Image
import numpy as np

model = get_kandinsky2('cuda', task_type='inpainting')
init_image = Image.open('image.jpg')
mask = np.ones((512, 512), dtype=np.float32)
mask[100:] =  0
images = model.generate_inpainting('Девушка в красном платье', init_image, mask, num_steps=50, denoised_type='dynamic_threshold', dynamic_threshold_v=99.5, sampler='ddim_sampler', ddim_eta=0.05, guidance_scale=10)

prompt: "Девушка в красном платье"

3. img2img

from kandinsky2 import get_kandinsky2
from PIL import Image

model = get_kandinsky2('cuda', task_type='img2img')
init_image = Image.open('image.jpg')
images = model.generate_img2img('кошка', init_image, strength=0.8, num_steps=50, denoised_type='dynamic_threshold', dynamic_threshold_v=99.5, sampler='ddim_sampler', ddim_eta=0.05, guidance_scale=10)

Authors

Arseniy Shakhmatov: Github, Blog
Anton Razzhigaev: Github, Blog
Aleksandr Nikolich: Github, Blog
Vladimir Arkhipkin: Github
Igor Pavlov: Github
Andrey Kuznetsov: Github
Denis Dimitrov: Github

Name		Name	Last commit message	Last commit date
Latest commit History 82 Commits
content		content
kandinsky2		kandinsky2
notebooks		notebooks
README.md		README.md
license		license
setup.py		setup.py

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

Kandinsky 2.0

Model architecture:

How to use:

1. text2img

2. inpainting

3. img2img

Authors

About

Releases

Packages

Languages

License

Truenya/Kandinsky-2.0

Folders and files

Latest commit

History

Repository files navigation

Kandinsky 2.0

Model architecture:

How to use:

1. text2img

2. inpainting

3. img2img

Authors

About

Resources

License

Stars

Watchers

Forks

Releases

Packages 0

Languages

Packages