Files
AnimateDiff/README.md

169 lines
5.8 KiB
Markdown
Raw Normal View History

2023-07-03 00:32:41 +08:00
# AnimateDiff
2023-07-03 01:09:13 +08:00
This repository is the official implementation of [AnimateDiff]().
2023-07-03 00:32:41 +08:00
**[AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning]()**
</br>
Yuwei Guo,
2023-07-03 01:13:53 +08:00
Ceyuan Yang*,
2023-07-03 00:32:41 +08:00
Anyi Rao,
Yaohui Wang,
Yu Qiao,
Dahua Lin,
2023-07-03 01:13:53 +08:00
Bo Dai
2023-07-09 23:37:00 +08:00
<p style="font-size: 0.8em; margin-top: -1em">*Corresponding Author</p>
2023-07-03 00:32:41 +08:00
2023-07-09 22:40:48 +08:00
[Arxiv Report]() | [Project Page](https://animatediff.github.io/)
2023-07-03 00:32:41 +08:00
2023-07-09 22:40:48 +08:00
## Todo
2023-07-09 22:41:48 +08:00
- [x] Code Release
2023-07-09 22:40:48 +08:00
- [ ] Arxiv Report
- [ ] GPU Memory Optimization
- [ ] Gradio Interface
<!-- - [ ] Online Demo -->
<!-- - [ ] Support WebUI -->
## Setup for Inference
### Prepare Environment
Our approach takes around 60 GB GPU memory to inference. NVIDIA A100 is recommanded.
```
git clone https://github.com/guoyww/animatediff.git
cd animatediff
conda create -n animatediff python=3.8
conda activate animatediff
pip install -r requirments.txt
```
### Download Base T2I & Motion Module Checkpoints
We provide two versions of our Motion Module, which are trained on stable-diffusion-v1-4 and finetuned on v1-5 seperately.
It's recommanded to try both of them for best results.
```
git lfs install
git clone https://huggingface.co/runwayml/stable-diffusion-v1-5 models/StableDiffusion/
bash download_bashscripts/0-MotionModule.sh
```
You may also directly download the motion module checkpoints from [Google Drive](https://drive.google.com/drive/folders/1EqLC65eR1-W-sGD0Im7fkED6c8GkiNFI?usp=sharing), then put them in `models/Motion_Module/` folder.
### Prepare Personalize T2I
Here we provide inference configs for 6 demo T2I on CivitAI.
You may run the following bash scripts to download these checkpoints.
```
bash download_bashscripts/1-ToonYou.sh
bash download_bashscripts/2-Lyriel.sh
bash download_bashscripts/3-RcnzCartoon.sh
bash download_bashscripts/4-MajicMix.sh
bash download_bashscripts/5-RealisticVision.sh
bash download_bashscripts/6-Tusun.sh
bash download_bashscripts/7-FilmVelvia.sh
bash download_bashscripts/8-GhibliBackground.sh
```
### Inference
After downloading the above peronalized T2I checkpoints, run the following commands to generate animations.
```
python -m scripts.animate --config configs/prompts/1-ToonYou.yaml
python -m scripts.animate --config configs/prompts/2-Lyriel.yaml
python -m scripts.animate --config configs/prompts/3-RcnzCartoon.yaml
python -m scripts.animate --config configs/prompts/4-MajicMix.yaml
python -m scripts.animate --config configs/prompts/5-RealisticVision.yaml
python -m scripts.animate --config configs/prompts/6-Tusun.yaml
python -m scripts.animate --config configs/prompts/7-FilmVelvia.yaml
python -m scripts.animate --config configs/prompts/8-GhibliBackground.yaml
```
2023-07-03 01:09:13 +08:00
<!-- ## Setup
2023-07-03 00:32:41 +08:00
Install the required packages:
```
pip install -r requirements.txt
```
## Inference
1. Download personalized Stable Diffusion (LoRA or DreamBooth are currently supported) from [CivitAI](https://civitai.com/) or Huggingface;
2. Download our motion module's pretrained weights from [this link]();
3. Edit config files in
```
configs/prompts/lora.yaml
```
4. Execute animation generation:
```
python -m scripts.animate --prompt configs/prompts/lora.yaml
2023-07-03 01:09:13 +08:00
``` -->
2023-07-09 22:40:48 +08:00
## Gallery
2023-07-09 23:34:13 +08:00
Here we demonstrate several best results we got in previous experiments.
2023-07-03 01:09:13 +08:00
<table class="center">
<tr>
<td><img src="__assets__/animations/model_01/01.gif"></td>
<td><img src="__assets__/animations/model_01/02.gif"></td>
<td><img src="__assets__/animations/model_01/03.gif"></td>
<td><img src="__assets__/animations/model_01/04.gif"></td>
</tr>
</table>
<p style="margin-left: 2em; margin-top: -1em">Model<a href="https://civitai.com/models/30240/toonyou">ToonYou</a></p>
<table>
<tr>
<td><img src="__assets__/animations/model_02/01.gif"></td>
<td><img src="__assets__/animations/model_02/02.gif"></td>
<td><img src="__assets__/animations/model_02/03.gif"></td>
<td><img src="__assets__/animations/model_02/04.gif"></td>
</tr>
</table>
<p style="margin-left: 2em; margin-top: -1em">Model<a href="https://civitai.com/models/4468/counterfeit-v30">Counterfeit V3.0</a></p>
<table>
<tr>
<td><img src="__assets__/animations/model_03/01.gif"></td>
<td><img src="__assets__/animations/model_03/02.gif"></td>
<td><img src="__assets__/animations/model_03/03.gif"></td>
<td><img src="__assets__/animations/model_03/04.gif"></td>
</tr>
</table>
<p style="margin-left: 2em; margin-top: -1em">Model<a href="https://civitai.com/models/4201/realistic-vision-v20">Realistic Vision V2.0</a></p>
<table>
<tr>
<td><img src="__assets__/animations/model_04/01.gif"></td>
<td><img src="__assets__/animations/model_04/02.gif"></td>
<td><img src="__assets__/animations/model_04/03.gif"></td>
<td><img src="__assets__/animations/model_04/04.gif"></td>
</tr>
</table>
<p style="margin-left: 2em; margin-top: -1em">Model <a href="https://civitai.com/models/43331/majicmix-realistic">majicMIX Realistic</a></p>
<table>
<tr>
<td><img src="__assets__/animations/model_05/01.gif"></td>
<td><img src="__assets__/animations/model_05/02.gif"></td>
<td><img src="__assets__/animations/model_05/03.gif"></td>
<td><img src="__assets__/animations/model_05/04.gif"></td>
</tr>
</table>
2023-07-09 23:49:59 +08:00
<p style="margin-left: 2em; margin-top: -1em">Model<a href="https://civitai.com/models/66347/rcnz-cartoon-3d">RCNZ Cartoon</a></p>
2023-07-03 01:09:13 +08:00
<table>
<tr>
<td><img src="__assets__/animations/model_06/01.gif"></td>
<td><img src="__assets__/animations/model_06/02.gif"></td>
<td><img src="__assets__/animations/model_06/03.gif"></td>
<td><img src="__assets__/animations/model_06/04.gif"></td>
</tr>
</table>
<p style="margin-left: 2em; margin-top: -1em">Model<a href="https://civitai.com/models/33208/filmgirl-film-grain-lora-and-loha">FilmVelvia</a></p>
2023-07-09 22:40:48 +08:00
## Citation
Coming soon.
## Acknowledgements
Codebase built upon [Tune-a-Video](https://github.com/showlab/Tune-A-Video).