|
--- |
|
license: apache-2.0 |
|
language: |
|
- en |
|
library_name: diffusers |
|
pipeline_tag: text-to-image |
|
--- |
|
|
|
# PhotoMaker Model Card |
|
|
|
<div align="center"> |
|
|
|
[**Project Page**](https://photo-maker.github.io/) **|** [**Paper (ArXiv)**](https://arxiv.org/abs/2312.04461) **|** [**Code**](https://github.com/TencentARC/PhotoMaker) |
|
|
|
[🤗 **Gradio demo (Realistic)**](https://huggingface.co/spaces/TencentARC/PhotoMaker) **|** [🤗 **Gradio demo (Stylization)**](https://huggingface.co/spaces/TencentARC/PhotoMaker-Style) |
|
|
|
</div> |
|
|
|
## Introduction |
|
|
|
<!-- Provide a quick summary of what the model is/does. --> |
|
Users can input one or several face photos, along with a text prompt, to receive a customized photo or painting within seconds (no training required!). Additionally, this model can be adapted to any base model based on SDXL or used in conjunction with other LoRA models. |
|
|
|
### Realistic results |
|
|
|
data:image/s3,"s3://crabby-images/c67b4/c67b4b67d97546a4fc9c7c54ebca58ea83433d98" alt="image/jpeg" |
|
|
|
data:image/s3,"s3://crabby-images/ac9f6/ac9f65c514c9a433c2d44ca0a0245c8a2ffff33f" alt="image/jpeg" |
|
|
|
### Stylization results |
|
|
|
data:image/s3,"s3://crabby-images/dc36e/dc36e7fdf7c1ed4f2d01637fc2be207faf755ad9" alt="image/jpeg" |
|
|
|
|
|
data:image/s3,"s3://crabby-images/8e90d/8e90df1b972010f26777190589eac29e035ceb3a" alt="image/jpeg" |
|
|
|
More results can be found in our [project page](https://photo-maker.github.io/) |
|
|
|
## Model Details |
|
|
|
### Model Description |
|
|
|
<!-- Provide a longer summary of what this model is. --> |
|
|
|
|
|
|
|
- **Developed by:** [More Information Needed] |
|
- **Funded by [optional]:** [More Information Needed] |
|
- **Shared by [optional]:** [More Information Needed] |
|
- **Model type:** [More Information Needed] |
|
- **Language(s) (NLP):** [More Information Needed] |
|
- **License:** [More Information Needed] |
|
- **Finetuned from model [optional]:** [More Information Needed] |
|
|
|
|
|
## Bias, Risks, and Limitations |
|
|
|
<!-- This section is meant to convey both technical and sociotechnical limitations. --> |
|
|
|
[More Information Needed] |
|
|
|
|
|
## Citation [optional] |
|
|
|
<!-- If there is a paper or blog post introducing the model, the APA and Bibtex information for that should go in this section. --> |
|
|
|
**BibTeX:** |
|
|
|
```bibtex |
|
@article{li2023photomaker, |
|
title={PhotoMaker: Customizing Realistic Human Photos via Stacked ID Embedding}, |
|
author={Li, Zhen and Cao, Mingdeng and Wang, Xintao and Qi, Zhongang and Cheng, Ming-Ming and Shan, Ying}, |
|
booktitle={arXiv preprint arxiv:2312.04461}, |
|
year={2023} |
|
} |
|
``` |