Just upload a few photos, then enter text prompts like "a photo of a woman playing sports wearing a baseball cap" to generate images of yourself in various scenarios and clone your face.
Uploading more photos will yield better results.
Adjust the weight of facial structure to get different generation results!
IP-Adapter-FaceID is fully compatible with existing controllable tools such as ControlNet and T2I-Adapter.
It is available for use in Comfy UI (via the ComfyUI_IPAdapter_plus plugin), and is also supported by the ControlNet plugin (ip-adapter-faceid_sd15) for SD Web UI.
Our method not only outperforms other approaches in image quality, but also generates images that align better with reference images.
Thanks to the decoupled cross-attention strategy, image prompts can be combined with text prompts to achieve multimodal image generation.
We use face ID embeddings from face recognition models instead of CLIP image embeddings, and also employ LoRA to improve ID consistency. IP-Adapter-FaceID can generate various styled images conditioned on a face using only text prompts.
A text-compatible image prompt adapter for text-to-image diffusion models. Use our proposed IP Adapter for diverse image synthesis, applicable to pre-trained text-to-image diffusion models and additional structural controllers.