Open-Source Projects

LiveAvatar

Real-time, streaming, infinite-length audio-driven avatar generation powered by a 14B diffusion model — 45 FPS on 5×H800, FP8 release runs on 48GB GPUs. ECCV 2026. HF #1 Paper of the Day, 2k+ stars.

EasyControl

Efficient and flexible condition-guided control for Diffusion Transformers (DiT). Plug-and-play Condition Injection LoRA, position-aware training, and KV-Cache-accelerated inference. ICCV 2025.

EasyControl Ghibli

One-click Ghibli-style portrait generation built on EasyControl. Ranked the #1 trending Space on Hugging Face.

PhotoDoodle

The first in-context architecture for image editing & customizable image editing.

StableMakeup

Robust real-world makeup transfer with diffusion models — from light to extremely heavy styles. SIGGRAPH 2025.

Stable-Hair

Real-world hair transfer via diffusion models, robustly handling diverse and complex hairstyles. AAAI 2025; V2 in IEEE TVCG 2026.

SSR-Encoder

Encoding selective subject representation for subject-driven generation — high-fidelity, query-guided personalization without test-time fine-tuning. CVPR 2024.