FreeMD: Training-free multi-domain text-to-image generation with any control.

Mingwen Shao1, Chang Liu2, Xiang Lv2

  • 1School of Computer Science and Technology, China University of Petroleum (East China), Qingdao, 266580, China; Artificial Intelligence Research Institute, Shenzhen University of Advanced Technology, Shenzhen, 518107, China.

Summary

FreeMD enhances text-to-image generation by decoupling text and structure controls. This novel method improves semantic alignment with prompts and ensures better structure consistency using multi-domain guidance.