Skip to content

FlowLess: Controlling Abstract Image Generation

· 0 citations · 64 references

TL;DR

A novel self-supervised framework that enables granular control over image generation through a visual abstraction set that provides a richer, more flexible paradigm for creative design compared to state-of-the-art baselines across diverse styles and compositions is introduced.

View source

Similar papers

#computer vision Preprint Sep 2026

Editable Visual Design

While diffusion base models such as GPT-Image-2 and Nano-Banana exhibit remarkable visual expressiveness, their end-to-end generation inherently yields flattened bitmaps with error-prone text, precluding layer-wise post-editing. Conversely, code-based visual generation via Coding Agents provides precise layout control...

Jun-Yan Ye, Wei Liu, Dongzhi Jiang et al. · 0 citations
Aug 2026

Pose-Star++: Semantic-Visual Understanding for Fine-Grained Fashion Image Editing.

Fashion image editing demands high-dimensional, fine-grained control to follow personalized, unpredictable natural-language instructions. Yet current methods are limited by a fundamental trade-off: fashion-specific approaches offer structural accuracy but lack semantic flexibility, while general text-driven editors are...

Yuran Dong, Bo Du, Mang Ye · 0 citations
Preprint Aug 2026

SI-Edit: Toward Sketch-Instruction Guided Local Image Editing with Pixel-Level Precision

An automated pipeline leveraging Multimodal Large Language Models (MLLMs) is developed to synthesize comprehensive quadruplets comprising original images, local geometric sketches, semantic instructions, and corresponding edited images, which uniquely enables collaborative spatial-semantic learning.

Weixin Ye, Wei Wang, Hong-Guang Zhu et al. · 0 citations
Preprint Aug 2026

Abstract4D: A Large-Scale Dataset and Framework for Understanding the Visual Language of Abstract Art

This work analyzes the semantic structure of abstract art through large-scale embedding visualization, uncovering how perceptual relationships organize artistic meaning, and establishes benchmark tasks for classification, cross-modal retrieval, and text-to-image generation to evaluate how AI models perceive and reprodu...

Hao-Wei Zhang, Yuanpei Zhao, Ji-Zhe Zhou et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Paint-Anything: Unified Any-Color Control for Image Generation and Editing

Paint-Anything is presented, which learns a shared hex-prompt interface for generation and editing through object-level color supervision, and introduces Any Color Benchmark (ACBench), comprising ACBench-T2I and ACBench-Edit, to measure object-level hex color fidelity across both tasks.

Ji Xie, Dewei Zhou, Xin-Yu Huang et al. · 0 citations
Open access 2026

StyleSmith: A Signature Style Transfer Framework With Parametric and Temporal Attention

Artistic style transfer, which renders content images in the visual style of a reference artwork while preserving semantic structure, is a fundamental problem in computer vision and AIGC, with broad applications in digital art, commercial design, and creative media. Despite progress driven by diffusion-based approaches...

Yu Cheng, Anucha Pangkesorn, Jia-Rui Wu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.