Creative Agents

Agentic image/video/design generation

10 articles

Creative Agents

Identity preservation remains a distinct challenge for generative image models

The authors systematically benchmark three approaches to identity preservation in generative image models: encoding identity in the input context, using trainable subject-specific parameters (like LoRA), or maintaining a persistent identity layer. Their results show that persistent identity layers consistently reduce identity degradation across iterative edits, small subject scales, and multi-subject compositions, while preserving image quality and instruction adherence.

7 Sept 2026
Creative Agents

Coding agents combine generated imagery with editable web-based visual layouts

Ye and colleagues present Editable Visual Design, a system for producing visual designs that remain structurally editable rather than being flattened into a single image. The agent generates isolated visual assets, assembles them with native HTML/CSS, and revises the result using feedback from rendered previews; the authors report successful applications to posters, infographics, and related formats.

5 Sept 2026
Creative Agents

An agentic framework that fills in missing context for real-world image generation

The authors propose Qwen-Image-Agent to address the context gap where user requests for image generation are often underspecified, implicit, or depend on up-to-date knowledge. The framework uses Context-Aware Planning to identify missing details and Context Grounding to acquire them from reasoning, search, memory, and feedback. On the newly introduced IA-Bench benchmark and two other datasets, it achieves state-of-the-art performance against strong baselines.

25 June 2026
Creative Agents

Unified visual-generation agentic model outperforms larger closed-source models

The authors propose VisionCreator, a native visual-generation agentic model that unifies Understanding, Thinking, Planning, and Creation (UTPC) capabilities. End-to-end trained on a new dataset and optimized via Progressive Specialization Training and Virtual Reinforcement Learning, VisionCreator-8B/32B models outperform larger closed-source models on the new VisGenBench benchmark across multiple dimensions.

3 Mar 2026