
Agent Sketches One Part at a Time
Researchers developed a multi-modal LLM agent that generates vector sketches part-by-part using multi-turn process-reward RL after supervised fine-tuning. They created the ControlSketch-Part dataset with rich part-level annotations via an automatic segmentation and labeling pipeline. This enables interpretable, controllable, and editable text-to-vector sketch generation with visual feedback.






