Aphanta: Diagnosing Task-Aligned Image-Edited Intermediates for Multimodal Reasoning Paper • 2608.26993 • Published 7 days ago • 5
Aphanta: Diagnosing Task-Aligned Image-Edited Intermediates for Multimodal Reasoning Paper • 2608.26993 • Published 7 days ago • 5
Aphanta: Diagnosing Task-Aligned Image-Edited Intermediates for Multimodal Reasoning Paper • 2608.26993 • Published 7 days ago • 5
WithEveryone: Unified Planning and Identity Grounding for Group Image Generation Paper • 2608.20336 • Published 14 days ago • 42
WithEveryone: Unified Planning and Identity Grounding for Group Image Generation Paper • 2608.20336 • Published 14 days ago • 42
Translation as a Bridging Action: Transferring Manipulation Skills from Humans to Robots Paper • 2606.28133 • Published Jun 26 • 40
LUCID: Learning Unified Control for Image Deflaring and Exposure Mastery in Nighttime Photography Paper • 2606.06901 • Published Jun 5 • 4
FreeStyle: Free Control of Style-Content Dual-Reference Generation from Community LoRA Mining Paper • 2606.20506 • Published Jun 18 • 28
FreeStyle: Free Control of Style-Content Dual-Reference Generation from Community LoRA Mining Paper • 2606.20506 • Published Jun 18 • 28
FreeStyle: Free Control of Style-Content Dual-Reference Generation from Community LoRA Mining Paper • 2606.20506 • Published Jun 18 • 28
ControlLight: Towards Controllable, Consistent, and Generalizable Low-Light Enhancement Paper • 2605.25569 • Published May 25 • 23
CutClaw: Agentic Hours-Long Video Editing via Music Synchronization Paper • 2603.29664 • Published Mar 31 • 51
GEditBench v2: A Human-Aligned Benchmark for General Image Editing Paper • 2603.28547 • Published Mar 30 • 32
GEditBench v2: A Human-Aligned Benchmark for General Image Editing Paper • 2603.28547 • Published Mar 30 • 32
PixelSmile: Toward Fine-Grained Facial Expression Editing Paper • 2603.25728 • Published Mar 26 • 118
RealRestorer: Towards Generalizable Real-World Image Restoration with Large-Scale Image Editing Models Paper • 2603.25502 • Published Mar 26 • 58