Search papers, labs, and topics across Lattice.
This paper introduces CamoDreamer, a novel context-decoupled generative paradigm for camouflage image generation (CIG) that addresses the limitations of existing methods by isolating object and background cues. By employing a Contrast-aware Contextual Bridge and Context-Decoupled Assimilation Streams, the approach effectively mitigates appearance discrepancies and background artifacts, resulting in improved camouflage fidelity. Experimental results show that CamoDreamer significantly outperforms traditional techniques while maintaining a lightweight architecture, highlighting its practical applicability in generating visually concealed objects.
CamoDreamer achieves unprecedented camouflage fidelity by decoupling object and background cues, eliminating the artifacts that plague existing methods.
Camouflage image generation (CIG) focuses on generating visually concealed objects that seamlessly blend into their backgrounds. Existing methods typically follow either background-guided paradigms that adapt object appearance via style transfer, or foreground-guided strategies that outpaint surrounding regions conditioned on object features. However, they still suffer from appearance discrepancy and background artifacts. We attribute these limitations to cross-context representation leakage, where object and background cues are entangled in a coupled conditional space, resulting in ambiguous control and degraded camouflage fidelity. To tackle this, we propose a new context-decoupled generative paradigm, termed CamoDreamer, which aims to isolate contextual conditional guidance and explicitly decouple latent camouflage features into coordinated object and background control streams. First, a Contrast-aware Contextual Bridge is designed to model cross-context discrepancies and construct contrast-aware dual conditional guidance. Second, Context-Decoupled Assimilation Streams are employed to separate generative interactions conditioned on the dual guidance, while facilitating background rendering with target-aware cues in the latent space. Finally, a Frequency-Adaptive Contextual Blend module integrates complementary high-frequency textures and low-frequency structures from decoupled features to improve holistic coherence. Extensive experiments demonstrate that CamoDreamer consistently outperforms existing methods with a substantial margin, while maintaining a relatively lightweight design.