JoLT: Joint Latent Trajectories for Context-Guided High-Resolution Tiled Generation
This study addresses the challenge of generating detail-rich, high-resolution images with text-to-image models by proposing a Joint Latent Trajectory method. The approach introduces a dual-stream synchronous denoising mechanism that dynamically integrates low-resolution layouts with high-resolution details through cross-branch information exchange during sampling, combined with a patch-wise generation strategy for coordinated control. Experimental results demonstrate that this method effectively resolves the structural-textural imbalance inherent in high-definition generation, producing visually appealing images with rich details. Significantly outperforming existing baselines, this work establishes a novel paradigm for high-resolution image synthesis.