Diffuse-XL
A text-to-image diffusion model with photographic fidelity.
News3
<!-- SC_OFF --><div class="md"><p>I'm proposing a way to handle massive context longer than a model's context window by
Papers11
Limited-angle digital breast tomosynthesis (DBT) reconstructs a volume from a few low-dose projections over a narrow arc
paperSteering Optimisation Trajectories in Diffusion Representation LearningWe study why diffusion autoencoders can achieve similar image quality while learning substantially different latent stru
paperFADRA: Frequency-Aware Diffusion with Residual Adaptation for Video Face RestorationVideo face restoration (VFR) aims to recover high-quality and temporally consistent facial details from severely degrade
paperFeature-Space Guided Diffusion for Realistic Ultrasound Image SynthesisConditional diffusion models can generate anatomically plausible medical ultrasound (US) images, but anatomical plausibi
paperIntermediate Text Representation Guided Text-to-Image Generation for Enhancing One-and-Only AlignmentText-to-image (T2I) diffusion models often fail to faithfully render explicit textual descriptions, instead defaulting t
paperErasing Without Collateral Damage: Precise Concept Removal in Diffusion ModelsTraining-free concept erasure is an attractive mechanism for controlling text-to-image diffusion models, but precise era
paperMonocular Avatar Reconstruction via Cascaded Diffusion Priors and UV-Space Differentiable ShadingReconstructing high-fidelity, relightable 3D avatars from a single in-the-wild image is a challenging ill-posed problem,
paperPointDiT: Pixel-Space Diffusion for Monocular Geometry EstimationState-of-the-art single-image 3D reconstruction methods often rely on complex hybrid architectures and loss functions, o
