OBS-Diff: Accurate Pruning For Diffusion Models In One-Shot - Takara TLDR

Large-scale text-to-image diffusion models, while powerful, suffer from
prohibitive computational cost. Existing one-shot network pruning methods can
hardly be directly applied to them due to the iterative denoising nature of
diffusion models. To bridge the gap, this paper presents OBS-Diff, a novel
one-shot pruning framework that enables accurate and training-free compression
of large-scale text-to-image diffusion models. Specifically, (i) OBS-Diff
revitalizes the classic Optimal Brain Surgeon (OBS), adapting it to the complex
architectures of modern diffusion models and supporting diverse pruning
granularity, including unstructured, N:M semi-structured, and structured (MHA
heads and FFN neurons) sparsity; (ii) To align the pruning criteria with the
iterative dynamics of the diffusion process, by examining the problem from an
error-accumulation perspective, we propose a novel timestep-aware Hessian
construction that incorporates a logarithmic-decrease weighting scheme,
assigning greater importance to earlier timesteps to mitigate potential error
accumulation; (iii) Furthermore, a computationally efficient group-wise
sequential pruning strategy is proposed to amortize the expensive calibration
process. Extensive experiments show that OBS-Diff achieves state-of-the-art
one-shot pruning for diffusion models, delivering inference acceleration with
minimal degradation in visual quality.

Source link

What's Hot

Risk of Overqualified Candidates | Recruiting News Network

Indian Techie Uses Perplexity’s Comet Browser To Complete Coursera AI Course In Seconds; CEO Aravind Srinivas Responds

DexNDM: Closing the Reality Gap for Dexterous In-Hand Rotation via Joint-Wise Neural Dynamics Model – Takara TLDR

OBS-Diff: Accurate Pruning For Diffusion Models in One-Shot – Takara TLDR

DexNDM: Closing the Reality Gap for Dexterous In-Hand Rotation via Joint-Wise Neural Dynamics Model – Takara TLDR

Agent Learning via Early Experience – Takara TLDR

SciVideoBench: Benchmarking Scientific Video Reasoning in Large Multimodal Models – Takara TLDR

Frieze to Launch Abu Dhabi Fair in November 2026

Jeff Koons Returns to Gagosian with First New York Show in Seven Years

Ancient Egyptian Iconography Found in Roman-Era Bathhouse in Turkey

London Gallery Harlesden High Street Goes to Mayfair For a Pop-up

Risk of Overqualified Candidates | Recruiting News Network

Indian Techie Uses Perplexity’s Comet Browser To Complete Coursera AI Course In Seconds; CEO Aravind Srinivas Responds

DexNDM: Closing the Reality Gap for Dexterous In-Hand Rotation via Joint-Wise Neural Dynamics Model – Takara TLDR

What's Hot

OBS-Diff: Accurate Pruning For Diffusion Models in One-Shot – Takara TLDR

Related Posts

Subscribe to Updates