[2603.04122] FastWave: Optimized Diffusion Model for Audio

[2603.04122] FastWave: Optimized Diffusion Model for Audio Super-Resolution

arXiv - Machine Learning March 05, 2026 3 min read

About this article

Abstract page for arXiv paper 2603.04122: FastWave: Optimized Diffusion Model for Audio Super-Resolution

Computer Science > Sound arXiv:2603.04122 (cs) [Submitted on 4 Mar 2026] Title:FastWave: Optimized Diffusion Model for Audio Super-Resolution Authors:Nikita Kuznetsov, Maksim Kaledin View a PDF of the paper titled FastWave: Optimized Diffusion Model for Audio Super-Resolution, by Nikita Kuznetsov and 1 other authors View PDF HTML (experimental) Abstract:Audio Super-Resolution is a set of techniques aimed at high-quality estimation of the given signal as if it would be sampled with higher sample rate. Among suggested methods there are diffusion and flow models (which are considered slower), generative adversarial networks (which are considered faster), however both approaches are currently presented by high-parametric networks, requiring high computational costs both for training and inference. We propose a solution to both these problems by re-considering the recent advances in the training of diffusion models and applying them to super-resolution from any to 48 kHz sample rate. Our approach shows better results than NU-Wave 2 and is comparable to state-of-the-art models. Our model called FastWave has around 50 GFLOPs of computational complexity and 1.3 M parameters and can be trained with less resources and significantly faster than the majority of recently proposed diffusion- and flow-based solutions. The code has been made publicly available. Subjects: Sound (cs.SD); Machine Learning (cs.LG) Cite as: arXiv:2603.04122 [cs.SD] (or arXiv:2603.04122v1 [cs.SD] for this ver...

Originally published on March 05, 2026. Curated by AI News.

Llms

CLI for Google AI Search (gai.google) — run AI-powered code/tech searches headlessly from your terminal

Google AI (gai.google) gives Gemini-powered answers for technical queries — think AI-enhanced search with code understanding. I built a C...

Reddit - Artificial Intelligence · 1 min · 29 minutes ago

Machine Learning

Big increase in the amount of people using AI to write their replies with AI

I find it interesting that we’ve all randomly decided to use the “-“ more often recently on reddit, and everyone’s grammar has drasticall...

Reddit - Artificial Intelligence · 1 min · 29 minutes ago

Machine Learning

[D] MXFP8 GEMM: Up to 99% of cuBLAS performance using CUDA + PTX

New blog post by Daniel Vega-Myhre (Meta/PyTorch) illustrating GEMM design for FP8, including deep-dives into all the constraints and des...

Reddit - Machine Learning · 1 min · about 1 hour ago

Machine Learning

IIT Delhi launches 8th batch of Advanced AI, ML, and DL online programme: Check who is eligible, applicat

News News: The Continuing Education Programme (CEP) at IIT Delhi has announced the launch of the 8th batch of its Advanced Certificate Pr...

AI News - General · 9 min · about 2 hours ago

[2603.04122] FastWave: Optimized Diffusion Model for Audio Super-Resolution

About this article

Related Articles

CLI for Google AI Search (gai.google) — run AI-powered code/tech searches headlessly from your terminal

Big increase in the amount of people using AI to write their replies with AI

[D] MXFP8 GEMM: Up to 99% of cuBLAS performance using CUDA + PTX

IIT Delhi launches 8th batch of Advanced AI, ML, and DL online programme: Check who is eligible, applicat

No comments

Stay updated with AI News