SCAPES: Lightweight 36M Parameter Generative Model for Environmental Sounds
September 7, 2026
SCAPES is a 36-million parameter autoregressive prior for synthesizing environmental textures using Flow Matching on continuous latent manifolds. It achieves high-fidelity audio synthesis with high-level semantic control while avoiding the constraints of discrete tokenization.
HOW THIS AFFECTS YOU
●
builderThis offers a computationally efficient alternative to massive audio generative models for texture synthesis.
●
designerYou can generate high-fidelity environmental audio textures with semantic control using a lightweight model.