LAION-BVD provides a massive 10-million-hour corpus of web video, audio, and image-text pairs. It is designed to support large-scale multimodal pre-training for video and audio models.
HOW THIS AFFECTS YOU
●
builderThis provides a massive open alternative to proprietary video datasets.
●
researcherYou can use this for large-scale multimodal pre-training.