A 10-Million-Hour Open Video Dataset for Multimodal Pre-training
LAION eV
non-profit
AI & ML interests
open multi-modal foundation models and datasets for their creation; scaling laws, model evaluation; fully local, sovereign model deployment, personalized assistants and open local agentic systems
Recent Activity
View all activity
Collection of models and dataset related to MixtureVitae, open and fully reproducible pretraining dataset built from permissive sources
The full collection of our EmoNet effort. More info available at: https://huggingface.co/blog/felfri/emonet
Releases related to Open-ψ (Open-Sci) Collective
Re-LAION-5B-research
OpenCLIP models trained on DataComp (https://huggingface.co/papers/2304.14108).
-
laion/CLIP-ViT-L-14-DataComp.XL-s13B-b90K
Zero-Shot Image Classification • Updated • 53.3k • 126 -
laion/CLIP-ViT-B-16-DataComp.XL-s13B-b90K
Zero-Shot Image Classification • Updated • 9.71k • 9 -
laion/CLIP-ViT-B-32-256x256-DataComp-s34B-b86K
Zero-Shot Image Classification • Updated • 2.14k • 9 -
laion/CLIP-ViT-B-16-DataComp.L-s1B-b8K
Zero-Shot Image Classification • Updated • 743 • 2
CLAP is to audio what CLIP is to image.
-
laion/larger_clap_general
Feature Extraction • Updated • 317k • • 53 -
laion/larger_clap_music_and_speech
Feature Extraction • Updated • 149k • • 42 -
laion/larger_clap_music
Feature Extraction • Updated • 38.2k • • 49 -
laion/clap-htsat-fused
Audio Classification • 0.2B • Updated • 8.28M • 126
27 Delphi midtraining endpoint checkpoints (K=0.20): 9 scales x 3 mixes. Dense Qwen3, Llama-3 tok. marin#6279.
models and datasets related to openthoughts 4 experiments
-
laion/openthoughts-4-code-qwen3-32b-annotated-32k_qwen3-1.7B_32k
2B • Updated • 18 -
laion/openthoughts-4-code-qwen3-32b-annotated-32k_qwen2.5-1.5B_32k
Text Generation • 2B • Updated • 108 • 1 -
laion/openthoughts-3-QwQ-32b-annotated-16k_qwen2.5-1.5B_16k
Text Generation • 2B • Updated • 23 -
laion/openthoughts-4-code-qwen3-32b-annotated-7k_qwen3-1.7B_10k
Text Generation • 2B • Updated • 15
openMaMMUT/openCLIP models trained on DataComp-1.4B, DFN-1.4B and Re-LAION-2B. Pre-trained models on various scales, incl. intermediate checkpoints
-
laion/openMaMMUT-ViT-L-14-DataComp-1.4B-s12.8B-b180K
Zero-Shot Image Classification • Updated • 10 • 6 -
Scaling Laws for Robust Comparison of Open Foundation Language-Vision Models and Datasets
Paper • 2506.04598 • Published • 7 -
laion/openMaMMUT-ViT-L-14-512x512-pt_datacomp1b-ft_DFN512x512-s293M-b32k
Zero-Shot Image Classification • Updated • 10 • 2 -
laion/scaling-laws-for-comparison
Updated • 2
Re-LAION-5B research safe
OpenCLIP models trained on LAION-2B
-
laion/CLIP-ViT-bigG-14-laion2B-39B-b160k
Zero-Shot Image Classification • Updated • 78.9k • 317 -
laion/CLIP-ViT-g-14-laion2B-s34B-b88K
Zero-Shot Image Classification • Updated • 3.6k • 28 -
laion/CLIP-ViT-g-14-laion2B-s12B-b42K
1B • Updated • 14k • 44 -
laion/CLIP-ViT-H-14-laion2B-s32B-b79K
Zero-Shot Image Classification • 1.0B • Updated • 505k • 471
A 10-Million-Hour Open Video Dataset for Multimodal Pre-training
27 Delphi midtraining endpoint checkpoints (K=0.20): 9 scales x 3 mixes. Dense Qwen3, Llama-3 tok. marin#6279.
Collection of models and dataset related to MixtureVitae, open and fully reproducible pretraining dataset built from permissive sources
models and datasets related to openthoughts 4 experiments
-
laion/openthoughts-4-code-qwen3-32b-annotated-32k_qwen3-1.7B_32k
2B • Updated • 18 -
laion/openthoughts-4-code-qwen3-32b-annotated-32k_qwen2.5-1.5B_32k
Text Generation • 2B • Updated • 108 • 1 -
laion/openthoughts-3-QwQ-32b-annotated-16k_qwen2.5-1.5B_16k
Text Generation • 2B • Updated • 23 -
laion/openthoughts-4-code-qwen3-32b-annotated-7k_qwen3-1.7B_10k
Text Generation • 2B • Updated • 15
The full collection of our EmoNet effort. More info available at: https://huggingface.co/blog/felfri/emonet
Releases related to Open-ψ (Open-Sci) Collective
openMaMMUT/openCLIP models trained on DataComp-1.4B, DFN-1.4B and Re-LAION-2B. Pre-trained models on various scales, incl. intermediate checkpoints
-
laion/openMaMMUT-ViT-L-14-DataComp-1.4B-s12.8B-b180K
Zero-Shot Image Classification • Updated • 10 • 6 -
Scaling Laws for Robust Comparison of Open Foundation Language-Vision Models and Datasets
Paper • 2506.04598 • Published • 7 -
laion/openMaMMUT-ViT-L-14-512x512-pt_datacomp1b-ft_DFN512x512-s293M-b32k
Zero-Shot Image Classification • Updated • 10 • 2 -
laion/scaling-laws-for-comparison
Updated • 2
Re-LAION-5B-research
Re-LAION-5B research safe
OpenCLIP models trained on DataComp (https://huggingface.co/papers/2304.14108).
-
laion/CLIP-ViT-L-14-DataComp.XL-s13B-b90K
Zero-Shot Image Classification • Updated • 53.3k • 126 -
laion/CLIP-ViT-B-16-DataComp.XL-s13B-b90K
Zero-Shot Image Classification • Updated • 9.71k • 9 -
laion/CLIP-ViT-B-32-256x256-DataComp-s34B-b86K
Zero-Shot Image Classification • Updated • 2.14k • 9 -
laion/CLIP-ViT-B-16-DataComp.L-s1B-b8K
Zero-Shot Image Classification • Updated • 743 • 2
OpenCLIP models trained on LAION-2B
-
laion/CLIP-ViT-bigG-14-laion2B-39B-b160k
Zero-Shot Image Classification • Updated • 78.9k • 317 -
laion/CLIP-ViT-g-14-laion2B-s34B-b88K
Zero-Shot Image Classification • Updated • 3.6k • 28 -
laion/CLIP-ViT-g-14-laion2B-s12B-b42K
1B • Updated • 14k • 44 -
laion/CLIP-ViT-H-14-laion2B-s32B-b79K
Zero-Shot Image Classification • 1.0B • Updated • 505k • 471
CLAP is to audio what CLIP is to image.
-
laion/larger_clap_general
Feature Extraction • Updated • 317k • • 53 -
laion/larger_clap_music_and_speech
Feature Extraction • Updated • 149k • • 42 -
laion/larger_clap_music
Feature Extraction • Updated • 38.2k • • 49 -
laion/clap-htsat-fused
Audio Classification • 0.2B • Updated • 8.28M • 126