
MODELS
by LAION
LAION CLAP model checkpoint for multimodal audio-and-language representation and audio-classification workflows.
laion/clap-htsat-fused is a LAION model checkpoint listed on Hugging Face for audio-classification. Hugging Face Transformers documents CLAP as a multimodal audio-and-language model and includes examples that instantiate ClapAudioModel and ClapProcessor from laion/clap-htsat-fused. The official LAION-AI CLAP repository describes the project as providing audio and text representations via Contrastive Language-Audio Pretraining and links the implementation to LAION and the CLAP paper.
Use when you need a CLAP checkpoint from LAION for audio-and-language representation tasks. Use with Hugging Face Transformers examples that load ClapAudioModel and ClapProcessor from laion/clap-htsat-fused. Evaluate for audio-classification pipelines where a Hugging Face-hosted CLAP model checkpoint is appropriate.
Last updated Jun 4, 2026