Model Radar
AI models directory for capability-aware discovery
Focused on open models for self-hosting, research, and product development.
Model index
3
Currently filtered to open models
Model categories
Start from the workload you care about most
Model files
Open-source models
Not sure which model to inspect? Filter by openness, category, and keyword, then open profiles for capabilities, licensing, use cases, and related tools.
OpenAI
Whisper Large v3
Whisper Large v3 is a state-of-the-art open-source speech-to-text model developed by OpenAI, designed for robust multilingual transcription and translation. It excels in production audio workflows, offering high accuracy across diverse languages and acoustic conditions. The model is self-hostable, customizable, and widely used in voice applications, meeting accessibility, and batch processing pipelines.
laion
laion/clap-htsat-fused · Hugging Face
CLAP HTSAT-Fused is a contrastive language-audio pretraining model that learns a shared embedding space between audio and natural language descriptions. It uses HTSAT as the audio encoder and RoBERTa as the text encoder, with a feature fusion mechanism, trained on LAION-Audio-630K. It supports zero-shot audio classification, retrieval, and feature extraction.
hexgrad
hexgrad/Kokoro-82M · Hugging Face
Kokoro-82M is an open-weight text-to-speech model with 82 million parameters, fine-tuned from StyleTTS2-LJSpeech. It supports English and Arabic, delivers quality comparable to larger models, and is released under Apache 2.0. With over 16.7 million downloads, it is designed for efficient deployment in production and personal projects.