Model Radar
AI models directory for capability-aware discovery
Track AI models by developer, parameter scale, license, and practical use case.
Model index
22
Coverage across model profiles and internal links
Model categories
Start from the workload you care about most
Model files
Latest AI models
Not sure which model to inspect? Filter by openness, category, and keyword, then open profiles for capabilities, licensing, use cases, and related tools.
Kunlun Tech
Tiangong AI
Tiangong AI is a comprehensive artificial intelligence platform developed by Kunlun Tech, designed primarily for Chinese-speaking users. It functions as an integrated ecosystem combining advanced AI-powered search capabilities with robust creative content generation tools. Unlike standalone large language models that focus solely on text completion or chat, Tiangong AI positions itself as a productivity suite that bridges the gap between information retrieval and content production. By leveraging large language models (LLMs) under the hood, it allows users to perform complex searches, synthesize real-time web data, and immediately transform those findings into structured articles, reports, or creative writing. The platform is accessible via web and mobile applications, emphasizing ease of use for professionals, students, and content creators operating within the Chinese digital landscape.
360
360 AI Search
360 AI Search is a proprietary, LLM-integrated search service developed by Qihoo 360. It merges traditional web indexing with generative AI to provide direct, conversational answers while citing source URLs. Optimized primarily for the Chinese market, it operates within the 360 ecosystem, including the 360 Browser, and is not available as an open-source model or downloadable package.
OpenAI
GPT-4.1
GPT-4.1 is a hosted large language model from OpenAI optimized for practical software, agent, and business workflows. It is designed for general reasoning, coding assistance, and research tasks, with an API-first deployment model that simplifies integration but ties operations to OpenAI's infrastructure.
Moonshot AI
Kimi
Kimi is a large language model developed by Moonshot AI, designed to handle extremely long context windows of up to 2 million tokens, making it suitable for processing extensive documents and research materials.
Tencent
Tencent Hunyuan
Tencent Hunyuan is a large language model family developed by Tencent, powering the company's AI assistant and cloud AI services. It is designed for general-purpose natural language understanding and generation, with capabilities including text completion, question answering, and dialogue. The model is available in various sizes, with the largest version having over 100 billion parameters. Hunyuan is deployed on Tencent Cloud and integrated into Tencent's ecosystem, such as WeChat and QQ. It is not open-source; access is primarily through Tencent's API or cloud platform.
ByteDance
Doubao
Doubao is a family of AI models developed by ByteDance, targeting daily Q&A, writing, and multimodal tasks in the Chinese market. It includes large language models (Doubao-Pro, Doubao-Lite) and a multimodal model (Doubao-Vision). The models are optimized for Chinese language understanding and generation, with capabilities in text, image, and audio processing. Doubao models are available via API and have been integrated into ByteDance's consumer products like Douyin and Feishu. The models are proprietary, with no open-source release or public deployment details confirmed.
Anthropic
Claude 3.7 Sonnet
Claude 3.7 Sonnet is a hosted AI model from Anthropic, optimized for coding, reasoning, and enterprise writing. It excels in long-context tasks and is often compared with top models for its reliability in code generation, document analysis, and research workflows.
Mistral AI
Mistral Large 2
Mistral Large 2 is a proprietary large language model developed by Mistral AI, designed for general-purpose reasoning, coding, and enterprise applications. It offers strong performance in multilingual tasks and long-context understanding, with deployment primarily through hosted APIs and cloud partners.
Meta
Llama 3.3 70B
Llama 3.3 70B is a large language model from Meta, designed for general assistant and reasoning tasks. It is part of the Llama family and is known for its strong performance, open weights, and extensive ecosystem support. The model is suitable for self-hosted or customizable AI stacks, making it a popular choice for teams that need a capable open model with fine-tuning and inference flexibility.
OpenAI
GPT-4o
GPT-4o is a multimodal frontier model from OpenAI that processes text, images, and audio with low latency, designed for general-purpose assistant and reasoning tasks. It balances broad capability with strong multimodal utility, making it suitable for product-facing use cases. Deployment is primarily via hosted APIs or ChatGPT, offering fast adoption but vendor dependency.
Gemini 2.0 Flash
Gemini 2.0 Flash is a fast, multimodal AI model from Google, optimized for responsive assistant experiences, reasoning, and agentic systems. It is designed for deployment within Google's managed ecosystem, offering low latency and strong alignment with Google Cloud and AI platform tools.
Mistral AI
Mixtral 8x22B Instruct
Mixtral 8x22B Instruct is a sparse mixture-of-experts (MoE) large language model developed by Mistral AI, designed for self-hosted deployment and fine-tuning. It balances high performance with open availability, making it a strong candidate for teams seeking control over their AI stack without relying on closed APIs.
Anthropic
Claude 3.5 Sonnet
Claude 3.5 Sonnet is a mid-tier large language model by Anthropic, balancing capability and cost for general reasoning, coding, and long-context tasks. It is commonly evaluated for its strong performance in code review, long-form analysis, and deliberate written output, with deployment via hosted endpoints or first-party chat surfaces.
Alibaba
Qwen2.5 72B Instruct
Qwen2.5 72B Instruct is an open-weight large language model developed by Alibaba Cloud, optimized for instruction-following tasks with strong bilingual support in Chinese and English. It is designed for self-hosted or customizable AI stacks, general assistant and reasoning workloads, and coding assistants. The model offers practical deployment flexibility through open-weight or partner stacks, balancing control, cost, and compliance.
Cohere
Command R+
Command R+ is an enterprise-focused large language model optimized for retrieval-augmented generation (RAG), reasoning, and document-grounded tasks. It is designed for teams that need strong retrieval capabilities without relying on the largest platform vendors, and is typically accessed via partner clouds or managed APIs.
Alibaba Cloud
Qwen
Qwen is a family of large language models developed by Alibaba Cloud, covering chat, code, and multimodal variants. Designed for enterprise deployment, Qwen models are available in various sizes (1.8B to 72B parameters) and support both open-source and commercial use under the Qwen License. They excel in natural language understanding, code generation, mathematical reasoning, and multimodal tasks, with options for local deployment, cloud API, and fine-tuning.
Zhipu AI
GLM-4
GLM-4 is the latest generation of the General Language Model (GLM) series developed by Zhipu AI. It is a bilingual (Chinese and English) large language model with 130 billion parameters, supporting both API access and open-source releases. The model excels in long-context understanding, multimodal tasks, and agent-based applications, with a context window of up to 128K tokens. It is designed for a wide range of NLP tasks including text generation, reasoning, and code synthesis.
Gemini 1.5 Pro
Gemini 1.5 Pro is a multimodal AI model from Google DeepMind, known for its exceptionally long context window (up to 2 million tokens) and strong reasoning capabilities. It is designed for complex tasks involving large documents, code, audio, video, and images, and is primarily accessed via Google Cloud's Vertex AI and the Gemini API.
Gemma 2 27B
Gemma 2 27B is an open-weight language model from Google, designed for self-hosted and customizable AI stacks. It offers strong performance for general assistant and reasoning tasks, backed by Google's ecosystem and available under a permissive license.
MiniMax
MiniMax ABAB
MiniMax ABAB is a family of large language models developed by MiniMax, a Chinese AI startup. The models are designed for chat, voice, and multimodal applications, offering API access for developers. The flagship model, ABAB, is a dense transformer model with 1.8 trillion parameters, trained on a large corpus of text and code. It supports long context windows (up to 256k tokens) and is optimized for reasoning, instruction following, and multilingual tasks. The model is available via API and has been integrated into various applications, including conversational AI and content generation.