DeepSeek-V3.1
DeepSeek-V3.1 is a 671B/37B hybrid reasoning MoE with 128K context and MIT weights; its original provider API generation…
Compare leading AI models by provider, capabilities, modality, pricing and practical use cases. Every profile includes plain-English guidance, official sources and technical facts.
DeepSeek-V3.1 is a 671B/37B hybrid reasoning MoE with 128K context and MIT weights; its original provider API generation…
Cohere Command R is a live 128K-context enterprise model for RAG, tool use and multilingual chat, priced at…
Qwen1.5-1.8B-Chat is a compact multilingual chat checkpoint with a 32K context window, designed for instruction-following and local text…

Phi-4 Multimodal Instruct is Microsoft's 5.6B MIT-licensed model for text, image and audio input with a 128K context…
DeepSeek-V4-Pro is an active 1M-context model with 384K max output, tools and JSON support, plus peak/off-peak API pricing.
DeepSeek-Coder-V2 Lite is a 16B-total, 2.4B-active coding MoE family with 128K context. Its old provider API identity has…
Sora 2 is OpenAI's legacy synced-audio video model, still available through the Videos API at $0.10 per generated…
DeepSeek LLM 67B is an early English/Chinese model family with public Base and Chat weights and a 4096-token…
Qwen1.5-72B-Chat is a large multilingual 32K-context chat checkpoint. It remains downloadable, but Qwen now points users to newer…
Qwen1.5-110B-Chat is Qwen's April 2024 110B chat checkpoint with 32K context, multilingual support and substantial multi-GPU deployment requirements.
Phi-3 Mini 128K Instruct is Microsoft's compact 3.8B instruction model with 128K context, October 2023 cutoff and MIT-licensed…
Mistral 7B is an Apache 2.0 Mistral v0.3 checkpoint with 32K context; its managed API generation was retired…
Mistral 7B Instruct v0.3 is an Apache 2.0 Mistral v0.3 checkpoint with 32K context; its managed API generation…
Llama 3.1 8B is Meta's pretrained base checkpoint with 128K context, December 2023 cutoff and Llama 3.1 Community…
DeepSeek-V2 is a 236B-total, 21B-active MoE model with 128K context. Its weights remain available, but the V2 API…

Gemini 1.5 Flash is a discontinued Gemini 1.5 model. Google shut the API model down on September 29,…
Gemma 2 is Google's open-weight text family with 2B, 9B and 27B variants for generation, summarization, question answering…
Llama 3.1 405B is Meta's pretrained base checkpoint with 128K context, December 2023 cutoff and Llama 3.1 Community…
GPT-4o is OpenAI's 128K-context text-and-image API model with 16,384 max output, function calling, structured outputs and current token…
DALL-E 3 is a deprecated OpenAI image model that has been removed from the API; OpenAI recommends GPT-Image-2…
Eleven v3 is ElevenLabs' expressive speech synthesis model with 70+ languages, audio tags, dialogue generation and public API…

Phi-3 Vision 128K Instruct is Microsoft's 4.2B MIT-licensed multimodal model with text and image input, 128K context and…
Suno v3.5 is Suno's Summer 2024 music model with improved song structure and up to four-minute initial generations.
Grok-1 is xAI's Apache-2.0 open-weight 314B MoE base model with an 8192-token context window and Q3 2023 training…