Llama 3.1 70B
Llama 3.1 70B is Meta's pretrained base checkpoint with 128K context, December 2023 cutoff and Llama 3.1 Community…
Compare leading AI models by provider, capabilities, modality, pricing and practical use cases. Every profile includes plain-English guidance, official sources and technical facts.
Llama 3.1 70B is Meta's pretrained base checkpoint with 128K context, December 2023 cutoff and Llama 3.1 Community…
CodeGemma is Google's Gemma-based coding family with 2B and 7B variants for code completion, generation and instruction-following programming…
Qwen1.5-1.8B is a compact 32K-context base model from Qwen, best suited to fine-tuning, research and custom model development…
Qwen1.5-1.8B-Chat is a compact multilingual chat checkpoint with a 32K context window, designed for instruction-following and local text…
DeepSeek-Coder-V2 Lite is a 16B-total, 2.4B-active coding MoE family with 128K context. Its old provider API identity has…
DeepSeek LLM 67B is an early English/Chinese model family with public Base and Chat weights and a 4096-token…
Qwen1.5-72B-Chat is a large multilingual 32K-context chat checkpoint. It remains downloadable, but Qwen now points users to newer…
Qwen1.5-110B-Chat is Qwen's April 2024 110B chat checkpoint with 32K context, multilingual support and substantial multi-GPU deployment requirements.
Mistral 7B is an Apache 2.0 Mistral v0.3 checkpoint with 32K context; its managed API generation was retired…
Mistral 7B Instruct v0.3 is an Apache 2.0 Mistral v0.3 checkpoint with 32K context; its managed API generation…
Llama 3.1 8B is Meta's pretrained base checkpoint with 128K context, December 2023 cutoff and Llama 3.1 Community…
DeepSeek-V2 is a 236B-total, 21B-active MoE model with 128K context. Its weights remain available, but the V2 API…
Gemma 2 is Google's open-weight text family with 2B, 9B and 27B variants for generation, summarization, question answering…
Llama 3.1 405B is Meta's pretrained base checkpoint with 128K context, December 2023 cutoff and Llama 3.1 Community…

Phi-3 Mini is Microsoft's 3.8B MIT-licensed family with 4K and 128K instruction variants for compact local and edge…
DeepSeek-Coder-V2 is a 236B-total, 21B-active coding MoE family with 128K context. Its old API line was merged into…
Gemma 4 is Google's active Apache-2.0 multimodal family with 128K–256K context, 140+ languages, reasoning and function calling.
Llama 3.1 is Meta's 8B, 70B and 405B family with 128K context, eight supported languages and downloadable weights.
Qwen1.5 is Alibaba's 2024 open-weight multilingual family with base/chat models, quantized variants and a uniform 32K context window.
DeepSeek-R1 is the January 2025 MIT-licensed reasoning model with 671B total, 37B active parameters and 128K context. Its…
Grok-1.5V was xAI's first multimodal Grok model for documents, charts, screenshots, diagrams and photographs; it is now a…