LLM Models [49]
Browse LLM AI models on Segmind. Compare pricing, latency, and capabilities.
MiniMax M3
Reason over 1M-token context for coding and agents.
Nemotron 3 Ultra
1M-token reasoning for coding agents and deep research.
GLM 5.2
1M-token open-weight LLM for long-horizon coding.
Gemini 2.5 Flash Lite
Fastest Gemini 2.5 model for high-volume text and vision tasks.
Gemini 3.1 Flash Lite
Ultra-fast, affordable LLM for high-volume AI pipelines.
Gemini 3 Flash
Frontier-class reasoning and multimodal AI at scale.
Gemini 3.1 Pro
Frontier reasoning across text, images, video, and code.
GPT 5.5
Frontier reasoning and coding with 1M-token context window.
Claude Opus 4.7
Anthropic's most capable AI model excelling at agentic coding, complex reasoning, and high-resolution vision with a 1M-token context window.
Qwen Flash
Fastest low-cost LLM with 1M context for high-volume tasks.
Qwen Plus
Mid-tier 1M context LLM for summarization and content tasks.
QVQ Max
Chain-of-thought visual reasoning for math, charts, and diagrams.
Qwen 3 VL Flash
Fast, affordable vision-language model with 262K context OCR.
Qwen 3 VL Plus
Powerful visual QA and document analysis from images.
Qwen 3 Coder Flash
High-volume code generation with 1M token context window.
Qwen 3 Coder Plus
Generates, debugs, and refactors entire codebases efficiently.
QwQ Plus
Deep chain-of-thought reasoning for math, code, and logic.
Qwen 3 Max
1T-parameter LLM with hybrid reasoning and 262K context.
Qwen 3.5 Plus
Multimodal 1M context AI for image, video, and text.
Qwen 3.5 Flash
Fast multimodal AI processing text, images, and video affordably.
OpenAI o3 Mini
Cost-efficient reasoning model for coding, math, and science.
OpenAI o3
Frontier reasoning model for complex coding, math, and science.
GPT 5.4 Nano
Flagship-class AI for classification and extraction tasks.
GPT 5.4 Mini
Fastest efficient model for coding and computer-use tasks.
GPT 5.4
Most powerful GPT for frontier reasoning and multimodal tasks.
GPT 5.1
Precise code review and developer workflow assistant.
GPT 5.2
Advanced reasoning with multimodal input for precise tasks.
Gemini 3 Pro
Autonomous multimodal AI for complex reasoning and coding.
Kimi K2 Instruct 0905
Deep contextual understanding and complex code generation.
Claude 4 Sonnet
Advanced coding and multi-step agentic reasoning model.
GPT 5 Nano
Ultra-fast LLM responses for real-time AI applications.
GPT 5 Mini
Rapid high-quality AI across text, images, and files.
Gemini 2.5 Flash
Multimodal AI with transparent reasoning, fast and affordable.
Gemini 2.5 PRO
Complex multimodal reasoning across diverse inputs and formats.
Claude 4.5 Sonnet
Claude Sonnet 4.5 empowers developers with advanced coding and reasoning for complex software solutions.
GPT 5
GPT-5 automates complex coding tasks with integrated tools for seamless software development and deployment.
O4 Mini
OpenAI o4-mini enhances decision-making by processing text and images with advanced reasoning capabilities.
Llama 4 Scout Instruct Basic
Unlock powerful multimodal AI with Llama 4 Scout basic, a 17 billion active parameters model offering leading text & image understanding.
Llama 4 Maverick Instruct Basic
Llama 4 Maverick Instruct Basic is a 400B parameter powerhouse with 128 experts for unparalleled text and image understanding.
Qwen2 VL 72B Instruct
Qwen2-VL-72B-Instruct is a state-of-the-art multimodal model excelling in image and video understanding, with advanced capabilities for text-based interaction.
DeepSeek Chat
DeepSeek V3 combines cutting-edge AI technology with practical usability. Featuring a 671B parameter architecture, enhanced reasoning capabilities, and lightning-fast processing, it sets new standards for open-source AI models.
DeepSeek R1
DeepSeek-R1 is a cutting-edge AI reasoning model that combines reinforcement learning with supervised fine-tuning. Excels in complex problem-solving, mathematics, and coding tasks.
Llama 3.1 70b
Meta developed and released the Meta Llama 3 family of large language models (LLMs), a collection of pretrained and instruction tuned generative text models in 8 and 70B sizes. The Llama 3 instruction tuned models are optimized for dialogue use cases and outperform many of the available open source chat models on common industry benchmarks.
Llama 3.1 8b
Meta developed and released the Meta Llama 3 family of large language models (LLMs), a collection of pretrained and instruction tuned generative text models in 8 and 70B sizes. The Llama 3 instruction tuned models are optimized for dialogue use cases and outperform many of the available open source chat models on common industry benchmarks.
GPT 4 turbo
GPT-4 outperforms both previous large language models and as of 2023, most state-of-the-art systems (which often have benchmark-specific training or hand-engineering). On the MMLU benchmark, an English-language suite of multiple-choice questions covering 57 subjects, GPT-4 not only outperforms existing models by a considerable margin in English, but also demonstrates strong performance in other languages. Currently points to gpt-4-turbo-2024-04-09.
GPT 4o
GPT-4o (“o” for “omni”) is our most advanced model. It is multimodal (accepting text or image inputs and outputting text), and it has the same high intelligence as GPT-4 Turbo but is much more efficient—it generates text 2x faster and is 50% cheaper. Additionally, GPT-4o has the best vision and performance across non-English languages of any of our models. GPT-4o is available in the OpenAI API to paying customers.
GPT 4
GPT-4 outperforms both previous large language models and as of 2023, most state-of-the-art systems (which often have benchmark-specific training or hand-engineering). On the MMLU benchmark, an English-language suite of multiple-choice questions covering 57 subjects, GPT-4 not only outperforms existing models by a considerable margin in English, but also demonstrates strong performance in other languages.
Mixtral 8x22b
Mistral MoE 8x22B Instruct v0.1 model with Sparse Mixture of Experts. Fine tuned for instruction following.
Llama 3 8b
Meta developed and released the Meta Llama 3 family of large language models (LLMs), a collection of pretrained and instruction tuned generative text models in 8 and 70B sizes. The Llama 3 instruction tuned models are optimized for dialogue use cases and outperform many of the available open source chat models on common industry benchmarks.