Spesifikasi, Harga & Fitur Model AI
Bandingkan 352 model AI dari 192 provider. Context window, pricing, reasoning, tool calling, dan multimodal — semua di satu tempat.
Hy3 preview
tencent · Hy
Tencent Hy reasoning model for coding, instruction following, and agent tasks
Hy3
tencent · Hy
Tencent Hy reasoning model for coding, instruction following, and agent tasks
Llama 3.3 Nemotron Super 49B v1
nvidia · nemotron
Nemotron model for efficient reasoning, coding, and specialized AI agents
Nemotron 3 Nano 30B A3B
nvidia · nemotron
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
Nemotron 3.5 Content Safety
nvidia · nemotron
Safety model for policy screening, moderation, and risk-aware routing workflows
Nemotron Cascade 2 30B A3B
nvidia · nemotron
Nemotron model for efficient reasoning, coding, and specialized AI agents
Nemotron 3 Content Safety
nvidia · nemotron
Safety model for policy screening, moderation, and risk-aware routing workflows
Llama 3.1 Nemotron 70B Instruct
nvidia · nemotron
Nemotron model for efficient reasoning, coding, and specialized AI agents
Nemotron Nano 12B v2 VL
nvidia · nemotron
Nemotron multimodal model for visual reasoning and agentic AI workflows
Nemotron VoiceChat
nvidia · nemotron
Nemotron multimodal model for visual reasoning and agentic AI workflows
Llama Nemotron Rerank VL 1B v2
nvidia · nemotron
Reranking model for improving retrieval quality in search and recommendation systems
Llama 3.3 Nemotron Super 49B v1.5
nvidia · nemotron
Nemotron model for efficient reasoning, coding, and specialized AI agents
Nemotron 3 Super 120B A12B
nvidia · nemotron
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
Nemotron 3 Nano Omni 30B A3B Reasoning
nvidia · nemotron
Open Nemotron omni model combining reasoning with text, vision, and audio
Nemotron 3.5 Lightning 30B A3B
nvidia · nemotron
Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads
Nemotron Content Safety Reasoning 4B
nvidia · nemotron
Safety model for policy screening, moderation, and risk-aware routing workflows
Nemotron 3 Ultra 550B A55B
nvidia · nemotron
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Mistral Nemotron
nvidia · nemotron
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
Llama 3.1 Nemotron Safety Guard 8B v3
nvidia · nemotron
Safety model for policy screening, moderation, and risk-aware routing workflows
Nemotron Nano 9B v2
nvidia · nemotron
Compact Nemotron model for efficient reasoning and deployable AI agents
Llama 3.1 Nemotron Ultra 253B
nvidia · nemotron
Flagship Nemotron model for high-throughput reasoning and complex agents
Nemotron Mini 4B Instruct
nvidia · nemotron
Compact Nemotron model for efficient reasoning and deployable AI agents
Llama Nemotron Embed VL 1B v2
nvidia · nemotron
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
Gemma-SEA-LION-v4-27B-IT
aisingapore · gemma
Gemma 3 27B tuned by AI Singapore for Southeast Asian languages and instruction following
Phi-4-mini
microsoft · phi
Compact Microsoft instruction model tuned for efficient coding assistance, reasoning, and low-latency agent tasks
MAI-Code-1-Flash
microsoft · mai
Microsoft coding model built for fast, efficient assistance in everyday developer workflows
MAI-Code-1.1-Flash
microsoft · mai
Microsoft coding model with native vision support, optimized for fast and efficient software development
DeepSeek-R1-Distill-Qwen-32B
deepseek · deepseek-thinking
R1 reasoning distilled into Qwen 2.5 32B for efficient open-weight step-by-step problem solving
DeepSeek-R1
deepseek · deepseek-thinking
Classic open reasoning model for transparent math, coding, and deliberate problem solving
DeepSeek V3 0324
deepseek · deepseek
March 2025 checkpoint of DeepSeek-V3 with improved reasoning and coding
DeepSeek V4 Flash
deepseek · deepseek-flash
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
DeepSeek V4 Pro 0813
deepseek · deepseek-thinking
DeepSeek V4 Pro snapshot with million-token context and support for thinking and non-thinking modes
DeepSeek-V3.1
deepseek · deepseek
Hybrid-reasoning DeepSeek model with thinking and non-thinking modes
DeepSeek V4 Pro
deepseek · deepseek-thinking
Open MoE flagship with million-token context for coding and long agent runs
DeepSeek-V3
deepseek · deepseek
Open DeepSeek MoE chat model for coding, math, and general reasoning
DeepSeek V3.2
deepseek · deepseek
Hybrid-reasoning DeepSeek model with thinking and non-thinking modes, sparse attention, and tool-use
DeepSeek Chat
deepseek · deepseek
DeepSeek chat model for instruction following, coding, and analysis
DeepSeek Reasoner
deepseek · deepseek-thinking
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
DeepSeek OCR 2
deepseek ·
High-accuracy OCR model for extracting text from documents, screenshots, receipts, and natural scenes
DeepSeek V4 Flash 0731
deepseek · deepseek-flash
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
Trinity Large Preview
arcee-ai · trinity
Lightly post-trained 398B MoE chat model for creative work, long-context prompts, and tool-using agents
Trinity Mini
arcee-ai · trinity
Reasoning-tuned 26B MoE model with 3B active parameters for agents, tools, and multi-step workloads
Trinity Large Thinking
arcee-ai · trinity
Reasoning-optimized 398B MoE agent model with extended thinking for long-horizon and multi-turn tool use
Trinity Nano Preview
arcee-ai · trinity
Experimental chat-tuned 6B MoE model with 1B active parameters for low-resource chat and instruction following
Gemini 3.1 Flash TTS Preview
google · gemini-flash
Low-latency speech generation with steerable prompts and expressive audio tags
Gemma 4 26B A4B IT
google · gemma
Open Gemma instruction model for efficient chat and self-hosted deployments
Nano Banana Pro
google · gemini-pro
Nano Banana Pro for higher-fidelity image generation and design-heavy edits
Nano Banana Pro
google · gemini-pro
Nano Banana Pro for higher-fidelity image generation and design-heavy edits
Gemini Flash-Lite Latest
google · gemini-flash-lite
Low-latency Gemini model for high-volume multimodal and agent workloads
Gemini 2.0 Flash
google · gemini-flash
Earlier Gemini Flash workhorse for responsive multimodal apps and tool use
Gemini 2.5 Flash TTS
google · gemini-flash
Speech generation model for controllable voice, narration, and audio delivery
Gemini 3.5 Flash Lite
google · gemini-flash-lite
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Nano Banana 2 Lite
google · gemini-flash-lite
Fastest, most cost-efficient Gemini image model for high-volume 1K generation and editing
Gemini 3.1 Pro Preview
google · gemini-pro
Reasoning-first Gemini preview for agentic coding and complex problem solving
Gemini 3.1 Pro Preview Custom Tools
google · gemini-pro
Advanced Gemini model for complex reasoning, coding, and multimodal analysis
Gemini Deep Research Preview
google · gemini-pro
Agentic model for autonomous multi-step research, synthesis, and cited reports
Gemini 2.5 Flash-Lite
google · gemini-flash-lite
Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents
Gemini Robotics-ER 1.6 Preview
google · gemini
Vision-language model for embodied reasoning: spatial understanding, task planning, and physical-world agentic robotics
Gemma 4 31B IT
google · gemma
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
Gemini 2.5 Computer Use Preview
google · gemini-pro
Specialized Gemini 2.5 model for browser-control agents that automate UI tasks
Lyria 3 Clip Preview
google · lyria
Music generation model for short 30-second clips, loops, and previews from text or image prompts
Veo 3.1 Fast Preview
google · veo
Video model for prompt-guided generation, editing, and motion workflows
Gemini Embedding 001
google · gemini
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
Gemini Flash Latest
google · gemini-flash
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Gemini 3.5 Flash
google · gemini-flash
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Gemini 3.5 Live Translate Preview
google · gemini-pro
Low-latency audio-to-audio model for real-time speech translation across 70+ languages
Veo 3.1 Lite Preview
google · veo
Video model for prompt-guided generation, editing, and motion workflows
Gemini Omni Flash Preview
google · gemini
Video generation and editing model for fast, conversational text- and image-to-video workflows
Gemini 3.1 Flash Lite
google · gemini-flash-lite
Low-latency Gemini model for high-volume multimodal and agent workloads
Gemini 2.5 Flash
google · gemini-flash
Fast Gemini workhorse for multimodal apps where latency and price matter
Gemini 3 Pro Preview
google · gemini-pro
Preview Gemini flagship for complex reasoning, coding, and rich multimodal prompts
Veo 3.1 Preview
google · veo
Video model for prompt-guided generation, editing, and motion workflows
Gemini 2.5 Pro TTS
google · gemini-pro
Speech generation model for controllable voice, narration, and audio delivery
Gemini Embedding 2
google · gemini
Multimodal embedding model mapping text, images, video, audio, and PDFs into a unified embedding space
Gemma 4 E4B IT
google · gemma
Open Gemma instruction model for efficient chat and self-hosted deployments
Gemma 4 E2B IT
google · gemma
Open Gemma instruction model for efficient chat and self-hosted deployments
Gemini 3.6 Flash
google · gemini-flash
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Nano Banana 2
google · gemini-flash
Image model for prompt-driven generation, editing, and visual design workflows
Nano Banana 2
google · gemini-flash
Image model for prompt-driven generation, editing, and visual design workflows
Gemini 2.0 Flash-Lite
google · gemini-flash-lite
Low-latency Gemini model for high-volume multimodal and agent workloads
Lyria 3 Pro Preview
google · lyria
Music generation model for full-length songs from text or images with vocals and structure
Gemini 3.1 Flash Lite Preview
google · gemini-flash-lite
Low-latency Gemini model for high-volume multimodal and agent workloads
Gemini 3 Flash Preview
google · gemini-flash
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs
Nano Banana
google · gemini-flash
Nano Banana image model for fast generation, edits, and character-consistent assets
Gemini 2.5 Pro
google · gemini-pro
Google's proven reasoning model for coding, math, and multimodal analysis
Deep Research Max Preview
google · gemini-pro
Maximum-comprehensiveness agentic researcher for multi-step investigation, synthesis, and cited reports
Gemini 3.1 Flash Live Preview
google · gemini-flash
High-quality, low-latency Live API model for real-time dialogue and voice-first AI applications
Gemini 3.7 Flash
google · gemini-flash
High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning
Llama-Guard-3-8B
meta · llama
Llama 3.1-based safety classifier for moderating prompts and model responses
Llama-3.3-70B-Instruct
meta · llama
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
Llama 4 Scout 17B Instruct
meta · llama
Open Llama with long-context vision for efficient multimodal agents
Muse Glimmer 30B
meta · muse
Muse Glimmer is a 30-billion-parameter open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark for always-on local agents, tool use, coding, and image understanding.
Llama-3.1-8B-Instruct
meta · llama
Compact open Llama model for lightweight chat, drafting, and self-hosting
Muse Spark 1.2
meta · muse
Muse Spark 1.2 is a coding-focused update to Muse Spark 1.1 with improvements in code generation, complex debugging, codebase understanding, and end-to-end developer workflows.
Muse Spark 1.1
meta · muse
Muse Spark is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration.
Llama-3.2-1B
meta · llama
Compact open Llama base model for lightweight and on-device use
Llama 4 Maverick 17B Instruct
meta · llama
Open multimodal Llama for strong reasoning with efficient everyday serving
Llama-3.2-3B
meta · llama
Small open Llama base model for lightweight text generation and self-hosting
Llama-3.2-11B-Vision-Instruct
meta · llama
Open multimodal Llama model for image understanding, captioning, and visual QA
ALLaM-2-7b
sdaia ·
ALLaM-2-7b instruction tuned model by SDAIA
Laguna M.1
poolside · laguna
Poolside's open-weight model for agentic coding and long-horizon work
Laguna S 2.1
poolside · laguna
Agentic coding model from Poolside in the XS size class for local deployment
Laguna XS 2.1
poolside · laguna
Agentic coding model from Poolside in the XS size class for local deployment
Laguna XS.2
poolside · laguna
Agentic coding model from Poolside in the XS size class for local deployment
Seed 2.0 Code
bytedance-seed · seed
ByteDance Seed coding model for multimodal software engineering and long-running agents
Seed 2.1 Turbo
bytedance-seed · seed
Faster ByteDance Seed 2.1 model for multimodal reasoning and latency-sensitive agent workflows
Seed 1.6
bytedance-seed · seed
ByteDance Seed model for long-context reasoning, instruction following, and tool-assisted tasks
Seed Character
bytedance-seed · seed
ByteDance Seed model optimized for character-driven dialogue and consistent conversational behavior
Seed 1.6 Vision
bytedance-seed · seed
ByteDance Seed multimodal model for image understanding, visual reasoning, and tool-assisted tasks
Seed 2.0 Pro
bytedance-seed · seed
Flagship ByteDance Seed 2.0 model for complex multimodal reasoning and long-horizon agent workflows
Seed Evolving
bytedance-seed · seed
Rolling ByteDance Seed model for rapidly updated reasoning, coding, and agent capabilities
Seed 2.1 Pro
bytedance-seed · seed
Flagship ByteDance Seed 2.1 model for complex multimodal reasoning, coding, and agents
Seed 1.6 Flash
bytedance-seed · seed
Low-latency ByteDance Seed model for high-throughput chat, extraction, and lightweight tool use
Seed 2.0 Mini
bytedance-seed · seed
Lightweight ByteDance Seed 2.0 model for low-latency multimodal reasoning and high-volume tasks
Seed 1.8
bytedance-seed · seed
ByteDance Seed model for multimodal reasoning, long-context analysis, and agent workflows
Seed 2.0 Lite
bytedance-seed · seed
Cost-efficient ByteDance Seed 2.0 model for production chat, analysis, and structured generation
Kimi K2.7 Code
moonshotai · kimi-k2
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
Kimi K3
moonshotai · kimi-k3
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Kimi K2 Thinking Turbo
moonshotai · kimi-thinking
Kimi reasoning model for long-horizon research, planning, and tool use
Kimi K2.5
moonshotai · kimi-k2
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
Kimi K2.7 Code Highspeed
moonshotai · kimi-k2
Lower-latency Kimi Code variant for interactive edits and coding-agent loops
Kimi K2 Thinking
moonshotai · kimi-thinking
Thinking Kimi model for slower research passes, planning, and hard technical questions
Kimi K2.6
moonshotai · kimi-k2
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
Claude Opus 4
anthropic · claude-opus
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Claude Opus 4.7
anthropic · claude-opus
Stronger Opus tier for advanced software work and high-stakes reasoning
Claude Mythos 5
anthropic · claude-mythos
Restricted Claude model for advanced cybersecurity and biology research workflows
Claude Opus 4.1
anthropic · claude-opus
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Claude Sonnet 3.5 v2
anthropic · claude-sonnet
Balanced Claude model for coding, analysis, agent workflows, and cost control
Claude Opus 4.8
anthropic · claude-opus
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
Claude Haiku 3.5
anthropic · claude-haiku
Fast Claude model for responsive assistance, classification, and lightweight agents
Claude Opus 4.1 (latest)
anthropic · claude-opus
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Claude Sonnet 5
anthropic · claude-sonnet
Everyday Claude agent model for coding, planning, browsing, and general work
Claude Sonnet 3.7
anthropic · claude-sonnet
Balanced Claude model for coding, analysis, agent workflows, and cost control
Claude Sonnet 4.5
anthropic · claude-sonnet
Balanced Claude model for coding, analysis, agent workflows, and cost control
Claude Opus 4.5 (latest)
anthropic · claude-opus
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Claude Sonnet 4.6
anthropic · claude-sonnet
Claude workhorse for coding agents, careful analysis, and production cost control
Claude Opus 4 (latest)
anthropic · claude-opus
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Claude Haiku 4.5
anthropic · claude-haiku
Fast Claude model for responsive assistance, classification, and lightweight agents
Claude Sonnet 4.5 (latest)
anthropic · claude-sonnet
Balanced Claude model for coding, analysis, agent workflows, and cost control
Claude Opus 4.5
anthropic · claude-opus
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Claude Fable 5
anthropic · claude-fable
Claude model for creative writing, analysis, and controlled agent workflows
Claude Sonnet 4
anthropic · claude-sonnet
Balanced Claude model for coding, analysis, agent workflows, and cost control
Claude Sonnet 4 (latest)
anthropic · claude-sonnet
Balanced Claude model for coding, analysis, agent workflows, and cost control
Claude Haiku 4.5 (latest)
anthropic · claude-haiku
Fast Claude lane for lightweight agents, office tasks, and responsive chat
Claude Opus 4.6
anthropic · claude-opus
High-end Claude for difficult coding, planning, and slower expert reasoning
Claude Haiku 3
anthropic · claude-haiku
Legacy model retained for compatibility with older integrations
Claude Opus 5
anthropic · claude-opus
Strongest Claude Opus model for coding, agents, and professional work
Aya Expanse 32B
cohere ·
Open multilingual model optimized for generation across 23 languages
Command A Reasoning
cohere · command-a
Cohere reasoning model for multilingual enterprise agents, tools, and complex workflows
Aya Vision 32B
cohere ·
Open multilingual vision model for OCR, visual reasoning, and image question answering
Command R+
cohere · command-r
Cohere's RAG workhorse for long-context enterprise search and tool use
Command A Translate
cohere · command-a
Translation model for multilingual conversion, localization, and cross-language workflows
Command R7B Arabic
cohere · command-r
Open Command R model optimized for Arabic enterprise chat, RAG, and cultural knowledge
Aya Expanse 8B
cohere ·
Compact open multilingual model optimized for generation across 23 languages
Command A Vision
cohere · command-a
Cohere vision model for multilingual document analysis, OCR, and image understanding
Command A Plus
cohere · command-a
Cohere's stronger command model for multilingual agents and enterprise workflows
Command A
cohere · command-a
Cohere command model for multilingual enterprise agents, tools, and chat
Command R7B
cohere · command-r
Cohere retrieval model for long-context chat and enterprise RAG workflows
Command R
cohere · command-r
Cohere retrieval model for long-context chat and enterprise RAG workflows
North Mini Code
cohere · north
Cohere coding model for practical software engineering and agentic edits
Aya Vision 8B
cohere ·
Compact open multilingual vision model for OCR and visual question answering
GPT-4o
openai · gpt
Omni-era GPT for multimodal chat, practical coding, and general assistants
GPT-Image-1.5
openai · gpt-image
Image model for prompt-driven generation, editing, and visual design workflows
GPT-5.3 Chat (latest)
openai · gpt
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
GPT-5 Nano
openai · gpt-nano
Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
GPT-4o mini
openai · gpt-mini
Small omni GPT for cheap multimodal assistance and production-scale traffic
GPT-3.5-turbo
openai · gpt
Compact GPT model for low-latency assistance and high-volume workloads
o1-pro
openai · o-pro
O-series reasoning model for hard analysis, math, coding, and planning
Whisper 3 Large
openai · whisper
Open Whisper checkpoint for robust multilingual transcription and captioning
GPT-5.5 Pro
openai · gpt-pro
Highest-accuracy GPT-5.5 tier for slower, precision-heavy reasoning and coding
GPT-4 Turbo
openai · gpt
Compact GPT model for low-latency assistance and high-volume workloads
GPT-4
openai · gpt
GPT model for general reasoning, writing, coding, and tool-assisted tasks
GPT-5 Chat (latest)
openai · gpt-codex
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
GPT-5.6 Sol
openai · gpt-sol
Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows
o3-pro
openai · o-pro
High-effort o3 tier for difficult technical reasoning and careful answers
o3-deep-research
openai · o
Research model for long-horizon investigation, synthesis, and analytical reports
GPT-5.4 Pro
openai · gpt-pro
More exact GPT-5.4 tier for demanding professional reasoning and agent tasks
GPT-5
openai · gpt
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
GPT-5.1 Codex mini
openai · gpt-codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
GPT-5-Codex
openai · gpt-codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
GPT-5 Mini
openai · gpt-mini
Small GPT-5 for responsive agents, coding help, and everyday automation
GPT-5.6 Luna
openai · gpt-luna
Cost-efficient GPT-5.6 model for fast, high-volume workloads
GPT-5.3 Codex
openai · gpt-codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
GPT OSS 120B
openai · gpt-oss
Open GPT reasoning model for self-hosted agents and controllable deployments
GPT-5.3 Codex Spark
openai · gpt-codex-spark
Coding-optimized GPT model for repository edits, reviews, and agentic software work
GPT-4.1 nano
openai · gpt-nano
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks
o1
openai · o
O-series reasoning model for hard analysis, math, coding, and planning
GPT-5.2
openai · gpt
Reliable GPT generation for broad coding, writing, and tool-assisted product work
GPT-Image-1
openai · gpt-image
OpenAI image model for production generation, edits, and brand-safe visual workflows
GPT-Realtime-2.1
openai · gpt
Realtime speech-to-speech model with configurable reasoning, tool use, and robust voice-agent behavior
GPT-5.4 mini
openai · gpt-mini
Strong small GPT for coding subagents, quick tool use, and high-volume work
GPT-4o (2024-05-13)
openai · gpt
GPT model for general reasoning, writing, coding, and tool-assisted tasks
o3
openai · o
Deliberate o-series reasoner for hard math, coding, and multi-step analysis
o4-mini-deep-research
openai · o-mini
Research model for long-horizon investigation, synthesis, and analytical reports
GPT-5 Pro
openai · gpt-pro
Higher-accuracy GPT-5 tier for tough analysis, coding reviews, and planning
GPT-5.5
openai · gpt
Default frontier GPT for coding, computer use, research, and knowledge work
GPT-5.5 Instant
openai ·
Compact GPT model for low-latency assistance and high-volume workloads
GPT-5.2 Codex
openai · gpt-codex
Code-specialist GPT for repository edits, reviews, and long-running software agents
GPT-4.1
openai · gpt
Long-lived GPT workhorse for coding, instruction following, and production apps
GPT-4o (2024-08-06)
openai · gpt
GPT model for general reasoning, writing, coding, and tool-assisted tasks
GPT-5.6 Terra
openai · gpt-terra
Balanced GPT-5.6 model for capable, cost-efficient everyday work
o4-mini
openai · o-mini
Fast o-series model for compact reasoning, coding, and tool use
GPT-5.4
openai · gpt
Agent-ready GPT for coding and computer-use workflows at a lower cost
o3-mini
openai · o-mini
Smaller o-series reasoner for economical coding, math, and planning tasks
GPT-4o (2024-11-20)
openai · gpt
GPT model for general reasoning, writing, coding, and tool-assisted tasks
GPT OSS Safeguard 120B
openai · gpt-oss
Safety model for policy screening, moderation, and risk-aware routing workflows
GPT Realtime Whisper
openai · whisper
Streaming speech-to-text model for low-latency transcript deltas from live audio
GPT-5.2 Chat
openai · gpt-codex
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
GPT-5.2 Pro
openai · gpt-pro
Higher-accuracy GPT-5.2 variant for tougher reasoning and review workflows
GPT-5.1 Chat
openai · gpt-codex
Chat-tuned GPT-5.1 for polished assistants, writing, and product conversations
GPT-Image-2
openai · gpt-image
Image model for prompt-driven generation, editing, and visual design workflows
GPT-5.1 Codex Max
openai · gpt-codex
Coding-optimized GPT model for repository edits, reviews, and agentic software work
GPT-5.1 Codex
openai · gpt-codex
Codex GPT for repository edits, code review, and practical software agents
GPT-4.1 mini
openai · gpt-mini
Affordable GPT-4.1 lane for fast coding help and structured extraction
GPT OSS 20B
openai · gpt-oss
Open GPT reasoning model for self-hosted agents and controllable deployments
Whisper Large v3 Turbo
openai · whisper
Speech transcription model for accurate audio-to-text and captioning workflows
GPT-5.4 nano
openai · gpt-nano
Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation
GPT-5.1
openai · gpt
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
Grok 4.6
xai · grok
xAI's frontier model for long-running agents, coding, knowledge work, and visual projects
Grok 4.5
xai · grok
xAI's Grok model for chat, coding, agentic tools, and lower hallucination risk
Grok 4.20 (Non-Reasoning)
xai · grok
Grok model for agentic tool use, reasoning, coding, and live assistance
Grok Imagine Video 1.5
xai · grok
Video model for image-to-video generation, editing, and extension workflows
Grok 4.1 Fast
xai · grok
xAI's fast agentic tool-calling model with a 2M context window; non-reasoning variant for low-latency responses
Grok 4.20 (Reasoning)
xai · grok
Reasoning Grok for document-heavy analysis and long-horizon tool use
Grok Imagine Image 2.0
xai · grok
Image model for prompt-driven generation, editing, and visual design workflows
Grok Build 0.1
xai · grok-build
Fast Grok coding model tuned for agentic engineering and iterative edits
Grok 4.3
xai · grok
xAI's default Grok for chat, coding, agentic tools, and lower hallucination risk
LongCat-2.0
meituan · longcat
Meituan LongCat-2.0, a reasoning model with tool calling and a 1M-token context window
Apertus 70B
swiss-ai ·
Fully open 70B multilingual LLM supporting 1800+ languages with 65K context. Trained on 15T tokens of compliant open data. Apache 2.0, EU AI Act compliant.
Apertus 8B
swiss-ai ·
Fully open 8B multilingual LLM supporting 1800+ languages with 65K context. Trained on compliant open data. Apache 2.0, EU AI Act compliant.
Step 3.5 Flash 2603
stepfun ·
StepFun flash model for efficient multimodal reasoning, coding, and tool use
Step 3.5 Flash
stepfun ·
StepFun flash lane for quick multimodal reasoning and coding assistance
Step 3.7 Flash
stepfun ·
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Sarvam 105B
sarvam · sarvam
Flagship Indian-language reasoning model for enterprise multilingual applications
Sarvam 30B
sarvam · sarvam
Efficient Indian-language reasoning model for chat, coding, and multilingual work
Trendyol Asure 12B
trendyol · gemma
Turkish-language multimodal instruct model built on Gemma 3 12B for e-commerce text, chat, and image-text tasks
Ornith 1.0 35B
deepreinforce · ornith
Large coding-reasoning model for agentic software tasks and RL search
Ornith 1.0 397B
deepreinforce · ornith
Large coding-reasoning model for agentic software tasks and RL search
Ornith 1.0 31B
deepreinforce · ornith
Open coding-reasoning model for repository tasks and self-improving agents
Ornith 1.0 9B
deepreinforce · ornith
Open coding-reasoning model for repository tasks and self-improving agents
MiMo-V2.5-Pro-UltraSpeed
xiaomi · mimo
MiMo pro model for strong multimodal reasoning and agent execution
MiMo-V2.5
xiaomi · mimo
Open MiMo model for multimodal coding agents and long-context automation
MiMo-V2-Pro
xiaomi · mimo
Earlier MiMo Pro model for multimodal agents, reasoning, and code tasks
MiMo-V2-Omni
xiaomi · mimo
MiMo omni model for text, image, video, audio, and agents
MiMo-V2-Flash
xiaomi · mimo
MiMo flash model for fast multimodal assistance and agent workflows
MiMo-V2.5-Pro
xiaomi · mimo
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
Qwen3-VL Plus
alibaba · qwen
Qwen vision-language model for visual reasoning, documents, and agent tasks
Qwen3.8 Max Preview
alibaba · qwen
Preview Qwen flagship for million-token multimodal reasoning and long-horizon agentic workflows
Qwen3-Coder 30B-A3B Instruct
alibaba · qwen
Smaller Qwen coder for efficient local agents and repo-level fixes
Qwen3.7 Max
alibaba · qwen
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
Qwen Turbo
alibaba · qwen
Efficient Qwen model for fast chat, extraction, and high-volume workloads
Qwen-Omni Turbo
alibaba · qwen
Qwen omni model for text, vision, audio, and multimodal agent tasks
Qwen-VL Max
alibaba · qwen
Qwen vision-language model for visual reasoning, documents, and agent tasks
Qwen3 32B
alibaba · qwen
Dense open Qwen model for self-hosted chat, reasoning, and coding
Qwen3 235B-A22B
alibaba · qwen
Large open Qwen MoE for multilingual reasoning, coding, and tool use
Qwen Max
alibaba · qwen
Flagship Qwen model for complex reasoning, coding, and agentic workflows
Qwen3.5 35B-A3B
alibaba · qwen
Qwen vision-language model for visual reasoning, documents, and agent tasks
Qwen3 VL 235B A22B Thinking
alibaba · qwen
Qwen vision-language thinking model for visual reasoning, documents, and agent tasks
Qwen3 30B A3B
alibaba · qwen
Sparse MoE Qwen model with 3B active parameters for efficient chat and reasoning
Qwen3.5 397B-A17B
alibaba · qwen
Large open Qwen multimodal MoE for visual agents and long technical tasks
Qwen3.5 Flash
alibaba · qwen
Qwen vision-language model for visual reasoning, documents, and agent tasks
Qwen2.5-Coder-0.5B
alibaba · qwen
Tiny open Qwen code model for lightweight completion and on-device coding
Qwen Plus
alibaba · qwen
Qwen instruction model for multilingual chat, reasoning, and tool use
Qwen3 Coder Next
alibaba · qwen
Open-weight Qwen coding model for agents, repository edits, and multi-turn tool use
Qwen3.5 Plus
alibaba · qwen
Qwen vision-language model for visual reasoning, documents, and agent tasks
Qwen2.5-VL 72B Instruct
alibaba · qwen
Qwen vision-language model for visual reasoning, documents, and agent tasks
Qwen3 Coder Flash
alibaba · qwen
Qwen coding model for software agents, repository edits, and code reasoning
Qwen3.6 Max Preview
alibaba · qwen
Flagship Qwen model for complex reasoning, coding, and agentic workflows
Qwen3.8 Max
alibaba · qwen
2.4-trillion-parameter MoE flagship for coding, professional work, multimodal understanding, and long-horizon agentic workflows
Qwen3.7 Plus
alibaba · qwen
Multimodal Qwen workhorse for long-context agents, visual inputs, and coding
QwQ 32B
alibaba · qwen
Open reasoning model from the Qwen team for math, coding, and step-by-step problem solving
Qwen3.7 Flash
alibaba · qwen
Lightweight multimodal Qwen model for high-throughput text, image, and video tasks
Qwen3-Next 80B-A3B (Thinking)
alibaba · qwen
Efficient Qwen thinking model for local reasoning, math, and coding agents
Qwen3-Coder 480B-A35B Instruct
alibaba · qwen
Open Qwen coding heavyweight for repository reasoning and agentic engineering
Qwen2.5-Coder-32B-Instruct
alibaba · qwen
Open coding-focused Qwen model for code generation, repair, and repository reasoning
Qwen Flash
alibaba · qwen
Efficient Qwen model for fast chat, extraction, and high-volume workloads
Qwen3 Max
alibaba · qwen
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
Qwen3.6 35B-A3B
alibaba · qwen
Open multimodal Qwen MoE for local agents that need vision, audio, and code
Qwen3 235B-A22B Instruct 2507
alibaba · qwen
Updated large open Qwen3 MoE instruct model for multilingual chat, coding, and tool use
Qwen3.6 27B
alibaba · qwen
Qwen vision-language model for visual reasoning, documents, and agent tasks
Qwen3-Next 80B-A3B Instruct
alibaba · qwen
Qwen instruction model for multilingual chat, reasoning, and tool use
Qwen3.8 27B
alibaba · qwen
Dense 27B vision-language model for coding, agent tasks, and image and video understanding
Qwen3.6 Plus
alibaba · qwen
Earlier Qwen multimodal workhorse for million-token agent and document tasks
Qwen3.5 9B
alibaba · qwen
Qwen instruction model for multilingual chat, reasoning, and tool use
Qwen3.5 27B
alibaba · qwen
Qwen vision-language model for visual reasoning, documents, and agent tasks
QwQ Plus
alibaba · qwen
Qwen reasoning model for deliberate problem solving, math, and coding
Qwen3.5 122B-A10B
alibaba · qwen
Qwen vision-language model for visual reasoning, documents, and agent tasks
Qwen3 Coder Plus
alibaba · qwen
Hosted Qwen coder for software agents, repo edits, and long-context code
Qwen3 VL 235B A22B Instruct
alibaba · qwen
Qwen vision-language instruct model for visual reasoning, documents, and agent tasks
Qwen-VL Plus
alibaba · qwen
Qwen vision-language model for visual reasoning, documents, and agent tasks
Qwen3.6 Flash
alibaba · qwen3.6
Qwen vision-language model for visual reasoning, documents, and agent tasks
Qwen3.8 2.4T A95B
alibaba · qwen
Open-weight sparse MoE (2.4T total, 95B active), the open-weight twin of Qwen3.8 Max for coding, research, complex reasoning, and agentic workflows
Fugu
sakana · fugu
Multi-agent model for routing expert agents across complex analytical tasks
Sakana Namazu
sakana · sakana-namazu
Japanese-specialized reasoning model based on Kimi K2.6 and tuned for Japanese language, culture, and business workflows
Fugu Ultra
sakana · fugu
Quality-first multi-agent model for hard research, analysis, and competitions
Sonar Deep Research
perplexity · sonar
Sonar search model for autonomous research and citation-backed long-form reports
Sonar
perplexity · sonar
Fast web-grounded Sonar for current answers, citations, and lightweight retrieval
Sonar Reasoning Pro
perplexity · sonar-reasoning
Web-grounded Sonar for multi-step research questions that need cited reasoning
Sonar Pro
perplexity · sonar-pro
Deeper Sonar search model with broader retrieval and stronger synthesis
GLM-4.6
zhipuai · glm
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
GLM-4.5-Flash
zhipuai · glm-flash
Efficient GLM model for fast reasoning, coding, and agent workflows
GLM-5V-Turbo
zhipuai · glm
Fast GLM vision model for screenshots, documents, and multimodal agent tasks
GLM-4.7
zhipuai · glm
Mature GLM model for dependable coding, reasoning, and structured agent tasks
GLM-4.5V
zhipuai · glm
GLM vision model for visual reasoning, documents, and multimodal agents
GLM-4.7-FlashX
zhipuai · glm-flash
Efficient GLM model for fast reasoning, coding, and agent workflows
GLM-5
zhipuai · glm
General GLM flagship for coding, analysis, and tool-heavy engineering workflows
GLM-5-Turbo
zhipuai · glm
Faster GLM-5 lane for coding agents that need lower latency
GLM-5.3
zhipuai · glm
Flagship GLM model for long-horizon coding, agents, and complex project delivery
GLM-5.2
zhipuai · glm
Open flagship GLM for long-horizon coding agents and million-token context work
GLM-4.6V
zhipuai · glm
GLM vision model for visual reasoning, documents, and multimodal agents
GLM-4.5-Air
zhipuai · glm-air
Lighter GLM-4.5 variant for fast coding assistance and cheaper agents
GLM-5.1
zhipuai · glm
Strong GLM coding model for agentic engineering, terminals, and repository generation
GLM-4.7-Flash
zhipuai · glm-flash
Budget GLM lane for fast coding help, routing, and everyday automation
GLM-4.5
zhipuai · glm
Hybrid-reasoning GLM release that made the 4.5 line broadly useful
Granite-4.0-H-Micro
ibm · granite
Compact open-weight hybrid Granite model for lightweight enterprise chat and tool calling
Granite-4.0-H-Small
ibm · granite
Open-weight hybrid model for enterprise chat, coding, retrieval-augmented generation, and tool-calling workloads
Inkling Small
thinkingmachines · ling
Multimodal MoE reasoning model (276B total, 12B active) for text, image, and audio
Inkling
thinkingmachines · ling
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Magistral Small
mistral · magistral
Open Mistral reasoning model for transparent step-by-step problem solving
Mistral Small (latest)
mistral · mistral-small
Efficient Mistral model for fast chat, extraction, and production assistants
Devstral Medium
mistral · devstral
Mistral coding agent model for repository tasks and software engineering workflows
Mistral Small 3.1 24B
mistral · mistral-small
Efficient multimodal model for instruction following, coding, reasoning, and function calling
Mistral Large 2.1
mistral · mistral-large
Flagship Mistral model for advanced reasoning, coding, and multilingual work
Mistral Nemo
mistral · mistral-nemo
Efficient Mistral-NVIDIA open model for multilingual chat and local deployment
Mistral Large (latest)
mistral · mistral-large
Flagship Mistral model for advanced reasoning, coding, and multilingual work
Ministral 8B Instruct
mistral · ministral
Efficient open Mistral edge model for on-device chat and function calling
Devstral Small
mistral · devstral
Mistral coding agent model for repository tasks and software engineering workflows
Devstral 2
mistral · devstral
Mistral's coding-agent model for repository work, terminal tasks, and software fixes
Mistral Small 4
mistral · mistral-small
Fast Mistral production model for chat, extraction, and cost-sensitive agents
Mistral Small 3.2
mistral · mistral-small
Efficient Mistral model for fast chat, extraction, and production assistants
Pixtral Large (latest)
mistral · pixtral
Mistral's larger vision model for document-heavy image understanding and chat
Mistral Large 3
mistral · mistral-large
Mistral's largest general model for enterprise agents, coding, and multilingual reasoning
Mistral Medium 3.5
mistral · mistral-medium
Balanced Mistral model for enterprise assistants, multilingual work, and tools
Mistral Medium 3
mistral · mistral-medium
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
Devstral 2 (latest)
mistral · devstral
Mistral coding agent model for repository tasks and software engineering workflows
Pixtral 12B
mistral · pixtral
Mistral vision-language model for image understanding and multimodal chat
Voxtral Small (latest)
mistral · voxtral
Instruct model with native audio input for speech understanding and tool use
Codestral-22B-v0.1
mistral · codestral
Open Mistral code model for fill-in-the-middle and 80+ programming languages
Codestral (latest)
mistral · codestral
Mistral code model for completions, refactors, and developer IDE workflows
Magistral Medium (latest)
mistral · magistral-medium
Mistral reasoning model for transparent analysis, math, and complex decisions
Mistral Medium (latest)
mistral · mistral-medium
Balanced Mistral model for enterprise assistants, multilingual work, and tools
Solar Pro 2
upstage · solar-pro
Flagship model for demanding analysis, coding, and production agent workflows
Solar Pro 4
upstage · solar-pro
Upstage's flagship model, specialized for agentic use
Solar Pro 3
upstage · solar-pro
Flagship model for demanding analysis, coding, and production agent workflows
MiniMax-M2.7
minimax · minimax
Open MiniMax flagship for coding agents, office automation, and complex environments
MiniMax-M2.7-highspeed
minimax · minimax
Low-latency M2.7 variant for interactive coding plans and agent loops
MiniMax-M2.1
minimax · minimax
Earlier MiniMax agent model for practical coding and productivity tasks
MiniMax-M2
minimax · minimax
Efficient open MiniMax model built for coding agents and tool-heavy workflows
MiniMax-M2.5-highspeed
minimax · minimax
High-speed MiniMax model for low-latency coding and agent workflows
MiniMax-M3
minimax · minimax
MiniMax multimodal model for long-context coding, perception, and agent planning
MiniMax-M2 Her
minimax · minimax
MiniMax M2 variant tuned for conversational and character-driven agent interactions
MiniMax-M2.5
minimax · minimax
Prior MiniMax coding model for agent workflows, office edits, and automation