210 results
Foundation models you can call or deploy
by Anthropic·Nov 2025
Anthropic's Claude Opus 4.5 for coding, agents, and computer use; introduces an API 'effort' parameter to trade off cost against capability.
by Google·Mar 2025
Flagship Gemini 2.5 reasoning model with a 1M-token context window and native multimodal input across text, audio, image, video, and full code repos.
by OpenAI·Aug 2025
OpenAI's reasoning-native flagship succeeding the GPT-4 family, with automatic routing between fast and deep-thinking modes; released August 2025 across ChatGPT tiers.
by OpenAI·Jan 2025
Smaller, faster, cheaper reasoning variant of o3 with three selectable effort levels (low/medium/high); released as a successor to o1-mini for cost-sensitive workloads.
by OpenAI·Apr 2025
OpenAI's compact reasoning model succeeding o3-mini, tuned for speed and cost; supports vision and tool use natively, released alongside o3 in April 2025.
by DeepSeek·Jan 2025
Open-weight reasoning MoE (671B total/37B active) with 128K context, MIT license; RL-trained o1-class model with strong chain-of-thought on math and code.
by xAI·Feb 2025
xAI's flagship reasoning LLM trained on the 100k+ GPU Colossus cluster; introduces DeepSearch and Think mode with real-time X data integration on the X platform.
by OpenAI·May 2020
OpenAI's 175B-parameter autoregressive language model introduced via research paper in 2020, notable for strong few-shot performance across NLP tasks without fine-tuning.
by Alibaba·Apr 2025
Qwen3 open-weight family (dense 0.6B-32B + MoE 30B-A3B/235B-A22B), Apache-2.0 license; multilingual hybrid thinking modes for strong coding and math.
by OpenAI·Dec 2024
Higher-compute variant of OpenAI's o1 reasoning model, exclusive to ChatGPT Pro subscribers; uses more inference-time compute for harder math and coding problems.
by Alibaba·Nov 2024
32B dense reasoning preview model with 32K context, Apache-2.0 license; tuned for math/code chain-of-thought with strong AIME, GPQA, and MATH-500 scores.
by Anthropic·Jun 2024
Mid-tier Claude 3.5 model from Anthropic for chat, coding, and long-context tasks; launched alongside the Artifacts feature on claude.ai.
by Google·Apr 2026
Google's Apache 2.0 open-weight model family, tuned for advanced reasoning and agentic workflows and sized to run on developer hardware.
by OpenAI·Feb 2025
OpenAI's largest pre-reasoning-era model, tuned for writing and raw capability with reduced hallucinations; served as the last GPT-4-family flagship before GPT-5.
by Microsoft·Dec 2024
Microsoft's 14B dense Phi-4 SLM trained on synthetic textbook data, MIT license, tuned for reasoning tasks and competitive with much larger models on math/code.
by Alibaba·Sep 2024
Qwen 2.5 dense LLM family (0.5B-72B) with 128K context, mostly Apache-2.0; pretrained on 18T tokens with Coder, Math and 29-language multilingual variants.
by Google·Nov 2025
Google's third-generation flagship multimodal model with a 1M-token context, agent-friendly reasoning, and native support for text, image, video, and audio.
by DeepSeek·Nov 2023
Open-source code LLM family (1.3B/6.7B/33B) with 16K context, DeepSeek license, trained on 2T tokens (87% code) for code generation and FIM infilling.
by Anthropic·Sep 2025
Anthropic's Claude Sonnet 4.5 mid-tier model for coding and agents; reached 77.2% on SWE-bench Verified. Priced $3/$15 per MTok.
by DeepSeek·Dec 2024
671B/37B-active MoE with MLA and MTP, 128K context, MIT license; trained on 14.8T tokens with strong coding and math at economical inference cost.
by OpenAI·Mar 2023
OpenAI's earlier flagship multimodal (text+image) LLM, predecessor to GPT-4o, with 8K/32K context windows and strong performance on professional and academic benchmarks.
by DeepSeek·Apr 2026
862B DeepSeek V4 Pro MoE with compressed sparse attention and ~1M-token context, MIT license; flagship V4 tier for frontier long-context reasoning.
by Z.ai·Dec 2025
358B MoE (~32B active) with 128K context, MIT license; interleaved thinking with preserved reasoning across turns for agentic coding and UI generation.
by OpenAI·May 2024
OpenAI's flagship omni-modal model accepting text, image, and audio input, with a 128K context window and real-time voice conversation via ChatGPT and the API.