196 results
Foundation models you can call or deploy
by Z.ai·Dec 2025
358B MoE (~32B active) with 128K context, MIT license; interleaved thinking with preserved reasoning across turns for agentic coding and UI generation.
by Meta·Apr 2025
Meta's Llama 4 Maverick MoE model with 17B active / 128 experts (400B total), 1M token context, natively multimodal, under the Llama 4 Community License.
by Meta·Dec 2023
Meta's open-weight safety classifier for filtering harmful LLM inputs/outputs; initial release was a Llama 2 7B instruction-tuned model, followed by v2, v3, and v4.
by xAI·Feb 2025
Smaller, cost-efficient sibling of Grok 3 launched alongside it; reasoning-focused variant priced near GPT-4o-mini and noted for strong coding performance.
by Google·Feb 2026
Incremental Gemini 3 Pro upgrade with stronger core reasoning, verified 77.1% on ARC-AGI-2, shipped in preview across AI Studio, Antigravity, and Vertex AI.
by Z.ai·Apr 2026
754B MoE (glm_moe_dsa) with DSA sparse attention, MIT license; iterative successor to GLM-5 for long-horizon agentic tasks with thousands of tool calls.
by Google·Jun 2025
Smallest, cheapest Gemini 2.5 model, tuned for high-volume translation and classification with a 1M-token context and toggleable thinking budgets.
by OpenAI·May 2020
OpenAI's 175B-parameter autoregressive language model introduced via research paper in 2020, notable for strong few-shot performance across NLP tasks without fine-tuning.
by Meta·Dec 2024
Meta's 70B dense Instruct-only open-weight LLM from Llama 3.3 with a 128K token context, released December 2024 under the Llama 3.3 Community License.
by Meta·Apr 2024
Meta's 8B dense open-weight LLM from the Llama 3 family with an 8K token context, released April 2024 in base and Instruct variants, Llama 3 Community License.
by Mistral AI·Oct 2024
Mistral's 8B dense edge LLM with 128k context and sliding-window attention; released with Ministral 3B under a Mistral Research License, weights research-only.
by Meta·Jul 2024
Meta's 8B dense open-weight LLM from Llama 3.1, context extended to 128K tokens from Llama 3's 8K, released July 2024 under the Llama 3.1 Community License.
by Microsoft·Feb 2025
Microsoft's compact 3.8B Phi-4 SLM designed for on-device inference; released alongside Phi-4-multimodal under an MIT license with 128K context support.
Fast, cost-efficient Gemini 2.5 model with hybrid reasoning that lets developers toggle thinking on or off and set thinking budgets per request.
by Microsoft·May 2024
Microsoft's 7B dense Phi-3 SLM with 8K and 128K context variants (MIT license), benchmarked as a mid-tier textbook-trained model between Phi-3 mini and medium.
by Alibaba·Aug 2025
Text-to-image diffusion foundation with high-fidelity Chinese/English text rendering, Apache-2.0 license; supports style transfer, editing, and multi-aspect.
Meta's flagship 405B dense open-weight LLM from Llama 3.1 with a 128K token context, released July 2024 as the first frontier-scale open-weights model, Llama 3.1 license.
by Anthropic·Oct 2024
Fast, low-cost model in Anthropic's Claude 3.5 family; text-only at launch, priced $0.80/$4 per MTok on the API. Model ID claude-3-5-haiku-20241022.
by Alibaba·Jun 2025
Unified multimodal understanding-and-generation preview from Alibaba's Qwen team; text-to-image and instruction-based image editing accessible via Qwen Chat.
by Microsoft·Apr 2024
Microsoft's 3.8B dense Phi-3 SLM with 4K or 128K context variants (MIT license), tuned to punch above its size on reasoning and instruction-following benchmarks.
by Anthropic·May 2025
Anthropic's Claude Opus 4, initial Claude 4 flagship for coding, reasoning, and agent workflows; retired June 2026 in favor of Opus 4.8.
by Mistral AI·Dec 2023
Mistral's sparse Mixture-of-Experts family; Mixtral 8x7B (46.7B total, ~12.9B active) under Apache 2.0, outperforming Llama 2 70B and GPT-3.5 on many benchmarks.
by EleutherAI·Jun 2021
Open-source 6-billion parameter autoregressive language model by EleutherAI, trained on The Pile; released with weights as an alternative to GPT-3.
by OpenAI·Feb 2019
OpenAI's 2019 transformer language model (up to 1.5B parameters), known for coherent zero-shot text generation; initially released in staged increments over misuse concerns.