196 results
Foundation models you can call or deploy
by Z.ai·Mar 2026
1B multimodal OCR model, MIT license; complex document understanding with layout grounding and multilingual text extraction for real-world docs.
by Anthropic·Aug 2025
Anthropic's Claude Opus 4.1, an upgrade to Opus 4 for agentic tasks and coding; scored 74.5% on SWE-bench Verified. Same pricing as Opus 4.
by Google·Dec 2024
Fast, low-cost Gemini 2.0 model with native multimodal output (images plus steerable multilingual text-to-speech) and built-in tool use including Search.
by Moonshot AI·Oct 2025
48B/3B-active hybrid linear-attention MoE (Kimi Delta Attention + MLA at 3:1), 1M context, MIT license; 6x faster decode via ~75% KV cache reduction.
by Google·May 2025
Enhanced parallel-thinking mode for Gemini 2.5 Pro announced at I/O 2025; an advanced version later reached gold-medal standard at IMO 2025.
by Anthropic·May 2025
Anthropic's Claude Sonnet 4, initial Claude 4 mid-tier model for coding and reasoning; launched alongside Opus 4, retired June 15, 2026.
by Meta·Apr 2024
Meta's 70B dense open-weight LLM from the Llama 3 family with an 8K token context, released April 2024 in base and Instruct variants, Llama 3 Community License.
by OpenAI·Nov 2022
OpenAI's GPT-3.5 model series (text-davinci-003, gpt-3.5-turbo) introduced in late 2022; the fine-tuned lineage that powered the initial ChatGPT launch.
by Z.ai·Jul 2025
355B/32B-active MoE with 128K context, MIT license; hybrid thinking model with toggleable reasoning modes for agentic/reasoning/coding (ARC) tasks.
by Meta·Jul 2024
Meta's Llama 3.1 open-weight family with 8B, 70B, and 405B dense variants extended to a 128K token context, released under the Llama 3.1 Community License.
by Meta·Apr 2025
Meta's Llama 4 Scout MoE model with 17B active / 16 experts (109B total), 10M token context, natively multimodal, under the Llama 4 Community License.
by DeepSeek·Dec 2025
685B MoE variant of V3.2 with DSA, MIT license; reasoning-only tier with gold-medal IMO 2025/IOI 2025 performance and no tool-calling support.
by Mistral AI·Mar 2025
Mistral's multimodal OCR API for document understanding; parses text, tables, equations, and images from PDFs and scans into structured markdown output.
by Microsoft·Aug 2024
Refreshed 3.8B Phi-3.5-mini-instruct SLM from Microsoft with a 128K context window and multilingual gains, released on Hugging Face under an MIT license.
by OpenAI·Aug 2025
OpenAI's smaller open-weight 20B-parameter reasoning model released under Apache 2.0, designed for on-device/self-hosted use with function calling and structured outputs.
by Mistral AI·Jul 2024
12B dense LLM co-developed with NVIDIA, 128k context, Apache 2.0; multilingual and quantization-aware, positioned as a drop-in replacement for Mistral 7B.
by Moonshot AI·Jul 2025
1T/32B-active MoE with 256K context and MLA attention, Modified MIT license; 384-expert (8/token) agentic-focused foundation model from Moonshot.
by Google·Mar 2026
Fastest and cheapest Gemini 3 series model, aimed at high-volume workloads with adjustable thinking levels, 1M-token context, and multimodal input.
OpenAI's open-weight reasoning model family (gpt-oss-120b and gpt-oss-20b), the company's first open-weight LLMs since GPT-2 in 2019, released for self-hosted deployment.
by xAI·Jul 2025
xAI's flagship model succeeding Grok 3; adds native tool use and real-time search, trained with massively scaled reinforcement learning on the 200k-GPU Colossus cluster.
by Google·Nov 2025
Google DeepMind's Gemini 3 Pro Image model for studio-quality generation and editing, with improved text rendering and better world-knowledge grounding.
by Z.ai·Feb 2026
744B/40B-active MoE (glm_moe_dsa) with DeepSeek Sparse Attention, MIT license; trained on 28.5T tokens for long-horizon agentic engineering tasks.
by Z.ai·Apr 2025
32B dense bilingual GLM-4 model with 32K context (YaRN extendable), MIT license; strong Chinese/English chat and function-calling foundation model.
by Mistral AI·Feb 2024
Mistral's small-tier LLM family; original API model launched with Mistral Large in 2024, later refreshed as the 24B Apache 2.0 open Mistral Small 3 and 3.1.