196 results
Foundation models you can call or deploy
by Google·Feb 2025
Higher-capability Gemini 2.0 model aimed at complex reasoning and coding tasks, launched as an experimental release in Google AI Studio and Vertex AI.
by Mistral AI·Sep 2024
Mistral's first multimodal vision-language model: 12B decoder plus 400M vision encoder, 128k context, Apache 2.0; ingests images of arbitrary sizes and counts.
by Meta·Jul 2024
Meta's 8B dense open-weight LLM from Llama 3.1, context extended to 128K tokens from Llama 3's 8K, released July 2024 under the Llama 3.1 Community License.
by Meta·Apr 2025
Meta's fourth-generation Llama family using MoE: Scout (17B active/16 experts) and Maverick (17B active/128 experts), natively multimodal, Llama 4 Community License.
by Z.ai·Mar 2026
Faster turbo variant of Z.ai's GLM-5, optimized for agentic coding workflows; MoE with sparse attention. Announced/rumored tier; release details limited.
by Alibaba·Jan 2025
Vision-language Qwen2.5 family (3B/7B/72B) with dynamic-resolution ViT and 32K context (YaRN extendable), qwen license; strong OCR and video grounding.
by Moonshot AI·Jan 2026
1T/32B-active MoE with 256K context and MoonViT vision, Modified MIT license; native multimodal with agent-swarm parallel sub-task orchestration.
by Meta·Sep 2024
Meta's Llama 3.2 open-weight family: 1B and 3B text models with 128K context plus 11B and 90B vision multimodal variants, under the Llama 3.2 Community License.
by Alibaba·Sep 2024
Qwen 2.5 dense LLM family (0.5B-72B) with 128K context, mostly Apache-2.0; pretrained on 18T tokens with Coder, Math and 29-language multilingual variants.
by OpenAI·May 2020
OpenAI's 175B-parameter autoregressive language model introduced via research paper in 2020, notable for strong few-shot performance across NLP tasks without fine-tuning.
by Meta·Dec 2023
Meta's open-weight safety classifier for filtering harmful LLM inputs/outputs; initial release was a Llama 2 7B instruction-tuned model, followed by v2, v3, and v4.
by Google·Feb 2026
Gemini 3.1 Flash Image model that combines Nano Banana Pro-level quality with Flash speed, adding subject consistency and production-ready output specs.
by Z.ai·Jan 2026
30B/3B-active MoE (glm4_moe_lite) with 128K context, MIT license; small/fast GLM-4.7 tier for local and agentic coding with speculative decoding.
by Microsoft·Apr 2025
Reasoning-tuned variant of Microsoft's Phi-4-mini SLM for math and step-by-step chain-of-thought, released with Phi-4-reasoning and Phi-4-reasoning-plus.
by OpenAI·Nov 2025
OpenAI's GPT-5.1 update with Instant and Thinking variants, delivering a warmer conversational tone and better instruction following on top of GPT-5's architecture.
by DeepSeek·Aug 2025
Hybrid MoE (671B/37B active) with toggleable thinking mode, 128K context, MIT license; unifies V3 chat and R1 reasoning in a single hybrid model.
by Mistral AI·Jun 2025
Mistral's 24B open-weight reasoning model under Apache 2.0; first Mistral reasoning system with transparent, domain-specific, multilingual chain-of-thought traces.
by Microsoft·Feb 2025
Microsoft's compact 3.8B Phi-4 SLM designed for on-device inference; released alongside Phi-4-multimodal under an MIT license with 128K context support.
by Meta·Apr 2024
Meta's 8B dense open-weight LLM from the Llama 3 family with an 8K token context, released April 2024 in base and Instruct variants, Llama 3 Community License.
by Moonshot AI·May 2026
1T/32B-active MoE with 256K context, MoonViT vision, Modified MIT license; coding-focused successor to K2.5 with 300-agent swarm and long-horizon coding.
by Mistral AI·May 2024
Mistral's 22B code-specialized dense LLM under Mistral Non-Production License; supports 80+ programming languages, fill-in-the-middle and code completion.
by OpenAI·Mar 2026
Smaller cost-optimized variant of GPT-5.4, available on the free tier via the Thinking feature and as an API model; ~4x pricier than GPT-5 equivalents in the OpenAI API.
by OpenAI·Jun 2025
Higher-compute variant of OpenAI's o3 reasoning model, available to ChatGPT Pro users and via API with tool access (web, files, Python) for high-reliability tasks.
by Mistral AI·May 2025
Mistral's commercial mid-tier LLM served via API; modern line launched with Mistral Medium 3 targeting enterprise SOTA at ~8x lower cost, later refreshed as 3.1 and 3.5.