402 results
by Google·May 2025
Enhanced parallel-thinking mode for Gemini 2.5 Pro announced at I/O 2025; an advanced version later reached gold-medal standard at IMO 2025.
by Meta·Apr 2024
Meta's 70B dense open-weight LLM from the Llama 3 family with an 8K token context, released April 2024 in base and Instruct variants, Llama 3 Community License.
by espirado·Mar 2026
Open-source MCP server that audits agent decisions via signal fidelity, pattern classification, reliability scoring, and authority gates before commit.
by Microsoft·Apr 2024
Microsoft's 3.8B dense Phi-3 SLM with 4K or 128K context variants (MIT license), tuned to punch above its size on reasoning and instruction-following benchmarks.
by DeepSeek·Nov 2023
Open-source code LLM family (1.3B/6.7B/33B) with 16K context, DeepSeek license, trained on 2T tokens (87% code) for code generation and FIM infilling.
by Anthropic·Oct 2024
Fast, low-cost model in Anthropic's Claude 3.5 family; text-only at launch, priced $0.80/$4 per MTok on the API. Model ID claude-3-5-haiku-20241022.
by Alibaba·Sep 2024
Qwen 2.5 dense LLM family (0.5B-72B) with 128K context, mostly Apache-2.0; pretrained on 18T tokens with Coder, Math and 29-language multilingual variants.
by OpenAI·May 2020
OpenAI's 175B-parameter autoregressive language model introduced via research paper in 2020, notable for strong few-shot performance across NLP tasks without fine-tuning.
by Mistral AI·Jun 2025
Mistral's 24B open-weight reasoning model under Apache 2.0; first Mistral reasoning system with transparent, domain-specific, multilingual chain-of-thought traces.
by OpenAI·Apr 2026
OpenAI's second-generation image generation model in the API and Codex, offering fixed sizes (1024x1024, 1024x1536, 1536x1024), stronger editing, and improved text rendering.
by Character.AI·Sep 2022
Consumer chatbot platform where users create and chat with user-defined persona bots; widely used for roleplay, companionship, and interactive fiction.
by Dakera·May 2026
Self-hosted MCP server as a single Rust binary providing persistent agent memory, vector/hybrid search, knowledge graph, and session management.
by Meta·Jul 2024
Meta's flagship 405B dense open-weight LLM from Llama 3.1 with a 128K token context, released July 2024 as the first frontier-scale open-weights model, Llama 3.1 license.
by Mistral AI·Sep 2023
Instruction-tuned variant of Mistral 7B released the same day under Apache 2.0; chat/instruction fine-tune outperforming Llama 2 13B Chat on standard benchmarks.
by Mistral AI·Sep 2024
Mistral's first multimodal vision-language model: 12B decoder plus 400M vision encoder, 128k context, Apache 2.0; ingests images of arbitrary sizes and counts.
by Microsoft·Feb 2025
Microsoft's compact 3.8B Phi-4 SLM designed for on-device inference; released alongside Phi-4-multimodal under an MIT license with 128K context support.
by DeepSeek·May 2024
236B/21B-active MoE with Multi-head Latent Attention, 128K context, DeepSeek license; pretrained on 8.1T tokens with 93% KV cache reduction over dense peers.
by Meta·Dec 2024
Meta's 70B dense Instruct-only open-weight LLM from Llama 3.3 with a 128K token context, released December 2024 under the Llama 3.3 Community License.
Meta's third-generation open-weight LLM family with 8B and 70B dense variants, 8K token context, base and Instruct, under the Llama 3 Community License.
by Anthropic·Jun 2026
Anthropic's Claude Mythos 5 frontier model for defensive cybersecurity, offered in limited availability to Project Glasswing partners; 1M-token context.
by Mistral AI·Oct 2024
Mistral's 8B dense edge LLM with 128k context and sliding-window attention; released with Ministral 3B under a Mistral Research License, weights research-only.
by Google·Feb 2025
Higher-capability Gemini 2.0 model aimed at complex reasoning and coding tasks, launched as an experimental release in Google AI Studio and Vertex AI.
by Perplexity·Jul 2025
Perplexity's agentic web browser with a built-in AI assistant that summarizes pages, fills forms, and executes multi-step browsing tasks on the user's behalf.
Meta's 8B dense open-weight LLM from Llama 3.1, context extended to 128K tokens from Llama 3's 8K, released July 2024 under the Llama 3.1 Community License.