442 results
by Meta·Jul 2024
Meta's flagship 405B dense open-weight LLM from Llama 3.1 with a 128K token context, released July 2024 as the first frontier-scale open-weights model, Llama 3.1 license.
by Microsoft·Apr 2024
Microsoft's 3.8B dense Phi-3 SLM with 4K or 128K context variants (MIT license), tuned to punch above its size on reasoning and instruction-following benchmarks.
by Z.ai·Mar 2026
1B multimodal OCR model, MIT license; complex document understanding with layout grounding and multilingual text extraction for real-world docs.
by Meta·Apr 2024
Meta's 70B dense open-weight LLM from the Llama 3 family with an 8K token context, released April 2024 in base and Instruct variants, Llama 3 Community License.
by OpenAI·Nov 2022
OpenAI's GPT-3.5 model series (text-davinci-003, gpt-3.5-turbo) introduced in late 2022; the fine-tuned lineage that powered the initial ChatGPT launch.
by Z.ai·Apr 2026
754B MoE (glm_moe_dsa) with DSA sparse attention, MIT license; iterative successor to GLM-5 for long-horizon agentic tasks with thousands of tool calls.
by Anthropic·May 2025
Anthropic's Claude Opus 4, initial Claude 4 flagship for coding, reasoning, and agent workflows; retired June 2026 in favor of Opus 4.8.
by DeepSeek·Sep 2025
685B/37B-active MoE with 256 experts, 128K context, MIT license; debuts DeepSeek Sparse Attention (DSA) for cheap long-context training and inference.
by Runway·May 2026
Runway's MCP server exposing its Gen-series video generation and editing tools to LLM agents for programmatic storyboarding, shot tests, and creative workflows.
by Anthropic·Aug 2025
Anthropic's Claude browser extension for Chrome enabling agentic browsing inside a tab; launched as a research preview limited to 1,000 Max-plan users.
by OpenAI·Feb 2019
OpenAI's 2019 transformer language model (up to 1.5B parameters), known for coherent zero-shot text generation; initially released in staged increments over misuse concerns.
by Mistral AI·Sep 2023
Mistral's 7.3B dense open-weight LLM under Apache 2.0; uses grouped-query and sliding-window attention, a widely fine-tuned baseline for local inference.
by Meta·Dec 2024
Meta's 70B dense Instruct-only open-weight LLM from Llama 3.3 with a 128K token context, released December 2024 under the Llama 3.3 Community License.
by OpenAI·Aug 2025
OpenAI's smaller open-weight 20B-parameter reasoning model released under Apache 2.0, designed for on-device/self-hosted use with function calling and structured outputs.
by OpenAI·Jun 2018
First Generative Pre-trained Transformer from OpenAI (2018), a 117M-parameter decoder-only model pre-trained on unlabeled text then fine-tuned for NLP tasks.
by Google·May 2025
Enhanced parallel-thinking mode for Gemini 2.5 Pro announced at I/O 2025; an advanced version later reached gold-medal standard at IMO 2025.
by Anthropic·Jun 2026
Anthropic's Claude Mythos 5 frontier model for defensive cybersecurity, offered in limited availability to Project Glasswing partners; 1M-token context.
Faster turbo variant of Z.ai's GLM-5, optimized for agentic coding workflows; MoE with sparse attention. Announced/rumored tier; release details limited.
Meta's Llama 3.3 70B Instruct-only dense open-weight LLM with 128K context, delivering 405B-class quality at 70B cost, under the Llama 3.3 Community License.
by Mistral AI·Dec 2025
Mistral's terminal-native coding CLI agent, introduced alongside Devstral 2; open-source, supports plugins, skills, hooks, and MCP servers for end-to-end code automation.
by OpenAI·Apr 2025
OpenAI's API-only GPT-4 update with improved instruction following, coding gains, and a 1M-token context window; positioned as a cost-efficient GPT-4o alternative.
by Mistral AI·Mar 2025
Mistral's multimodal OCR API for document understanding; parses text, tables, equations, and images from PDFs and scans into structured markdown output.
OpenAI's second-generation reasoning model with tool use, delivering gains on math, coding, and science benchmarks over o1; released alongside o4-mini in April 2025.
by Google·Dec 2023
Smallest Gemini model, designed for on-device inference in Android (starting on Pixel 8 Pro) and Chrome, enabling local AI features without cloud calls.