442 results
by xAI·Apr 2026
No public release from xAI; the current post-Grok 4 line is Grok 4.1 (Nov 2025) and Grok 4.1 Fast (Nov 2025), with no Grok 4.3 or Grok 4.20 announced.
by AWS·Aug 2026
A managed registry for discovering, governing, and reusing agents, MCP servers, tools, skills, and custom agentic resources.
by Open Source·Mar 2025
Open-source MCP server that lets any LLM control Blender 3D via a Python addon, exposing scene manipulation, modeling, and rendering commands.
by Independent·Feb 2026
AI-facing directory of Claude Skills and MCP servers where agents can browse, test, and rate entries; awards a blue badge to those that run successfully.
by Google·Dec 2025
Parallel-reasoning mode built on Gemini 3 Pro for hard math, science, and logic problems; available to Google AI Ultra subscribers in the Gemini app.
by OpenAI·Dec 2025
OpenAI's image model succeeding gpt-image-1 with more precise edits, better instruction following, and improved rendering of dense small text; in API and ChatGPT.
by Moonshot AI·Oct 2025
48B/3B-active hybrid linear-attention MoE (Kimi Delta Attention + MLA at 3:1), 1M context, MIT license; 6x faster decode via ~75% KV cache reduction.
by Mistral AI·Jul 2024
12B dense LLM co-developed with NVIDIA, 128k context, Apache 2.0; multilingual and quantization-aware, positioned as a drop-in replacement for Mistral 7B.
by DeepSeek·Aug 2025
Hybrid MoE (671B/37B active) with toggleable thinking mode, 128K context, MIT license; unifies V3 chat and R1 reasoning in a single hybrid model.
by Google·Sep 2026
Google's generally available Flash model for long-horizon software engineering, autonomous agents, and complex enterprise workflows.
by Bright Data·Apr 2025
MCP server from Bright Data giving AI agents web search, scraping, and browser automation via Bright Data's unlocker/proxy infrastructure.
by DeepSeek·Jan 2026
3B vision-language OCR model (deepseek_vl_v2) with dynamic-resolution tiling and layout grounding, Apache-2.0; second-gen Visual Causal Flow architecture.
by Z.ai·Apr 2025
32B dense bilingual GLM-4 model with 32K context (YaRN extendable), MIT license; strong Chinese/English chat and function-calling foundation model.
by EleutherAI·Jun 2021
Open-source 6-billion parameter autoregressive language model by EleutherAI, trained on The Pile; released with weights as an alternative to GPT-3.
by Meta·Dec 2023
Meta's open-weight safety classifier for filtering harmful LLM inputs/outputs; initial release was a Llama 2 7B instruction-tuned model, followed by v2, v3, and v4.
by Notion·Apr 2025
Notion's Model Context Protocol server exposing workspace read/write tools to LLM agents so they can query and edit pages, databases, and comments.
by Microsoft·Aug 2024
Refreshed 3.8B Phi-3.5-mini-instruct SLM from Microsoft with a 128K context window and multilingual gains, released on Hugging Face under an MIT license.
Google's music generation model for creating full-length songs from text or image prompts through the Gemini API.
by Anthropic·Aug 2025
Anthropic's Claude Opus 4.1, an upgrade to Opus 4 for agentic tasks and coding; scored 74.5% on SWE-bench Verified. Same pricing as Opus 4.
by DeepSeek·May 2024
236B/21B-active MoE with Multi-head Latent Attention, 128K context, DeepSeek license; pretrained on 8.1T tokens with 93% KV cache reduction over dense peers.
by OpenAI·Nov 2025
OpenAI's GPT-5.1 update with Instant and Thinking variants, delivering a warmer conversational tone and better instruction following on top of GPT-5's architecture.
by Moonshot AI·Nov 2025
1T/32B-active MoE with 256K context, native INT4 QAT, Modified MIT license; deep-thinking variant with 200-300 sequential tool calls for agentic tasks.
by Vinkius Labs·Feb 2026
Open-source TypeScript framework by Vinkius Labs for building secure MCP servers with a presenter layer that normalizes tool outputs for agents.
by Microsoft·May 2024
Microsoft's 7B dense Phi-3 SLM with 8K and 128K context variants (MIT license), benchmarked as a mid-tier textbook-trained model between Phi-3 mini and medium.