AITrending catalog
Models
Models and everything being built on top of them.
53 models indexed
53 models shown
Qwen3.8 27B
Qwen's dense 27B open-weight model for multilingual reasoning and generation.
Qwen3.8 Flash Next
A fast Qwen3.8 checkpoint built for efficient agent and assistant workloads.
Qwen3.6 35B-A3B
A Qwen mixture-of-experts checkpoint balancing total capacity with lower active parameters.
DeepSeek V4 Pro
DeepSeek's open V4 Pro checkpoint for advanced general reasoning and generation.
DeepSeek V4.1 Flash
An MIT-licensed DeepSeek V4.1 checkpoint for fast multimodal text generation.
DeepSeek V4 Flash
A faster open DeepSeek V4 checkpoint aimed at efficient inference.
Phi-4
Microsoft's compact open language model for reasoning and resource-conscious deployment.
Voxtral Mini Realtime
Mistral AI's compact realtime speech model for low-latency transcription workflows.
Laya
An Apache-2.0 System One model for classification, routing, scoring and guardrails.
NVIDIA Nemotron 3 Super
NVIDIA's 120B mixture-of-experts Nemotron checkpoint for open reasoning workloads.
Qwen Image 2.1
Qwen's Diffusers-compatible model for text-to-image generation and image editing.
Llama 4 Scout
Meta's open-weight Llama 4 Scout mixture-of-experts instruction model.
Phi-4 Mini Instruct
A smaller instruction-tuned Phi-4 checkpoint for local and edge-friendly workloads.
NVIDIA Nemotron 3.5 Lightning
An optimized 30B mixture-of-experts Nemotron release for accelerated inference.
Ministral 3 14B
A compact Mistral instruction model sized for controlled and local deployment.
Mistral Medium 3.5
Mistral AI's 128B-class open model for capable general-purpose inference.
Llama 4 Maverick
Meta's larger Llama 4 Maverick mixture-of-experts instruction checkpoint.
Mistral Small 4
A 119B mixture-of-experts Mistral checkpoint for efficient general workloads.
NVIDIA Nemotron 3 Embed 8B
An NVIDIA embedding model for retrieval, semantic search and enterprise indexing.
NVIDIA Kumo Tabular
NVIDIA's tabular foundation models predict classifications or numeric values from labeled examples, with three sizes spanning 28M to 215M parameters.
Fara 1.5 4B
A compact Microsoft research checkpoint focused on agentic computer-use tasks.
Holo4 27B
H Company's 27B agent model combines graphical interfaces, code, MCP and APIs; its downloadable checkpoint carries a noncommercial license.
Claude Sonnet 5.5
Anthropic's September 28 Sonnet release targets faster coding and knowledge work at the same base token prices as Sonnet 5.
Claude
Anthropic's model family for reasoning, coding, analysis and agent workflows.
GPT-6.1 Sol
OpenAI's September 29 model for coding, computer use and professional work, with standard API pricing of $2 input and $10 output per million tokens.
GLM-5.3
Z.ai's open-weight model assessed for cyber capabilities by NIST and Anthropic.
Jev
TypeSafe AI's System One model for fast, typed, confidence-aware decisions.
Amazon Nova 2 Lite
Amazon's fast and cost-conscious Nova 2 model for broad Bedrock workloads.
Amazon Nova Pro
A higher-capability Amazon Nova model for complex multimodal enterprise workloads.
Amazon Nova 2 Sonic
Amazon's next-generation speech model for conversational and realtime voice experiences.
Claude Haiku 4.5
A fast Claude tier for high-volume tasks where latency matters more than maximum depth.
Command A Plus
Cohere's higher-capability Command A tier for enterprise generation and tool use.
Claude Opus 5
Anthropic's highest-capability Claude 5 tier for difficult reasoning and agent workflows.
Claude Sonnet 5
Anthropic's balanced Claude 5 tier for production coding, analysis and tool use.
Command A Reasoning
A Command A model specialized for deliberate reasoning in enterprise workflows.
Command A Vision
Cohere's multimodal Command A model for enterprise image and document understanding.
Gemini 3.1 Pro
A Gemini Pro model for complex multimodal reasoning and production applications.
Gemini 3.5 Flash
A fast Gemini model for multimodal, high-throughput and interactive applications.
Gemini 3.5 Flash-Lite
A lightweight Gemini tier designed for cost-sensitive, high-volume inference.
GPT-5.6 Luna
The faster and more economical tier in OpenAI's GPT-5.6 agentic model family.
GPT-5.6 Sol
A frontier agentic coding model in OpenAI's GPT-5.6 family for developers.
GPT-5.6 Terra
OpenAI's balanced GPT-5.6 model for everyday agentic coding and general workloads.
GPT-6 Astra
OpenAI's frontier model for complex reasoning, coding and demanding agentic work.
GPT Realtime 2.1
OpenAI's realtime model family for low-latency voice and interactive multimodal sessions.
Grok 4.7
A current Grok capability tier listed in xAI's official model catalog.
Grok Code Fast 1
xAI's coding-focused Grok model for fast software-development workflows.
Grok 4.20
An xAI Grok family with reasoning, non-reasoning and multi-agent variants.
DeepSeek-R1
Reasoning model release that made open reasoning systems a central developer conversation.
Gemini 3.8 Flash
Google DeepMind's Gemini 3.8 Flash and 3.8 Flash Cyber, announced September 2026.
Gemini 3.8 Live
Google DeepMind's Gemini 3.8 Live and 3.8 Live Extended Thinking.
Gemma 3
Google's open model family designed for capable, efficient multimodal applications.
GPT-OSS 120B
Open-weight reasoning model intended to run in environments controlled by developers.
Qwen3
Alibaba's open model family with reasoning and general-purpose variants for developers.
No matching models.