Browse all available AI models across providers
50 of 50 models
Alibaba
qwen3.7-plus
Alibaba
Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-con
ByteDance
256K, balanced performance+cost
ByteDance
256K, vision+tools+reasoning, flagship
OpenAI
The gpt-audio model is OpenAIs first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is p
OpenAI
A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Input is priced at $0.60 per million...
DeepSeek
Reasoning, chain-of-thought, math/code/science
DeepSeek
Latest MoE flagship, 1M context
Full-length songs are priced at $0.08 per song. Lyria 3 is Googles family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz...
AI music generation, composition+arrangement
MoonshotAI
200K context, 100-agent cluster, previous flagship
Mistral AI
Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class text performance. It excels at speech transcription, translati
xAI
Latest, multi-agent capable
xAI
Grok 4.20 Multi-Agent is a variant of xAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesi
Cohere
Efficient RAG LLM
Cohere
Enterprise RAG LLM, tool use, multilingual
Baidu
Fast, cost-effective
Huawei
Domestic chip sovereign, industrial verticals
SenseTime
CV leader, embodied AI, 30+ industrial scenes
iFlytek
Voice interaction leader, education/medical/office
01.AI
Chinese+English bilingual multimodal
Baichuan
Chinese LLM, medical/legal vertical focus
StepFun
Step 3.5 Flash is StepFuns most capable open-source foundation model. Built on a sparse Mixture of Experts (MoE) architecture, it selectively activates only 11B of its 196B parameters per token....
StepFun
Fast, multimodal terminal agent
Shengshu
VFX+AI sound, 5s 1080p
Tencent
Hunyuan-A13B is a 13B active parameter Mixture-of-Experts (MoE) language model developed by Tencent, with a total parameter count of 80B and support for reasoning via Chain-of-Thought. It offers compe
Amazon
Balanced performance
Microsoft
WizardLM-2 8x22B is Microsoft AIs most advanced Wizard model. It demonstrates highly competitive performance compared to leading proprietary models, and it consistently outperforms all existing state
Stability AI
44.1kHz stereo, 3-min music gen
Stability AI
8B flagship, open weights
Black Forest Labs
Professional
Black Forest Labs
Fastest, Apache 2.0 open
AI21 Labs
Jamba Large 1.7 is the latest model in the Jamba open family, offering improvements in grounding, instruction-following, and overall efficiency. Built on a hybrid SSM-Transformer architecture with a 2
AI21 Labs
Mamba-Transformer hybrid, 262K
Perplexity
Search-augmented, 200K, vision
Perplexity
Multi-step CoT + search
Reka AI
Reka Edge is an extremely efficient 7B multimodal vision-language model that accepts image/video+text inputs and generates text outputs. This model is optimized specifically to deliver industry-leadin
Reka AI
Reka Flash 3 is a general-purpose, instruction-tuned large language model with 21 billion parameters, developed by Reka. It excels at general chat, coding tasks, instruction-following, and function ca
Inflection AI
Inflection 3 Pi powers Inflections [Pi](https://pi.ai) chatbot, including backstory, emotional intelligence, productivity, and safety. It has access to recent news, and excels in scenarios like custo
Inflection AI
Inflection 3 Productivity is optimized for following instructions. It is better for tasks requiring JSON output or precise adherence to provided guidelines. It has access to recent news. For emotional
Liquid AI
LFM2.5-1.2B-Instruct is a compact, high-performance instruction-tuned model built for fast on-device AI. It delivers strong chat quality in a 1.2B parameter footprint, with efficient edge inference an
Liquid AI
LFM2.5-1.2B-Thinking is a lightweight reasoning-focused model optimized for agentic tasks, data extraction, and RAG—while still running comfortably on edge devices. It supports long context (up to 32K
Databricks
132B MoE open-source
Snowflake
Dense-MoE hybrid
Shanghai AI Lab
Chinese multimodal open-source flagship
Shanghai AI Lab
Vision-language, image/video/document
ModelBest
On-device, extreme compression, Sparse
IBM
IBM Apache 2.0, 12 languages, FIM
Poolside
Laguna M.1 is the flagship coding agent model from [Poolside](https://poolside.ai), optimized for complex software engineering tasks. Designed for agentic coding workflows, it supports tool calling an
Poolside
Laguna XS.2 is the second-generation model in the XS size class from [Poolside](https://poolside.ai), their efficient coding agent series. It combines tool calling and reasoning capabilities with a co