WritingmateWritingmate

Multimodal AI Models

AI Models

350+ AI Models in Writingmate All-in-One AI Platform

Explore and compare the best AI models from OpenAI, Anthropic, Google, and more.

Multimodal

220 models

Muse Spark 1.2 Contributor

Meta

BasicTextReasoning

Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost.

View details

DeepSeek V4 Flash Vision Exp

DeepSeek

BasicTextReasoning

DeepSeek V4 Flash Vision Exp is a model by DeepSeek.

View details

Ox Alpha

Stealth

BasicTextReasoning

Ox Alpha is a reasoning model designed for coding, sustained agentic work, and production workloads.

View details

FLUX Video Upscale

Black Forest Labs

ProVideo

FLUX Video Upscale is a video upscaling model from Black Forest Labs.

View details

Qwen3.8 27B

Qwen

ProTextReasoning

Qwen3.8 27B is an open-weight dense vision-language model from Qwen.

View details

Seedream 5.0 Lite

ByteDance

ProImage

Seedream 5.0 Lite is an image generation model from ByteDance Seed.

View details

Gemini 3.7 Flash

Google

ProTextReasoning

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning.

View details

Seedream 5.0 Pro

ByteDance

ProImage

Seedream 5.0 Pro is an image generation and editing model from ByteDance Seed.

View details

Seedance 2.0 Mini

ByteDance

ProVideo

Seedance 2.0 Mini is a video generation model from ByteDance.

View details

Seed 2.1 Turbo

ByteDance

ProTextReasoning

Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows.

View details

Seed-2.0-Code

ByteDance

ProTextReasoning

Seed 2.0 Code is a model from ByteDance Seed optimized for agentic coding.

View details

Grok 4.6

xAI

ProTextReasoning

Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

View details

Grok Imagine Image 2.0

xAI

ProImage

Grok Imagine Image 2.0 is an image generation and editing model from xAI.

View details

Sakana Namazu

Sakana AI

ProTextReasoning

Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts.

View details

Muse Glimmer 30B

Meta

ProTextReasoning

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware.

View details

Seedance 2.5

ByteDance

ProVideo

Seedance 2.5 is a video generation model from ByteDance.

View details

Muse Spark 1.2

Meta

ProTextReasoning

Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks.

View details

Qwen Image 3

Qwen

ProImage

Qwen Image 3 is a unified image generation and editing model from Qwen.

View details

Qwen Image 3 Pro

Qwen

ProImage

Qwen Image 3 Pro is an image generation and editing model from Qwen.

View details

FLUX.3 Video

Black Forest Labs

ProVideo

FLUX.3 Video is a video generation model from Black Forest Labs.

View details

Qwen3.8 Max

Qwen

ProTextReasoning

Qwen3.8 Max is the flagship model in Alibaba's Qwen3.8 series, the general-availability successor to the Qwen3.8 Max Preview.

View details

Inkling Small

Thinking Machines

ProTextReasoning

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total.

View details

H3

MiniMax

ProVideo

MiniMax H3 is a lightweight, open-weights video generation model from MiniMax.

View details

Aleph 2.0

Runway

ProVideo

Runway Aleph 2.0 is an in-context video editing model from Runway.

View details

Gen-4.5

Runway

ProVideo

Runway Gen-4.5 is a video generation model from Runway for text-to-video and image-to-video workflows.

View details

Qwen3.7 Flash

Qwen

BasicTextReasoning

Qwen3.7 Flash is a vision-language reasoning model from Alibaba.

View details

Claude Opus 5

Anthropic

UltimateTextReasoning

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work.

View details

MAI-Image-2.5 Pro

Microsoft

ProImage

Microsoft's MAI-Image-2.5 is a high-quality image generation model available via Azure AI Foundry.

View details

Gemini 3.6 Flash

Google

ProTextReasoning

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development.

View details

Gemini 3.5 Flash Lite

Google

BasicTextReasoning

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities.

View details

Krea 2 Large

Krea

ProImage

Krea 2 Large is Krea's high-capability image generation model, more than twice the size of Krea 2 Medium.

View details

Krea 2 Medium

Krea

ProImage

Krea 2 Medium is Krea's balanced, cost-efficient image generation model and a practical starting point for a broad range of use cases.

View details

Krea 2 Medium Turbo

Krea

ProImage

Krea 2 Medium Turbo is a distilled, speed-focused variant of Krea 2 Medium from Krea.

View details

Grok Imagine Video 1.5

xAI

ProVideo

Grok Imagine Video 1.5 is a video generation model from SpaceXAI.

View details

Inkling

Thinking Machines

ProTextReasoning

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total.

View details

Kimi K3

Moonshot AI

ProTextReasoning

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI.

View details

Muse Spark 1.1

Meta

ProTextReasoning

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks.

View details

GPT-5.6 Luna Pro

OpenAI

BasicTextReasoning

GPT-5.6 Luna Pro is the same underlying model as GPT-5.6 Luna, served with reasoning.mode set to pro for higher-quality responses on complex tasks.

View details

GPT-5.6 Luna

OpenAI

BasicTextReasoning

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series.

View details

GPT-5.6 Terra Pro

OpenAI

ProTextReasoning

GPT-5.6 Terra Pro is the same underlying model as GPT-5.6 Terra, served with reasoning.mode set to pro for higher-quality responses on complex tasks.

View details

GPT-5.6 Terra

OpenAI

ProTextReasoning

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier.

View details

GPT-5.6 Sol Pro

OpenAI

ProTextReasoning

GPT-5.6 Sol Pro is the same underlying model as GPT-5.6 Sol, served with reasoning.mode set to pro for higher-quality responses on complex tasks.

View details

GPT-5.6 Sol

OpenAI

ProTextReasoning

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series.

View details

Grok 4.5

xAI

ProTextReasoning

Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.

View details

Claude Sonnet 5

Anthropic

ProTextReasoning

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work.

View details

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)

Google

BasicImageTextReasoning

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for high-velocity developer pipelines and rapid-fire visual exploration.

View details

Nex-N2-Mini

Nex AGI

BasicTextReasoning

Nex-N2-Mini is an open-source agentic mixture-of-experts model from Nex AGI, the smaller sibling in the Nex-N2 series.

View details

Fugu Ultra

Sakana AI

UltimateTextReasoning

Fugu Ultra is the higher-performance model in Sakana AI's Fugu family.

View details

HappyHorse 1.1

Alibaba

ProVideo

HappyHorse 1.1 is a video generation model from Alibaba.

View details

GPT Image 1 Mini

OpenAI

ProImage

A cost-efficient variant of GPT Image 1 for high-quality image generation at reduced latency and cost via OpenAI's dedicated Images API.

View details

HappyHorse 1.0

Alibaba

ProVideo

HappyHorse 1.0 is a video generation model from Alibaba.

View details

Nano Banana 2 (Gemini 3.1 Flash Image)

Google

ProImageTextReasoning

Gemini 3.1 Flash Image, a.k.a.

View details

Nano Banana Pro (Gemini 3 Pro Image)

Google

ProImageTextReasoning

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro.

View details

Kimi K2.7 Code

Moonshot AI

ProTextReasoning

MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts.

View details

Nex-N2-Pro

Nex AGI

BasicTextReasoning

Nex-N2-Pro is an agentic mixture-of-experts model from Nex AGI, with 17B active parameters out of 397B total.

View details

Riverflow V2.5 Pro

Sourceful

ProImageReasoning

Riverflow V2.5 Pro is the most powerful variant of Sourceful's Riverflow 2.5 lineup, best for top-tier control and quality-sensitive outputs.

View details

Riverflow V2.5 Fast

Sourceful

ProImageReasoning

Riverflow V2.5 Fast is the speed-optimized variant of Sourceful's Riverflow 2.5 lineup, best for production deployments and latency-critical workflows.

View details

Qwen3.7 Plus

Qwen

ProTextReasoning

Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series.

View details

MAI-Image-2.5

Microsoft

ProImage

Microsoft's MAI-Image-2.5 is a high-quality image generation model available via Azure AI Foundry.

View details

MiniMax M3

MiniMax

BasicTextReasoning

MiniMax-M3 is a multimodal foundation model from MiniMax.

View details

Step 3.7 Flash

StepFun

BasicTextReasoning

Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model.

View details

Claude Opus 4.8

Anthropic

UltimateTextReasoning

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family.

View details

Grok Build 0.1

xAI

ProTextReasoning

Grok Build 0.1 is SpaceXAI’s fast coding model trained specifically for agentic software engineering workflows.

View details

Gemini 3.5 Flash

Google

ProTextReasoning

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed.

View details

Grok Imagine Video

xAI

ProVideo

Grok Imagine Video is SpaceXAI's fast, text-, image-, and reference-conditioned video generation model.

View details

Grok Imagine Image Quality

xAI

ProImage

Grok Imagine Image Quality is SpaceXAI's fast, high-fidelity image generation and editing model.

View details

Recraft V4.1 Pro Vector

Recraft

ProImage

Recraft V4.1 Pro Vector is the vector (SVG) variant of Recraft V4.1 Pro, tuned for high aesthetics.

View details

Recraft V4.1 Vector

Recraft

ProImage

Recraft V4.1 Vector is the vector (SVG) variant of Recraft V4.1, tuned for high aesthetics.

View details

Recraft V4.1 Utility Pro

Recraft

ProImage

Recraft V4.1 Utility Pro is a general-purpose image generation model from Recraft.

View details

Recraft V4.1 Utility

Recraft

ProImage

Recraft V4.1 Utility is a general-purpose image generation model from Recraft.

View details

Recraft V4.1 Pro

Recraft

ProImage

Recraft V4.1 Pro is an image generation model from Recraft tuned for high aesthetics.

View details

Recraft V4.1

Recraft

ProImage

Recraft V4.1 is an image generation model from Recraft tuned for high aesthetics.

View details

Recraft V4 Pro Vector

Recraft

ProImage

Recraft V4 Pro Vector is the vector (SVG) variant of Recraft V4 Pro.

View details

Recraft V4 Vector

Recraft

ProImage

Recraft V4 Vector is the vector (SVG) variant of Recraft V4.

View details

Perceptron Mk1

Perceptron

BasicTextReasoning

Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning.

View details

Recraft V4 Pro

Recraft

ProImage

Recraft V4 Pro is an image generation model from Recraft.

View details

Recraft V4

Recraft

ProImage

Recraft V4 is an image generation model from Recraft.

View details

Recraft V3

Recraft

ProImage

Recraft V3 is an image generation model from Recraft.

View details

Gemini 3.1 Flash Lite

Google

BasicTextReasoning

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads.

View details

Grok 4.3

xAI

ProTextReasoning

Grok 4.3 is a reasoning model from SpaceXAI.

View details

Mistral Medium 3.5

Mistral AI

ProTextReasoning

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI.

View details

Video v3.0 Pro

KwaiVGI

ProVideo

Kling v3.0 Pro is Kuaishou's premium video generation model, offering higher visual quality than the Standard tier.

View details

Video v3.0 Standard

KwaiVGI

ProVideo

Kling v3.0 Standard is a video generation model from Kuaishou.

View details

Qwen3.5 Plus 2026-04-20

Qwen

BasicTextReasoning

Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba.

View details

Qwen3.6 Flash

Qwen

BasicTextReasoning

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series.

View details

Qwen3.6 35B A3B

Qwen

BasicTextReasoning

Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token.

View details

Qwen3.6 27B

Qwen

ProTextReasoning

Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026.

View details

GPT-5.5

OpenAI

UltimateTextReasoning

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks.

View details

Veo 3.1 Fast

Google

ProVideo

Google's mid-tier video generation model balancing speed and quality.

View details

Veo 3.1 Lite

Google

ProVideo

Google's most cost-effective video generation model, designed for high-volume applications and rapid iteration.

View details

MiMo-V2.5

Xiaomi

BasicTextReasoning

MiMo-V2.5 is a native omnimodal model by Xiaomi.

View details

Video O1

KwaiVGI

ProVideo

Kling Video O1 is a video generation model from Kuaishou.

View details

Hailuo 2.3

MiniMax

ProVideo

Hailuo 2.3 is a video generation model from MiniMax.

View details

Kimi K2.6

Moonshot AI

ProTextReasoning

Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration.

View details

Claude Opus 4.7

Anthropic

UltimateTextReasoning

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents.

View details

Wan 2.7

Alibaba

ProVideo

Wan 2.7 is a video generation model from Alibaba.

View details

Seedance 2.0

ByteDance

ProVideo

Seedance 2.0 is a video generation model from ByteDance.

View details

Seedance 2.0 Fast

ByteDance

ProVideo

Seedance 2.0 Fast is a video generation model from ByteDance.

View details

Gemma 4 26B A4B

Google

BasicTextReasoning

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind.

View details

Gemma 4 31B

Google

BasicTextReasoning

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output.

View details

Qwen3.6 Plus

Qwen

ProTextReasoning

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference.

View details

GLM 5V Turbo

Z.AI

ProTextReasoning

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks.

View details

Grok 4.20 Multi-Agent

xAI

ProTextReasoning

Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows.

View details

Grok 4.20

xAI

ProTextReasoning

Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities.

View details

Lyria 3 Pro Preview

Google

BasicText

Full-length songs are priced at $0.08 per song.

View details

Lyria 3 Clip Preview

Google

BasicText

30 second duration clips are priced at $0.04 per clip.

View details

Wan 2.6

Alibaba

ProVideo

Alibaba's most advanced video generation model, supporting over 10 visual creation capabilities in a unified system.

View details

Seedance 1.5 Pro

ByteDance

ProVideo

ByteDance's next-generation audio-visual generation model with a 4.5B parameter Dual-Branch Diffusion Transformer architecture.

View details

Sora 2 Pro

OpenAI

ProVideo

OpenAI's flagship video generation model, delivering production-quality video with physics-accurate motion, synchronized audio, and world-state persistence across shots.

View details

Veo 3.1

Google

ProVideo

Google's state-of-the-art video generation model, built for maximum visual fidelity in final production cuts.

View details

Reka Edge

Reka AI

BasicTextReasoning

Reka Edge is an extremely efficient 7B multimodal vision-language model that accepts image/video+text inputs and generates text outputs.

View details

GPT-5.4 Nano

OpenAI

BasicTextReasoning

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks.

View details

GPT-5.4 Mini

OpenAI

ProTextReasoning

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads.

View details

Mistral Small 4

Mistral AI

BasicTextReasoning

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system.

View details

Seed-2.0-Lite

ByteDance

BasicTextReasoning

Seed-2.0-Lite is a model by ByteDance.

View details

Qwen3.5-9B

Qwen

BasicTextReasoning

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture.

View details

GPT-5.4

OpenAI

ProTextReasoning

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system.

View details

Gemini 3.1 Flash Lite Preview

Google

BasicTextReasoning

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases.

View details

Seed-2.0-Mini

ByteDance

BasicTextReasoning

Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment.

View details

Nano Banana 2 (Gemini 3.1 Flash Image Preview)

Google

ProImageTextReasoning

Gemini 3.1 Flash Image Preview, a.k.a.

View details

Qwen3.5-35B-A3B

Qwen

BasicTextReasoning

The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency.

View details

Qwen3.5-27B

Qwen

BasicTextReasoning

The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance.

View details

Qwen3.5-122B-A10B

Qwen

BasicTextReasoning

The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency.

View details

Qwen3.5-Flash

Qwen

BasicTextReasoning

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency.

View details

Gemini 3.1 Pro Preview Custom Tools

Google

ProTextReasoning

Gemini 3.1 Pro Preview Custom Tools is a model by Google.

View details

GPT-5.3-Codex

OpenAI

ProTextReasoning

GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2.

View details

Gemini 3.1 Pro Preview

Google

ProTextReasoning

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows.

View details

Claude Sonnet 4.6

Anthropic

ProTextReasoning

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work.

View details

Qwen3.5 Plus 2026-02-15

Qwen

BasicTextReasoning

The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency.

View details

Qwen3.5 397B A17B

Qwen

ProTextReasoning

The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency.

View details

Claude Opus 4.6

Anthropic

UltimateTextReasoning

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks.

View details

Riverflow V2 Pro

Sourceful

ProImage

Riverflow V2 Pro is the most powerful variant of Sourceful's Riverflow 2.0 lineup, best for top-tier control and perfect text rendering.

View details

Riverflow V2 Fast

Sourceful

ProImage

Riverflow V2 Fast is the fastest variant of Sourceful's Riverflow 2.0 lineup, best for production deployments and latency-critical workflows.

View details

Kimi K2.5

Moonshot AI

ProTextReasoning

Kimi K2.5 is Moonshot AI's native multimodal model, delivering state-of-the-art visual coding capability and a self-directed agent swarm paradigm.

View details

GPT Audio

OpenAI

ProText

The gpt-audio model is OpenAI's first generally available audio model.

View details

GPT Audio Mini

OpenAI

ProText

A cost-efficient version of GPT Audio.

View details

FLUX.2 Klein 4B

Black Forest Labs

ProImage

FLUX.2 [klein] 4B is the fastest and most cost-effective model in the FLUX.2 family, optimized for high-throughput use cases while maintaining excellent image quality.

View details

GPT-5.2-Codex

OpenAI

ProTextReasoning

GPT-5.2-Codex is an upgraded version of GPT-5.1-Codex optimized for software engineering and coding workflows.

View details

Seedream 4.5

ByteDance

ProImage

Seedream 4.5 is the latest in-house image generation model developed by ByteDance.

View details

Seed 1.6 Flash

ByteDance

BasicTextReasoning

Seed 1.6 Flash is an ultra-fast multimodal deep thinking model by ByteDance Seed, supporting both text and visual understanding.

View details

Seed 1.6

ByteDance

BasicTextReasoning

Seed 1.6 is a general-purpose model released by the ByteDance Seed team.

View details

Gemini 3 Flash Preview

Google

ProTextReasoning

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance.

View details

FLUX.2 Max

Black Forest Labs

ProImage

FLUX.2 [max] is the new top-tier image model from Black Forest Labs, pushing image quality, prompt understanding, and editing consistency to the highest level yet.

View details

GPT-5.2 Chat

OpenAI

ProText

GPT-5.2 Chat (AKA Instant) is the fast, lightweight member of the 5.2 family, optimized for low-latency chat while retaining strong general intelligence.

View details

GPT-5.2

OpenAI

ProTextReasoning

GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1.

View details

GLM 4.6V

Z.AI

BasicTextReasoning

GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media.

View details

GPT-5.1-Codex-Max

OpenAI

ProTextReasoning

GPT-5.1-Codex-Max is OpenAI’s latest agentic coding model, designed for long-running, high-context software development tasks.

View details

Nova 2 Lite

Amazon

BasicTextReasoning

Nova 2 Lite is a fast, cost-effective reasoning model for everyday workloads that can process text, images, and videos to generate text.

View details

Ministral 3 14B 2512

Mistral AI

BasicText

The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart.

View details

Ministral 3 8B 2512

Mistral AI

BasicText

A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.

View details

Ministral 3 3B 2512

Mistral AI

BasicText

The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities.

View details

Mistral Large 3 2512

Mistral AI

ProText

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.

View details

FLUX.2 Flex

Black Forest Labs

ProImage

FLUX.2 [flex] excels at rendering complex text, typography, and fine details, and supports multi-reference editing in the same unified architecture.

View details

FLUX.2 Pro

Black Forest Labs

ProImage

A high-end image generation and editing model focused on frontier-level visual quality and reliability.

View details

Claude Opus 4.5

Anthropic

UltimateTextReasoning

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use.

View details

Nano Banana Pro (Gemini 3 Pro Image Preview)

Google

ProImageTextReasoning

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro.

View details

GPT-5.1

OpenAI

ProTextReasoning

GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5.

View details

GPT-5.1-Codex

OpenAI

ProTextReasoning

GPT-5.1-Codex is a specialized version of GPT-5.1 optimized for software engineering and coding workflows.

View details

GPT-5.1-Codex-Mini

OpenAI

BasicTextReasoning

GPT-5.1-Codex-Mini is a smaller and faster version of GPT-5.1-Codex.

View details

Nova Premier 1.0

Amazon

ProText

Amazon Nova Premier is the most capable of Amazon’s multimodal models for complex reasoning tasks and for use as the best teacher for distilling custom models.

View details

Sonar Pro Search

Perplexity

ProTextReasoning

Exclusively available on the OpenRouter API, Sonar Pro's new Pro Search mode is Perplexity's most advanced agentic search system.

View details

Voxtral Small 24B 2507

Mistral AI

BasicText

Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class text performance.

View details

Qwen3 VL 32B Instruct

Qwen

BasicText

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video.

View details

GPT-5 Image Mini

OpenAI

ProImageTextReasoning

GPT-5 Image Mini combines OpenAI's advanced language capabilities, powered by GPT-5 Mini, with GPT Image 1 Mini for efficient image generation.

View details

Claude Haiku 4.5

Anthropic

ProTextReasoning

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models.

View details

Qwen3 VL 8B Thinking

Qwen

BasicTextReasoning

Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences.

View details

Qwen3 VL 8B Instruct

Qwen

BasicText

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video.

View details

Nano Banana (Gemini 2.5 Flash Image)

Google

BasicImageText

Gemini 2.5 Flash Image, a.k.a.

View details

Qwen3 VL 30B A3B Thinking

Qwen

BasicTextReasoning

Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos.

View details

Qwen3 VL 30B A3B Instruct

Qwen

BasicText

Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos.

View details

Claude Sonnet 4.5

Anthropic

ProTextReasoning

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows.

View details

Qwen3 VL 235B A22B Thinking

Qwen

ProTextReasoning

Qwen3-VL-235B-A22B Thinking is a multimodal model that unifies strong text generation with visual understanding across images and video.

View details

Qwen3 VL 235B A22B Instruct

Qwen

BasicText

Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual understanding across images and video.

View details

Mistral Medium 3.1

Mistral AI

ProText

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost.

View details

GLM 4.5V

Z.AI

ProTextReasoning

GLM-4.5V is a vision-language foundation model for multimodal agent applications.

View details

GPT-5

OpenAI

ProTextReasoning

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience.

View details

GPT-5 Mini

OpenAI

BasicTextReasoning

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks.

View details

GPT-5 Nano

OpenAI

BasicTextReasoning

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments.

View details

Codestral 2508

Mistral AI

BasicText

Mistral's cutting-edge language model for coding released end of July 2025.

View details

UI-TARS 7B

ByteDance

BasicText

UI-TARS-1.5 is a multimodal vision-language agent optimized for GUI-based environments, including desktop interfaces, web browsers, mobile systems, and games.

View details

Gemini 2.5 Flash Lite

Google

BasicTextReasoning

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency.

View details

ERNIE 4.5 VL 424B A47B

Baidu

ProTextReasoning

ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token.

View details

Mistral Small 3.2 24B

Mistral AI

BasicText

Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling.

View details

Gemini 2.5 Flash

Google

BasicTextReasoning

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks.

View details

Gemini 2.5 Pro

Google

ProTextReasoning

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks.

View details

Gemini 2.5 Pro Preview 06-05

Google

ProTextReasoning

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks.

View details

Claude Sonnet 4

Anthropic

ProTextReasoning

Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability.

View details

Mistral Medium 3

Mistral AI

ProText

Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost.

View details

Gemini 2.5 Pro Preview 05-06

Google

ProTextReasoning

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks.

View details

Llama Guard 4 12B

Meta

BasicText

Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification.

View details

o4 Mini High

OpenAI

ProTextReasoning

OpenAI o4-mini-high is the same model as o4-mini with reasoning_effort set to high.

View details

o3

OpenAI

ProTextReasoning

o3 is a well-rounded and powerful model across domains.

View details

o4 Mini

OpenAI

ProTextReasoning

OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities.

View details

GPT-4.1

OpenAI

ProText

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning.

View details

GPT-4.1 Mini

OpenAI

ProText

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost.

View details

GPT-4.1 Nano

OpenAI

BasicText

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series.

View details

Llama 4 Maverick

Meta

BasicText

Llama 4 Maverick is a model by Meta.

View details

Llama 4 Scout

Meta

BasicText

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B.

View details

Mistral Small 3.1 24B

Mistral AI

ProText

Mistral Small 3.1 24B Instruct is an upgraded variant of Mistral Small 3 (2501), featuring 24 billion parameters with advanced multimodal capabilities.

View details

Gemma 3 4B

Google

BasicText

Gemma 3 introduces multimodality, supporting vision-language input and text outputs.

View details

Gemma 3 12B

Google

BasicText

Gemma 3 introduces multimodality, supporting vision-language input and text outputs.

View details

Gemma 3 27B

Google

BasicText

Gemma 3 introduces multimodality, supporting vision-language input and text outputs.

View details

Sonar Reasoning Pro

Perplexity

ProTextReasoning

Note: Sonar Pro pricing includes Perplexity search pricing.

View details

Sonar Pro

Perplexity

ProText

Note: Sonar Pro pricing includes Perplexity search pricing.

View details

Saba

Mistral AI

BasicText

Mistral Saba is a 24B-parameter language model specifically designed for the Middle East and South Asia, delivering accurate and contextually relevant responses while maintaining efficient performance.

View details

o3 Mini High

OpenAI

ProTextReasoning

OpenAI o3-mini-high is the same model as o3-mini with reasoning_effort set to high.

View details

Qwen2.5 VL 72B Instruct

Qwen

ProText

Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects.

View details

o3 Mini

OpenAI

ProTextReasoning

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding.

View details

Sonar

Perplexity

ProText

Sonar is lightweight, affordable, fast, and simple to use — now featuring citations and the ability to customize sources.

View details

MiniMax-01

MiniMax

BasicText

MiniMax-01 is a combines MiniMax-Text-01 for text generation and MiniMax-VL-01 for image understanding.

View details

Nova Lite 1.0

Amazon

BasicText

Amazon Nova Lite 1.0 is a very low-cost multimodal model from Amazon that focused on fast processing of image, video, and text inputs to generate text output.

View details

Nova Pro 1.0

Amazon

ProText

Amazon Nova Pro 1.0 is a capable multimodal model from Amazon focused on providing a combination of accuracy, speed, and cost for a wide range of tasks.

View details

GPT-4o (2024-11-20)

OpenAI

ProText

The 2024-11-20 version of GPT-4o offers a leveled-up creative writing ability with more natural, engaging, and tailored writing to improve relevance & readability.

View details

Mistral Large 2407

Mistral AI

ProText

This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407).

View details

GPT-4o (2024-08-06)

OpenAI

ProText

The 2024-08-06 version of GPT-4o offers improved performance in structured outputs, with the ability to supply a JSON schema in the respone_format.

View details

GPT-4o-mini

OpenAI

BasicText

GPT-4o mini is OpenAI's newest model after GPT-4 Omni, supporting both text and image inputs with text outputs.

View details

GPT-4o-mini (2024-07-18)

OpenAI

BasicText

GPT-4o mini is OpenAI's newest model after GPT-4 Omni, supporting both text and image inputs with text outputs.

View details

GPT-4o

OpenAI

ProText

GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs.

View details

Mixtral 8x22B Instruct

Mistral AI

ProText

Mistral's official instruct fine-tuned version of Mixtral 8x22B.

View details

Mistral Large

Mistral AI

ProText

This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407).

View details

Access all models in one platform

GPT-5, Claude, Gemini, Sora, FLUX, and 350+ more AI models - all in one subscription.