WritingmateWritingmate

AI model reference · 2026

MiniMax M3

MiniMax: MiniMax M3 logoMiniMaxModel details, evidence, and API facts

MiniMax-M3 is a multimodal foundation model from MiniMax.

Text generation
Image input
Reasoning
Provider endpoint accepts tool parameters
Context window
1,048,576 tokens
Maximum output
512,000 tokens
Base provider API input
$0.30 / 1M tokens
Base provider API rate
Published evidence
4 benchmarks, 2 Arena results

Overview

About MiniMax M3

A concise catalog overview. Technical limits and published evaluation evidence are listed separately below.

MiniMax-M3 is a multimodal foundation model from MiniMax.

text
multimodal
reasoning

Published evaluations

Benchmark results

Only results mapped to this exact model and backed by a named source appear here. Different evaluation protocols are not treated as interchangeable.

Source-backed benchmark results for MiniMax M3
BenchmarkScoreEvaluation detailsSource
OSWorld-VerifiedA verified computer-use benchmark in which multimodal agents operate desktop applications and are graded from the resulting environment state.75.2%Mean task rewardVersion: OSWorld-Verified v1 · Split: Up to 361 tasks, excluding eight Google Drive tasks where required · Harness: Official OSWorld unified environment · Attempts: One run · Tools: GUI interactionProtocol note: Only 100-step general-model rows are included in this normalized snapshot. Individual evaluated-task counts can differ when a task could not run.; Run date: 2026-06-07; Evaluated tasks: 358Methodology OSWorldBenchmark owner
GPQA DiamondThe highest-quality subset of Graduate-Level Google-Proof Q&A, designed to test expert-level scientific reasoning in biology, physics, and chemistry.92.9%AccuracyVersion: GPQA Diamond 198 · Split: 198 Diamond questions · Harness: Artificial Analysis independent evaluation · Tools: Not disclosed on the score pageMethodology Artificial AnalysisIndependently reproduced
Humanity's Last ExamA 2,500-question expert-level benchmark spanning dozens of academic fields. Tool-assisted and no-tools results are separate protocols and must not be merged.39.0%AccuracyVersion: May 2025 text-only revision · Split: 2,158 text-only questions from the 2,500-question May 2025 revision · Attempts: pass@1 · Tools: No browser or retrieval tools · Evaluator: LLM equality checker with numerical toleranceMethodology Artificial AnalysisIndependently reproduced
LiveBenchA contamination-resistant benchmark refreshed on a fixed release cadence. Scores from different LiveBench releases must never be compared as the same protocol.67.3%Mean of category averagesVersion: LiveBench 2026-06-25 · Split: 2026-06-25 release, 23 tasks across seven categoriesProtocol note: Overall is the mean of category averages. This protocol is not comparable with the 2026-01-08 release.Methodology LiveBenchBenchmark owner

Human preference

Arena results

Arena scores come from blind human preference votes. They are reported separately from task benchmarks and are not used as substitutes for missing benchmark results.

Arena results for MiniMax M3
CategoryScoreRankVotesStatus
Arena TextBlind human preference score for text responses.1,444±5#7237,312Published
Arena WebDevBlind human preference score for code and web development.1,490±7#3410,053Published

Technical reference

Specifications and API facts

Technical limits and reference provider prices come from the current catalog unless a reviewed external source is listed. Missing values stay marked as unavailable.

Model
Provider
MiniMax
Release date
Not available in reviewed sources
Catalog added
May 31, 2026
Context window
1,048,576 tokens
Maximum output
512,000 tokens
Knowledge cutoff
Not available in reviewed sources
License
Not available in reviewed sources
Capabilities and API
Input types
Text, Image, Video
Output types
Text
Reasoning
Supported
Provider endpoint tool parameters
Accepts tool parameters
Base provider API input
$0.30 / 1M tokens
Base provider cached input
$0.06 / 1M tokens
Base provider API output
$1.20 / 1M tokens

These are reference provider API rates, not Writingmate checkout charges. Access in Writingmate follows the allowances of your Writingmate plan.

Provenance

Sources and update status

Source links are attached to the facts and evaluations they support. Catalog-only values are not presented as independently verified claims.

Try MiniMax M3 in Writingmate

Use this model, then compare its response with other AI models in the same workspace.