Compare models
HEAD TO HEAD

MiniMax-M3 vs Muse Spark 1.3

Published test results, prices, limits and features, side by side. Share the link to show someone exactly this comparison.

  1. MiniMax-M3MiniMax

At a glance

  • Capability index

    • MiniMax-M3146.9
    • Muse Spark 1.3156.8

    Muse Spark 1.3 is 9.8 points higher.

  • Price per 1M tokens

    • MiniMax-M3$0.525
    • Muse Spark 1.3$2.00

    MiniMax-M3 is 3.8 times cheaper.

  • Context window

    • MiniMax-M31M tokens
    • Muse Spark 1.31M tokens

    They take in about the same.

  • Released

    • MiniMax-M3Jun 1, 2026
    • Muse Spark 1.3Sep 2, 2026

    Muse Spark 1.3 is 3 months newer.

Where each scores higher

On the 17 tests both have taken, each on its own scale. A gap of a point or two can sit within a test’s margin of error.

MiniMax-M3

No shared test where it scores higher.

Muse Spark 1.3

  • LiveBench overall+14.3 points
  • Arena text+54 points
  • LiveBench coding+12.9 points
  • Arena WebDev+175 points
  • Arena coding+44 points
  • LiveBench mathematics+19.0 points
  • Mock AIME 2024–2025+28.1 points
  • Arena math+77 points
  • LiveBench agentic coding+23.4 points
  • LiveBench language+6.0 points
  • LiveBench instructions+20.5 points
  • Arena creative writing+53 points
  • Arena instruction following+51 points
  • LiveBench data analysis+3.4 points
  • LiveBench reasoning+15.2 points
  • Chess puzzles+24.0 points
  • Arena hard prompts+55 points

4 more tests have results for only one of them; the table below lists every result.

Everything side by side

Comparison of MiniMax-M3 vs Muse Spark 1.3
MeasureMiniMax-M3MiniMaxMuse Spark 1.3Meta AI
Model
MakerMiniMaxMeta AI
Sold here byMiniMax (minimax.io)DevPass (LLM Gateway)
API model IDMiniMax-M3muse-spark-1.3
ReleasedJun 1, 2026Sep 2, 2026
Knowledge cutoffJan 2025Not reported
WeightsOpen: downloadableClosed: hosted access only
Developer’s countryChinaUnited States
Hosts selling it56including MiniMax directly11
Test results
Capability index146.9#68 of 274156.8 (best of these models)#19 of 274
LiveBench overall67.3%#61 of 6681.6% (best of these models)#7 of 66
Arena text1440#96 of 4131494 (best of these models)#9 of 413
LiveBench coding68.2%#64 of 6681.1% (best of these models)#17 of 66
Arena WebDev1482#58 of 1381657 (best of these models)#14 of 138
Arena coding1494#86 of 4081539 (best of these models)#11 of 408
LiveBench mathematics77.0%#65 of 6696.0% (best of these models)#12 of 66
Mock AIME 2024–202571.1%#133 of 29799.2% (best of these models)#15 of 297
Arena math1432#102 of 3961509 (best of these models)#9 of 396
FrontierMath Tiers 1–3Not tested74.4%#19 of 114
FrontierMath Tier 4Not tested46.3%#25 of 70
LiveBench agentic coding40.7%#60 of 6664.1% (best of these models)#9 of 66
CursorBenchNot tested41.6%#22 of 62
LiveBench language76.8%#45 of 6682.8% (best of these models)#24 of 66
LiveBench instructions57.5%#61 of 6678.0% (best of these models)#4 of 66
Arena creative writing1406#99 of 4111459 (best of these models)#28 of 411
Arena instruction following1434#88 of 4131486 (best of these models)#16 of 413
LiveBench data analysis76.2%#34 of 6679.6% (best of these models)#12 of 66
LiveBench reasoning74.5%#61 of 6689.7% (best of these models)#15 of 66
GPQA Diamond90.9%#34 of 319Not tested
Chess puzzles14.0%#113 of 22738.0% (best of these models)#32 of 227
Arena hard prompts1463#90 of 4131517 (best of these models)#10 of 413
Price per million tokens
Input$0.30 (best of these models)$1.25
Output$1.20 (best of these models)$4.25
Blended, 3 input to 1 output$0.525 (best of these models)$2.00
Cached input$0.06 (best of these models)$0.15
Long requests$0.60 in, $2.40 outabove 512K tokensNo separate price reported
The maker’s own priceThis listingNot sold directly
Limits
Context window1M tokens1M tokens
Max input512K tokens1M tokens (best of these models)
Max output512K tokens (best of these models)131.1K tokens
Features
Readsimages, text and videoaudio, images, PDFs, text and video
Producestexttext
ReasoningYeseffort low, medium, high; can be switched off; thinking budget can be setYeseffort minimal, low, medium, high, xhigh
Tool callingYesYes
Structured outputYes, such as JSONYes, such as JSON
File attachmentsYesYes
Temperature settingSupportedSupported

Bold marks the better value in each row. Prices are each model’s own list price, or the middle price across hosts when the maker does not sell it, blended as three input tokens for every output token. Test results as published by LiveBench (Apache-2.0), Epoch AI (CC BY 4.0), Cursor CursorBench, compiled by Epoch AI (CC BY 4.0), Arena (CC BY 4.0); model details and prices from models.dev (MIT). Confirm pricing with the provider before use.