← comparisons · head to head

Muse Spark 1.3 vs GLM-5.3

Muse Spark 1.3 scores higher than GLM-5.3 on all 6 public leaderboards that rate both. On the ArtificialWatch Index, which takes the median of each model's percentile across the boards that score it, Muse Spark 1.3 is at 81.4 and GLM-5.3 at 42.4 (out of 100).

Scores as published by each board; table composed September 25, 2026.

Get the next launch alert, free →

Leaderboard by leaderboard

BoardMuse Spark 1.3GLM-5.3Higher
LMArena
Elo (community votes)
1,658.21,621.8Muse Spark 1.3
LiveBench
global average (0-100)
81.676.1Muse Spark 1.3
SimpleBench
AVG@5 % (private set)
81.866.2Muse Spark 1.3
LLM Stats
LLM Stats Score · composite (0-100), their rendered top 15
53.852.1Muse Spark 1.3
Vals AI
Vals Index accuracy % · weighted finance + coding tasks
64.557Muse Spark 1.3
Artificial Analysis
intelligence index · measured runs only
48.144.8Muse Spark 1.3

Raw scores are each board's own scale, so compare within a row, not across rows. Full table: the bench index.

Muse Spark 1.3 at a glance

Developer
Meta
Status
Live since September 2, 2026
API price
$1.25 input / $4.25 output per 1M tokens · Meta, muse-spark-1.3
Context window
1.05M tokens

Live — Sep 2 · $1.25/$4.25 per 1M and 1M ctx, identical to 1.2 · Meta: ~20% fewer tool calls, ~25% fewer tokens vs 1.2 · open weights on the roadmap, undated.

Muse Spark 1.3 is out — same price as 1.2, and the radar called it in 61 seconds

2026-09-03 · Meta released Muse Spark 1.3 today. Its own words: "We're excited to release Muse Spark 1.3, which delivers improved performance across agentic and coding tasks." The numbers Meta gives are efficiency, not capability: ~20% fewer tool calls and ~25% fewer tokens than 1.2 on coding work, fewer interaction turns, less…

Full Muse Spark 1.3 tracker →

GLM-5.3 at a glance

Developer
Zhipu AI
Status
Live since August 14, 2026
API price
$1.40 input / $4.40 output per 1M tokens · Z.AI, glm-5.3
Context window
1M tokens

Live - same 743B base as 5.2, every gain from post-training · tops CyberGym · open weights on Hugging Face (zai-org/GLM-5.3, 141 safetensors files, public when we read it Sep 25; repo created Aug 25).

GLM-5.3 tops CyberGym and trails the frontier on exploitation — and we missed it because the sweep had no eyes on models.dev

2026-08-14 · Z.ai shipped GLM-5.3 today, and we did not catch it. The entry was armed and the pattern was right — the sweep simply had no source that carried the id. That is now fixed, and the fix is the more useful half of this post.

Full GLM-5.3 tracker →

Muse Spark 1.3 vs GLM-5.3: quick answers

Which is better, Muse Spark 1.3 or GLM-5.3?

It depends on what you measure. Muse Spark 1.3 scores higher than GLM-5.3 on all 6 public leaderboards that rate both. Muse Spark 1.3's widest lead is on SimpleBench. On the ArtificialWatch Index, which takes the median of each model's percentile across the boards that score it, Muse Spark 1.3 is at 81.4 and GLM-5.3 at 42.4 (out of 100).

Which is newer, Muse Spark 1.3 or GLM-5.3?

Muse Spark 1.3. It went live on September 2, 2026, 19 days after GLM-5.3 (August 14, 2026).

Which is cheaper, Muse Spark 1.3 or GLM-5.3?

Muse Spark 1.3 is cheaper on both input and output. On each vendor's own API, Muse Spark 1.3 is $1.25 input / $4.25 output per million tokens and GLM-5.3 is $1.40 input / $4.40 output per million tokens. Prices as published in the models.dev catalog, as of 2026-09-25.

Which has the bigger context window, Muse Spark 1.3 or GLM-5.3?

Muse Spark 1.3, at 1.05M tokens against GLM-5.3's 1M, as each vendor lists it.

Who makes Muse Spark 1.3 and GLM-5.3?

Muse Spark 1.3 is from Meta; GLM-5.3 is from Zhipu AI.

Where do these numbers come from?

Each score is the leaderboard's own published figure, read by ArtificialWatch (table composed September 25, 2026). Boards measure different things, so raw scores are only comparable within a row.

More comparisons

GPT-6 Astra vs Claude Fable 5.1GPT-6 Astra vs GPT-5.6GPT-6 Astra vs Claude Opus 5Claude Fable 5.1 vs Claude Fable 5Claude Fable 5.1 vs Claude Opus 5Claude Opus 5 vs GPT-5.6Claude Opus 5 vs Claude Opus 4.8Claude Opus 5 vs Claude Fable 5Claude Opus 5 vs Kimi K3Kimi K3 vs Claude Fable 5Kimi K3 vs GLM-5.2Kimi K3 vs GPT-5.6DeepSeek V4 Pro vs GLM-5.2DeepSeek V4 Pro vs Kimi K3Grok 4.7 vs GPT-6 AstraGrok 4.7 vs Claude Opus 5Grok 4.7 vs Claude Fable 5.1Qwen 3.8 Max vs Claude Opus 5Qwen 3.8 Max vs Claude Fable 5Qwen 3.8 Max vs Kimi K3Qwen 3.8 Max vs DeepSeek V4 ProQwen 3.8 Max vs GLM-5.2Qwen 3.8 Max vs GPT-5.6GPT-6 Astra vs Claude Fable 5Claude Fable 5 vs Claude Opus 4.8Claude Fable 5 vs Claude Sonnet 5Claude Sonnet 5 vs Claude Opus 4.8Claude Sonnet 5 vs Claude Opus 5GLM-5.2 vs Claude Opus 4.8GLM-5.2 vs Claude Fable 5GLM-5.2 vs GPT-5.5Gemini 3.5 Flash vs Gemini 3.1 ProGemini 3.5 Flash vs Claude Sonnet 5GPT-6 Astra vs Claude Opus 4.8Claude Fable 5.1 vs Claude Opus 4.8Kimi K3 vs Claude Opus 4.8DeepSeek V4 Pro vs Claude Opus 4.8DeepSeek V4 Pro vs GPT-5.5DeepSeek V4 Pro vs Claude Fable 5GLM-5.3 vs Kimi K3GLM-5.3 vs GLM-5.2GLM-5.3 vs GPT-6 AstraGrok 4.6 vs Claude Opus 5Grok 4.6 vs Claude Fable 5Grok 4.6 vs GPT-5.6Grok 4.6 vs Grok 4.5Grok 4.6 vs Claude Sonnet 5Gemini 3.8 Flash vs Gemini 3.1 ProGemini 3.8 Flash vs Claude Opus 5Muse Spark 1.3 vs Gemini 3.8 FlashGPT-6 Sol vs GPT-6 AstraGPT-6 Sol vs Claude Opus 5GPT-6 Sol vs Claude Fable 5GPT-6 Luna vs GPT-6 AstraGPT-6 Sol vs GPT-6 LunaClaude Fable 5 vs GPT-5.6Claude Opus 4.8 vs GPT-5.5GPT-5.6 vs Claude Opus 4.8GPT-5.6 vs Claude Sonnet 5Claude Sonnet 5 vs GPT-5.5Claude Fable 5 vs GPT-5.5DeepSeek V4 Pro vs Claude Opus 5Qwen 3.8 Max vs GLM-5.3Qwen 3.8 Max vs Claude Opus 4.8GLM-5.3 Flash vs GLM-5.3GLM-5.3 Flash vs GLM-5.2MiniMax M3 vs Kimi K3MiniMax M3 vs DeepSeek V4 ProMiniMax M3 vs GLM-5.2MiniMax M3 vs GPT-5.5
sweeping every 60 seconds

Hear the minute the next one drops. Free.

A free Chrome push the moment a new model answers on a public API, an email about 15 minutes later. Texts or a phone call on paid plans.

  • Free forever
  • No card
  • Unsubscribe in one click

Also: the bench index · API pricing · the wire