Compare/ERNIE 5.0 Thinking Preview vs GLM-4.5V (Non-reasoning)

ERNIE 5.0 Thinking PreviewvsGLM-4.5V (Non-reasoning)

Side-by-side comparison of pricing, 12 benchmarks, and generation speed.

Baidu

ERNIE 5.0 Thinking Preview

Input
Output
Speed
TTFT
Z AI

GLM-4.5V (Non-reasoning)

Input
$0.6/M
Output
$1.8/M
Speed
92 tok/s
TTFT
1.85s

Winner by Category

Cheaper
GLM-4.5V (Non-reasoning)
Faster (tok/s)
GLM-4.5V (Non-reasoning)
Lower Latency
GLM-4.5V (Non-reasoning)
Benchmarks (1-0)
ERNIE 5.0 Thinking Preview

Pricing Comparison

MetricERNIE 5.0 Thinking PreviewGLM-4.5V (Non-reasoning)
Input ($/M tokens)$0.6
Output ($/M tokens)$1.8
Cost for 1M input + 100K output tokens:
GLM-4.5V (Non-reasoning)$0.78

Speed Comparison

Output Speed (tokens/s) — higher is better
ERNIE 5.0 Thinking Preview
GLM-4.5V (Non-reasoning)
92 tok/s
Time to First Token (seconds) — lower is better
ERNIE 5.0 Thinking Preview
GLM-4.5V (Non-reasoning)
1.85s

Editorial Analysis

Verdict. ERNIE 5.0 Thinking Preview wins the overall benchmark matchup 1–0 across 1 overlapping categories, but raw benchmark score is only one input to the decision.

Pricing. Pricing varies significantly between these models — check the table above for the exact per-token rates. Many production workloads actually surface input-token cost (retrieval-augmented prompts, code-context windows), so factor both directions.

Strengths. ERNIE 5.0 Thinking Preview is strongest on Intelligence Index (22.3). GLM-4.5V (Non-reasoning) leads on Intelligence Index (6.8).

Speed. Speed data is incomplete for this pair; benchmark and price should decide.

Provider. Baidu and Z AI sell to overlapping but distinct developer audiences: Baidu tends to ship frontier reasoning models with premium positioning, while Z AI often prices more aggressively. Your existing vendor relationships, billing, and SLA preferences may matter as much as the raw numbers above.

Recommendation. Both models have legitimate use cases — the right answer depends on whether you are optimizing for benchmark ceiling, latency, or unit cost. Start with the cheaper / faster model, evaluate against your specific task, and only switch if the upgrade shows a meaningful lift.

Benchmark Comparison

Data from Artificial Analysis API — 12 benchmarks

Intelligence Index
22.36.8
Coding Index
Math Index
GPQA Diamond
MMLU-Pro
LiveCodeBench
AIME 2025
MATH-500
Humanity's Last Exam
SciCode
IFBench
TerminalBench
ERNIE 5.0 Thinking Preview1 wins
0 winsGLM-4.5V (Non-reasoning)

Frequently Asked Questions

Which is cheaper, ERNIE 5.0 Thinking Preview or GLM-4.5V (Non-reasoning)?

GLM-4.5V (Non-reasoning) is cheaper overall. Its blended price (3:1 input/output ratio) is $0.90/M tokens vs $—/M for ERNIE 5.0 Thinking Preview.

Which model performs better on benchmarks?

ERNIE 5.0 Thinking Preview wins 1 out of 12 benchmarks compared to 0 for GLM-4.5V (Non-reasoning). See the detailed benchmark chart above for per-category results.

Which is faster for real-time applications?

GLM-4.5V (Non-reasoning) generates tokens faster at 92 tok/s vs — tok/s. However, GLM-4.5V (Non-reasoning) has lower time-to-first-token (1.85s vs —s).

When should I use ERNIE 5.0 Thinking Preview vs GLM-4.5V (Non-reasoning)?

Choose based on your priorities: GLM-4.5V (Non-reasoning) for lower cost, ERNIE 5.0 Thinking Preview for stronger benchmark performance, and GLM-4.5V (Non-reasoning) for faster generation. For latency-sensitive apps, check the TTFT comparison above.