Tag

AI Models

A collection of 56 posts

GLM

GLM-5.3-Flash: Specs, Pricing and MIT Open Weights

Z.ai's GLM-5.3-Flash is the model that ran anonymously as Ox Alpha: a 320B-total, 18B-active natively multimodal MoE with a 1M-token context and MIT open weights. Here are the verified specs, pricing and benchmarks.

· 11 min read
AI Models

Ox Alpha Was GLM-5.3-Flash: Specs, Price, Access

Ox Alpha was Z.ai's GLM-5.3-Flash, confirmed on 26 August 2026. The free stealth preview has ended and the OpenRouter listing is gone. Verified specs, MIT-licensed weights, real pricing, and how to use the model now.

· 13 min read
Kimi

Kimi K3 vs Claude Opus 5: Which Should You Use? (2026)

Claude Opus 5 leads on independently verified coding benchmarks; Kimi K3 costs 40% less and ships open weights. A head-to-head on benchmarks, cost, context, openness and speed — with an explicit decision rule.

· 11 min read
Kimi

Kimi K3 Pricing: API Costs, Plans & Real Bills (2026)

Kimi K3 costs $3 per 1M input tokens, $0.30 on a cache hit, and $15 per 1M output — flat across the full 1M-token context. Here are the verified rates, the rate-limit tiers, three worked cost examples, and how it prices against Claude Opus 5, GPT-5.6 and DeepSeek.

· 10 min read
Qwen

Qwen 3.8 vs Qwen 3.6: Same Architecture, +14 Points (2026)

The config.json diff between Qwen 3.8-27B and Qwen 3.6-27B is empty — same layers, same hidden size, same vocab. Every gain came from post-training. Here's what actually changed, what the upgrade costs you, and who should stay on 3.6.

· 7 min read
DeepSeek

DeepSeek V4 vs Kimi K3: Two Open Giants (2026)

DeepSeek V4 is 3.5-6x cheaper per token than Kimi K3 after the August 2026 price change. A cost-per-task comparison of two open-weight giants: pricing table, worked monthly bill, and exactly when K3's native vision and front-end coding earn the premium.

· 10 min read
DeepSeek

DeepSeek V4-Flash vs Claude Opus 5: The Real Cost Gap (2026)

DeepSeek V4-Flash costs $0.14/$0.28 vs Claude Opus 5 at $5/$25 - 36x cheaper input and 89x cheaper output today, and 18x/29x blended once DeepSeek's August 16 price rise lands. Worked monthly bills, benchmarks, and how to route bulk work cheap.

· 10 min read