Blog

Read about our latest announcements.
DeepSeek V4 Pro vs Flash: A Buyer's Decision Guide
Models

DeepSeek V4 Pro vs Flash: A Buyer's Decision Guide

Compare DeepSeek V4 Pro and Flash across benchmarks, reasoning, coding, speed, context length, and API pricing to choose the right model.

Aug 7, 2026 18 min read
DeepSeek V4 Flash Benchmarks: Scores, Speed, Pricing and How It Compares
Models

DeepSeek V4 Flash Benchmarks: Scores, Speed, Pricing and How It Compares

Explore DeepSeek V4 Flash benchmarks, SWE-bench and GPQA scores, API pricing, inference speed, hardware requirements, and comparisons with leading models.

Aug 6, 2026 18 min read
DeepSeek V4 Flash API Pricing: Rates, Cost Calculator and Provider Comparison
Models

DeepSeek V4 Flash API Pricing: Rates, Cost Calculator and Provider Comparison

Compare DeepSeek V4 Flash API pricing, cache rates, provider costs, free access, and worked cost examples for chat, coding, and reasoning workloads.

Aug 5, 2026 16 min read
Introducing Organization: A Better Way for Teams to Build on ZenMux
Product

Introducing Organization: A Better Way for Teams to Build on ZenMux

Manage team AI usage with ZenMux Organizations: individual API keys, a shared PAYG balance, member limits, centralized billing, logs, and usage visibility.

Aug 4, 2026 3 min read
Why Does an AI Model Say It Is Another Model?
Case Studies

Why Does an AI Model Say It Is Another Model?

ZenMux asked 27 LLMs “Who are you?” 29,700 times to explore why AI models sometimes misidentify themselves — and why self-identification is not reliable proof of model routing.

Jul 31, 2026 11 min read
Ling-3.0-flash API Now Available on ZenMux: Built for Agentic Coding
Models

Ling-3.0-flash API Now Available on ZenMux: Built for Agentic Coding

Use the Ling-3.0-flash API on ZenMux for agentic coding, long-horizon tasks, and tool workflows, with 256K context and a practical API example.

Jul 24, 2026 8 min read
Token Economics: What Happens When We Price Every "Eastern Model" to Match DeepSeek?
Research

Token Economics: What Happens When We Price Every "Eastern Model" to Match DeepSeek?

Token Economics: What Happens When We Price Every "Eastern Model" to Match DeepSeek?

Jun 23, 2026 21 min read
ZenMux Token Economics: When Model Prices Meet, What Decides the Winner?
Events

ZenMux Token Economics: When Model Prices Meet, What Decides the Winner?

ZenMux Token Economics brings 10+ popular AI models to DeepSeek-level pricing. Test them on real work and let real usage decide.

Jun 23, 2026 7 min read
Claude Agent SDK with ZenMux: Cut Agent Loop Costs, Keep Loops Running
Models

Claude Agent SDK with ZenMux: Cut Agent Loop Costs, Keep Loops Running

Point the Claude Agent SDK at ZenMux to access 200+ models, route phases of your agent loop to cheaper models, and keep loops running through provider outages — all behind one API key.

Jun 11, 2026 15 min read
We Spent 1,000 Dollars Asking 27 Frontier LLMs From 16 Vendors One Question 29,700 Times: "Who Are You?"
Research

We Spent 1,000 Dollars Asking 27 Frontier LLMs From 16 Vendors One Question 29,700 Times: "Who Are You?"

We spent about 1,000 dollars asking 27 frontier LLMs from 16 vendors the same question 29,700 times: “Who are you?” Most of the time they knew. But in 7.1% of answers they claimed another vendor’s identity—mostly Anthropic, OpenAI, or Google—and that confusion spiked with language and de-branding prompts. This isn’t proof of distillation; it’s an identity echo in the training data. Full data and tools are open source.

Jun 5, 2026 19 min read
An Experimental Validation Plan for the Feedback That "ZenMux Has Worse Cache Hit Rates Than OpenRouter"
Technology

An Experimental Validation Plan for the Feedback That "ZenMux Has Worse Cache Hit Rates Than OpenRouter"

This article describes an experimental framework for comparing cache hit rates between ZenMux and OpenRouter. By fixing the model, provider, and prompt sequence, and by measuring cached_tokens and token_hit_rate across multi-turn conversations, the experiment shows that there is no significant difference in cache hit rate between the two platforms under the tested conditions.

Apr 27, 2026 11 min read
Debugging a Claude Code Freeze
Technology

Debugging a Claude Code Freeze

Recently, several users reported that Claude Code would freeze when calling models through ZenMux — after sending a prompt, nothing appeared on screen, and the only option was to force quit. As a model aggregation platform, we needed to figure out whether this was a ZenMux issue or something else entirely.

Mar 29, 2026 5 min read