DOSDOS
Pricing
Get started

Providers

Models (25)

Featured Models

DOS.AINew
LLM

DOS.AI Auto

Smart routing - automatically picks the best model for your request. Free for simple tasks, paid models for complex ones.

Smart Router128K context
autodynamic pricing
DOS.AINew
LLM

DOS.AI

Ultra-efficient MoE model — 35B total, 3B active parameters. Fast inference at near-8B cost with 70B-class quality.

35B MoE (3B active)128K context
$0.07 / $0.50per 1M tokens
DeepSeekNew
LLM

DeepSeek V4 Flash

1M-context fast tier replacing DeepSeek V3

MoE (fast tier)1M context
$0.15 / $0.29per 1M tokens
GoogleNew
LLM

Gemini 3.1 Pro

Google's most advanced reasoning model for complex tasks

Pro1M context
$2.10 / $12.60per 1M tokens
OpenAINew
LLM

GPT-5.5

OpenAI flagship — replaces GPT-5.4 at top tier with native reasoning

Flagship1M context
$5.25 / $31.50per 1M tokens
AnthropicNew
LLM

Claude Opus 4.8

Opus1M context
$5.25 / $26.25per 1M tokens

All Models

DOS.AINew
LLM

DOS.AI Auto

Smart routing - automatically picks the best model for your request. Free for simple tasks, paid models for complex ones.

Smart Router128K context
autodynamic pricing
DOS.AINew
LLM

DOS.AI

Ultra-efficient MoE model — 35B total, 3B active parameters. Fast inference at near-8B cost with 70B-class quality.

35B MoE (3B active)128K context
$0.07 / $0.50per 1M tokens
DeepSeekNew
LLM

DeepSeek V4 Flash

1M-context fast tier replacing DeepSeek V3

MoE (fast tier)1M context
$0.15 / $0.29per 1M tokens
OpenAINew
LLM

GPT-5.4 Nano

Cheapest GPT-5.4-class model for simple high-volume tasks

Nano400K context
$0.21 / $1.31per 1M tokens
GoogleNew
LLM

Gemini 3.1 Flash-Lite

Fastest and most cost-efficient Gemini 3 model

Flash-Lite1M context
$0.26 / $1.58per 1M tokens
Google
LLM

Gemini 3.5 Flash-Lite

Fast, cost-effective Flash-Lite (GA). 1M context. Closest Gemini peer to dos-ai (Qwen3.6-35B-A3B). [PROMO] Free when used BY a DOSClaw agent on a paid plan (agent traffic only - direct API usage is billed normally). Limited-time.

1M context
$0.32 / $2.63per 1M tokens
Google
LLM

Gemini 2.5 Flash

Previous-generation Flash, best price-performance. GA and stable.

1M context
$0.32 / $2.63per 1M tokens
QwenNew
LLM

Qwen 3.7 Plus

Qwen 3.7 Plus - multimodal (text+image), 1M context

Plus1M context
$0.42 / $1.68per 1M tokens
GoogleNew
LLM

Gemini 3.1 Flash Live

Real-time voice and dialogue model

Flash Live1M context
$0.79 / $4.73per 1M tokens
OpenAI
LLM

GPT-5.4 Mini

Strong mini model for coding, computer use, and sub-agents

Mini400K context
$0.79 / $4.73per 1M tokens
Anthropic
LLM

Claude Haiku 4.5

Fastest and most compact Claude model

Haiku200K context
$1.05 / $5.25per 1M tokens
xAINew
LLM

Grok 4.3

xAI flagship reasoning model - 1M context, text+image (replaces grok-4.1-fast)

Flagship1M context
$1.31 / $2.63per 1M tokens
GoogleNew
LLM

Gemini 3.5 Flash

Latest Gemini Flash - frontier performance, standard tier

Flash1M context
$1.58 / $9.45per 1M tokens
Google
LLM

Gemini 3.6 Flash

Google latest balanced Flash model (launched 2026-07-21). 1M context, 64k output.

1M context
$1.58 / $7.88per 1M tokens
DeepSeekNew
LLM

DeepSeek V4 Pro

Near-frontier quality at ~1/6 the cost of Opus 4.7 / GPT-5.5

MoE (pro tier)1M context
$1.83 / $3.65per 1M tokens
GoogleNew
LLM

Gemini 3.1 Pro

Google's most advanced reasoning model for complex tasks

Pro1M context
$2.10 / $12.60per 1M tokens
xAINew
LLM

Grok 4.20

xAI flagship reasoning model with 2M context

Flagship2M context
$2.10 / $6.30per 1M tokens
QwenNew
LLM

Qwen 3.7 Max

Qwen 3.7 Max tier - 1M context (standard price; OpenRouter shows a 50%-off promo)

Max1M context
$2.63 / $7.88per 1M tokens
AnthropicNew
LLM

Claude Sonnet 4.6

Fast, intelligent model for everyday tasks

Sonnet200K context
$3.15 / $15.75per 1M tokens
OpenAINew
LLM

GPT-5.5

OpenAI flagship — replaces GPT-5.4 at top tier with native reasoning

Flagship1M context
$5.25 / $31.50per 1M tokens
AnthropicNew
LLM

Claude Opus 4.8

Opus1M context
$5.25 / $26.25per 1M tokens
MiniMaxNew
Audio

MiniMax Music 2.5+

Music generation with vocals or instrumental. ~3 min per track. Price: per 1000 seconds of audio.

Music gen
$50.00 / $0.00per 1000 sec
AlibabaNew
Video

Wan 2.7 Text-to-Video

Video generation from text prompt via Alibaba Wan 2.7. Duration 2-15s, 1080P, native audio. Pricing: per 1000 seconds.

Text-to-Video
$100.00 / $0.00per 1000 sec
AlibabaNew
Video

Wan 2.7 Image-to-Video

Video generation from image + text prompt via Alibaba Wan 2.7. Pricing: per 1000 seconds.

Image-to-Video
$100.00 / $0.00per 1000 sec
Alibaba

Wan 2.2 Text-to-Image Flash

Fast text-to-image generation via Alibaba Wan 2.2. Pricing: per 1000 generated images.

$25.00 / $0.00per 1M tokens

Ready to get started?

Start building with $10 in free credits. No credit card required.

Start building for freeRead the docs
DOSDOS

AI infrastructure for everyone. Inference, agents, and safety - all in one platform.

Product

  • Models
  • Pricing
  • API Inference
  • DOSClaw

Developers

  • Documentation
  • API Reference
  • Status

DOS Ecosystem

  • DOSafe
  • DOS.Me
  • DOScan
  • DOSwap
  • MetaDOS

Company

  • About
  • Contact
  • Careers
  • Privacy
  • Terms

© 2026 All rights reserved.