Qwen3 30B A3B
Qwen3-30B-A3B is a smaller Mixture-of-Experts (MoE) model from the Qwen3 series by Alibaba, with 30.5 billion total parameters and 3.3 billion activated parameters. Features hybrid thinking/non-thinking modes, support for 119 languages, and enhanced agent capabilities. It aims to outperform previous models like QwQ-32B while using significantly fewer activated parameters.
| Provider | Status | Input | Output | Limits | Uptime | Speed | Notes |
|---|---|---|---|---|---|---|---|
| DeepInfra | available | $0.12/Mtok | $0.50/Mtok | 41K tokens context | 100.0% | 210 ms p50 TTFT | fp8 |
| Alibaba | available | $0.13/Mtok | $0.52/Mtok | 131K tokens context | 100.0% | 449 ms p50 TTFT | |
| NextBit | available | $0.14/Mtok | $0.55/Mtok | 33K tokens context | — | 1,105 ms p50 TTFT | fp8 |