Significantly improved performance on reasoning tasks, including logical reasoning, mathematics, science, coding, and academic benchmarks that typically require human expertise.
Markedly better general capabilities, such as instruction following, tool usage, text generation, and alignment with human preferences.
Enhanced 256K long-context understanding capabilities.
Pricing
- Input Tokens: $0.1028 /M tokens
- Output Tokens: $0.4112 /M tokens
Input Modalities
Providers
StreamLake streamlake-qwen3-30b-a3b-instruct-2507
Pricing$0.1028$0.4112
Context128K
Max output0
Latency-
Throughput-
Uptime
0.00% uptime 2 days ago
100.00% uptime yesterday
0.00% uptime today
Performance for qwen3-30b-a3b-instruct-2507
Uptime is the percentage of requests that succeeded over the past 72 hours. AIHubMix continuously monitors every provider and automatically retries with the next-best provider when one returns an error or responds too slowly; Latency is total round-trip time (lower is better); Throughput is how fast the model writes (tokens per second, higher is better).
Uptime
Loading...
Latency
Loading...
Throughput
Loading...
Try this model
Python