Qwen3 30B A3B Instruct 2507
Qwen logo

Qwen3 30B A3B Instruct 2507

qwen3-30b-a3b-instruct-2507llms.txt
Qwen
Significantly improved performance on reasoning tasks, including logical reasoning, mathematics, science, coding, and academic benchmarks that typically require human expertise. Markedly better general capabilities, such as instruction following, tool usage, text generation, and alignment with human preferences. Enhanced 256K long-context understanding capabilities.

Pricing

  • Input Tokens: $0.1028 /M tokens
  • Output Tokens: $0.4112 /M tokens

Input Modalities

    Providers

    StreamLake streamlake-qwen3-30b-a3b-instruct-2507
    Pricing$0.1028$0.4112
    Context128K
    Max output0
    Latency-
    Throughput-
    Uptime
    0.00% uptime 2 days ago
    100.00% uptime yesterday
    0.00% uptime today

    Performance for qwen3-30b-a3b-instruct-2507

    Uptime is the percentage of requests that succeeded over the past 72 hours. AIHubMix continuously monitors every provider and automatically retries with the next-best provider when one returns an error or responds too slowly; Latency is total round-trip time (lower is better); Throughput is how fast the model writes (tokens per second, higher is better).

    Uptime
    Loading...
    Latency
    Loading...
    Throughput
    Loading...

    Try this model

    Python
    import os
    from openai import OpenAI
    
    client = OpenAI(
        api_key=os.environ["AIHUBMIX_API_KEY"],
        base_url="https://aihubmix.com/v1",
    )
    
    response = client.chat.completions.create(
        model="qwen3-30b-a3b-instruct-2507",
        messages=[
          {
            "role": "user",
            "content": "Hello, how are you?"
          }
        ],
        max_tokens=1024,
        stream=False,
    )
    
    print(response.choices[0].message.content)

    Frequently asked questions

    What is Qwen3 30B A3B Instruct 2507?

    Significantly improved performance on reasoning tasks, including logical reasoning, mathematics, science, coding, and academic benchmarks that typically require human expertise. Markedly better general capabilities, such as instruction following, tool usage, text generation, and alignment with human preferences. Enhanced 256K long-context understanding capabilities.