intern-ai Models

2 modelsGeneral models free to startUp to 262K context

Usage

Last 8 days · 2026-09-17 to 2026-09-24

Tokens

165M

Requests

2.6K

Models in use

2 of 2

Tokens per day, stacked by model

017.2M34.3M09-1709-242026-09-17 — 15,706,180 tokens agents-a1-free: 14,089,025 intern-s2-free: 1,617,1552026-09-18 — 3,576,235 tokens agents-a1-free: 2,953,825 intern-s2-free: 622,4102026-09-19 — 17,839,115 tokens agents-a1-free: 15,166,360 intern-s2-free: 2,672,7552026-09-20 — 33,468,945 tokens agents-a1-free: 21,653,980 intern-s2-free: 11,814,9652026-09-21 — 27,518,885 tokens agents-a1-free: 18,668,960 intern-s2-free: 8,849,9252026-09-22 — 22,102,960 tokens agents-a1-free: 15,644,935 intern-s2-free: 6,458,0252026-09-23 — 34,333,025 tokens agents-a1-free: 22,842,460 intern-s2-free: 11,490,5652026-09-24 — 10,597,820 tokens agents-a1-free: 8,861,780 intern-s2-free: 1,736,040
  • agents-a1-free
  • intern-s2-free

Which models that traffic went to

  1. Agents A1 (free)72.6%120M
  2. Intern S2 (free)27.4%45.3M

Share of 165M tokens.

The two views disagree on purpose: a model can take a large share of the calls and a small share of the tokens — many short requests — or the reverse. Which one matters depends on whether your cost is driven by call volume or by prompt length. Measured on AIHubMix over the last 8 days, counting the 2 model IDs listed on this page; traffic routed through upstream-specific IDs that are not in the public catalog is not included.

All 2 intern-ai Models

Open in model list
intern-ai models on AIHubMix with input and output modalities, context length, maximum output, price per million tokens including cache read and cache write rates, and measured throughput and latency.
Modalities
agents-a1-freeTakes text, vision, returns text.262KFreeFree/M
intern-s2-freeTakes text, vision, returns text.262KFreeFree/M

Prices are USD per million tokens; cache read and cache write are the rates for prompt-cache hits and for writing a prompt into the cache. Throughput and latency are measured on AIHubMix — the same figures the model detail page shows — not vendor claims. A dash means the catalog does not publish that field for that model, which is not the same as the model not supporting it.

intern-ai on AIHubMix

Which intern-ai model should I start with?

agents-a1-free is free on input — the cheapest entry here that declares a token price, and it carries a 262K context. Move up to intern-s2-free when answer quality matters more than cost.

Why are there several entries for the same model?

Because each row is a route you can call, not a model release. Some IDs name an upstream (azure-, alicloud-, cc-), and some differ only in capitalisation, kept so older integrations keep working.

The catalog does not carry a field saying which of those a given row is, so this page does not sort them into buckets it would have to invent. Every row shows that route’s own price, context and speed — compare those directly, and open a model to see the upstreams that serve it.

Do I need a separate intern-ai account?

No. One AIHubMix key covers every model on this page, and switching between them is a change to the model string — billing, rate limits, and logs stay in one place.

Start calling intern-ai in one line

One key, one endpoint, 908 models across 41 model authors.