intern-ai Models

2 modelsGeneral models free to startUp to 262K context

All 2 intern-ai Models

Open in model list
intern-ai models on AIHubMix with input and output modalities, context length, maximum output, price per million tokens including cache read and cache write rates, and measured throughput and latency.
Modalities
agents-a1-freeTakes text, vision, returns text.262KFreeFree/M
intern-s2-freeTakes text, vision, returns text.262KFreeFree/M

Prices are USD per million tokens; cache read and cache write are the rates for prompt-cache hits and for writing a prompt into the cache. Throughput and latency are measured on AIHubMix — the same figures the model detail page shows — not vendor claims. A dash means the catalog does not publish that field for that model, which is not the same as the model not supporting it.

intern-ai on AIHubMix

Which intern-ai model should I start with?

agents-a1-free is free on input — the cheapest entry here that declares a token price, and it carries a 262K context. Move up to intern-s2-free when answer quality matters more than cost.

Why are there several entries for the same model?

Because each row is a route you can call, not a model release. Some IDs name an upstream (azure-, alicloud-, cc-), and some differ only in capitalisation, kept so older integrations keep working.

The catalog does not carry a field saying which of those a given row is, so this page does not sort them into buckets it would have to invent. Every row shows that route’s own price, context and speed — compare those directly, and open a model to see the upstreams that serve it.

Do I need a separate intern-ai account?

No. One AIHubMix key covers every model on this page, and switching between them is a change to the model string — billing, rate limits, and logs stay in one place.

Start calling intern-ai in one line

One key, one endpoint, 883 models across 40 model authors.