How does Ling 3.1 Flash rank among AI models?
Ling 3.1 Flash holds ModelCap Index position #75 of 280 ranked language models as of 2 October 2026, with an Index score of 59.8 (interval 46.2–73.5); evidence: modeled · succession. The Index combines public benchmark boards with published uncertainty; the methodology page explains the weighting.
How much does Ling 3.1 Flash cost per 1M tokens?
Ling 3.1 Flash is listed at Free per 1M input tokens and Free per 1M output tokens as of 2 October 2026. Prices come from the OpenRouter catalogue and re-observe on every refresh. For 1,000 requests using 2,000 input and 500 output tokens each, the listed rates imply Free total (2M input + 0.5M output). This excludes caching, batch discounts, prompt-length tiers, tool charges and retries; verify the chosen endpoint's rate and limits before budgeting.
What is the context window of Ling 3.1 Flash?
Ling 3.1 Flash has a published context window of 262K tokens (262,144) and a maximum output of 33K tokens. Individual providers can serve less than the published maximum; the providers table lists each endpoint's own limit.
Which API providers serve Ling 3.1 Flash?
One provider serves Ling 3.1 Flash through OpenRouter as of 2 October 2026: Novita. Each provider's price, context limit, quantization and measured uptime are in the providers table above.