Ling-3.0-flash-VL
Ling-3.0-flash-VL is an Ant Group open-weights reasoning model. It is not yet ranked: 4 independent results so far, and the index needs at least five across two categories. Output speed of 143 tokens per second puts it in the fastest quarter, with a first answer in 15.8 s.
- Blended price
- –
- Output speed
- 143 tok/s
- measured by Artificial Analysis
- First answer
- 16 s
- Context
- 262k
- Knowledge62
- Multimodal63
- Long context66
- Composite60
Benchmark results
One column per reasoning effort. Rank is among every configuration of every model on that benchmark. Hover a score for the run it came from.
Reasoning
| Benchmark | default | Source |
|---|---|---|
| GPQA Diamond (AA)not in index | 86.2%#113 | Artificial Analysis |
| Humanity's Last Exam (AA)not in index | 22.0%#153 | Artificial Analysis |
Coding
| Benchmark | default | Source |
|---|---|---|
| SciCode (AA)not in index | 44.2%#98 | Artificial Analysis |
Knowledge
| Benchmark | default | Source |
|---|---|---|
| AA-Omniscience | -4.5#116 | Artificial Analysis |
Multimodal
| Benchmark | default | Source |
|---|---|---|
| MMMU-Pro | 79.0%#47 | Artificial Analysis |
Long context
| Benchmark | default | Source |
|---|---|---|
| AA-LCR | 78.3%#78 | Artificial Analysis |
Composite
| Benchmark | default | Source |
|---|---|---|
| AA Intelligence Index | 24.8#123 | Artificial Analysis |
What a task costs
Estimates from list price, output speed and time to first answer for the best configuration. “With caching” assumes three-quarters of the input is served from the prompt cache. Reasoning tokens are not modelled.
| Workload | Tokens in / out | Cost | With caching | Time |
|---|---|---|---|---|
| Chat reply | 400 / 300 | – | – | 17.9 s |
| Summarise a 30-page report | 12,000 / 600 | – | – | 20.0 s |
| Code edit | 6,000 / 1,500 | – | – | 26.3 s |
| Agentic coding session | 60,000 / 4,000 | – | – | 43.8 s |
| Structured extraction | 2,000 / 200 | – | – | 17.2 s |
See also
Data as of 10 Sept 2026. Compare with another model.