Ling 3.0 Flash VL

Inclusionai🇨🇳 China
active
Context window131K tokens
Input / 1M tokens$0.06
Output / 1M tokens$0.18

Version History

3.0-flash-vlminor

Ling 3.0 Flash VL extends the text-only Ling 3.0 Flash MoE model with native image and video understanding, while InclusionAI claims further improvements to underlying language capabilities. The model retains the 124B total / 5.5B active parameter MoE architecture and 131K context window.

Benchmark Scores

Full leaderboard →
86.2%
GPQA
145.0 tokens_per_sec
Speed (tok/s)

Coverage

model releaseInclusionai

InclusionAI Releases Ling 3.0 Flash VL, Adding Vision to Its 124B MoE Model

InclusionAI has released Ling 3.0 Flash VL, a vision-language extension of its 124B total-parameter, 5.5B active Mixture-of-Experts model. The model adds native image and video understanding, supports a 131K token context window, and is priced at $0.06 per 1M input tokens and $0.18 per 1M output tokens via OpenRouter.

2 min read