Ling 3.0 Flash VL

by InclusionAI Intelligent Coding Reasoning Agentic Open Weight Multimodal Image Input Video Input Long Context

Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), strengthening its language capabilities and adding native visual perception.

Choose a model to compare against Ling 3.0 Flash VL

Specifications

Specifications for Ling 3.0 Flash VL
LabInclusionAI
Context window262.1K tokens
Input price $0.02/1M
Output price $0.06/1M
ReleaseSep 2026

Strengths

  • 124B MoE with 5.5B active parameters
  • Native image and video understanding
  • 262K-token context window
  • Supports reasoning and tool-use parameters

Weaknesses

  • Free-tier availability may have rate limits
  • No published benchmark record in the catalog yet

Best for

  • Video and image analysis
  • Long-context document work
  • Multimodal coding and research

In Depth: Ling 3.0 Flash VL

Summary

Ling 3.0 Flash VL is an AI model from InclusionAI.

Released in Sep 2026. It takes text, image, and video input and produces text output, with a context window of 262.1K tokens. Input pricing is $0.02 per 1M tokens and output is $0.06 per 1M tokens on OpenRouter.

Sources & Further Reading

Related Models