Inclusionai: Ling 3.0 Flash VL

by Inclusionai intelligent coding reasoning agentic open-weight multimodal image-input video-input long-context

Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual... 124B MoE with 5.5B active parameters Native image and video understanding 262K-token context window Supports reasoning and tool-use parameters Free-tier availability may have rate limits No published benchmark record in the catalog yet Video and image analysis Long-context document work Multimodal coding and research

Choose a model to compare against Ling 3.0 Flash VL

Specifications

Specifications for Ling 3.0 Flash VL
LabInclusionai
Context window131,072
Input price $0.06/1M
Output price $0.18/1M
Release2026-09-01 00:00:00

Strengths

  • 124B MoE with 5.5B active parameters
  • Native image and video understanding
  • 262K-token context window
  • Supports reasoning and tool-use parameters

Weaknesses

  • Free-tier availability may have rate limits
  • No published benchmark record in the catalog yet

Best for

  • Video and image analysis
  • Long-context document work
  • Multimodal coding and research

In Depth: Ling 3.0 Flash VL

Summary

Ling 3.0 Flash VL is an AI model from Inclusionai.

Released 2026-09-01 00:00:00. It supports Text, image, video input and produces Text output, with a context window of 131.1K tokens. Input pricing is $0.06 per 1M tokens and output is $0.18 per 1M tokens on OpenRouter.

Sources & Further Reading

Related Models