Ling 3.0 Flash VL
ling
Native multimodal hybrid MoE model for image and video understanding, reasoning, and visual agents
ReasoningTool CallingAttachmentsOpen WeightsMIT
Context Window
262K
Max Output
33K
Temperature
Yes
Open Weights
Yes
Knowledge Cutoff
N/A
Released
2026-09-10
Last Updated
2026-09-10
License
MIT
Modalities
Input:TextImageVideo
→Output:Text
Available Providers (5)
| Provider | Input /1M | Output /1M | Cache Read /1M | Cache Write /1M | Reasoning | Status |
|---|---|---|---|---|---|---|
| Rp 1.073 | Rp 3.219 | Rp 215 | — | effort | — | |
| Rp 1.342 | Rp 3.934 | Rp 269 | — | effort | — | |
| Rp 1.342 | Rp 3.934 | Rp 269 | — | toggle | — | |
| Rp 376 | Rp 1.102 | Rp 76 | — | toggle | — | |
| Rp 1.342 | Rp 3.934 | Rp 269 | — | toggle | — |