MODELS / INCLUSIONAI

Ling 3.0 Flash VL

Ling 3.0 Flash VL builds on Ling 3.0 Flash with stronger language capabilities, native visual perception, and visual agent capabilities. It supports text, image, and video inputs with text output, reasoning, and function calling.

TEXTREASONINGREASONINGTOOL-USEVIDEO-INPUTVISION
256K
CONTEXT WINDOW
32K
MAX OUTPUT TOKENS
$0.075
INPUT PRICE / 1M TOKENS
$0.015
OUTPUT PRICE / 1M TOKENS
QUICKSTART

Call it in one request.

OpenAI-compatible: change the base URL, and the setup is complete.

curl https://enterprise.blackbox.ai/v1/chat/completions \
  -H "Authorization: Bearer $BLACKBOX_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "blackboxai/inclusionai/ling-3.0-flash-vl",
    "messages": [{ "role": "user", "content": "Hello!" }],
    "stream": true
  }'
SPECIFICATIONS

The full spec sheet.

Context, pricing, modalities, routing, and deployment: everything Ling 3.0 Flash VL supports, in one table.

MODEL ID
blackboxai/inclusionai/ling-3.0-flash-vl
DEVELOPED BY
Inclusionai
MODEL TYPE
Text
RELEASED
Sep 8, 2026
CONTEXT WINDOW
256K
MAX OUTPUT TOKENS
32K
INPUT MODALITIES
text, image, video
OUTPUT MODALITIES
text
INPUT PRICE
$0.075 / 1M tokens
OUTPUT PRICE
$0.015 / 1M tokens
CACHE READ
$0.22 / 1M tokens
REASONING
Supported
DATA RETENTION
Standard retention
DEDICATED DEPLOYMENT
Not available (closed-weight model)
SUPPORTED PARAMETERS
max_tokenstemperaturestoptoolstool_choicereasoninginclude_reasoning