infery
← All models

Qwen3 VL 235B A22B Instruct

qwen3-vl-235b-a22b-instruct

Chat & Textby Qwen

Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual understanding across images and video. 262,144 token context window, maximum output of 16,384 tokens. Higher uptime with 5 providers. Includes independent benchmarks from Artificial Analysis.

chatstreamingvisiontoolsjson mode

Details

Accepts
text
Context window
262K tokens

Pricing

Input
26.25 cr / 1M tokens
Output
238 cr / 1M tokens
Cached input
12.5 cr / 1M tokens

Prices in credits (1 credit = $0.01).

Data schema

approximate

Input

FieldTypeDescription
promptrequiredstringThe text prompt.

Output

FieldTypeDescription
textstringGenerated text (choices[0].message.content).