← All models
Perceptron Mk1
Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning.** It accepts image and video inputs paired with natural language queries, and produces detailed visual understanding responses, either structured or natural language. 32,768 token context window, maximum output of 8,192 tokens.
chatstreamingvision
Details
- Accepts
- text
- Context window
- 33K tokens
Pricing
- Input
- 18.75 cr / 1M tokens
- Output
- 188 cr / 1M tokens
Prices in credits (1 credit = $0.01).
Data schema
approximateInput
| Field | Type | Description |
|---|---|---|
| promptrequired | string | The text prompt. |
Output
| Field | Type | Description |
|---|---|---|
| text | string | Generated text (choices[0].message.content). |