P
Compare
Perceptron Mk1
Perceptron · Multimodal · Released May 2026
A vision-language model from Perceptron designed to process images and videos alongside natural language queries for visual understanding tasks.
- Strengths
- Handles multimodal inputs including video and still images, enabling detailed analysis of visual content and spatial reasoning across temporal sequences.
- Best for
- Applications requiring video analysis, embodied reasoning, or detailed visual understanding from image and video inputs paired with natural language instructions.
- Limitations
- Optimized for vision and embodied reasoning; may not match specialized text-only models on pure language tasks unrelated to visual content.
Input / 1M
$0.15
Output / 1M
$1.50
Cached input / 1M
—
Context window
32K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $0.15 | $1.50 | — | Imported from OpenRouter | openrouter.ai |