Vision Task Picks
Models supporting image input for OCR, visual analysis, diagrams, and multimodal workflows.
Selection Rule: Models supporting image input (modality tags include vision or image), excluding removed models, sorted by output price ascending.
| Rank | Model | Provider | Context Window | Input / 1M | Output / 1M | Action |
|---|---|---|---|---|---|---|
| #1 | Dots3-Note Preview (free) | Dots Studio | 512,000 tokens | $0.00 | $0.00 | View Details |
| #2 | Free Models Router | OpenRouter | 200,000 tokens | $0.00 | $0.00 | View Details |
| #3 | Gemma 4 26B A4B (free) | 262,144 tokens | $0.00 | $0.00 | View Details | |
| #4 | Gemma 4 31B (free) | 262,144 tokens | $0.00 | $0.00 | View Details | |
| #5 | Inkling (free) | Thinking Machines | 1,048,576 tokens | $0.00 | $0.00 | View Details |
| #6 | Inkling Small (free) | Thinking Machines | 1,048,576 tokens | $0.00 | $0.00 | View Details |
| #7 | Lyria 3 Clip Preview | 1,048,576 tokens | $0.00 | $0.00 | View Details | |
| #8 | Lyria 3 Pro Preview | 1,048,576 tokens | $0.00 | $0.00 | View Details | |
| #9 | Nemotron 3 Nano Omni (free) | NVIDIA | 256,000 tokens | $0.00 | $0.00 | View Details |
| #10 | Nemotron 3.5 Content Safety (free) | NVIDIA | 128,000 tokens | $0.00 | $0.00 | View Details |