Spotlight: Pricing, Context Window & Benchmarks
by Arcee AI
Spotlight is a 7‑billion‑parameter vision‑language model derived from Qwen 2.5‑VL and fine‑tuned by Arcee AI for tight image‑text grounding tasks. It offers a 32 k‑token context window, enabling rich multimodal conversations that combine lengthy documents with one or more images. Training emphasized fast inference on consumer GPUs while retaining strong captioning, visual‐question‑answering, and diagram‑analysis accuracy. As a result, Spotlight slots neatly into agent workflows where screenshots, charts or UI mock‑ups need to be interpreted on the fly. Early benchmarks show it matching or out‑scoring larger VLMs such as LLaVA‑1.6 13 B on popular VQA and POPE alignment tests.
What you can do with Spotlight
Everyday Q&A and clear explanations
Writing help (emails, posts, summaries)
Idea generation and brainstorming
Learning support with step-by-step guidance
Benchmarks not available
This model isn't listed on Artificial Analysis yet. Showing OpenRouter specs below.
| Metric | Value |
|---|---|
| Provider | Arcee AI |
| Context Window | 131,072 tokens |
| Input Price | $0.18/1M tokens |
| Output Price | $0.18/1M tokens |
| Release Date | May 5, 2025 |
| Modalities | image, text |
| Capabilities | Vision |
Compare Spotlight to other models
See how it stacks up on price, quality, and overall performance.
Frequently asked questions
What is Spotlight good for?
Use Spotlight for everyday tasks like writing, summarizing, brainstorming, and getting clear explanations.
How much does Spotlight cost?
Pricing is based on usage. Current rates are $0.18/1M tokens for input and $0.18/1M tokens for output.
Can I try Spotlight for free?
Yes. You can start a chat instantly and test the model before deciding on a plan.
Does Spotlight support images or audio?
Spotlight can understand images.
Similar models
Pricing, context, and capability data are sourced from OpenRouter.
Compare Models
Select a model to compare with Spotlight