Nova Lite v1 API Benchmarks, Pricing & Provider Data
Compare Nova Lite v1 with another model
Choose a model to open its comparison page.
Amazon Nova Lite 1.0 is a very low-cost multimodal model from Amazon that focused on fast processing of image, video, and text inputs to generate text output. Amazon Nova Lite...
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Возможности
Эндпоинты OpenRouter
2 эндпоинтовСторонние данные эндпоинтов OpenRouter показаны отдельно от измерений LMSpeed. Некоторые 30-минутные live-метрики появляются только после синхронизации с OpenRouter API key.
| Provider endpoint | Вход | Вывод | Доступность 1 день | Задержка 30м | Пропускная способность 30м | Контекст / вывод |
|---|---|---|---|---|---|---|
Amazon Bedrock amazon-bedrock | $0.060/M | $0.240/M | 100.0% | — | — | undefined токенов / undefined токенов |
Amazon Bedrock amazon-bedrock/eu-west-1 | $0.060/M | $0.240/M | 84.1% | — | — | undefined токенов / undefined токенов |
Альтернативы и похожие модели
GPT-5.4
gpt-5-4
OpenAI GPT-5.4 extends the GPT-5 family with stronger instruction following, deeper tool use, and improved performance on coding, math, and long-document analysis.
GPT-5.3 Codex
gpt-5-3-codex
OpenAI GPT-5.3 Codex is a code-specialized variant in the GPT-5 series, optimized for code generation, debugging, and software development tasks.
GPT-5.2
gpt-5-2
OpenAI GPT-5.2 is a GPT-5 series model emphasizing advanced reasoning, multimodal understanding, and high-quality outputs for complex enterprise workloads.
GPT-5.4 Mini
gpt-5-4-mini
OpenAI GPT-5.4 Mini is a compact language model in the GPT-5 series, optimized for quick responses and high throughput.
Claude Opus 4.6
claude-opus-4-6
Anthropic Claude Opus 4.6 is the most capable Claude Opus tier, optimized for complex analysis, long-horizon coding, and high-stakes enterprise reasoning workloads.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
