Skip to main content
Kaino.dev
Discover
Evals
News
Academics
Insights
Kaino.dev

Discover, evaluate, and compare AI tools, models, and agents.

Explore

  • Discover
  • Evaluations
  • News
  • Academics
  • Insights

Community

  • Twitter
  • YouTube
  • Instagram
Privacy PolicyTerms of Service

© 2026 Kaino.dev. All rights reserved.

Version 1.1.0
Amazon Nova 2 Lite logo
model evaluation

Amazon Nova 2 Lite

Amazon

Cost-efficient multimodal reasoning model on Amazon Bedrock for everyday automation, document processing, customer support, and agentic AI applications.

modelsource:aws.amazon.comamazonawsamazon-novaamazon-bedrockmultimodalreasoning
78.6KAINO SCORERecommended
Evaluated Jul 31, 202610 reviews
Website Docs

Scorecard

PricingMultimodalCostDev expTechnicalSpeedCodingReasoningRiskAdoption
  • Cost effectiveness88
  • Multimodal & I/O84
  • Speed & availability84
  • Developer experience83
  • Technical capability80
  • Coding & agentic77
  • Reasoning & knowledge77
  • Risk & evidence74
  • Pricing clarity72
  • Adoption signal67

Kainotomic evaluation

Nova 2 Lite is a capable cost-oriented multimodal reasoning model rather than a frontier general-purpose leader. Amazon reports multimodal input, optional extended thinking, and a 1M-token context window; Artificial Analysis lists 145.5 output tokens/s, 1.23s TTFT, and $0.30/$2.50 per million input/output tokens. This supports stronger cost and speed scores than Grok 3, while its documented Bedrock integration supports a developer experience near Kimi K2.5. Coding evidence is material but qualified: Amazon’s report gives 53.6% SWE-Bench Verified under a standard setup and 64.5% with inference-time scaling plus internal agent scaffolding, alongside 71.0 on a specified LiveCodeBench v5 subset. That places it above low-evidence or weaker coding options such as Hermes 4.3 36B, but below Kimi K2.5, Grok 3, GPT-5.5, and Claude Opus 4.8, which have substantially higher calibrated coding/agentic scores. Its broad multimodal and long-context positioning is credible, but does not match Gemini 3.5 Flash’s stronger multimodal anchor. Pricing clarity is limited because the supplied official material does not provide a price table; the price figures come from Artificial Analysis. Public adoption and preference evidence is also immature: neither DeepSWE nor the checked Arena-Hard leaderboard lists Nova 2 Lite, and no Terminal-Bench/Aider result was supplied. Scores therefore credit official documentation, Amazon-reported benchmark detail, and independent performance/price tracking, while discounting vendor-reported results and unavailable external replications.

Strengths

  • Low reported token pricing and fast Artificial Analysis measurements.
  • Multimodal inputs, optional extended thinking, and 1M-token context.
  • Documented SWE-Bench Verified and LiveCodeBench results.
  • Bedrock availability is practical for AWS-native deployment.

Caveats

  • The higher 64.5% SWE-Bench Verified result depends on inference-time scaling and internal agentic scaffolding.
  • Official supplied sources do not establish transparent list pricing.
  • No public Arena-Hard or DeepSWE listing was found.
  • No supplied Terminal-Bench or Aider result supports workflow-agent performance.
Amazon Nova 2 Lite · Evaluations · Kaino