LiteLLMModels

Nova Micro API Pricing

As of October 10, 2026, Nova Micro costs $0.035 per 1M input tokens and $0.14 per 1M output tokens from Amazon Bedrock, the lowest of 2 providers, with cache reads at $0.00875 per 1M tokens. It has a 128,000-token context window and returns up to 10,000 tokens per response.

By AmazonModel ID amazon.nova-micro-v1:0Prices updated Checked
Input
$0.035 / 1M tokens
Output
$0.14 / 1M tokens
Cache read
$0.00875 / 1M tokens
Context window
128,000 tokens
Max output
10,000 tokens
Providers
2
Added to LiteLLM
December 2024

Nova Micro pricing by provider

6 listings across 2 providers. Token prices are in US dollars per 1M tokens. The highlighted row is the price quoted above.

ProviderModel name in LiteLLMInputOutputCache readContext
Amazon Bedrockamazon.nova-micro-v1:0$0.035$0.14$0.00875128K
Vercel AI Gatewayvercel_ai_gateway/amazon/nova-micro$0.035$0.14–128K
Amazon BedrockAPAC cross-regionapac.amazon.nova-micro-v1:0$0.037$0.148$0.00925128K
Amazon Bedrockus-gov-west-1bedrock/us-gov-west-1/amazon.nova-micro-v1:0$0.042$0.168$0.0105128K
Amazon BedrockEU cross-regioneu.amazon.nova-micro-v1:0$0.046$0.184$0.0115128K
Amazon BedrockUS cross-regionus.amazon.nova-micro-v1:0$0.035$0.14$0.00875128K

What Nova Micro costs in practice

1,000 requests with 2,000 input and 500 output tokens each

Amazon Bedrock$0.14
Vercel AI Gateway$0.14

100 long-document requests with 100,000 input and 2,000 output tokens each

Amazon Bedrock$0.378
Vercel AI Gateway$0.378

Nova Micro features

Use Nova Micro with LiteLLM

Call Nova Micro through the LiteLLM Python SDK, or put it behind the LiteLLM proxy and send OpenAI-format requests to /v1/chat/completions. LiteLLM tracks the cost of every request at these prices.

Python SDK
from litellm import completion

response = completion(
    model="amazon.nova-micro-v1:0",
    messages=[{"role": "user", "content": "Hello!"}],
)
Proxy config.yaml
model_list:
  - model_name: nova-micro
    litellm_params:
      model: amazon.nova-micro-v1:0
      aws_region_name: os.environ/AWS_REGION_NAME
Call the proxy
curl http://0.0.0.0:4000/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer sk-1234" \
  -d '{
    "model": "nova-micro",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'

Nova Micro FAQ

How much does Nova Micro cost?

Nova Micro costs $0.035 per 1M input tokens and $0.14 per 1M output tokens from Amazon Bedrock, the lowest of 2 providers, with cache reads at $0.00875 per 1M tokens.

What is the context window of Nova Micro?

Nova Micro has a 128,000-token context window and returns up to 10,000 tokens per response.

Which providers offer Nova Micro?

Nova Micro is available from 2 providers through LiteLLM: Amazon Bedrock and Vercel AI Gateway.

How do I call Nova Micro with an OpenAI-compatible API?

Use the LiteLLM Python SDK with model="amazon.nova-micro-v1:0", or add the model to the LiteLLM proxy's config.yaml and send requests to /v1/chat/completions from any OpenAI client. LiteLLM tracks the cost of every request at the prices on this page.

What features does Nova Micro support?

Nova Micro supports function calling, structured output and prompt caching.

When did LiteLLM add Nova Micro?

Nova Micro was added to LiteLLM's model price file in December 2024.

Sources

Prices come from LiteLLM's open-source model_prices_and_context_window.json, which LiteLLM uses to calculate the cost of every request.

LiteLLM docs: Amazon Bedrock, Vercel AI Gateway.

Provider pages: Amazon Bedrock, Vercel AI Gateway.