ModelsIBM Granite 4.0 Nano 1.5B
IBM Granite 4.0 Nano 1.5B
by IBM
Local/Edge
Pricing
Input
Self-hosted
per 1M tokens
Output
Self-hosted
per 1M tokens
Cached
N/A
per 1M tokens
Note: Released Oct 2025. Mamba-2/transformer hybrid, Apache 2.0
Context & Output
Context Window128K tokens
Max Output32K tokens
Latency
Fast
Capabilities
Multimodal
Streaming
Function Calling
Prompt Caching
Key Strengths
What makes this model stand out
Runs in browser
70% memory reduction
ISO 42001
Similar Models in Local/Edge Tier
Other models with similar pricing and performance characteristics
IBM Granite 4.0 Nano 350M
IBM
Input:Self-hosted/M
Context:128K tokens
Phi-4
Microsoft
Input:Self-hosted/M
Context:16K tokens
Phi-4-multimodal
Microsoft
Input:Self-hosted/M
Context:16K tokens