Gemini Nano Banana 2.1 is an upgraded multimodal reasoning image model for production-grade image generation and editing, with stronger prompt adherence, multi-reference editing, and improved visual fidelity. It understands text, images, video, and PDF inputs, supports interleaved text and image generation, and can reason over complex spatial and compositional instructions before generating images. The model also supports grounding with Google Search and Google Image Search for better long-tail entity recognition, visual understanding, and output accuracy.
Providers #
Route requests across multiple providers. Copy a provider slug to set your preference.
$1.5
/ M tokens $7.5
/ M tokens Read:
-/ M tokens
Write:
-/ M tokens128K12.4s90.8tps
Uptime #
24hours Direct request success rate on AI Gateway and per-provider.
Throughput #
24hours P50 throughput on live AI Gateway traffic, in tokens per second (TPS).
Latency #
24hours P50 time to first token (TTFT) on live AI Gateway traffic, in milliseconds.
Activity #
Token volume and request traffic to this model over time.
Benchmarks #
Scores on standardized evaluations. Higher percentages are better — and rank percentile shows
Metrics sourced fromArtificial Analysis
Apps #
Public apps that send the most traffic to this model. Good signal for what real production workloads look like — and a hint at which use cases this model is best suited for. View All
Related Models #
More models from Google