Skip to content

Compare models

Put up to 4 models beside each other — token prices, context windows, capabilities and provider, from the same catalogue the model pages read.

  1. Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)GoogleRemove
  2. GPT-6 Sol ProOpenAIRemove
  3. Perceptron Mk1.5PerceptronRemove
gemini-3.1-flash-lite-image vs gpt-6-sol-pro vs perceptron-mk1.5
AttributeNano Banana 2 Lite (Gemini 3.1 Flash Lite Image)gemini-3.1-flash-lite-imageGPT-6 Sol Progpt-6-sol-proPerceptron Mk1.5perceptron-mk1.5
Pricing
Request$0.7754 / request——
BillingPay Per Request——
Cache Write (5m)Not applicable$2.00 / 1M$0.15 / 1M
Cache Write (1h)Not applicable$2.00 / 1M$0.15 / 1M
Cache ReadNot applicable$2.00 / 1M$0.15 / 1M
Input—$2.00 / 1M$0.15 / 1M
Output—$10.00 / 1M$1.50 / 1M
Web Search—$0 / 1M$0 / 1M
Context
Max context66K1.1M36.9K
Max outputN/AN/AN/A
Capabilities
VisionYesYesNo
Function CallingYesYesNo
JSON ModeYesYesNo
StreamingNoYesYes
Catalogue
ProviderGoogleOpenAIPerceptron
Categoryimagechatchat
Charge typePay Per RequestPay As You GoPay As You Go
Released——2026-09-25
Description
SummaryNano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is Google's fastest and most cost-efficient multimodal image generation model, designed for high-throughput visual workflows and real-time applications. It supports text-to-image generation, image editing, and multi-image composition through a unified API, while also producing text outputs alongside images. Delivering image generation in approximately 4 seconds, it combines fast inference with strong character consistency, precise editing, and real-world knowledge. The model generates 1K-resolution images across 14 aspect ratios and embeds an invisible SynthID watermark in all outputs. Optimized for the best balance of quality, speed, and cost, Nano Banana 2 Lite is ideal for prototyping, developer pipelines, and large-scale visual content generation.GPT-6 Sol Pro uses the same underlying model as GPT-6 Sol, but runs with reasoning.mode set to pro for higher-quality responses on complex and demanding tasks. Optimized for deeper reasoning, greater accuracy, and more reliable multi-step execution, it is particularly well suited for agentic coding, long-horizon software engineering, professional analysis, and complex automated workflows where solution quality takes priority over latency and cost.Perceptron Mk1.5 chat model that accepts audio input. Audio and text input are billed per token.