Skip to content

Compare models

Put up to 4 models beside each other — token prices, context windows, capabilities and provider, from the same catalogue the model pages read.

  1. Universal-3.5 ProAssemblyAIRemove
  2. Gemma 4 31B (Free)GoogleRemove
  3. Perceptron Mk1.5PerceptronRemove
universal-3-5-pro vs gemma-4-31b-it:free vs perceptron-mk1.5
AttributeUniversal-3.5 Prouniversal-3-5-proGemma 4 31B (Free)gemma-4-31b-it:freePerceptron Mk1.5perceptron-mk1.5
Pricing
Input— Not priced per input token$0 / 1M$0.15 / 1M
Output— Not priced per output token$0 / 1M$1.50 / 1M
Cache Write (5m)Not applicable—$0.15 / 1M
Cache Write (1h)Not applicable—$0.15 / 1M
Cache ReadNot applicable$0 / 1M$0.15 / 1M
Web Search$0 / 1M—$0 / 1M
Cache Write—$0 / 1M—
Context
Max contextN/A262.1K36.9K
Max outputN/AN/AN/A
Capabilities
VisionNoYesNo
Function CallingNoYesNo
JSON ModeNoYesNo
StreamingNoYesYes
Catalogue
ProviderAssemblyAIGooglePerceptron
Categoryvoicechatchat
Charge typePay As You GoFreePay As You Go
Released2026-09-22—2026-09-25
Description
SummaryAssemblyAI Universal-3.5 Pro speech-to-text. Billed per second of audio.Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model, supporting text and image inputs with text outputs. It features a 256K token context window, configurable thinking/reasoning modes, native function calling, and broad multilingual support across 140+ languages. The model delivers strong performance in coding, reasoning, and document understanding, making it well suited for developer workflows, multilingual applications, and structured knowledge tasks.Perceptron Mk1.5 chat model that accepts audio input. Audio and text input are billed per token.