Compare models
Put up to 4 models beside each other — token prices, context windows, capabilities and provider, from the same catalogue the model pages read.
- Transcribe 1 ProFish AudioRemove
- Voxtral Small 24B 2507 STTMistral AIRemove
- S1Fish AudioRemove
- Grok 4.7SpaceXAIRemove
4 is the maximum. Remove one to add another.
| Attribute | Transcribe 1 Protranscribe-1-pro | Voxtral Small 24B 2507 STTvoxtral-small-24b-2507-stt | S1s1 | Grok 4.7grok-4.7 |
|---|---|---|---|---|
| Pricing | ||||
| Input | $0 / 1M | $0 / 1M | $0 / 1M | $1.60 / 1M |
| Output | $0 / 1M | $0 / 1M | $0 / 1M | $4.80 / 1M |
| Cache Write (5m) | Not applicable | Not applicable | Not applicable | $1.60 / 1M |
| Cache Write (1h) | Not applicable | Not applicable | Not applicable | $1.60 / 1M |
| Cache Read | Not applicable | Not applicable | Not applicable | $1.60 / 1M |
| Web Search | $0 / 1M | $0 / 1M | $0 / 1M | $0 / 1M |
| Context | ||||
| Max context | N/A | N/A | N/A | 500K |
| Max output | N/A | N/A | N/A | N/A |
| Capabilities | ||||
| Vision | No | No | No | Yes |
| Function Calling | No | No | No | Yes |
| JSON Mode | No | No | No | Yes |
| Streaming | No | No | No | Yes |
| Catalogue | ||||
| Provider | Fish Audio | Mistral AI | Fish Audio | SpaceXAI |
| Category | voice | voice | voice | chat |
| Charge type | Pay As You Go | Pay As You Go | Pay As You Go | Pay As You Go |
| Released | 2026-09-24 | 2026-08-13 | 2026-07-29 | — |
| Description | ||||
| Summary | Fish Audio Transcribe 1 Pro speech-to-text with speaker labels; transcripts include speaker tags such as <|speaker:0|>. Billed per second of audio. | Mistral AI Voxtral Small (24B) speech-to-text. Billed per second of audio. | Fish Audio S1 text-to-speech. Billed per UTF-8 byte of input text. | Grok 4.7 is SpaceXAI's flagship model for coding, agentic workflows, and professional knowledge work, succeeding Grok 4.6. It is particularly strong at long-running software engineering, self-verification, and long-context execution, while improving capabilities in document drafting, presentations, and other professional tasks. Trained with extended reinforcement learning focused on multi-hour problems, Grok 4.7 is optimized for sustained, complex task execution and natively supports the Grok Bot harness for conversational workflows. It also introduces an enhanced safeguard stack designed to combine strong jailbreak resistance with low refusal rates for legitimate technical work. Reported benchmark results use xhigh reasoning effort. |