Skip to content
  • Models
  • Rankings
  • Ori
Sign Up
Sign Up
OpenRouterOpenRouter
© 2026 OpenRouter, Inc

Product

  • Chat
  • Rankings
  • Benchmarks
  • Apps
  • Discover
  • Models
  • Collections
  • Providers
  • Tools
  • Pricing
  • Business
  • Enterprise
  • Labs

Company

  • About
  • Blog
  • Careers
    Hiring
  • Privacy
  • Terms of Service
  • Trust Center
  • Support
  • Works With OR
  • Data
  • Brand

Developer

  • Documentation
  • API Reference
  • Developer Platform
  • Status
  • AI Site Map

Connect

  • Discord
  • GitHub
  • LinkedIn
  • X
  • YouTube
Favicon for fish-audio

Fish Audio

Access 6 Fish Audio models through the OpenRouter unified API including Transcribe 1 Pro, Transcribe 1, and S1. Compare pricing, context windows, benchmarks, and capabilities between different Fish Audio models.

Fish Audio tokens processed on OpenRouter

  • Favicon for fish-audio
    Fish Audio: Transcribe 1 ProTranscribe 1 Pro
    2.45M characters

    Transcribe 1 Pro is a speech-to-text model from Fish Audio tuned for interviews, meetings, and podcasts. It labels speakers with inline speaker markers, preserves emotion and vocal-event cues such as [laughter], detects language automatically, and can return timestamped word-level segments.

    by fish-audioSep 24, 2026$0.0001/second
  • Favicon for fish-audio
    Fish Audio: Transcribe 1Transcribe 1
    4.11M characters

    Transcribe 1 is a speech-to-text model from Fish Audio. It is suited for audio transcription with automatic language detection and can return timestamped word-level segments when alignment details are requested.

    by fish-audioJul 29, 2026$0.0001/second
  • Favicon for fish-audio
    Fish Audio: S1S1
    55K tokens

    S1 is a multilingual text-to-speech model from Fish Audio. It is suited for voice applications that need broad emotional expression, using parenthetical controls to guide speaking style across its supported languages.

    by fish-audioJul 29, 2026$15/M UTF-8 bytes
  • Favicon for fish-audio
    Fish Audio: S2 ProS2 Pro
    128K tokens

    S2 Pro is a multilingual text-to-speech model from Fish Audio. It is suited for expressive narration and multi-speaker dialogue, with natural-language controls for speaking style and emotion.

    by fish-audioJul 29, 2026$15/M UTF-8 bytes
  • Favicon for fish-audio
    Fish Audio: S2.1 Pro Free (free)S2.1 Pro Free (free)Free variant
    11M tokens

    S2.1 Pro Free is the no-cost variant of Fish Audio S2.1 Pro, intended for testing, prototyping, and low-volume applications. It provides the same synthesis capabilities without production latency or availability guarantees.

    by fish-audioJul 29, 2026$0/M input tokens$0/M output tokens
  • Favicon for fish-audio
    Fish Audio: S2.1 ProS2.1 Pro
    5.19M tokens

    S2.1 Pro is a production-oriented text-to-speech model from Fish Audio. It is suited for multilingual voice applications, expressive narration, and dialogue synthesis, with open-ended natural-language controls for speaking style and emotion.

    by fish-audioJul 29, 2026$15/M UTF-8 bytes

Frequently asked questions

OpenRouter serves 6 Fish Audio models behind one OpenAI-compatible API. Create an OpenRouter API key, point your client at https://openrouter.ai/api/v1, and set the model to an ID such as fish-audio/transcribe-1-pro. The quickstart has request examples for every supported SDK.

S2.1 Pro Free (free) is available at no cost through the OpenRouter API.

Transcribe 1 Pro is the most recently added Fish Audio model on OpenRouter, listed on September 24, 2026.