fishaudio_s1-mini
fishaudio · View on Hugging Face ↗
Get this model
Seeders: 1 · Leechers: 0
Observed 2026-09-02T13:56:39Z via announce.aitorrent.org:7070.
Model card
The complete upstream card, rendered from this payload's README.md — the same hash-verified bytes the torrent distributes. Images and off-site links are removed; the original card on Hugging Face carries them.
tags:
- text-to-speech license: cc-by-nc-sa-4.0 language:
- zh
- en
- de
- ja
- fr
- es
- ko
- ar
- nl
- ru
- it
- pl
- pt pipeline_tag: text-to-speech inference: false extra_gated_prompt: >- You agree to not use the model to generate contents that violate DMCA or local laws. extra_gated_fields: Country: country Specific date: date_picker I agree to use this model for non-commercial use ONLY: checkbox
FishAudio S1
FishAudio S1 is a leading text-to-speech (TTS) model trained on more than 2 million hours of audio data in multiple languages.
Supported languages:
- English (en)
- Chinese (zh)
- Japanese (ja)
- German (de)
- French (fr)
- Spanish (es)
- Korean (ko)
- Arabic (ar)
- Russian (ru)
- Dutch (nl)
- Italian (it)
- Polish (pl)
- Portuguese (pt)
Please refer to Fish Speech Github for more info. Demo available at Fish Audio Playground. Visit the Fish Audio website for blog & tech report.
Emotion and Tone Support
FishAudio S1 supports a variety of emotional, tone, and special markers to enhance speech synthesis:
1. Emotional markers: (angry) (sad) (disdainful) (excited) (surprised) (satisfied) (unhappy) (anxious) (hysterical) (delighted) (scared) (worried) (indifferent) (upset) (impatient) (nervous) (guilty) (scornful) (frustrated) (depressed) (panicked) (furious) (empathetic) (embarrassed) (reluctant) (disgusted) (keen) (moved) (proud) (relaxed) (grateful) (confident) (interested) (curious) (confused) (joyful) (disapproving) (negative) (denying) (astonished) (serious) (sarcastic) (conciliative) (comforting) (sincere) (sneering) (hesitating) (yielding) (painful) (awkward) (amused)
2. Tone markers: (in a hurry tone) (shouting) (screaming) (whispering) (soft tone)
3. Special markers: (laughing) (chuckling) (sobbing) (crying loudly) (sighing) (panting) (groaning) (crowd laughing) (background laughter) (audience laughing)
Special markers with corresponding onomatopoeia:
- Laughing: Ha,ha,ha
- Chuckling: Hmm,hmm
Model Variants and Performance
FishAudio S1 includes the following models:
- S1 (4B, proprietary): The full-sized model.
- S1-mini (0.5B): A distilled version of S1.
Both S1 and S1-mini incorporate online Reinforcement Learning from Human Feedback (RLHF).
Seed TTS Eval Metrics (English, auto eval, based on OpenAI gpt-4o-transcribe, speaker distance using Revai/pyannote-wespeaker-voxceleb-resnet34-LM):
- S1:
- WER (Word Error Rate): 0.008
- CER (Character Error Rate): 0.004
- Distance: 0.332
- S1-mini:
- WER (Word Error Rate): 0.011
- CER (Character Error Rate): 0.005
- Distance: 0.380
License
This model is permissively licensed under the CC-BY-NC-SA-4.0 license.
Magnet link
Opens the swarm directly in your torrent client — no file download needed. Copy-paste works too:
magnet:?xt=urn:btih:2be65499e8cfc3da5f0e3da6bcefdef73a401efd&dn=fishaudio_s1-miniOpen magnet in torrent client · infohash 2be65499e8cfc3da5f0e3da6bcefdef73a401efd
Files & hashes
| Path | Size | sha1 | sha256 |
|---|---|---|---|
| README.md | 2.8 KB (2,858 B) | 4d8dd883137ac1c72bcb88eaa5bce321864ad2aa | 3b9a9399924a718b9ded7869ef58565679ff9c8ac436635886a43a72f3efb89a |
| codec.pth | 1.74 GB (1,871,099,728 B) | 776e95b0c9ddd89e11a5a356910ed2f200300bec | 74fc41c5a7151c6f350af8bd7e5d6e3accfcc7f3dfbfac23afd35af07052bb2f |
| config.json | 844 B (844 B) | 53320aac97d86ee9a5017e8b321024c879c7d9c7 | 0e46c32fd452751d48a553283e16a30819ea7321b32b9bf9e4a8b9c9cc02fa50 |
| model.pth | 1.62 GB (1,735,122,974 B) | 11c8252f4dabfc9fcce22cfc5af54a6c09f1af61 | 9e59be7dc6714040dce3cde1f41e730c2f0daa5339785b1cd3b60041208c35e6 |
| special_tokens.json | 123.3 KB (126,275 B) | 954449504ac94d1a89b31aa1bc914f07f898b6bb | efc4b254afdf5c3898dc26bd2b7791a9cae993747d92a3176cf9ce4ccfe40526 |
| tokenizer.tiktoken | 2.4 MB (2,561,218 B) | 9b9b0e0416d84d7c88333eb261c77e5fe2d7f7be | b2b1b8dfb5cc5f024bafc373121c6aba3f66f9a5a0269e243470a1de16a33186 |
Cite this release
- Canonical URL
- https://aiseedbank.org/models/fishaudio_s1-mini/
- Slug
- fishaudio_s1-mini
- Infohash
- 2be65499e8cfc3da5f0e3da6bcefdef73a401efd
- License
- cc-by-nc-sa-4.0
- Signing key fingerprint
- 85a3b32c3712427b
Every file carries a locally computed sha256 — verify a download against the signed sums: fishaudio_s1-mini.SHA256SUMS (+ minisign signature).
Provenance
| Upstream repository | fishaudio/s1-mini |
|---|---|
| Revision (pinned) | f4b445029346701e082b60bb63fcc2d1bb17a0e2 |
| Fetched at | 2026-09-02T13:40:50Z |
| License at fetch | cc-by-nc-sa-4.0 |
| Snapshot tool | huggingface · seedbank 0.1.0 |
Trackers
- udp://announce.aitorrent.org:6969/announce
- http://announce.aitorrent.org:7070/announce
- udp://announce2.aitorrent.org:6970/announce
- http://announce2.aitorrent.org:7071/announce
- udp://tracker.opentrackr.org:1337/announce
- udp://open.demonii.com:1337/announce
- udp://open.stealth.si:80/announce
- udp://exodus.desync.com:6969/announce
- udp://tracker.torrent.eu.org:451/announce
✓ verified · rehash-vs-hf-metadata at 2026-09-02T13:41:02Z
cc-by-nc-sa-4.0non-commercial use only3.36 GB (3,608,913,897 bytes)dual_artext-to-speech13 languages (zh, en, de …)