AI SeedbankHelp preserve open and free AI for humanity's future

← All models

MIT_ast-finetuned-audioset-10-10-0.4593

MIT · View on Hugging Face ↗

Get this model

Download TorrentMagnet Link

Seeders: · Leechers:

Model card

The complete upstream card, rendered from this payload's README.md — the same hash-verified bytes the torrent distributes. Images and off-site links are removed; the original card on Hugging Face carries them.


license: bsd-3-clause tags:

  • audio-classification

Audio Spectrogram Transformer (fine-tuned on AudioSet)

Audio Spectrogram Transformer (AST) model fine-tuned on AudioSet. It was introduced in the paper AST: Audio Spectrogram Transformer by Gong et al. and first released in this repository.

Disclaimer: The team releasing Audio Spectrogram Transformer did not write a model card for this model so this model card has been written by the Hugging Face team.

Model description

The Audio Spectrogram Transformer is equivalent to ViT, but applied on audio. Audio is first turned into an image (as a spectrogram), after which a Vision Transformer is applied. The model gets state-of-the-art results on several audio classification benchmarks.

Usage

You can use the raw model for classifying audio into one of the AudioSet classes. See the documentation for more info.

Magnet link

Opens the swarm directly in your torrent client — no file download needed. Copy-paste works too:

magnet:?xt=urn:btih:73fd5baba8460b356d350e01d47cd8d94a2d1978&dn=MIT_ast-finetuned-audioset-10-10-0.4593

Open magnet in torrent client · infohash 73fd5baba8460b356d350e01d47cd8d94a2d1978

Files & hashes

PathSizesha1sha256
README.md1.1 KB (1,165 B)fc4ece98b629cab4731db057bebdc4240fda9239dbc8ce1fc5abd1635d073640d8cb032901bfe264a5e06507f363f10a6112cfc8
config.json26.1 KB (26,763 B)9b822c630bc1a61312ac702cca4d8fb5a524e729a93d525511d77e8ecc933d09674b85099815bbbb417c228a4edd655e252fb9ff
model.safetensors330.4 MB (346,404,948 B)e37828c3e66ca97046aa500f342570dfe3879463ae0c1e2ad4e1381d851fa9bf298ba13ebc9c5a914cdee2dbe427a6583869924d
preprocessor_config.json297 B (297 B)d93e9e561c374604096fd5868c97e08130638f728d04ba5a9c6fca5d39d0de2b1fd05ecf79deb589fbba279728bbebac39934231
pytorch_model.bin330.4 MB (346,445,675 B)c729af4d47777d87f6614c90df3cabafb12686c59ca280ba0276a0d2243f7dd561c72b10cc7c3955a2b1f12e42e93983f24b8ca4

Cite this release

Canonical URL
https://aiseedbank.org/models/MIT_ast-finetuned-audioset-10-10-0.4593/
Slug
MIT_ast-finetuned-audioset-10-10-0.4593
Infohash
73fd5baba8460b356d350e01d47cd8d94a2d1978
License
bsd-3-clause
Signing key fingerprint
85a3b32c3712427b

Every file carries a locally computed sha256 — verify a download against the signed sums: MIT_ast-finetuned-audioset-10-10-0.4593.SHA256SUMS (+ minisign signature).

Provenance

Upstream repositoryMIT/ast-finetuned-audioset-10-10-0.4593
Revision (pinned)f826b80d28226b62986cc218e5cec390b1096902
Fetched at2026-09-03T19:02:59Z
License at fetchbsd-3-clause
Snapshot toolhuggingface · seedbank 0.1.0

Trackers

✓ verified · rehash-vs-hf-metadata at 2026-09-03T19:03:07Z

bsd-3-clause660.8 MB (692,878,848 bytes)transformerspytorchsafetensorsaudio-spectrogram-transformeraudio-classificationendpoints_compatiblepaper: 2104.01778