AI SeedbankHelp preserve open and free AI for humanity's future

← All models

HauhauCS_Gemma4-31B-QAT-Uncensored-HauhauCS-Balanced-MTP

HauhauCS · View on Hugging Face ↗

Get this model

Download TorrentMagnet Link

Seeders: 1 · Leechers: 0

Observed 2026-09-01T16:02:52Z via announce.aitorrent.org:7070.

Model card

The complete upstream card, rendered from this payload's README.md — the same hash-verified bytes the torrent distributes. Images and off-site links are removed; the original card on Hugging Face carries them.


license: gemma tags:

  • uncensored
  • gemma4
  • gguf
  • vision
  • multimodal
  • agentic
  • coding
  • creative-writing
  • roleplay
  • rp
  • conversational language:
  • en pipeline_tag: image-text-to-text base_model: google/gemma-4-31B-it

Gemma4-31B-QAT-Uncensored-HauhauCS-Balanced-MTP

Join the Discord for updates, roadmaps, projects, or just to chat.

Gemma4-31B (QAT) uncensored by HauhauCS. 0/465 Refusals*

About

No changes to datasets or capabilities — fully functional, 100% of what the original authors intended, just without the refusals. Built from the official QAT weights, so the 4-bit quant stays close to full-precision quality.

Balanced

The Balanced variant (recommended — 99%+ of users will be happy here) uses optimized full uncensoring tuned especially for agentic coding, reasoning, creative writing and reliability-critical tasks. It reasons before answering and stays dependable and on-instruction. An Aggressive variant, for cases where Balanced still deflects too much, after current testing is not required.

~53% faster with MTP

Ships with an MTP (multi-token-prediction) draft head for speculative decoding — roughly 53% faster generation with identical output (the model verifies every drafted token, so quality is unchanged — pure speed). This release is tuned to pair well with the included MTP head.

llama.cpp:

llama-server \
  -m Gemma4-31B-QAT-Uncensored-HauhauCS-Balanced-Q4_K_M.gguf \
  -md mtp-gemma-4-31B-it.gguf --spec-type draft-mtp \
  -ngl 99 -fa on

Note: the MTP speedup was currently tested by me through llama.cpp (llama-server / llama-cli).

Downloads

File Type Size
Gemma4-31B-QAT-Uncensored-HauhauCS-Balanced-Q4_K_M.gguf Q4_K_M (text) 18.7 GB
mmproj-Gemma4-31B-QAT-Uncensored-HauhauCS-Balanced-BF16.gguf mmproj (vision) 1.2 GB
mtp-gemma-4-31B-it.gguf MTP speculative drafter 280 MB

Why only Q4_K_M? Gemma 4 is quantization-aware-trained for ~4-bit, so Q4_K_M is the sweet spot — higher-precision quants add size with no real quality gain. Carefully quantized for best quality at 4-bit.

Vision

Load the mmproj alongside the model for image input:

llama-server -m Gemma4-31B-QAT-Uncensored-HauhauCS-Balanced-Q4_K_M.gguf \
  --mmproj mmproj-Gemma4-31B-QAT-Uncensored-HauhauCS-Balanced-BF16.gguf -ngl 99 -fa on

Recommended sampling

These are dialed in specifically for this HauhauCS build — use them for the intended behaviour and quality:

  • temperature 0.6
  • top_k 64
  • top_p 0.9
  • min_p 0.05
  • repeat_penalty 1.1

This release is tuned end-to-end as its own thing; the settings above are part of that and aren't the stock Gemma defaults.

Specs

  • 31B dense · 256K (262144) context
  • Vision (image input) via mmproj
  • Based on Gemma 4 31B by Google DeepMind

Compatibility

  • Works with llama.cpp, LM Studio, Jan, koboldcpp, and other GGUF runtimes.
  • Multi-GPU + LM Studio: I've personally noticed Gemma 4 can crash under LM Studio's tensor-split mode — use a single GPU (layer-split or priority order) for this model.

Acknowledgements

  • Google DeepMind — Gemma 4.
  • The included mtp-gemma-4-31B-it.gguf speculative draft head comes from Unsloth's Gemma 4 release — many thanks to the Unsloth team for it.

* Tested with both automated and manual refusal benchmarks — none have been found in standard use. A small number of edge-case prompts deflect on the first ask but comply on a re-ask or strategic framing. If you hit one that's actually obstructive to your use case, join the Discord and flag it so I can work on it in a future revision.

Magnet link

Opens the swarm directly in your torrent client — no file download needed. Copy-paste works too:

magnet:?xt=urn:btih:641c74b5579b6da72a85c962abb1ab45c361f445&dn=HauhauCS_Gemma4-31B-QAT-Uncensored-HauhauCS-Balanced-MTP

Open magnet in torrent client · infohash 641c74b5579b6da72a85c962abb1ab45c361f445

Files & hashes

PathSizesha1sha256
Gemma4-31B-QAT-Uncensored-HauhauCS-Balanced-Q4_K_M.gguf17.40 GB (18,687,062,176 B)71667f9e601a4b914a98425c59150b731f6e15d260d661dbd1f1ee07469fc7db
README.md3.8 KB (3,848 B)3ad4df55ca25d91c46b91ab98152354c89ec27ef744efcebca6f22cc49343c9c4f03ba62f0ded80586f79ec29dcf57223ca075b6
mmproj-Gemma4-31B-QAT-Uncensored-HauhauCS-Balanced-BF16.gguf1.12 GB (1,200,726,016 B)7bef0d0fb3e85fc2941ec5f1c375febf3742645f158132a43ced557093aea841
mtp-gemma-4-31B-it.gguf267.0 MB (279,954,368 B)b5c4e583fc5982439080114bbc1b7edaec361f9d4c9193d6bed606a3de401b62

Cite this release

Canonical URL
https://aiseedbank.org/models/HauhauCS_Gemma4-31B-QAT-Uncensored-HauhauCS-Balanced-MTP/
Slug
HauhauCS_Gemma4-31B-QAT-Uncensored-HauhauCS-Balanced-MTP
Infohash
641c74b5579b6da72a85c962abb1ab45c361f445
License
gemma
Signing key fingerprint
85a3b32c3712427b

Every file carries a locally computed sha256 — verify a download against the signed sums: HauhauCS_Gemma4-31B-QAT-Uncensored-HauhauCS-Balanced-MTP.SHA256SUMS (+ minisign signature).

Provenance

Upstream repositoryHauhauCS/Gemma4-31B-QAT-Uncensored-HauhauCS-Balanced-MTP
Revision (pinned)9654466e82d83f5ebfe1518a369bc5900873abb1
Fetched at2026-08-31T23:24:27Z
License at fetchgemma
Snapshot toolhuggingface · seedbank 0.1.0

Trackers

✓ verified · rehash-vs-hf-metadata at 2026-08-31T23:28:19Z

gemma18.78 GB (20,167,746,408 bytes)ggufuncensoredgemma4visionmultimodalagenticcodingcreative-writingroleplayconversationalimage-text-to-textendpoints_compatible2 languages (rp, en)