Help preserve open and free AI for humanity's future

← All models

Qwen_Qwen3.8-27B

Qwen · View on Hugging Face ↗

Qwen3.8 27B open LLM from the current Qwen generation.

✓ verified · rehash-vs-hf-metadata at 2026-08-20T02:58:07Z

apache-2.051.77 GB (55,586,113,293 bytes)transformerssafetensorsqwen3_5image-text-to-textconversationaleval-resultsendpoints_compatible

Get this model

Download Qwen_Qwen3.8-27B.torrent

Recommended — the .torrent carries the webseed url-list, so your client can fall back to plain HTTPS if the swarm is thin. See/verify for the full download + verification walkthrough.

Model card

The complete upstream card, rendered from this payload's README.md — the same hash-verified bytes the torrent distributes. Images and off-site links are removed; the original card on Hugging Face carries them.


library_name: transformers license: apache-2.0 pipeline_tag: image-text-to-text

Qwen3.8-27B

[!Note] This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format.

These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, TokenSpeed, etc.

[!Tip] For users seeking managed, scalable inference without infrastructure maintenance, the official Qwen API service is provided by Qwen Cloud. In particular, Qwen3.8-27B will be available as a hosted version with more production features, e.g., 1M context length by default, official built-in tools. For more information, please refer to the Qwen3.8-27B Overview. The service is coming soon. Stay tuned for updates.

Following the widespread community adoption of the Qwen3.5 and Qwen3.6 series, we are pleased to introduce Qwen3.8, the most capable generation in the Qwen open-model family to date.

Built on the architectural foundation of Qwen3.5, Qwen3.8 delivers substantial gains across coding, professional work, research, and long-horizon agentic tasks. Qwen3.8-27B brings these advances to a compact, deployment-friendly dense model: a native vision-language model that understands images and videos, with flexible thinking control, designed to carry complex, multi-step tasks through to completion with greater reliability.

Qwen3.8 Highlights

Qwen3.8-27B features the following enhancements:

  • Core Capabilities: Comprehensive improvements across coding, professional work, research, and long-horizon agentic tasks.
  • Agent Execution: Stronger autonomous planning and better handling of environment feedback, leading to more reliable end-to-end task completion.
  • Downstream Compatibility: Broader support for popular harnesses and development tools, making it easier to integrate into your existing stack.
  • Flexible Thinking Control: Thinking mode is on by default and can be disabled per request; reasoning depth can be tuned with reasoning_effort, and reasoning context from historical messages is retained via preserve_thinking.
  • Vision-Language Understanding: Native support for image and video understanding, from STEM diagrams and documents to hour-scale videos.

Model Overview

  • Type: Causal Language Model with Vision Encoder
  • Training Stage: Pre-training & Post-training
  • Language Model
    • Number of Parameters: 27B
    • Hidden Dimension: 5120
    • Token Embedding: 248,320 (Padded)
    • Number of Layers: 64
    • Hidden Layout: 16 × (3 × (Gated DeltaNet → FFN) → 1 × (Gated Attention → FFN))
    • Gated DeltaNet:
      • Number of Linear Attention Heads: 48 for V and 16 for QK
      • Head Dimension: 128
    • Gated Attention:
      • Number of Attention Heads: 24 for Q and 4 for KV
      • Head Dimension: 256
      • Rotary Position Embedding Dimension: 64
    • Feed Forward Network:
      • Intermediate Dimension: 17,408
    • LM Output: 248,320 (Padded)
    • MTP (Multi-Token Prediction): trained with multiple steps
  • Context Length: 262,144 natively and extensible up to 1,000,000 tokens.

Benchmark Results

Text Performance

Qwen3.8-27BQwen3.6-27BQwen3.7-PlusMuse Glimmer-30BOpus4.6 Max
Coding
Agentic terminal codingTerminal Bench 2.1 (Terminus) 73.0 63.4 64.0 51.7 78.2
Agentic codingSWE-bench Pro 61.7 53.5 57.6 51.2 53.4
Repo-level code generationNL2Repo-Bench 42.3 36.2 41.1 -- 47.6
Agentic codingDeepSWE 1.1 42.2 13.3 14.2 -- --
Software engineeringQwenSWEBench 79.0 49.3 59.2 -- 63.8
Agent
Long-horizon office workCoWorkBench 70.7 61.0 65.1 -- 68.2
Professional job tasksJobBench 33.4 21.8 27.6 -- --
Frontier agentic tasksAgents' Last Exam Pass@120.4Score42.9 [email protected] [email protected] -- --
General
Instruction followingIFBench 79.5 69.1 79.1 77.0 62.5
Scientific reasoningGPQA Diamond 89.2 87.8 90.3 83.5 91.3
Multidisciplinary reasoningHLE 30.8 24.0 34.7 22.0 40.0
Competitive codingLiveCodeBench v6 90.3 83.9 89.6 -- 88.8
  1. SWE-bench Pro: Except for Opus4.6 Max, which uses the officially reported score, all models are evaluated with the Claude Code harness at temp=1.0, top_p=0.95, and a 256K context window. Problematic tasks were corrected, and all baseline models were re-evaluated on the refined benchmark.
  2. NL2Repo-Bench: Evaluated with the Claude Code harness. To prevent reward hacking, we disable Bash commands that attempt to access the specific repository, such as pip download, pip install, and git clone.
  3. DeepSWE 1.1: Evaluated with the Claude Code harness at temp=1.0, top_p=0.95, and a 256K context window.
  4. QwenSWEBench: In-house coding benchmark for evaluating models' software engineering capabilities. Evaluated with the Claude Code harness. Reporting avg@3 with an 8-hour timeout, max_tokens=32,768, temperature=1.0, and a 256K context window.
  5. CoWorkBench: In-house cowork benchmark for evaluating long-horizon tasks across computer science, finance, law, medical, and other productivity domains.
  6. HLE: Judged by GPT-4o.
  7. The best result in each row is shown in bold.
  8. Empty cells (--) indicate that results are not yet available or not applicable.

VL Performance

Qwen3.8-27BQwen3.6-27BQwen3.7-PlusMuse Glimmer-30BOpus4.6 Max
Agentic Multimodal Intelligence
Computer useOSWorld-Verified84.363.973.365.972.7
Browser useWebArena-Verified64.848.855.3----
Mobile useAndroidWorld81.970.381.0--62.0
Application recreationRecreationBench47.129.830.2----
Multimodal tool useClawEval-MMPass@357.4Average56.9[email protected]Pass@357.4Average60.1--[email protected]
Multimodal software engineeringSWE-MM38.625.730.0--27.1
Visual web developmentVision2Web62.945.042.1----
General Multimodal Intelligence
Visual math problem solvingMathVisionWithout CI90.0With CI94.6Without CI85.1Without CI90.3--Without CI65.5
General visual reasoningBabyVisionWithout CI65.7With CI85.6Without CI28.9Without CI64.7With CI70.4--Without CI12.6
Scientific chart analysisCharXiv (RQ)Without CI83.7With CI90.2Without CI78.4Without CI85.8With CI85.978.8Without CI66.0
Document intelligenceOmniDocBench 1.591.189.491.475.886.6
Real-world perceptionRealWorldQA85.984.186.9--73.9
Embodied intelligenceERQA65.562.569.8--40.8
  1. MathVision, BabyVision, and CharXiv (RQ): Where both settings are available, cells report “Without CI” and “With CI” separately; otherwise, only the available setting is shown. A small number of incorrect ground-truth annotations in MathVision and CharXiv (RQ) were corrected following manual verification, and all reported scores on those benchmarks were computed using the corrected annotations.
  2. MathVision: Qwen3.8-27B is evaluated using the fixed prompt: “Please reason step by step, and put your final answer within \boxed{}.” For the remaining models, we report the higher score from two prompt variants—one with and one without the \boxed{} formatting requirement.
  3. WebArena-Verified: Scores are computed with the official WebArena-Verified grader under the OSWorld scaffold.
  4. RecreationBench: An in-house, long-horizon application-recreation benchmark designed to evaluate hybrid-agent capabilities across five platforms: desktop (Ubuntu, macOS, and Windows), mobile (Android), and the web.
  5. ClawEval-MM: Scores are reported as “Pass@3 / average score.” Pass@3 is the percentage of tasks passed in at least one of three trials; the average score is the mean benchmark score across the three trials.
  6. Vision2Web: Scores are averaged across the frontend, webpage, and website categories. Evaluations use the Claude Code harness and are judged by gpt-5.4-2026-03-05.
  7. SWE-MM: Scores are evaluated on the Claude Code harness using the public dev split of SWE-bench Multimodal, with the modifications described in Appendix 8.3 of the Claude Opus 4.7 system card.
  8. Empty cells (--) indicate that results are not yet available or not applicable.

Quickstart

For streamlined integration, we recommend using Qwen3.8 via APIs.

Serving Qwen3.8

[!Important] Inference efficiency and throughput vary significantly across frameworks. We recommend using the latest framework versions to ensure optimal performance and compatibility. For production workloads or high-throughput scenarios, dedicated serving engines such as SGLang, vLLM, or TokenSpeed are recommended.

Qwen3.8 can be deployed with popular inference frameworks, e.g.:

  • SGLang: Qwen3.8 Cookbook
  • vLLM: Qwen3.8 Recipe
  • TokenSpeed: Qwen3.8 Recipe

API Usage

[!Important] Qwen3.8 models operate in thinking mode by default, generating thinking content signified by <think>\n...</think>\n\n before producing the final response. To disable thinking content and obtain a direct response, refer to the examples here.

[!Tip] We recommend using the following sets of sampling parameters for generation:

  • Thinking Mode: temperature=1.0, top_p=0.95, top_k=20, min_p=0.0, presence_penalty=0.0, repetition_penalty=1.0
  • Instruct (or non-thinking) mode: temperature=0.7, top_p=0.80, top_k=20, min_p=0.0, presence_penalty=1.5, repetition_penalty=1.0

Please note that the support for sampling parameters varies according to inference frameworks.

Qwen3.8 comes with official support for reasoning_effort, which can be used to adjust reasoning depth and control cost:

  • xhigh (default): for complex tasks demanding thorough analysis
  • medium: balancing accuracy and speed
  • low: efficient reasoning optimizing for speed and cost

In addition, preserve_thinking is enabled by default for all workloads for the best out-of-the-box experience. To disable preserved thinking, refer to the examples here.

[!Tip] In multi-turn agentic tasks, lower reasoning effort does not always reduce overall task completion time. Although it may produce faster per-turn responses, it can also lead to insufficient analysis, more failures, and repeated retries, which may increase total latency and token consumption.

Chat Completions API

The Chat Completions API can be used with most inference frameworks, as well as Qwen Cloud. Before starting, make sure the OpenAI Python SDK is installed and the API key and the API base URL are configured, e.g.:

pip install -U openai

# Set the following accordingly
export OPENAI_BASE_URL='your-base-url'
export OPENAI_API_KEY='your-api-key'
Text-Only Input
from openai import OpenAI
# Configured by environment variables
client = OpenAI()

messages = [{"role": "user", "content": "Write a Python function to merge two sorted linked lists."}]

completion = client.chat.completions.create(
    model="Qwen/Qwen3.8-27B",
    messages=messages,
    extra_body={
        "chat_template_kwargs": {
            "enable_thinking": True,  # on by default
            "preserve_thinking": True, # on by default
        },
    },
    reasoning_effort="xhigh",  # xhigh by default; supported levels are xhigh, medium, and low
    stream=True,
    stream_options={"include_usage": True},
)

reasoning_content = ""
answer_content = ""
is_answering = False
print("\n" + "=" * 20 + "Reasoning" + "=" * 20 + "\n")

for chunk in completion:
    if not chunk.choices:
        print("\nUsage:")
        print(chunk.usage)
        continue

    delta = chunk.choices[0].delta

    if hasattr(delta, "reasoning_content") and delta.reasoning_content is not None:
        if not is_answering:
            print(delta.reasoning_content, end="", flush=True)
        reasoning_content += delta.reasoning_content
    elif hasattr(delta, "reasoning") and delta.reasoning is not None:
        if not is_answering:
            print(delta.reasoning, end="", flush=True)
        reasoning_content += delta.reasoning

    if hasattr(delta, "content") and delta.content:
        if not is_answering:
            print("\n" + "=" * 20 + "Answer" + "=" * 20 + "\n")
            is_answering = True
        print(delta.content, end="", flush=True)
        answer_content += delta.content

messages.append({
    "role": "assistant",
    "content": answer_content,
    "reasoning_content": reasoning_content,
    "reasoning": reasoning_content,
})
Image Input
from openai import OpenAI
# Configured by environment variables
client = OpenAI()

messages = [
    {
        "role": "user",
        "content": [
            {
                "type": "image_url",
                "image_url": {
                    "url": "https://qianwen-res.oss-accelerate.aliyuncs.com/Qwen3.5/demo/CI_Demo/mathv-1327.jpg"
                }
            },
            {
                "type": "text",
                "text": "The centres of the four illustrated circles are in the corners of the square. The two big circles touch each other and also the two little circles. With which factor do you have to multiply the radii of the little circles to obtain the radius of the big circles?\nChoices:\n(A) $\\frac{2}{9}$\n(B) $\\sqrt{5}$\n(C) $0.8 \\cdot \\pi$\n(D) 2.5\n(E) $1+\\sqrt{2}$"
            }
        ]
    }
]

chat_response = client.chat.completions.create(
    model="Qwen/Qwen3.8-27B",
    messages=messages,
)
print("Chat response:", chat_response)
Video Input
from openai import OpenAI
# Configured by environment variables
client = OpenAI()

messages = [
    {
        "role": "user",
        "content": [
            {
                "type": "video_url",
                "video_url": {
                    "url": "https://qianwen-res.oss-accelerate.aliyuncs.com/Qwen3.5/demo/video/N1cdUjctpG8.mp4"
                }
            },
            {
                "type": "text",
                "text": "How many porcelain jars were discovered in the niches located in the primary chamber of the tomb?"
            }
        ]
    }
]

chat_response = client.chat.completions.create(
    model="Qwen/Qwen3.8-27B",
    messages=messages,
)

# When vLLM is launched with `--media-io-kwargs '{"video": {"num_frames": -1}}'`,
# video frame sampling can be configured via `extra_body` (e.g., by setting `fps`).
# This feature is currently supported only in vLLM.
#
# By default, `fps=2` and `do_sample_frames=True`.
# With `do_sample_frames=True`, you can customize the `fps` value to set your desired video sampling rate.
# chat_response = client.chat.completions.create(
#     model="Qwen/Qwen3.8-27B",
#     messages=messages,
#     extra_body={
#         "mm_processor_kwargs": {"fps": 2, "do_sample_frames": True},
#     }, 
# )

print("Chat response:", chat_response)
Instruct (or Non-Thinking) Mode

Qwen3.8-27B will think by default before responding. You can obtain a direct response from the model without thinking by configuring the API parameters. For example,

from openai import OpenAI
# Configured by environment variables
client = OpenAI()

messages = [
    {
        "role": "user",
        "content": [
            {
                "type": "image_url",
                "image_url": {
                    "url": "https://qianwen-res.oss-accelerate.aliyuncs.com/Qwen3.5/demo/RealWorld/RealWorld-04.png"
                }
            },
            {
                "type": "text",
                "text": "Where is this?"
            }
        ]
    }
]

chat_response = client.chat.completions.create(
    model="Qwen/Qwen3.8-27B",
    messages=messages,
    temperature=0.7,
    top_p=0.8,
    presence_penalty=1.5,
    extra_body={
        "top_k": 20,
        "chat_template_kwargs": {"enable_thinking": False},
    }, 
)
print("Chat response:", chat_response)

[!Note] If you are using APIs from Qwen Cloud, in addition to changing model, please use "enable_thinking": False instead of "chat_template_kwargs": {"enable_thinking": False}.

Disable Preserved Thinking

By default, Qwen3.8 retains thinking blocks from all historical messages, maintaining a complete reasoning trace across the conversation. This behavior, known as preserved thinking, ensures full context continuity and is especially beneficial for agent scenarios where decision consistency and reduced redundant reasoning are critical. It also improves KV cache utilization, optimizing inference efficiency in both thinking and non-thinking modes.

If you prefer to retain only the thinking blocks from the latest user message, you can disable this behavior by setting preserve_thinking to False:

from openai import OpenAI

# Configured by environment variables
client = OpenAI()
messages = [...]
chat_response = client.chat.completions.create(
    model="Qwen/Qwen3.8-27B",
    messages=messages,
    extra_body={
        "chat_template_kwargs": {"preserve_thinking": False},
    },
)
print("Chat response:", chat_response)

[!Note] If you are using APIs from Qwen Cloud, in addition to changing model, please use "preserve_thinking": False directly instead of wrapping it in chat_template_kwargs.

Best Practices

To achieve optimal performance, we recommend the following settings:

  1. Sampling Parameters: We suggest using the following sets of sampling parameters:

    • Thinking Mode: temperature=1.0, top_p=0.95, top_k=20, min_p=0.0, presence_penalty=0.0, repetition_penalty=1.0
    • Instruct (or non-thinking) mode: temperature=0.7, top_p=0.80, top_k=20, min_p=0.0, presence_penalty=1.5, repetition_penalty=1.0

    For supported frameworks, you can adjust the presence_penalty parameter between 0 and 2 to reduce endless repetition. However, using a higher value may occasionally result in language mixing and a slight decrease in model performance.

  2. Adequate Output Length: To optimize performance on agentic tasks, we recommend allocating sufficient output length to allow the model to generate detailed and comprehensive responses. For frameworks that support separate token limits for internal reasoning and final outputs, we suggest the following configuration within the 1M context length:

    • Reasoning Content: Set the maximum output length to 262,144 tokens.
    • Final Response: Set the maximum output length to 131,072 tokens.

    These settings provide the necessary capacity for complex reasoning while ensuring ample space for high-quality final deliverables.

  3. Processing Ultra-Long Texts: Qwen3.8-27B natively supports context lengths of up to 262,144 tokens. For long-horizon tasks where the total length (including both input and output) exceeds this limit, we recommend using RoPE scaling techniques to handle long texts effectively, e.g., YaRN.

    YaRN is currently supported by several inference frameworks, e.g., vLLM, SGLang, and TokenSpeed. In general, there are two approaches to enabling YaRN for supported frameworks:

    • Modifying the model configuration file:

      In the config.json file, change the rope_parameters fields in text_config to:

      {
          "mrope_interleaved": true,
          "mrope_section": [
              11,
              11,
              10
          ],
          "rope_type": "yarn",
          "rope_theta": 10000000,
          "partial_rotary_factor": 0.25,
          "factor": 4.0,
          "original_max_position_embeddings": 262144,
      }
      
    • Passing command line arguments:

      For vLLM, you can use

      VLLM_ALLOW_LONG_MAX_MODEL_LEN=1 vllm serve ... --hf-overrides '{"text_config": {"rope_parameters": {"mrope_interleaved": true, "mrope_section": [11, 11, 10], "rope_type": "yarn", "rope_theta": 10000000, "partial_rotary_factor": 0.25, "factor": 4.0, "original_max_position_embeddings": 262144}}}' --max-model-len 1000000  
      

      For SGLang, you can use

      SGLANG_ALLOW_OVERWRITE_LONGER_CONTEXT_LEN=1 python -m sglang.launch_server ... --json-model-override-args '{"text_config": {"rope_parameters": {"mrope_interleaved": true, "mrope_section": [11, 11, 10], "rope_type": "yarn", "rope_theta": 10000000, "partial_rotary_factor": 0.25, "factor": 4.0, "original_max_position_embeddings": 262144}}}' --context-length 1000000
      

      For TokenSpeed, you can use

      TOKENSPEED_ALLOW_OVERWRITE_LONGER_CONTEXT_LEN=1 tokenspeed serve ... --hf-overrides '{"text_config": {"rope_parameters": {"mrope_interleaved": true, "mrope_section": [11, 11, 10], "rope_type": "yarn", "rope_theta": 10000000, "partial_rotary_factor": 0.25, "factor": 4.0, "original_max_position_embeddings": 262144}}}' --max-model-len 1000000  
      

    [!NOTE] All the notable open-source frameworks implement static YaRN, which means the scaling factor remains constant regardless of input length, potentially impacting performance on shorter texts. We advise modifying the rope_parameters configuration only when processing long contexts is required. It is also recommended to modify the factor as needed. For example, if the typical context length for your application is 524,288 tokens, it would be better to set factor as 2.0.

  4. Long Video Understanding: To optimize inference efficiency for plain text and images, the size parameter in the released video_preprocessor_config.json is conservatively configured. It is recommended to set the longest_edge parameter in the video_preprocessor_config file to 469,762,048 (corresponding to 224k video tokens) to enable higher frame-rate sampling for hour-scale videos and thereby achieve superior performance. For example,

    {"longest_edge": 469762048, "shortest_edge": 4096}
    

    Alternatively, override the default values via engine startup parameters. For implementation details, refer to: vLLM / SGLang.

Citation

If you find our work helpful, feel free to give us a cite.

@misc{qwen38,
    title = {{Qwen3.8-Max}: A New Bar for Coding and Cowork},
    url = {https://qwen.ai/blog?id=qwen3.8},
    author = {{Qwen Team}},
    month = {August},
    year = {2026}
}

Magnet link (secondary — no webseeds)

Opens the swarm directly, but carries no webseed url-list. Prefer the.torrent download above — HTTP fallback seeds ride inside it.

magnet:?xt=urn:btih:05c43c64933e1e54f135137166b11c0be832d22f&dn=Qwen_Qwen3.8-27B

Open magnet in torrent client · infohash 05c43c64933e1e54f135137166b11c0be832d22f

Files & hashes

PathSizeMethodHash
LICENSE11.3 KB (11,544 B)sha1-git-blobf938136e3adacfd92be087f6e113b5d6d97f678f
README.md63.5 KB (65,012 B)sha1-git-blobbc8aa0e396cd029c21cd773cca21830d0ded28ec
chat_template.jinja8.7 KB (8,952 B)sha1-git-blobc0c686f9c38d70d179fb7b5f5aa7530bc913dda3
config.json4.2 KB (4,312 B)sha1-git-blob706cebd746c4b6f2b1d1f892630867acfdfd3df8
crc32.txt238 B (238 B)sha1-git-blob6de5ee6a0c6596744baee911af0a5cdcb8d99a1e
generation_config.json202 B (202 B)sha1-git-blob023756cfadf88e5bf69eefeee3e172f38c448d64
merges.txt3.2 MB (3,353,259 B)sha1-git-bloba494e019ca1502219fd0128658b979e5f05ae8e8
model-00001-of-00018.safetensors3.69 GB (3,966,730,552 B)sha256-lfsba0ce20aae489ad196733da5064bcdf159a1fe84f53336648196e1ebb7751b1c
model-00002-of-00018.safetensors2.83 GB (3,043,080,328 B)sha256-lfs06a148c01bfbe3faa14a5f184a7ff29a706f7ae1c8b2705d2058e26d17a001fb
model-00003-of-00018.safetensors2.37 GB (2,542,796,952 B)sha256-lfs2e1bf62cbcd406eaa64b60d10353e1f0ef4039d0976e56f05cabe953454f9968
model-00004-of-00018.safetensors3.72 GB (3,988,973,152 B)sha256-lfs511e34063187882659753c4d93f3859f93c019fd438d8813071921c81d9a3f1a
model-00005-of-00018.safetensors1.96 GB (2,099,339,864 B)sha256-lfs635cb53446dc74f219740fc59e18b774f877b803b9722e289ca62575a6efa701
model-00006-of-00018.safetensors3.71 GB (3,979,553,696 B)sha256-lfs0bc5214fac607f0e6cc92eec3789d4b8559410ef9fce66621ba8158e8410dae0
model-00007-of-00018.safetensors1.96 GB (2,108,759,344 B)sha256-lfs80b0c49033e9a0d5762562aa12f4acdb7f54da586f3d0110f28c48d91cf07892
model-00008-of-00018.safetensors3.71 GB (3,979,553,696 B)sha256-lfs7192c5b66185d3592927daabee1cc19e6f6e0ce75988ee20e824b624765fda79
model-00009-of-00018.safetensors1.96 GB (2,108,759,344 B)sha256-lfsaf3c48cc37af44f3db6ae0579baf019180d48d9c527caa0a1f03ff85813a56d8
model-00010-of-00018.safetensors3.71 GB (3,979,553,696 B)sha256-lfs163490a76f3bea3a40855b7efc04ce6d27afaf1a34f0bbde495b9491f76457c9
model-00011-of-00018.safetensors1.96 GB (2,108,759,344 B)sha256-lfs5f3ae1b948aeee39da77aec558e8236cd65fe4d7cb7686a76bb007acc563c6d8
model-00012-of-00018.safetensors3.71 GB (3,979,553,696 B)sha256-lfsa3de1c7114677a8f5ac5c4892c90e8238ea5c1e2038c80e757dfc87c3902ca55
model-00013-of-00018.safetensors1.96 GB (2,108,759,344 B)sha256-lfs06ab79a41f74c9c5cb734816feb0c7fc364104b227165ee7391231e1155aa02a
model-00014-of-00018.safetensors3.71 GB (3,979,553,696 B)sha256-lfs4138ed94603065ba884bbcadedb04d7718bb40117e85e6f5c6fc5b9c05b7a85b
model-00015-of-00018.safetensors1.96 GB (2,108,759,344 B)sha256-lfs69224e27b9de4e7dbf6fc936c6eaae08447bda3b80a6c31a871ab451173afd22
model-00016-of-00018.safetensors3.71 GB (3,979,564,040 B)sha256-lfs73cb9a1089fb6155cb648609478d6633be8a5c7d9ca5a05bc8925ce8a553cefe
model-00017-of-00018.safetensors1.96 GB (2,108,759,344 B)sha256-lfsbeb51f01056142ac4984bd800507b0dd0fd18de57f8e9ef6ea41d1a3598983a8
model-00018-of-00018.safetensors3.16 GB (3,392,197,344 B)sha256-lfs1d3479509e21494658f9b64d317f5ea8e55c4025d28c702d6c4d0b356ce8ea06
model.safetensors.index.json109.6 KB (112,216 B)sha1-git-blobda35e3c564457dface7d138f0b6cac284ff8958c
preprocessor_config.json390 B (390 B)sha1-git-blob2ea84a437d448ff71b08df68fdd949d5cc4ebb64
tokenizer.json12.2 MB (12,809,320 B)sha256-lfs0997f410c57a1f4e53b09e4be8f4a172d90edd9564368fb0847030937229b9f3
tokenizer_config.json17.5 KB (17,928 B)sha1-git-blob5de744b3fca2129d7186979ae47c06be33903243
video_preprocessor_config.json385 B (385 B)sha1-git-blob3ba673a5ad7d4d13f54155ecd38b2a94a6dac8fe
vocab.json6.4 MB (6,722,759 B)sha1-git-blob0aa0ce0658d60ac4a5d609f4eadb0e8e43514176

Provenance

Upstream repositoryQwen/Qwen3.8-27B
Revision (pinned)1d4bf0f2ff6012fd82039f2fa52739d0dd7c60c0
Fetched at2026-08-20T02:02:59Z
License at fetchapache-2.0
Snapshot toolhuggingface · seedbank 0.1.0

Trackers

Webseeds