LLMBOTTLENECK.COM
Hugging Face repository / answered by DeepSeek-V3.2-Exp

deepseek-ai/DeepSeek-V3.2-Speciale: what it needs to run

deepseek-ai/DeepSeek-V3.2-Speciale is a fine-tune of deepseek-ai/DeepSeek-V3.2-Exp-Base. Its configuration matches DeepSeek-V3.2-Exp's on every field that decides memory, so DeepSeek-V3.2-Exp's requirements are this repository's.

Answered by

Geometrypublished data. Both configurations are published files, read at the revisions named below and compared field by field. Not a test that the model loads.

DeepSeek-V3.2-Exp (deepseek-ai/DeepSeek-V3.2-Exp) — every compared field is stated by both and equal.

Before you use those figures

  • The declared base is deepseek-ai/DeepSeek-V3.2-Exp-Base; the catalogue holds its sibling deepseek-ai/DeepSeek-V3.2-Exp, which is the same architecture.

DeepSeek-V3.2-Exp: weight-file size by format

FormatSizeRangeBasis
FP161371.4 GB959.6 GB – 1384.5 GBsize range
Q8_0728.8 GB509.8 GB – 737.3 GBsize range
Q6_K562.9 GB393.6 GB – 570.1 GBsize range
Q5_K_M489.6 GB329.8 GB – 570.1 GBsize range
Q5_0478.3 GB329.8 GB – 570.1 GBsize range
Q4_K_M420.7 GB269.9 GB – 570.1 GBsize range
Q4_0398.8 GB269.9 GB – 570.1 GBsize range
Q3_K_M342.8 GB206.2 GB – 570.1 GBsize range
Q2_K271.4 GB157.4 GB – 570.1 GBsize range

These are DeepSeek-V3.2-Exp’s sizes. A model with the same geometry quantizes to the same size, to within any difference in parameter count noted above. Which devices hold each, and how fast they run it, is on DeepSeek-V3.2-Exp’s page.

What was read, and where

  1. deepseek-ai/DeepSeek-V3.2-Speciale at revision c562883eda91 is tagged by its publisher as a fine-tune of deepseek-ai/DeepSeek-V3.2-Exp-Base.
  2. deepseek-ai/DeepSeek-V3.2-Speciale/config.json at c562883eda91 was compared with deepseek-ai/DeepSeek-V3.2-Exp/config.json at 194c67e12b1b, the revision the catalogue pins: layers and their pattern, widths, head counts, experts, latent and recurrent dimensions, vocabulary, tied embeddings and any vision or audio tower.
Licence, as tagged
mit
Parameters in its safetensors index
685B
Task, as tagged
text-generation
Last changed on Hugging Face
2025-12-01

Publishing this model? A badge for its card

The memory badge for this repository

[![Memory to run this model, from llmbottleneck.com](https://llmbottleneck.com/badge/deepseek-ai/DeepSeek-V3.2-Speciale)](https://llmbottleneck.com/hf/deepseek-ai/DeepSeek-V3.2-Speciale)

It states the memory to run the model at Q4_K_M with 8,192 tokens of context — built on this repository’s own Q4_K_M file when it publishes one — and links back to this page, where the evidence is.

Resolve another repository

The same answer as JSON: GET /v1/resolve?repo=deepseek-ai/DeepSeek-V3.2-Speciale, with a free key from /keys.

llmbottleneck
catalogue 2026-10-03models 327devices 135