LLMBOTTLENECK.COM
Hugging Face repository / answered by Mistral-7B-v0.1

HuggingFaceH4/zephyr-7b-beta: what it needs to run

HuggingFaceH4/zephyr-7b-beta is a fine-tune of mistralai/Mistral-7B-v0.1. Its configuration matches Mistral-7B-v0.1's on every field that decides memory, so Mistral-7B-v0.1's requirements are this repository's.

Answered by

Geometrypublished data. Both configurations are published files, read at the revisions named below and compared field by field. Not a test that the model loads.

Mistral-7B-v0.1 (mistralai/Mistral-7B-v0.1) — every compared field is stated by both and equal.

Mistral-7B-v0.1: weight-file size by format

FormatSizeRangeBasis
FP1614.5 GB14.2 GB – 14.6 GBreconstructed size
Q8_07.70 GB7.56 GB – 7.74 GBreconstructed size
Q6_K5.94 GB5.83 GB – 5.97 GBreconstructed size
Q5_K_M5.13 GB5.02 GB – 5.16 GBreconstructed size
Q5_05.00 GB4.89 GB – 5.02 GBreconstructed size
Q4_K_M4.37 GB4.26 GB – 4.40 GBreconstructed size
Q4_04.11 GB4.00 GB – 4.14 GBreconstructed size
Q3_K_M3.56 GB3.31 GB – 3.81 GBreconstructed size
Q2_K2.65 GB2.54 GB – 2.75 GBreconstructed size

These are Mistral-7B-v0.1’s sizes. A model with the same geometry quantizes to the same size, to within any difference in parameter count noted above. Which devices hold each, and how fast they run it, is on Mistral-7B-v0.1’s page.

What was read, and where

  1. HuggingFaceH4/zephyr-7b-beta at revision 892b3d7a7b1c is tagged by its publisher as a fine-tune of mistralai/Mistral-7B-v0.1.
  2. HuggingFaceH4/zephyr-7b-beta/config.json at 892b3d7a7b1c was compared with mistralai/Mistral-7B-v0.1/config.json at 27d67f1b5f57, the revision the catalogue pins: layers and their pattern, widths, head counts, experts, latent and recurrent dimensions, vocabulary, tied embeddings and any vision or audio tower.

Matched through a checked library default rather than a stated value: head_dim.

Licence, as tagged
mit
Parameters in its safetensors index
7.2B
Task, as tagged
text-generation
Last changed on Hugging Face
2024-10-16

Publishing this model? A badge for its card

The memory badge for this repository

[![Memory to run this model, from llmbottleneck.com](https://llmbottleneck.com/badge/HuggingFaceH4/zephyr-7b-beta)](https://llmbottleneck.com/hf/HuggingFaceH4/zephyr-7b-beta)

It states the memory to run the model at Q4_K_M with 8,192 tokens of context — built on this repository’s own Q4_K_M file when it publishes one — and links back to this page, where the evidence is.

Resolve another repository

The same answer as JSON: GET /v1/resolve?repo=HuggingFaceH4/zephyr-7b-beta, with a free key from /keys.

llmbottleneck
catalogue 2026-10-03models 327devices 135