LLMBOTTLENECK.COM
farbodtavakkoli / gemma4_text

OTel-LLM-E4B-IT

Parameter count not published.

Architecturepublished data

Architecture available · weight sizes unavailable

Weight sizeno data

This snapshot holds OTel-LLM-E4B-IT’s architecture but no weight size for any format, so the calculator cannot answer for it and it has no hardware table.

Why: Neither a published artifact nor a published parameter count exists for this model at its pinned revision. A size is never inferred from the model’s name.

The coverage page explains what the catalogue does not size, and why.

Under the hood

Architecture, read from the publisher’s file

The numbers every figure above is computed from, with the file they came from.

Architecture

✓ Architecture read from the published config.json

Retrieved 2026-09-01 at pinned commit 12ffb1ef5812.

Architecture
gemma4_text
Layers
42
Hidden size
2,560
Attention heads
8
KV heads
2
Head dimension
256
Feed-forward width
10,240
Vocabulary
262,144
Context ceiling
131,072
Sliding window
512

Where the memory goes

tokenembedding× 42 decoder blocksGrouped-query attention8 query · 2 KV headswindow 512Feed-forwardone networkall activeoutputprojectiongrows with contextfixed per token

Grouped-query attention shares each key/value head across 4 query heads, so the KV cache is 25% of what multi-head attention would need at the same context.

This is a multimodal checkpoint (vision + audio); the diagram and memory sizing above cover the text decoder. Encoder/audio/image activations are not included in the KV or weight figures.

More from farbodtavakkoli

Citing this page

LLM Bottleneck. “OTel-LLM-E4B-IT VRAM and hardware requirements.” Architecture from farbodtavakkoli/OTel-LLM-E4B-IT at revision 12ffb1ef5812, retrieved 2026-09-01. https://llmbottleneck.com/models/farbodtavakkoli-otel-llm-e4b-it

Every figure above is either the published value or a reconstruction whose measured error is on the accuracy page.

llmbottleneck
catalogue 2026-10-03models 327devices 135