Jamba 1.7 Large

AI21 Labs · ai21/ai21-jamba-large-1-7

ProviderAI21 Labs
Familyjamba
Typellm-chat
Statusactive
Released2025-07-02
Updated2026-02-02
Parameters398.6B
Open weightsyes

What it fits on

Ordered by predicted speed. Bandwidth sets decode rate; memory decides whether it runs at all.

Every figure here is computed, not measured. Fit is weights at each quantisation against device memory, with a 25% allowance for the KV cache, activations and the OS. Decode rate is the memory-bandwidth roofline at 70% efficiency. Nobody has run this model on these devices.
DeviceMemoryBandwidthBest qualityWeightstok/sFastest
NVIDIA GB200 Grace Blackwell Superchip372.0 GB16000.0 GB/sq5249.1 GB~45.0~56.2 q4
AMD Instinct MI355X288.0 GB8000.0 GB/sq4199.28 GB~28.1~28.1 q4
Apple M3 Ultra512.0 GB819.0 GB/sq6298.92 GB~1.9~2.9 q4

Where it runs

Taken from the card's availability section.

PlatformId on that platform
Hugging Faceai21labs/AI21-Jamba-Large-1.7gated

Verified benchmark evidence

Each row was checked against its source by a reviewer, and carries the model identifier as actually evaluated — which is not always the same as this card's.
BenchmarkEvaluated asScoreEvidence dateSource kind
aa_lcrJamba 1.7 Large19.33%2026-09-10 evaluatedindependent evaluatorsource
critptJamba 1.7 Large0.0%2026-09-10 evaluatedindependent evaluatorsource
gpqa_diamondJamba 1.7 Large38.99%2026-09-10 evaluatedindependent evaluatorsource

Reported benchmark scores

This card reports no benchmark scores yet.

BenchmarkCatalogue standingScore

Data

This card as JSON · See it in the graph · Edit on GitHub