LG AI Research · lgai-exaone/k-exaone-2-0-750b-a37b
Provider
LG AI Research
Family
k-exaone
Type
llm-chat
Status
active
Released
2026-07-29
Updated
2026-08-07
Parameters
749.4B
Open weights
yes
What it fits on
Ordered by predicted speed. Bandwidth sets decode rate; memory decides whether it runs at all.
Every figure here is computed, not measured. Fit is weights at each quantisation against device memory, with a 25% allowance for the KV cache, activations and the OS. Decode rate is the memory-bandwidth roofline at 70% efficiency. Nobody has run this model on these devices. This is a mixture-of-experts model and no card carries active-parameter counts, so these predictions use total parameters and understate the real speed.
Device
Memory
Bandwidth
Best quality
Weights
tok/s
Fastest
Apple M3 Ultra
512.0 GB
819.0 GB/s
q4
374.68 GB
~1.5
~1.5 q4
Where it runs
Taken from the card's availability section.
Platform
Id on that platform
Hugging Face
LGAI-EXAONE/K-EXAONE-2.0-750B-A37B
Verified benchmark evidence
Each row was checked against its source by a reviewer, and carries the model identifier as actually evaluated — which is not always the same as this card's.