// RUN โ€” FITS

Can the Huawei Ascend 950DT run GenSLM?

Yes โ€” here's how fast and up to what context, with real measured numbers.

GenSLM on the Huawei Ascend 950DT

QuantFits?Max contextSpeedBeyond context
FP16 yes 309k 52-63.2 t/s ~0.5-0.6 (slow)
Q8 yes 404k 104-126.4 t/s ~1-1.3 (slow)
Q4 yes 446k 185.7-225.7 t/s ~1.9-2.3 (slow)

Estimate (memory-bound). Beyond "max context" the KV cache spills to system RAM offload (depends on your system RAM) and speed drops to the "beyond" figure.

FAQ

Can the Huawei Ascend 950DT run GenSLM?โ–ถ

Yes. At Q4 it generates ~185.7-225.7 tok/s and fits up to 446k context; beyond that it spills to system RAM offload (depends on your system RAM) and slows to ~1.9-2.3 tok/s.

How much context fits?โ–ถ

At Q4, up to about 446k tokens stay in fast memory; longer context spills to system RAM offload (depends on your system RAM) and throughput collapses.