// RUN โ€” FITS

Can the FuriosaAI RNGD run Prithvi WxC?

Yes โ€” here's how fast and up to what context, with real measured numbers.

Prithvi WxC on the FuriosaAI RNGD

QuantFits?Max contextSpeedBeyond context
FP16 yes 581k 212-257.6 t/s ~5.7-6.9 (slow)
Q8 yes 616k 423.9-515.2 t/s ~11.3-13.7 (slow)
Q4 yes 631k 757-920 t/s ~20.2-24.5 (slow)

Estimate (memory-bound). Beyond "max context" the KV cache spills to system RAM offload (depends on your system RAM) and speed drops to the "beyond" figure.

FAQ

Can the FuriosaAI RNGD run Prithvi WxC?โ–ถ

Yes. At Q4 it generates ~757-920 tok/s and fits up to 631k context; beyond that it spills to system RAM offload (depends on your system RAM) and slows to ~20.2-24.5 tok/s.

How much context fits?โ–ถ

At Q4, up to about 631k tokens stay in fast memory; longer context spills to system RAM offload (depends on your system RAM) and throughput collapses.