top | item 47193717

(no title)

mirekrusin | 1 day ago

2x RTX 4090, Q8, 256k context, 110 t/s

discuss

order

instagib|1 day ago

1 4090, Qwen3.5-35B-A3B-UD-MXFP4_MOE, 64k context, 122 t/s. Llama.cpp

mirekrusin|19 hours ago

I believe it's mentioned that MXFP4 performs surprisingly bad, you may want to try other Q4s.