Prefill a 284B model on Nvidia. Decode it on Apple Silicon. Over plain 10GbE
A prefill/decode disaggregation setup running DeepSeek-V4-Flash — a 284B total / 13B active model with 256 routed experts — bridged a 700,630-token cold prompt end to end in 11 minutes 26 seconds over plain 10 gigabit Et…