this is the most recent stuff I can find: Reddit
everything more recent I find just ends up linking back to that post. which like, sure, I get it, that was where folks pointed out the PR for TP SYCL that got merged into llama.cpp. I just want to know if folks around here are running it and whether or not you’ve seen improvements in performance over the past two months, especially on everyone’s favorite, Qwen3.6/8-27B. I am slowly leaning in the direction of selling my two 16GB 5060Tis and purchasing a couple of Arc Pro B65s (if I can find them) in order to have the VRAM to run that model specifically at a reasonable quant without quantizing kv.