HW/FW security researcher & Demoscene elder.
I started having arguments online back on Fidonet and Usenet. I'm too tired to care now.
HW/FW security researcher & Demoscene elder.
I started having arguments online back on Fidonet and Usenet. I'm too tired to care now.
llama rpc exists and works really well
Qwen 27B Q4 at usable speed on 16GB VRAM
My llama-server suddenly started error 400 on the chat template - this fixed it
GPU bifurcation - options
Opencode llama-server prefill/generation stats plugin
North Mini Code v1.0 - a Qwen 3.6 35B MoE alternative
Don't skimp on the quant when using MoE
Permanently Deleted
Are you still here Proton?
All quants updated with correct template now