Pixelfed
- Posts
- 16
- Comments
- 42
- Joined
- 4 yr. ago
- Posts
- 16
- Comments
- 42
- Joined
- 4 yr. ago
LocalLLaMA @sh.itjust.works Gemma4 12b released with "unified" approach to multi-modality
LocalLLaMA @sh.itjust.works llama.cpp: don't sleep on --split-mode tensor
LocalLLaMA @sh.itjust.works Gemma 4 is here
LocalLLaMA @sh.itjust.works Smaller qwen3.5 models released
LocalLLaMA @sh.itjust.works Qwen3-Coder-Next
LocalLLaMA @sh.itjust.works Relevance of GPU driver version for inference performance
LocalLLaMA @sh.itjust.works Magistral-Small-2509 by Mistral has been released
LocalLLaMA @sh.itjust.works Qwen3-Next with 80b-a3b parameters is out
LocalLLaMA @sh.itjust.works ExLlamaV3 adds tensor parallelism support
LocalLLaMA @sh.itjust.works New, promising MoE model "Hunyuan" by Tencent
LocalLLaMA @sh.itjust.works Do you quantize models yourself?
Selfhosted @lemmy.world Any experience with Pangolin?
Technology @lemmy.world More than 140 Kenya Facebook moderators diagnosed with severe PTSD
Selfhosted @lemmy.world Chaining routers and GUA IPv6 addresses
Selfhosted @lemmy.world Any of you have a self-hosted AI "hub"? (e.g. for LLM, stable-diffusion, ...)
Selfhosted @lemmy.world Migrated my self-hosted Nextcloud to AIO and I absolutely love it







Hugging Face used the incident to promote running open-weights LLMs on-prem, which a large part of their business focusses on.
Their blog post