Ollama uses ROCm whereas llama.cpp uses Vulkan compute. Which one will perform better depends on many factors, but Vulkan compute should be easier to setup.
- Posts
- 8
- Comments
- 193
- Joined
- 1 yr. ago
- Posts
- 8
- Comments
- 193
- Joined
- 1 yr. ago
Dsub2000 and Tempo are active FOSS alternatives.