Skip to main content

Local LLM stack

Qwen3.8-27B on two GPUs with llama.cpp, llama-swap and Unsloth Studio. What worked, what did not.

12 entries