• 3 Posts
  • 1.29K Comments
Joined 3 years ago
cake
Cake day: June 4th, 2023

help-circle






  • I have a Strix Halo as well and I have it configured to dynamically allocate everything, so I have run some 100GB+ models as well. The OS itself needs way under a GB without UI so you can get pretty close to the 128GB.

    Currently on Fedora, running llama-swap to start the llama.cpp toolboxes by kyuz0, audio.cpp and ComfyUI.

    Qwen3.8-27b is my favorite model right now for most things, still playing around with Qwen3.8-Flash-Next but not quite there yet.

    Gemma4-31b is also really good with languages and natural writing but I prefer Qwen3.8 for anything programming or logical.