feat(local-inference): replace ollama with llama-swap + llama.cpp on zix790prors

- Add local-inference NixOS role using llama-swap (from nixpkgs-unstable) with llama.cpp (CUDA-enabled, from nixpkgs-unstable) - Serves Qwen3.6-35B-A3B via HuggingFace auto-download with --cpu-moe - Add nixosSpecialArgs for nixpkgs-unstable module access - Configure opencode with llama-local provider pointing to zix790prors:8080 - Update gptel from Ollama backend to OpenAI-compatible llama-swap backend - Remove ollama service from zix790prors
2026-04-16 15:20:37 -07:00
parent d16c8aa67e
commit 10efafd92e
7 changed files with 165 additions and 11 deletions
--- a/home/roles/base/default.nix
+++ b/home/roles/base/default.nix
@@ -99,6 +99,10 @@ in
      };
    };

+    xdg.configFile."opencode/opencode.json" = {
+      source = ./opencode-config.json;
+    };
+
    # Note: modules must be imported at top-level home config
  };
 }