feat: GPU detection for LLM server (NVIDIA/AMD/Intel/Vulkan + CPU fallback)

- Added start-llama-server.bat with cross-brand GPU auto-detection:
  * Detection 1: nvidia-smi for NVIDIA
  * Detection 2: WMIC for AMD, Intel Arc, Intel HD/Iris/UHD
  * Detection 3: vulkaninfo as final fallback probe
  * Falls back to CPU-only build when no GPU found
- Downloaded llama-server b9947 builds:
  * Vulkan build (91 MB) - supports all GPU brands
  * CPU build (45 MB) - fallback in cpu/ subdir
- Downloaded Qwen2.5-Coder-1.5B Q4_K_M model (1.1 GB)
- Fixed run-hermes.bat polling: replaces fixed 5s timeout with
  curl-based readiness polling (up to 30 attempts, 1s apart)
- All .bat files verified: ASCII text, CRLF line terminators
- Added .gitignore for build artifacts and large binaries
This commit is contained in:
2026-07-10 00:58:56 -04:00
parent ecda528b90
commit 4e0ab94ee2
14 changed files with 4353 additions and 66 deletions
+4 -4
View File
@@ -155,10 +155,10 @@ run_diagnostics() {
info "Stress tests disabled (enable in config or set RUN_STRESS_TESTS=1)"
fi
# Phase 5: Backup/restore (manual activation)
if [ -n "${RUN_BACKUP}" ]; then
run_script "${HERMES_DIR}/scripts/backup.py" "05-backup.txt"
fi
# Phase 5: Backup/restore (manual activation) — stub placeholder
# if [ -n "${RUN_BACKUP}" ]; then
# run_script "${HERMES_DIR}/scripts/backup.py" "05-backup.txt"
# fi
echo "───────────────────────────────────────────────────────" >> "${LOG_FILE}"
info "Diagnostic pipeline complete."