ESP-IDF CI / Host Tests (Unity) (push) Successful in 1m8s
CI / firmware-native (push) Successful in 2m57s
Rust Protection Tests / Cargo test (host) (push) Failing after 3m21s
ESP-IDF CI / ESP-IDF Build (v5.4) (push) Failing after 6m55s
ESP-IDF CI / Memory Budget Gate (push) Has been skipped
qa-cicd-environments / qa-kxkm-s3-build (push) Successful in 8m53s
qa-cicd-environments / qa-sim-host (push) Successful in 2m2s
qa-cicd-environments / qa-kxkm-s3-memory-budget (push) Successful in 11m17s
Context: the project archive (KXKM_Batterie_Parallelator-main) had no git history locally; a fresh repository is needed to host it on git.saillant.cc (electron/KXKM_Batterie_Parallelator). Approach: initialize a new repo on branch main, stage the archive content, and harden .gitignore before the first commit. Changes: - Import the full project tree: firmware/, firmware-idf/, firmware-rs/, iosApp/, kxkm-bmu-app/, kxkm-api/, hardware/, docs/, specs/, scripts/, models/, tests/ - Keep project dotfiles tracked despite the trailing '.*' ignore rule: .github/, .claude/, .superpowers/, .gitattributes, .markdownlint.json - Extend .gitignore: firmware/src/credentials.h (local secrets, template kept), kxkm-bmu-app/**/build/ (66 MB compiled iOS framework), .remember/ (session data) Impact: the project can now be maintained on the self-hosted Gitea forge with a clean, secret-free initial history.
42 lines
1.1 KiB
Docker
42 lines
1.1 KiB
Docker
# soh-llm: Qwen2.5-7B diagnostic inference service
|
|
# Runs on kxkm-ai (RTX 4090 24 GB)
|
|
|
|
FROM nvidia/cuda:12.4.1-runtime-ubuntu22.04
|
|
|
|
ENV DEBIAN_FRONTEND=noninteractive
|
|
ENV PYTHONUNBUFFERED=1
|
|
|
|
# System dependencies
|
|
RUN apt-get update && apt-get install -y --no-install-recommends \
|
|
python3.12 python3.12-venv python3-pip git curl \
|
|
&& rm -rf /var/lib/apt/lists/*
|
|
|
|
# Create app user
|
|
RUN useradd -m -s /bin/bash llm
|
|
USER llm
|
|
WORKDIR /app
|
|
|
|
# Python dependencies
|
|
COPY requirements.txt .
|
|
RUN python3.12 -m pip install --user --no-cache-dir -r requirements.txt
|
|
|
|
# Application code
|
|
COPY config.py prompt_template.py inference_server.py diagnostic_api.py ./
|
|
|
|
# Model volume mount point
|
|
VOLUME /models
|
|
|
|
# Environment defaults
|
|
ENV LLM_BASE_MODEL=Qwen/Qwen2.5-7B
|
|
ENV LLM_LORA_ADAPTER=/models/qwen-bmu-diag/lora-adapter
|
|
ENV LLM_API_HOST=0.0.0.0
|
|
ENV LLM_API_PORT=8401
|
|
ENV INFLUXDB_URL=http://influxdb:8086
|
|
|
|
EXPOSE 8401
|
|
|
|
HEALTHCHECK --interval=30s --timeout=10s --retries=3 \
|
|
CMD curl -f http://localhost:8401/health || exit 1
|
|
|
|
CMD ["python3.12", "-m", "uvicorn", "diagnostic_api:app", "--host", "0.0.0.0", "--port", "8401", "--workers", "1"]
|