Files
L'électron rare f55093d6fe
ESP-IDF CI / Host Tests (Unity) (push) Successful in 1m8s
CI / firmware-native (push) Successful in 2m57s
Rust Protection Tests / Cargo test (host) (push) Failing after 3m21s
ESP-IDF CI / ESP-IDF Build (v5.4) (push) Failing after 6m55s
ESP-IDF CI / Memory Budget Gate (push) Has been skipped
qa-cicd-environments / qa-kxkm-s3-build (push) Successful in 8m53s
qa-cicd-environments / qa-sim-host (push) Successful in 2m2s
qa-cicd-environments / qa-kxkm-s3-memory-budget (push) Successful in 11m17s
chore: import KXKM Batterie Parallelator
Context: the project archive (KXKM_Batterie_Parallelator-main) had
no git history locally; a fresh repository is needed to host it on
git.saillant.cc (electron/KXKM_Batterie_Parallelator).

Approach: initialize a new repo on branch main, stage the archive
content, and harden .gitignore before the first commit.

Changes:
- Import the full project tree: firmware/, firmware-idf/,
  firmware-rs/, iosApp/, kxkm-bmu-app/, kxkm-api/, hardware/,
  docs/, specs/, scripts/, models/, tests/
- Keep project dotfiles tracked despite the trailing '.*' ignore
  rule: .github/, .claude/, .superpowers/, .gitattributes,
  .markdownlint.json
- Extend .gitignore: firmware/src/credentials.h (local secrets,
  template kept), kxkm-bmu-app/**/build/ (66 MB compiled iOS
  framework), .remember/ (session data)

Impact: the project can now be maintained on the self-hosted Gitea
forge with a clean, secret-free initial history.
2026-07-04 12:32:28 +02:00

42 lines
1.1 KiB
Docker

# soh-llm: Qwen2.5-7B diagnostic inference service
# Runs on kxkm-ai (RTX 4090 24 GB)
FROM nvidia/cuda:12.4.1-runtime-ubuntu22.04
ENV DEBIAN_FRONTEND=noninteractive
ENV PYTHONUNBUFFERED=1
# System dependencies
RUN apt-get update && apt-get install -y --no-install-recommends \
python3.12 python3.12-venv python3-pip git curl \
&& rm -rf /var/lib/apt/lists/*
# Create app user
RUN useradd -m -s /bin/bash llm
USER llm
WORKDIR /app
# Python dependencies
COPY requirements.txt .
RUN python3.12 -m pip install --user --no-cache-dir -r requirements.txt
# Application code
COPY config.py prompt_template.py inference_server.py diagnostic_api.py ./
# Model volume mount point
VOLUME /models
# Environment defaults
ENV LLM_BASE_MODEL=Qwen/Qwen2.5-7B
ENV LLM_LORA_ADAPTER=/models/qwen-bmu-diag/lora-adapter
ENV LLM_API_HOST=0.0.0.0
ENV LLM_API_PORT=8401
ENV INFLUXDB_URL=http://influxdb:8086
EXPOSE 8401
HEALTHCHECK --interval=30s --timeout=10s --retries=3 \
CMD curl -f http://localhost:8401/health || exit 1
CMD ["python3.12", "-m", "uvicorn", "diagnostic_api:app", "--host", "0.0.0.0", "--port", "8401", "--workers", "1"]