Infographie de l'architecture IA cible : infrastructure matérielle, plateforme logicielle, stack IA, services et écosystème, avec caractéristiques clés et flux d'accès
Overview — target architecture of the platform (visual summary). The layer-by-layer detail and the status of each building block are below.

Overview of the platform's guiding thread, organized as under Proxmox: a hardware base, a library of templates, a permanent infrastructure (AI nodes and services), AI consumers (agents and automation), and a shared ZFS storage. The status of each building block distinguishes what is in service, what is proven on the test bench, and what is planned.

Active — in service on the Proxmox target Validated — proven on test bench / PoC, porting planned Planned — not yet created
1Hardware base & hypervisor— the machine
Serveur
AMD Ryzen 9 9950X3D · 192 Go DDR5
GPU
RTX 5090 · exclusive (1 VM GPU à la fois)
Stockage
ZFS rpool · miroir NVMe (~1,75 Tio)
Hyperviseur
Proxmox VE 9.2.2
↓
2ZFS Storage— shared rpool
rpool/models
Modèles IA
rpool/rag
Documents & index RAG
rpool/config
Configurations
rpool/backups
vzdump · sur le même pool (pas une sauvegarde indépendante)
rpool/nextcloud-data
Données Nextcloud
rpool/data
Disques des VM Proxmox
↓
3VM Templates— clonable bases
Templates 110 → 131
W11 · Ubuntu Desktop / Server · ± CUDA
140 Debian-Server
Template Debian · planifié
↓
4Permanent infrastructure— AI nodes & services
210 · IA-Core-CUDA
Ubuntu · Ollama · Open WebUI · RAG · modèles/config
260 · W11-CUDA
Alternative · banc de POC / validation rapide
215 · Services communs
PostgreSQL / pgvector · backend du RAG
205 · Nextcloud
Fichiers auto-hébergés (Docker + MariaDB)
200 · ReverseProxy
Nginx · TLS · Certbot · Fail2ban — publication HTTPS active
4bAI software building blocks— served by the nodes above
Ollama
Inférence LLM local · VM210
ComfyUI
Génération d'image (RealVisXL/SDXL) · VM210 (Linux) — distinct du POC Windows antérieur (VM260, Flux)
Open WebUI
Interface chat + RAG · VM210
RAG
VM210 · pgvector sur VM215 · Qdrant (repli)
Whisper
Transcription — validé · portage planifié sur VM210
↓
5AI consumers— agents & automation
300 · n8n
Orchestration de workflows · VM300 (autodémarrage)
301 · Hermes + OpenCode
Station agentique Linux (Ubuntu) · inférence déléguée à VM210 · Hermes (qwen3.6:35b) + OpenCode (qwen3-coder:30b) validés
302 · Station Dev IA (W11)
Poste W11 · VSCodium · Python 3.13 · Node.js 24 · DeepSeek Harness (préliminaire) · inférence GPU distante · validé
SOVEREIGNTY
Data and models hosted locally, no cloud
SECURITY
VM isolation · controlled exposure
MODULARITY
Decoupled services, replaceable block by block
SCALABILITY
Adding VMs, models and services on demand

Related pages

References & Sources

CategoryResourceURL
PositioningGeneral presentation of the platform — approach and purposeFiche interne « presentation-generale »
ArchitectureVM naming & ZFS datasets (adopted) · milestones J0 → J13Suivi interne
Maturity levelsActive · Validated (test bench) · Planned—
Content of this pageShared under CC BY-SA 4.0creativecommons.org/licenses/by-sa/4.0