DGX Spark MLOps Stack — Dual-Node LLM Pipeline

Built an end-to-end LLM cycle for internal use on two NVIDIA DGX Spark nodes (Production + R&D) — ClearML 10-step pipeline, NGC training containers, AWQ quantization, vLLM inference, an OpenAI-compatible router, and an operations agent that drives the stack from chat.

Architecture