I let Codex CLI drive the whole thing from spinning up a fresh VM on a Proxmox node
to passing through the GPU & updating drivers to setting up Hermes + llama.cpp inference
and putting the model online to giving me access through Discord Amazing what agents can do for us now
AI agent fully deploys GPU inference server via Discord
By
–
