vllm-warden
Self-hosted, OpenAI-compatible LLM inference with a guided setup wizard. Deploy any HuggingFace model in minutes — no command-line tuning required.
Self-hosted, OpenAI-compatible LLM inference with a guided setup wizard. Deploy any HuggingFace model in minutes — no command-line tuning required.