Deploy
A catalog model, or any vLLM-servable Hugging Face repo.
Catalog model
Any HF model
Model
Select a catalog model
GPU
defaults to the model's recommended GPU
Select a GPU
Deploy
non-blocking; a daemon drives it (run
gpu up
)