Work
MODEL-RUNNER / VLLM
High-throughput GPU serving for open models.
model-runnerfreeself-host
MACHINE-READABLE
GET /api/v1/resources/vllm