Language models
vLLM
vLLM inference server. OpenAI-compatible /v1/chat/completions endpoint.
Details
Image
vllm/vllm-openai:latest
Ports
8000/http, 22/tcp
Setup
3-4 minutes
Difficulty
advanced
Features
OpenAI-compatible API, PagedAttention, Continuous batching
Compatible GPUs
Start
locked
Hold any amount of $GPU to start a treasury-funded session. Every session is capped at 1 hour.
Runtime 1 hour (fixed)
GPU NVIDIA GeForce RTX 4090
App vLLM
Ports 8000/http,22/tcp
Cost ~$0.44 covered if eligible
Use your public key only. Create a key locally with ssh-keygen -t ed25519, then paste the contents of ~/.ssh/id_ed25519.pub.
Connect a wallet on Robinhood Chain.
My sessions & SSH access