Instancias dedicadas de GPU con JupyterLab, ComfyUI, servicio vLLM y plantillas de terminal web de un solo clic. Facturado por segundo, solo mientras la instancia está en funcionamiento, a velocidades de $0.65/hr.
La facturación se realiza por segundo a la tarifa horaria indicada, solo mientras la instancia está en funcionamiento. Haz clic en una GPU para ver las especificaciones completas, la disponibilidad y las opciones de despliegue con un solo clic.
Billing is per second at the listed hourly rate, and only while the instance is running. The rate is locked in when you deploy, and stopping or destroying the instance stops the charge.
One-click templates cover JupyterLab notebooks, ComfyUI, vLLM model serving (bring a Hugging Face model id), and a browser web terminal. You connect through the authenticated EmpirioLabs connect endpoint or call the workload through /v1/gpu/connect/{instance_id}/{path} on the API.
Yes. Everything the dashboard does is also available through the API: deploy, stop, and destroy instances under /v1/gpu on api.empiriolabs.ai, and reach the running workload through the connect endpoint. The full reference is in the GPU Cloud docs.
Runtime storage targets range from 100 to 300 GB with a 150 GB default, bundled into the displayed hourly price.
Create an EmpirioLabs account, open GPU Cloud in the dashboard, pick a GPU and template, and deploy. Billing is pay-as-you-go credits.
Check out our pricing or reach out if you want your own model deployed on our stack.