Resources
Documentation, tutorials, and community resources to help you get the most out of GPU cloud computing.
Documentation
Complete reference for every endpoint, with request parameters, response schemas, and code examples.
Complete REST API documentation. GPU instances, inference endpoints, volumes, SSH keys, billing — every endpoint with request/response schemas.
Create an account, generate an API key, launch your first GPU instance, and SSH in — all in under 5 minutes. Step-by-step walkthrough for new users.
JWT-based auth with access and refresh tokens. API key authentication for programmatic access. Rate limits, token rotation, and security best practices.
Guides & Tutorials
Practical tutorials for deploying models, fine-tuning, and building production AI systems on GPU cloud.
Complete guide to running Moonshot AI's Kimi K2.5 (1T parameter MoE) locally on Ubuntu with Ollama, llama.cpp, or vLLM. Covers quantization, hardware requirements, and benchmarks.
Step-by-step guide to fine-tuning Meta LLaMA 3 on your custom dataset using LoRA adapters on a single A100 or H100 GPU. Includes dataset preparation and evaluation.
Set up distributed training across multiple GPU nodes using DeepSpeed ZeRO-3. Configure InfiniBand networking, NCCL, and checkpoint saving for large model training.
Run vLLM as a production inference server on gpuLabs. OpenAI-compatible API, continuous batching, PagedAttention, and auto-scaling configuration.
End-to-end retrieval-augmented generation: embed documents with BGE, index in a vector database, and serve answers with a fine-tuned LLM on dedicated GPU.
Batch image generation with SDXL on high-VRAM GPUs. API setup, prompt engineering, ControlNet integration, and scaling to thousands of images per hour.
Quick Links
Community & Support
Connect with other GPU cloud users, browse open-source examples, or contact our engineering team directly.
Join 2,000+ GPU cloud users. Get help with deployments, share configs, and get notified about GPU availability and new features.
Open-source example notebooks, deployment scripts, Docker configurations, and integration templates for common AI frameworks.
Create support tickets directly from your dashboard. Our engineering team responds within 2 hours during business hours. Priority support for enterprise accounts.
Integrations
Start with the API docs, follow a tutorial, or launch your first GPU instance.