ModelRefs / RunPod — AI Glossary
RunPod — AI Glossary
A GPU cloud marketplace offering on-demand and spot GPU rentals for LLM training, fine-tuning, and inference at competitive pricing.
Overview
RunPod provides bare-metal-like GPU pods (H100, A100, RTX 4090) billed per second with persistent volumes. Serverless endpoints allow deploying custom container images with autoscaling. Community cloud (consumer GPUs) is the lowest-cost option; secure cloud uses data-center hardware. Popular for LoRA fine-tuning and hosting Stable Diffusion.
Reference details
| Topic | infrastructure |
|---|---|
| Last reviewed | 2026-06-24 |
Related terms
Commonly confused with
Raw GPU rental rather than a managed inference service: you get the machine and run whatever you like on it. That is the opposite end of the spectrum from a hosted API, and the practical difference is who handles serving, scaling and failure — here, you do. Spot and on-demand pricing is the other axis, and spot capacity can be reclaimed mid-run, which matters for training more than for inference.
Continue your research
Use these connected ModelRefs sections to compare alternatives, inspect implementation paths, and review the evidence and governance boundaries relevant to RunPod — AI Glossary.
Frequently asked questions
What is RunPod?
A GPU cloud marketplace offering on-demand and spot GPU rentals for LLM training, fine-tuning, and inference at competitive pricing.
What concepts are related to RunPod?
Closely related concepts include modal, replicate, serverless inference.