ModelRefs / RunPod — AI Glossary

RunPod — AI Glossary

A GPU cloud marketplace offering on-demand and spot GPU rentals for LLM training, fine-tuning, and inference at competitive pricing.

Overview

RunPod provides bare-metal-like GPU pods (H100, A100, RTX 4090) billed per second with persistent volumes. Serverless endpoints allow deploying custom container images with autoscaling. Community cloud (consumer GPUs) is the lowest-cost option; secure cloud uses data-center hardware. Popular for LoRA fine-tuning and hosting Stable Diffusion.

Reference details

Topicinfrastructure
Last reviewed2026-06-24

Commonly confused with

Raw GPU rental rather than a managed inference service: you get the machine and run whatever you like on it. That is the opposite end of the spectrum from a hosted API, and the practical difference is who handles serving, scaling and failure — here, you do. Spot and on-demand pricing is the other axis, and spot capacity can be reclaimed mid-run, which matters for training more than for inference.

Continue your research

Use these connected ModelRefs sections to compare alternatives, inspect implementation paths, and review the evidence and governance boundaries relevant to RunPod — AI Glossary.

Frequently asked questions

What is RunPod?

A GPU cloud marketplace offering on-demand and spot GPU rentals for LLM training, fine-tuning, and inference at competitive pricing.

What concepts are related to RunPod?

Closely related concepts include modal, replicate, serverless inference.