Replicate

Run thousands of open-source AI models by the second

replicate.com · Per-second GPU billing

Replicate runs open-source models (image, video, audio, LLMs) on managed GPUs, billed per second — the long tail of AI capability without infrastructure.

Capabilities

Capabilities

  • Thousands of community models, custom model deploys (Cog), per-second GPU billing

Best for

  • Niche model needs (restoration, music, specific fine-tunes) in clones

Not for

  • Mainstream chat AI (first-party APIs are cheaper and faster)