Skip to main content

AI for model hosting

Deploy and serve custom models on managed infrastructure.

#ToolCategoryPricingVisit
1Modal

Serverless Python and GPU cloud — deploy ML models, batch jobs, and APIs with decorator syntax

LLM FrameworksFreemiumVisit
2Baseten

ML model deployment platform — deploy any model as a production API in minutes

LLM FrameworksFreemiumVisit
3Predibase

Managed fine-tuning and LoRA serving — production-grade custom LLM deployment

LLM FrameworksFreemiumVisit
4Fireworks AI

Ultra-fast serverless inference for open-source LLMs — Llama, Mixtral, and SDXL at speed

LLM FrameworksFreemiumVisit
5Groq

Ultra-fast LLM inference API — run Llama, Mixtral, and Gemma at 500+ tokens/second on custom LPU hardware

LLM FrameworksFreemiumVisit

See also