Cerebrium

Cerebrium runs voice agents, LLM inference, and image or video pipelines on autoscaling serverless GPUs. You point it at a Python entry point or Dockerfil

Cerebrium