Companies

Cerebrium is a serverless infrastructure platform for AI applications [3] headquartered in New York City [12]. The company was accepted into Y Combinator's Winter 2022 batch [6]. Cerebrium offers serverless GPUs with low cold start times and over 12 varieties of GPU chips [4], enabling customers to run large-scale batch jobs and real-time voice applications [4]. The platform is used by teams at Tavus, CivitAI, and Twilio [4]. According to the company, customers typically experience 40 percent cost savings compared to traditional cloud providers and can scale models to more than 10,000 requests per minute with minimal engineering overhead [4]. The company has a team of 4 people [1] and maintains active status [7].

About

  • Cerebrium is a serverless infrastructure platform for AI applications. We make it easier for companies to build and deploy AI based applications. We offer Serverless GPU's with low cold start times, over 12 varieties of GPU chips, allow you to run large scale batch jobs, run realtime voice applications and much more. We are used by the teams at Tavus, CivitAI, Twilio and many more.  Customers typically experience 40% in cost savings when compared to using traditional cloud providers and can scale models to more than 10K requests per minute with minimal engineering overhead. [4]
  • Cerebrium is a serverless infrastructure platform for AI applications. We make it easier for companies to build and deploy AI based applications. We offer Serverless GPU's with low cold start times, over 12 varieties of GPU chips, allow you to run large scale batch jobs, run realtime voice applications and much more. We are used by the teams at Tavus, CivitAI, Twilio and many more.  Customers typically experience 40% in cost savings when compared to using traditional cloud providers and can scale models to more than 10K requests per minute with minimal engineering overhead.[10]
  • Cerebrium is a serverless infrastructure platform for AI applications. We make it easier for companies to build and deploy AI based applications. We offer Serverless GPU's with low cold start times, over 12 varieties of GPU chips, allow you to run large scale batch jobs, run realtime voice applications and much more. We are used by the teams at Tavus, CivitAI, Twilio and many more.  Customers typically experience 40% in cost savings when compared to using traditional cloud providers and can scale models to more than 10K requests per minute with minimal engineering overhead. [19]

Something wrong or missing? Send a correction. Fixed within 24 hours.