CosmicAC Logo

Get started

Deploy CosmicAC on your host machine, then create your first GPU Container Job or Managed Inference Job.

To start using CosmicAC, set up your deployment, install the CLI, and create your first job.

Set up CosmicAC

Deploy the CosmicAC stack on your host machine. After deployment, CosmicAC connects to your GPU Kubernetes cluster.

Complete these steps in order.

  1. Prepare the cluster: confirm that the cluster meets the Kubernetes, GPU, virtualization, storage, and registry requirements. See Requirements.
  2. Deploy the stack: deploy CosmicAC on your host machine with Docker Compose. See Deploy CosmicAC.
  3. Set up recommended model configurations: add a recommended model configuration for each model you plan to serve. See Set up recommended model configurations.

Install the CLI

To create and manage jobs from your terminal, install the CosmicAC CLI.

Create your first job

Create a job with the CLI or in the web interface.

GPU Container Job

You can also create a GPU Container Job in the web interface.

Managed Inference Job

You can also create a vLLM Managed Inference Job or a Parakeet Managed Inference Job in the web interface.

Call a Managed Inference endpoint

Create an API key, then send requests to a vLLM endpoint or transcribe audio with a Parakeet endpoint.

On this page