Get started
Deploy CosmicAC on your host machine, then create your first GPU Container Job or Managed Inference Job.
To start using CosmicAC, set up your deployment, install the CLI, and create your first job.
Set up CosmicAC
Deploy the CosmicAC stack on your host machine. After deployment, CosmicAC connects to your GPU Kubernetes cluster.
Complete these steps in order.
- Prepare the cluster: confirm that the cluster meets the Kubernetes, GPU, virtualization, storage, and registry requirements. See Requirements.
- Deploy the stack: deploy CosmicAC on your host machine with Docker Compose. See Deploy CosmicAC.
- Set up recommended model configurations: add a recommended model configuration for each model you plan to serve. See Set up recommended model configurations.
Install the CLI
To create and manage jobs from your terminal, install the CosmicAC CLI.
Create your first job
Create a job with the CLI or in the web interface.
GPU Container Job
You can also create a GPU Container Job in the web interface.
Managed Inference Job
You can also create a vLLM Managed Inference Job or a Parakeet Managed Inference Job in the web interface.
Call a Managed Inference endpoint
Create an API key, then send requests to a vLLM endpoint or transcribe audio with a Parakeet endpoint.