Demos
Video

CoreWeave ARENA: Serverless and Dedicated Inference

Play video

CoreWeave offers two distinct managed inference execution paths. With Serverless Inference, you pick a model from the catalog and pay per token—CoreWeave manages the services. Dedicated Inference is pay-per-GPU-hour and offers more flexibility: bring your own model weights and choose your GPU and runtime.

CoreWeave ARENA lets you assess real workload performance and cost on our purpose-built AI cloud before you commit to production. In this demo, you’ll see how to try both Serverless Inference and Dedicated Inference in ARENA.