Back to full agenda
Demo ARENA

CoreWeave Inference: From Model Training to Production-Scale AI Deployment

Session
900
Day & Date
Thursday, Oct. 1, 2026
Time
Location
Expo Theater

About this session

The live demonstration will showcase the end-to-end workflow of deploying and serving AI models using CoreWeave Inference platforms. The demo begins with a brief walkthrough of how trained models transition into production inference environments, followed by deploying a sample generative AI model across different CoreWeave inference options. Attendees will see: Serverless Inference for rapid deployment and automatic scaling of inference workloads; Dedicated Inference for consistent low-latency performance and predictable throughput; and Inference on CKS (CoreWeave Kubernetes Service) demonstrating containerized deployment, orchestration, and infrastructure customization for enterprise-grade AI applications. The demo will include real-time inference requests, scaling behavior, API interactions, monitoring insights, and performance comparisons across deployment types. The objective is to provide attendees with a practical understanding of how CoreWeave enables scalable, production-ready AI inference while balancing performance, flexibility, and operational efficiency.

Share this session

Get in the room. San Francisco, September 29.

Fully Connected 2026 is where the engineers, leaders, and operators running AI in production come together for three days of depth, access, and real conversation.