This guide outlines AMAX's services for deploying on-premises AI infrastructure. It details the shift from cloud-based solutions to on-prem systems driven by rising costs, data security concerns, and the need for predictable performance. The document presents various deployment options, including turnkey data center solutions, colocation services, and on-site hosting, emphasizing the importance of integrated compute, cooling, and cluster monitoring. Additionally, it describes the design of GPU clusters based on NVIDIA reference architectures to optimize performance for AI and HPC workloads. The guide also addresses power, space, and cooling considerations necessary for high-density AI deployments, highlighting AMAX's expertise in managing the full deployment process. This includes installation, configuration, and optimization of systems to ensure efficient operation and minimal disruption during the transition from cloud to on-premises infrastructure.