Rackspace Technology
San Antonio, Texas, United States
10,000+ Employees
AI Model Serving Specialist
India - Remote, IN
Posted on Feb 26, 2026
Responsibilities
Enable enterprise customers to operationalize AI workloads by deploying and optimizing model-serving platforms (e.g., NVIDIA Triton, vLLM, KServe) within Rackspace’s Private Cloud and Hybrid environments. This role bridges AI engineering and platform operations, ensuring secure, scalable, and cost-efficient inference services.
Requirements
Hands-on experience with NVIDIA Triton, vLLM, or similar serving stacks. Strong knowledge of Kubernetes, GPU scheduling, and CUDA/MIG.Familiarity with VMware VCF9, NSX-T networking, and vSAN storage classes. Proficiency in Python and containerization (Docker).Understanding of observability stacks (Prometheus, Grafana) and FinOps principles.
Exposure to RAG architectures, vector DBs, and secure multi-tenant environments. Excellent problem-solving and customer-facing communication skills.
About Rackspace Technology
Rackspace Technology is a leading end-to-end multicloud technology solutions company. With a history dating back to 1998, Rackspace has been a trailblazer in helping businesses navigate and optimize their cloud environments. The company's comprehensive suite of services, including cloud management, security, and application modernization, has positioned Rackspace as a trusted partner for organizations looking to harness the power of the cloud for innovation and growth.
Job overview
Employment type
Full Time
Experience level
Mid Level
Work type
Remote
Department
Private Cloud - Delivery
Posted
on Feb 26, 2026
Financials
Type
public
Annual Revenue
--
Market Cap
$433.0M+
Subscribe to Refer Me premium to get referred to this role at Rackspace Technology!
