Already filled

Don't miss the next one. Get matching roles delivered to your inbox.

Kubernetes / OpenShift AI Platform Engineer

Job summary

Chandler

Work model

Hybrid · 2 days home
1 month ago
Job description

Overview

We are seeking a Kubernetes / OpenShift AI Platform Engineer to design, build, and optimize enterprise-scale infrastructure supporting advanced AI/ML workloads. This role sits at the intersection of platform engineering, DevOps, and AI infrastructure, enabling model development, training, and real-time inference in a highly regulated environment.

You will work cross-functionally with AI/ML engineers, data scientists, DevOps, and infrastructure teams to deliver scalable, secure, and high-performance AI platforms.

Key Responsibilities

  • Design and manage Kubernetes and OpenShift clusters at enterprise scale
  • Build and optimize infrastructure for AI/ML model training and inference workloads
  • Develop automation for deployment, configuration, patching, and platform operations using Python
  • Support GPU-enabled workloads and high-performance compute environments
  • Implement and maintain CI/CD pipelines, GitOps workflows, and infrastructure-as-code (Terraform)
  • Ensure platform reliability, scalability, and performance optimization
  • Implement security best practices including RBAC, network policies, and secrets management
  • Enable observability through Prometheus, Grafana, and logging frameworks
  • Collaborate with engineering teams to standardize and streamline AI platform environments

Required Qualifications

  • 5-7+ years of experience with Kubernetes (production environments)
  • Strong experience with Red Hat OpenShift in enterprise environments
  • 5-7+ years of hands-on experience with Docker and containerization technologies
  • Strong proficiency in Python for automation and platform engineering
  • Solid experience working in Linux environments (systems, networking, storage)
  • Experience with AWS or other cloud platforms
  • Hands-on experience with Terraform and CI/CD tools (e.g., Jenkins)
  • Experience supporting AI/ML platforms, model deployment pipelines, or similar workloads

Preferred Qualifications

  • Experience with AI/ML frameworks such as PyTorch, TensorFlow, Triton Inference Server, and vLLM
  • Experience with agentic AI systems or intelligent agents
  • Familiarity with Kubernetes Operators, Helm, and GitOps practices
  • Strong understanding of observability (Prometheus, Grafana) and security models (SCCs, RBAC)

Work Environment & Benefits

  • Location: Chandler, AZ (Hybrid - On-site 3 Days a Week)
  • Pay: $70.00 - $89.00/hr
  • Benefits: Medical, dental, vision, 401(k), disability, and more (eligibility requirements apply)
  • Contract: This is a contract position.

About TEKsystems

We're partners in transformation. We help clients activate ideas and solutions to take advantage of a new world of opportunity. We are a team of 80,000 strong, working with over 6,000 clients, including 80% of the Fortune 500.

TEKsystems is an equal opportunity employer.