Design, architect, and build a multi-tenant bare-metal OpenShift cluster; define networking, storage, ingress, monitoring, and logging; produce architecture docs and SOPs; lead migration of on-prem and Azure workloads; enforce RBAC, tenant isolation, quotas and compliance; optimize GPU workloads; own platform stability, upgrades, and troubleshooting.
Role Overview
We are building a high-performance, multi-tenant OpenShift cluster on bare metal in our AI-optimized data center. Our goal is to offer OpenShift as a Service to private AI-focused clients who wish to host their compute-intensive workloads in a scalable, secure, and isolated environment.
We’re looking for a hands-on Red Hat OpenShift Engineer who can design, architect, and implement this platform from scratch using industry best practices. Once the cluster is built, this role will also lead the migration of existing on-prem and Azure workloads to the new OpenShift environment.
Key Responsibilities
- Design, architect, and build a multi-tenant OpenShift cluster on bare metal
- Configure and maintain all aspects of the OpenShift platform for high availability, scalability, and security
- Define and implement networking, storage, ingress, monitoring, and logging
- Develop detailed architecture documents, blueprints, and SOPs
- Lead and execute migration of existing workloads from on-premise and Azure environments to OpenShift
- Ensure smooth onboarding for multiple AI-focused client tenants with isolated environments
- Support DevOps team with advanced Linux/OpenShift troubleshooting
- Implement and enforce RBAC, tenant isolation, resource quotas, and compliance controls
- Optimize performance for AI-heavy workloads running on GPU-enabled infrastructure
- Own operational stability, platform upgrades, and monitoring
Required Skills & Experience
- 5+ years of deep hands-on experience with Red Hat OpenShift and Kubernetes
- Proven experience designing, building, and managing bare metal OpenShift clusters
- Solid understanding of Linux internals, container runtimes, networking, and troubleshooting
- Experience with application migration from both on-premise and Azure environments into OpenShift
- Strong experience with multi-tenancy architecture, including workload isolation and security
- Familiarity with storage (CSI), networking (CNI), and service mesh implementations
- Proficiency in monitoring and observability tools (e.g., Prometheus, Grafana, ELK)
- Experience with Infrastructure as Code (Ansible, Terraform) and CI/CD automation
- Strong documentation and communication skills
Certifications (Required)
- Red Hat Certified Specialist in OpenShift Administration
- Red Hat Certified Engineer (RHCE) or equivalent Linux certification
Nice to Have
- Familiarity with AI/ML compute environments (e.g., GPU workloads, NVIDIA operators)
- Experience with hybrid cloud or edge computing models
- Exposure to enterprise-grade security, compliance, and policy enforcement
Similar Jobs
Artificial Intelligence • Hardware • Information Technology • Machine Learning
Leads SSD assembly equipment engineering to improve tool performance, availability, yield, quality, reliability, and cost. Establishes SPC and FD controls, preventive and predictive maintenance programs, spare inventory systems, productivity initiatives, equipment capability roadmaps, supplier collaboration, and benchmarking. Builds continuous improvement teams and develops personnel toward a Center of Excellence model. Requires an engineering degree, at least eight years of PCB assembly engineering experience, leadership, analytical problem-solving, and AI application capabilities.
Top Skills:
Artificial IntelligenceCpk Process Capability AnalysisFailure Detection (Fd) MethodologyPcb Assembly EquipmentPredictive MaintenancePreventive MaintenanceSsd Assembly EquipmentStatistical Process Control (Spc)
Artificial Intelligence • Hardware • Information Technology • Machine Learning
Administers and supports enterprise physical security technologies, including Lenel OnGuard access control and Milestone XProtect video management systems. Troubleshoots integrated hardware, software, database, network, CCTV, and visitor management issues; supports investigations through system data analysis. Collaborates with IT, Facilities, Engineering, Security Operations, vendors, and integrators on upgrades and modernization projects while developing documentation, standardizing systems, and improving lifecycle management and operational processes.
Top Skills:
Access Control SystemsActive DirectoryCctvGenerative AiLenel OnguardMilestone XprotectSQL ServerTcp/Ip NetworkingVideo AnalyticsVideo SurveillanceWindows Server
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Leads incentive compensation strategy, quota-setting, recognition programs, analytics, governance, implementation, and continuous improvement across Pfizer’s international markets. Partners with business leaders and cross-functional teams to design compliant, motivating, and financially responsible programs, deliver data-driven recommendations, oversee payouts and operational timelines, and maintain audit readiness. The role requires strong quantitative analysis, commercial operations expertise, project management, stakeholder influence, and experience supporting pharmaceutical or healthcare organizations.
What you need to know about the Chennai Tech Scene
To locals, it's no secret that South India is leading the charge in big data infrastructure. While the environmental impact of data centers has long been a concern, emerging hubs like Chennai are favored by companies seeking ready access to renewable energy resources, which provide more sustainable and cost-effective solutions. As a result, Chennai, along with neighboring Bengaluru and Hyderabad, is poised for significant growth, with a projected 65 percent increase in data center capacity over the next decade.
.jpeg)
