Design, architect, and build a multi-tenant bare-metal OpenShift cluster; define networking, storage, ingress, monitoring, and logging; produce architecture docs and SOPs; lead migration of on-prem and Azure workloads; enforce RBAC, tenant isolation, quotas and compliance; optimize GPU workloads; own platform stability, upgrades, and troubleshooting.
Role Overview
We are building a high-performance, multi-tenant OpenShift cluster on bare metal in our AI-optimized data center. Our goal is to offer OpenShift as a Service to private AI-focused clients who wish to host their compute-intensive workloads in a scalable, secure, and isolated environment.
We’re looking for a hands-on Red Hat OpenShift Engineer who can design, architect, and implement this platform from scratch using industry best practices. Once the cluster is built, this role will also lead the migration of existing on-prem and Azure workloads to the new OpenShift environment.
Key Responsibilities
- Design, architect, and build a multi-tenant OpenShift cluster on bare metal
- Configure and maintain all aspects of the OpenShift platform for high availability, scalability, and security
- Define and implement networking, storage, ingress, monitoring, and logging
- Develop detailed architecture documents, blueprints, and SOPs
- Lead and execute migration of existing workloads from on-premise and Azure environments to OpenShift
- Ensure smooth onboarding for multiple AI-focused client tenants with isolated environments
- Support DevOps team with advanced Linux/OpenShift troubleshooting
- Implement and enforce RBAC, tenant isolation, resource quotas, and compliance controls
- Optimize performance for AI-heavy workloads running on GPU-enabled infrastructure
- Own operational stability, platform upgrades, and monitoring
Required Skills & Experience
- 5+ years of deep hands-on experience with Red Hat OpenShift and Kubernetes
- Proven experience designing, building, and managing bare metal OpenShift clusters
- Solid understanding of Linux internals, container runtimes, networking, and troubleshooting
- Experience with application migration from both on-premise and Azure environments into OpenShift
- Strong experience with multi-tenancy architecture, including workload isolation and security
- Familiarity with storage (CSI), networking (CNI), and service mesh implementations
- Proficiency in monitoring and observability tools (e.g., Prometheus, Grafana, ELK)
- Experience with Infrastructure as Code (Ansible, Terraform) and CI/CD automation
- Strong documentation and communication skills
Certifications (Required)
- Red Hat Certified Specialist in OpenShift Administration
- Red Hat Certified Engineer (RHCE) or equivalent Linux certification
Nice to Have
- Familiarity with AI/ML compute environments (e.g., GPU workloads, NVIDIA operators)
- Experience with hybrid cloud or edge computing models
- Exposure to enterprise-grade security, compliance, and policy enforcement
Similar Jobs
Artificial Intelligence • Big Data • Logistics • Machine Learning • Software • Transportation
Own the Over-the-Road visibility product portfolio across truckload, LTL, parcel, and last mile. Analyze adoption, engagement, retention, tracking quality, and ETA accuracy using SQL and product analytics tools. Conduct customer discovery, prioritize roadmaps, build AI-assisted prototypes, and partner with engineering, design, customer success, sales, and marketing to ship measurable product improvements.
Top Skills:
AmplitudeClaudeCursorPendoSQL
Fintech • Professional Services • Consulting • Energy • Financial Services • Cybersecurity • Generative AI
Lead end-to-end transformation delivery and Agile ways-of-working improvements. Build and operationalize agentic AI capabilities, including Claude-style skills, workflows, playbooks, and reusable assets. Identify and incubate AI use cases that automate or improve requirements, planning, governance, reporting, testing, documentation, RAID management, and decision support. Collaborate with business analysts, project managers, PMO, IT managers, and transformation teams.
Top Skills:
Agentic AiAgileClaudeConfluenceJIRA
Cloud • Information Technology • Productivity • Security • Software • App development • Automation
Design, build, and operate Atlassian’s foundational data platform, including Flink streaming ingestion and MPP query systems such as StarRocks. The role owns architecture, production reliability, performance, cost, data quality, observability, code reviews, complex problem resolution, cross-team initiatives, and mentoring. Candidates need strong distributed-systems and data-engineering expertise, coding skills in Java, Scala, or Python, cloud infrastructure experience, and operational experience with high-scale platforms.
Top Skills:
Apache DorisApache FlinkApache PinotSparkAWSGCPJavaPythonScalaStarrocksTrino
What you need to know about the Mumbai Tech Scene
From haggling for the best price at Chor Bazaar to the bustle of Crawford Market, the energy of Mumbai's traditional markets is a key part of the city's charm. And while these markets will always have their place, the city also boasts a thriving e-commerce scene, ranking among the largest in the region. Driven by online sales in everything from snacks to licensed sports merchandise to children's apparel, the local industry is worth billions, with companies actively recruiting to meet the demands of continued growth.


