Photon Logo

Photon

Senior GCP Cloud Engineer - IN

Posted Yesterday
Be an Early Applicant
Remote
Hiring Remotely in India
Senior level
Remote
Hiring Remotely in India
Senior level
Designs, governs, and supports secure, scalable GCP platforms across multiple environments. Leads cloud architecture, Terraform-based infrastructure automation, Kubernetes and container deployments, networking, reliability engineering, monitoring, incident response, disaster recovery, and production readiness reviews. Establishes cloud standards, security controls, SLIs and SLOs, and mentors engineering teams. Collaborates with architecture, development, QA, security, networking, and offshore teams to improve automation, performance, availability, and operational excellence.
The summary above was generated by AI

Senior GCP Cloud Engineer 

Role Overview 

We are seeking a highly skilled Senior GCP Cloud Engineer to design, implement, govern, and support scalable, secure, highly available, and automated cloud platforms on Google Cloud Platform (GCP). The role requires deep technical expertise in GCP services, Infrastructure as Code (IaC), cloud reliability engineering, networking, containerization, and production operations. 

The ideal candidate will act as a senior technical leader responsible for defining reference architectures, driving platform reliability, governing cloud standards across multiple projects, and mentoring engineering teams while collaborating closely with architects, development teams, and stakeholders. 

Key Responsibilities 

Cloud Infrastructure & Platform Engineering 

  • Design, implement, and govern scalable, secure, and highly available cloud infrastructure across GCP environments (Development, Testing, UAT, Pre-Production, Production). 

  • Define and maintain reusable reference architectures and platform standards for GCP services and cloud-native deployments. 

  • Build enterprise-grade architectures using GCP services such as: 

  • GKE, Cloud Run, Cloud Functions 

  • BigQuery, Bigtable, Cloud SQL, Spanner 

  • Cloud Storage, Pub/Sub, Memorystore, Artifact Registry 

  • Configure and manage networking components including: 

  • VPCs, Shared VPC, Private Service Connect, VPC Peering 

  • Load Balancers, Cloud DNS, Cloud Armor, Firewall Rules 

  • VPN and Interconnect connectivity solutions 

  • Drive standardization and consistency of cloud platform implementations across multiple projects and teams. 

Reliability Engineering & Production Operations 

  • Lead reliability engineering initiatives to improve platform stability, resiliency, scalability, and fault tolerance. 

  • Establish and enforce operational best practices for monitoring, alerting, incident management, and disaster recovery. 

  • Conduct production readiness reviews (PRRs) for new applications, services, and platform changes prior to production deployments. 

  • Perform capacity planning and infrastructure sizing to ensure platform scalability and performance under varying workloads. 

  • Define and track platform SLIs, SLOs, and operational health metrics. 

  • Lead root cause analysis (RCA) and resolution of critical production incidents and complex infrastructure issues. 

  • Drive continuous improvement initiatives focused on reliability, availability, performance, and operational excellence. 

Infrastructure as Code (IaC) 

  • Develop, enhance, and maintain Terraform modules and reusable infrastructure templates. 

  • Ensure consistency, reliability, and compliance across environments using Infrastructure as Code best practices. 

  • Drive automation for infrastructure provisioning, configuration management, and deployment processes. 

  • Collaborate with platform and Terraform engineering teams to improve reusable IaC standards and governance. 

 

Containerization & Microservices 

  • Build and manage containerized applications using Docker. 

  • Deploy and manage workloads on Kubernetes (GKE) and Cloud Run platforms. 

  • Support cloud-native and microservices-based application architectures. 

  • Collaborate with application teams to optimize container orchestration, deployment reliability, and scalability. 

 

Monitoring, Troubleshooting & Incident Management 

  • Monitor platform health using Cloud Monitoring, Cloud Logging, Dynatrace, and related observability tools. 

  • Troubleshoot and resolve complex infrastructure, networking, deployment, and performance issues. 

  • Act as the senior technical escalation point for critical GCP platform related challenges. 

  • Support application teams in diagnosing platform-related production issues and performance bottlenecks. 

Security, Governance & Compliance 

  • Implement and enforce IAM policies, security controls, and governance standards across GCP environments. 

  • Ensure secure networking, data protection, and compliance best practices are consistently followed. 

  • Govern cloud resource usage, security posture, and platform standards across multiple projects and environments. 

  • Collaborate with security and compliance teams to support audits, risk remediation, and governance initiatives. 

Technical Leadership 

  • Provide technical leadership and mentorship to junior cloud and DevOps engineers. 

  • Guide engineering teams on GCP best practices, cloud-native architectures, automation strategies, and operational excellence. 

  • Participate actively in technical planning discussions. 

  • Maintain and improve technical documentation, architectural standards, and operational runbooks. 

 

Collaboration & Agile Delivery 

  • Participate actively in Agile/Scrum ceremonies (stand-ups, sprint planning, retrospectives). 

  • Collaborate with:  

  • Platform leads and architects 

  • Development and QA teams 

  • Security and networking teams 

  • Onsite and offshore stakeholders 

  • Maintain and update technical documentation in Confluence and other tools. 

Continuous Improvement & Innovation 

  • Stay up to date with the latest GCP innovations, cloud-native technologies, DevOps practices, and industry trends. 

  • Identify opportunities for automation, optimization, performance tuning, and cost efficiency. 

  • Drive innovation in cloud platform engineering, observability, reliability engineering, and deployment automation. 

Minimum Qualifications 

  • Bachelor’s or Master’s degree in Computer Science, Information Technology, or a related field. 

  • 6+ years of overall IT experience with minimum 4+ years of hands-on GCP experience. 

  • Strong hands-on experience with: 

  • Google Cloud Platform (GCP) 

  • Terraform (Infrastructure as Code) 

  • Docker & containerization 

  • Kubernetes (GKE) 

  • CI/CD tools such as Jenkins and SonarQube 

  • Scripting (Bash/Shell, Groovy) 

  • Strong understanding of: 

  • GCP architecture and cloud-native services 

  • Reliability engineering and production operations 

  • Networking (VPC, VPN, Interconnect, PSC) 

  • Security, IAM, and governance 

  • Monitoring, logging, and observability 

  • Familiarity with Atlassian tools (JIRA, Confluence, Bitbucket). 

  • Strong analytical, troubleshooting, communication, and leadership skills. 

 

Preferred / Nice-to-Have Skills 

  • GCP certifications: 

  • Professional Cloud Architect 

  • Professional Cloud DevOps Engineer 

  • Experience with: 

  • Apigee API management 

  • Microservices architecture 

  • Hybrid cloud environments (On-Premises + GCP) 

  • Reliability engineering / SRE practices 

  • Production readiness review processes 

  • Capacity planning and performance engineering 

  • Exposure to: 

  • Akamai integrations 

  • Dynatrace and advanced observability platforms 

  • GitOps and deployment automation tools 

  • Knowledge of build tools such as NPM and Gradle. 

Key Competencies 

  • Strong hands-on technical expertise 

  • Cloud architecture and platform engineering leadership 

  • Reliability engineering and operational excellence mindset 

  • Advanced troubleshooting and incident management skills 

  • Automation and DevOps orientation 

  • Technical mentorship and cross-team collaboration 

  • Strategic thinking and problem-solving ability 

  • Clear communication and documentation skills 

  • Adaptability in fast-paced enterprise environments 

Similar Jobs

An Hour Ago
In-Office or Remote
Senior level
Senior level
Cloud • Information Technology • Productivity • Security • Software • App development • Automation
Design, build, and operate Atlassian’s foundational data platform, including Flink streaming ingestion and MPP query systems such as StarRocks. The role owns architecture, production reliability, performance, cost, data quality, observability, code reviews, complex problem resolution, cross-team initiatives, and mentoring. Candidates need strong distributed-systems and data-engineering expertise, coding skills in Java, Scala, or Python, cloud infrastructure experience, and operational experience with high-scale platforms.
Top Skills: Apache DorisApache FlinkApache PinotSparkAWSGCPJavaPythonScalaStarrocksTrino
An Hour Ago
In-Office or Remote
Mid level
Mid level
Cloud • Information Technology • Productivity • Security • Software • App development • Automation
Build full-stack software features, REST APIs, customer-facing React components, backend monitoring, CI/CD pipelines, content management integrations, and reliable services. The role requires collaboration with engineering, UX, and visual design teams, plus experience with JavaScript/TypeScript, Java, React, Node.js, cloud platforms, REST integrations, enterprise content management systems, and related server technologies.
Top Skills: AWSCi/CdCSSDockerEnterprise Content Management SystemsEs6ExpressFreemarkerGitHTMLJavaJavaScriptNginxNode.jsNpmReactRest ApisRustSassTypescript
Yesterday
Remote
Gujarat, IND
Senior level
Senior level
Artificial Intelligence • Hardware • Information Technology • Machine Learning
Lead equipment engineering for assembly manufacturing to improve safety, quality, reliability, productivity and cost. Drive equipment strategy, lifecycle management, automation, Lean/Six Sigma improvements, and AI-enabled predictive maintenance and analytics. Build team capability, deploy Industry 4.0 solutions, and partner across functions to meet OEE, downtime, and productivity targets.
Top Skills: AIApcCopilot TechnologiesData AnalyticsDigital TwinFdcGenerative AiIndustry 4.0Intelligent AutomationIotMachine LearningManufacturing IntelligenceMesPredictive MaintenanceSpcStatistical Process Control

What you need to know about the Mumbai Tech Scene

From haggling for the best price at Chor Bazaar to the bustle of Crawford Market, the energy of Mumbai's traditional markets is a key part of the city's charm. And while these markets will always have their place, the city also boasts a thriving e-commerce scene, ranking among the largest in the region. Driven by online sales in everything from snacks to licensed sports merchandise to children's apparel, the local industry is worth billions, with companies actively recruiting to meet the demands of continued growth.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account