Photon Logo

Photon

Site Reliability Engineer - IN

Reposted One Month Ago
Be an Early Applicant
Remote
Hiring Remotely in India
Mid level
Remote
Hiring Remotely in India
Mid level
Maintain and monitor production systems to ensure uptime and performance. Build automation and infrastructure tooling, analyze metrics, support large-scale distributed applications, partner with development teams on testing and releases, and drive reliability improvements and capacity planning.
The summary above was generated by AI

SRE Engineer is responsible for ensuring website uptime, optimizing performance, and maintaining security of the production application. This role involves monitoring site reliability, addressing technical issues, automating maintenance tasks, and collaborating with cross-functional teams to meet business objectives. 

Responsibilities 

Run the production environment by monitoring availability and taking a holistic view of system health 

Build software and systems to manage platform infrastructure and applications 

Improve reliability, quality, and time-to-market of our suite of software solutions 

Measure and optimize system performance, with an eye toward pushing our capabilities forward, getting ahead of customer needs, and innovating for continual improvement 

Provide primary operational support and engineering for multiple large-scale distributed software applications 

Gather and analyze metrics from operating systems as well as applications to assist in performance tuning and fault finding 

Partner with development teams to improve services through rigorous testing and release procedures 

Participate in system design consulting, platform management, and capacity planning 

Create sustainable systems and services through automation and uplifts 

Balance feature development speed and reliability with well-defined service-level objectives 

Required skills and qualifications 

Bachelor’s degree (or equivalent) in computer science or related discipline 

Experience in SRE, DevOps, or similar roles. 

Expertise in monitoring tools, infrastructure management, and automation. 

Strong problem-solving skills and a collaborative mindset. 

Proactive approach to identifying problems, performance bottlenecks, and areas for improvement 

Similar Jobs

22 Days Ago
Remote or Hybrid
Senior level
Senior level
Digital Media • eCommerce • Gaming • Mobile • News + Entertainment
Lead reliability, scalability, observability, automation, infrastructure, disaster recovery, and security initiatives for Crunchyroll’s cloud-native data platforms. Establish SRE practices including SLIs, SLOs, error budgets, incident management, and postmortems. Operate Kubernetes and GCP environments, implement Infrastructure as Code, optimize capacity and performance, and drive vulnerability remediation, penetration-testing support, and cloud platform security.
Top Skills: Ci/CdDatadogGCPGoGrafanaIdentity And Access ManagementInfrastructure As CodeJavaKubernetesLinuxOpentelemetryOwasp Top 10PrometheusPythonShellTerraform
Yesterday
Remote
India
Senior level
Senior level
HR Tech • Legal Tech • Software • Consulting
Own and evolve Mitratech’s AWS DevOps and platform engineering practice, including multi-platform CI/CD, Terraform and AWS CDK infrastructure as code, AWS compute, networking, databases, security controls, and AI/analytics operations. Design reusable pipelines, enforce quality and security gates, manage multi-account environments, improve automation reliability, mentor engineers, and participate in on-call support.
Top Skills: Account Factory For TerraformAlbAmazon LinuxAnsibleAuroraAWSAws CdkAws ConfigAws OrganizationsAzureBedrockBitbucket PipelinesCfn-GuardCheckovCloudtrailControl TowerDirect ConnectEc2Ec2 Image BuilderEcrEcs FargateEksEventbridgeGithub ActionsGuarddutyHashicorp VaultIam Access AnalyzerJenkinsLambdaLinearbLinuxMacieNlbOpentofuPackerPrivatelinkPythonQuicksightRdsRedshiftRhelRoute 53S3Secrets ManagerSecurity HubSleuthSnsSqsSsm Patch ManagerStep FunctionsTerraformTransit GatewayTypescriptUbuntuVpcVpn
6 Days Ago
Remote
IND
Senior level
Senior level
eCommerce • Marketing Tech • Design • SEO
Designs and operates cloud-native infrastructure, automated CI/CD pipelines, infrastructure-as-code, Kubernetes platforms, observability systems, and secure configuration management. The role focuses on high availability, disaster recovery, cluster scaling, fault tolerance, incident response, and root-cause analysis. Candidates need substantial DevOps or systems engineering experience, advanced Kubernetes and Terraform expertise, Linux and scripting knowledge, and a mandatory AWS, Azure, or Kubernetes certification.
Top Skills: Amazon EksArgocdAWSAws Secrets ManagerAzureAzure AksBashDastDatadogElk StackFlaggerGithub ActionsGitlab CiGitopsGoGoogle GkeGrafanaHashicorp VaultIstioJenkinsKubernetesLinuxOpentofuPrometheusPythonSastTerraform

What you need to know about the Mumbai Tech Scene

From haggling for the best price at Chor Bazaar to the bustle of Crawford Market, the energy of Mumbai's traditional markets is a key part of the city's charm. And while these markets will always have their place, the city also boasts a thriving e-commerce scene, ranking among the largest in the region. Driven by online sales in everything from snacks to licensed sports merchandise to children's apparel, the local industry is worth billions, with companies actively recruiting to meet the demands of continued growth.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account