Sign up to save this job, get alerts, and apply with an optimized CV.

Senior Site Reliability Engineer - Cloud Platform

Full Time Remote Senior

Job description

Ideal Candidate •5+ years of experience in Site Reliability Engineering, Platform Engineering, Infrastructure Engineering, or a similar role supporting production AWS environments. •Strong Python development skills with experience building automation, services, or platform tooling. •Hands-on experience with Infrastructure as Code, including CloudFormation and/or AWS CDK. •Experience operating core AWS services including IAM, VPC networking, EC2, Lambda, and managed storage or database services. •Strong Linux systems administration, troubleshooting, incident response, and production operations experience. You might also have... •Experience operating AWS environments across multiple accounts or AWS Organisations. •Experience with AWS networking technologies such as Transit Gateway, IPAM, or BYOIP.Cloud cost optimisation or FinOps experience. •Experience with CI/CD pipelines and GitOps-style deployment practices. •Familiarity with AI-assisted engineering workflows and tooling. •Exposure to policy-as-code, compliance automation, or cloud governance tooling. We encourage you to apply even if your experience or skillset doesn’t align perfectly with every requirement. We value a wide range of backgrounds and transferable skills, and we are excited to support learning and growth. Job Description Global Compute builds and operates the core cloud infrastructure that engineering teams rely on every day. We provision and manage AWS accounts across the company, operate the network backbone that connects them, and maintain the security guardrails that keep those environments safe, compliant, and scalable. We believe reliability is an engineering challenge, not an operations task. We automate repetitive work, build for scale before it becomes a problem, and invest heavily in observability to identify issues before they impact the business. What you'll get to do... •Operate and scale AWS production infrastructure, owning the health of services that provision, secure, and manage accounts across GoDaddy AWS organisations. •Design, build, and maintain cloud platform capabilities using Python, CloudFormation, AWS CDK, and automation-first practices. •Drive cost optimisation initiatives that improve efficiency and deliver measurable business impact.Improve observability through monitoring, alerting, dashboards, and operational tooling. •Participate in on-call rotations, lead incident response efforts, and drive long-term reliability improvements through blameless post-incident reviews. •Support strategic AWS initiatives across networking, identity, governance, and multi-account architecture. •Review code and designs, contribute documentation and operational runbooks, and mentor fellow engineers. •Leverage AI-assisted tooling to improve engineering productivity, accelerate automation, and reduce operational toil. null Company Description At GoDaddy, the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely. Remote: This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings.

Sign up to apply

Create a free account to apply for this job and get access to:

  • AI-powered CV optimization for this specific job
  • Save jobs and create custom alerts
  • See your CV match score for each job

Company information

Company
GoDaddy
Location
Full time
Romania
Posted
9 hours ago

Find similar jobs

Explore more opportunities like this one.

Interested in this position?

Create your free account and tailor your CV to match this job.