Sign up to save this job, get alerts, and apply with an optimized CV.
Senior Site Reliability Engineer - Cloud Platform
Full Time
Remote
Senior
Job description
Ideal Candidate •5+ years of experience in Site Reliability Engineering, Platform Engineering, Infrastructure Engineering, or a similar role supporting production AWS environments. •Strong Python development skills with experience building automation, services, or platform tooling. •Hands-on experience with Infrastructure as Code, including CloudFormation and/or AWS CDK. •Experience operating core AWS services including IAM, VPC networking, EC2, Lambda, and managed storage or database services. •Strong Linux systems administration, troubleshooting, incident response, and production operations experience. You might also have... •Experience operating AWS environments across multiple accounts or AWS Organisations. •Experience with AWS networking technologies such as Transit Gateway, IPAM, or BYOIP.Cloud cost optimisation or FinOps experience. •Experience with CI/CD pipelines and GitOps-style deployment practices. •Familiarity with AI-assisted engineering workflows and tooling. •Exposure to policy-as-code, compliance automation, or cloud governance tooling. We encourage you to apply even if your experience or skillset doesn’t align perfectly with every requirement. We value a wide range of backgrounds and transferable skills, and we are excited to support learning and growth. Job Description Global Compute builds and operates the core cloud infrastructure that engineering teams rely on every day. We provision and manage AWS accounts across the company, operate the network backbone that connects them, and maintain the security guardrails that keep those environments safe, compliant, and scalable. We believe reliability is an engineering challenge, not an operations task. We automate repetitive work, build for scale before it becomes a problem, and invest heavily in observability to identify issues before they impact the business. What you'll get to do... •Operate and scale AWS production infrastructure, owning the health of services that provision, secure, and manage accounts across GoDaddy AWS organisations. •Design, build, and maintain cloud platform capabilities using Python, CloudFormation, AWS CDK, and automation-first practices. •Drive cost optimisation initiatives that improve efficiency and deliver measurable business impact.Improve observability through monitoring, alerting, dashboards, and operational tooling. •Participate in on-call rotations, lead incident response efforts, and drive long-term reliability improvements through blameless post-incident reviews. •Support strategic AWS initiatives across networking, identity, governance, and multi-account architecture. •Review code and designs, contribute documentation and operational runbooks, and mentor fellow engineers. •Leverage AI-assisted tooling to improve engineering productivity, accelerate automation, and reduce operational toil. null Company Description At GoDaddy, the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely. Remote: This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings.
Required skills
documentation
mentoring
monitoring
incident response
governance
python
aws
troubleshooting
automation
ec2
lambda
ci/cd pipelines
infrastructure as code
observability
networking
gitops
identity
iam
alerting
finops
cloudformation
dashboards
platform engineering
production operations
policy-as-code
transit gateway
site reliability engineering
aws cdk
linux systems administration
ai-assisted engineering
cloud governance
infrastructure engineering
on-call
ipam
compliance automation
cloud cost optimisation
runbooks
engineering productivity
vpc networking
operational toil
operational tooling
multi-account architecture
aws organisations
byoip
blameless post-incident reviews
Sign up to apply
Create a free account to apply for this job and get access to:
- AI-powered CV optimization for this specific job
- Save jobs and create custom alerts
- See your CV match score for each job
Company information
- Company
- GoDaddy
- Location
-
Full time
Romania - Posted
- 9 hours ago
Interested in this position?
Create your free account and tailor your CV to match this job.