MediSolution
2 months ago
Principal Systems Engineer
Sign up to save this job, get alerts, and apply with an optimized CV.
Company information
- Company
- MediSolution
- Location
- Singapore Singapore
- Posted
- 2 months ago
Job description
As a Site Reliability Engineer (SRE) at Altera, you will be responsible for ensuring the reliability, scalability, and performance of our hosted healthcare platforms. This role blends software and systems engineering to enhance service availability, automate operations, and improve the customer experience. You will act as a technical leader in monitoring, troubleshooting, incident response, and continuous improvement across our cloud and hybrid environments.
Key Responsibilities- Maintain and improve the reliability, availability, and performance of production environments.
- Lead the investigation and resolution of complex application, database, and infrastructure issues.
- Participate in incident management, conduct root cause analysis (RCA), and contribute to post-incident reviews to prevent future occurrences.
- Define and measure Service Level Indicators (SLIs) and Objectives (SLOs) to meet service commitments.
- Develop proactive monitoring and alerting strategies to identify and resolve issues before they impact customers.
- Automate operational tasks using scripting and Infrastructure-as-Code (IaC) to improve efficiency.
- Partner with engineering and cloud teams to refine deployment, monitoring, and support processes.
- Provide technical leadership during major incidents and act as a key escalation point for critical issues.
- 7+ years of experience supporting enterprise applications, infrastructure, or cloud environments.
- Strong experience with APM tools such as LogicMonitor, AppDynamics, Azure Monitor, SentryOne, Dynatrace, Datadog, or New Relic.
- Deep knowledge of Windows Server administration, IIS, .NET applications, Windows Clustering, MSMQ, Event Logs, and PerfMon.
- Strong SQL Server experience including performance tuning, query optimization, blocking analysis, and Always On Availability Groups.
- Experience with Azure cloud environments and solid understanding of networking fundamentals (DNS, TCP/IP, load balancing, firewalls).
- Familiarity with ServiceNow (or other ITSM platforms) and ITIL principles.
- Preferred: scripting with PowerShell, Python, or similar; Infrastructure as Code (Terraform, ARM Templates, Bicep); CI/CD pipelines and deployment automation (Azure DevOps, GitHub Actions); Kubernetes and containerized workloads; implementing SLOs, SLIs, and Error Budgets; healthcare technology background.
- Bachelor's Degree in Computer Science, Information Technology, or Engineering preferred, or equivalent professional experience.
Remote position within the United States. Participate in an on‑call rotation to support a 24x7 healthcare environment. Occasional after‑hours work required for activations, upgrades, and major incidents. No travel required.
Benefits- Competitive compensation and benefits package.
Altera is an Equal Opportunity/Affirmative Action Employer. We consider applicants without regard to race, color, religion, age, national origin, ancestry, ethnicity, gender, gender identity, gender expression, sexual orientation, marital status, veteran status, disability, genetic information, citizenship status, or membership in any other group protected by local law.
Required skills
- engineering
- customer experience
- monitoring
- incident response
- reliability
- support processes
- information technology
- python
- kubernetes
- troubleshooting
- cloud
- github actions
- sre
- powershell
- terraform
- datadog
- azure devops
- sql server
- msmq
- cloud environments
- dns
- technical leadership
- disability
- ci/cd pipelines
- site reliability engineer
- computer science
- professional experience
- tcp/ip
- root cause analysis
- scripting
- dynatrace
- firewalls
- load balancing
- servicenow
- performance
- new relic
- incident management
- performance tuning
- scalability
- activations
- color
- iac
- compensation
- slis
- networking fundamentals
- benefits package
- engineering teams
- infrastructure-as-code
- rca
- post-incident reviews
- bicep
- appdynamics
- iis
- healthcare technology
- gender
- equal opportunity employer
- hybrid environments
- query optimization
- deployment automation
- azure monitor
- upgrades
- infrastructure support
- escalation point
- arm templates
- slos
- apm tools
- .net applications
- remote position
- cloud teams
- logicmonitor
- enterprise applications
- production environments
- event logs
- age
- always on availability groups
- service availability
- on-call rotation
- religion
- veteran status
- genetic information
- error budgets
- major incidents
- deployment processes
- sexual orientation
- perfmon
- critical issues
- race
- marital status
- itil principles
- windows server administration
- containerized workloads
- itsm platforms
- affirmative action
- healthcare platforms
- service level objectives
- application issues
- citizenship status
- ethnicity
- gender identity
- proactive monitoring
- national origin
- ancestry
- after-hours work
- gender expression
- operations automation
- database issues
- infrastructure issues
- service level indicators
- service commitments
- alerting strategies
- operational tasks automation
- sentryone
- windows clustering
- blocking analysis
- azure cloud environments
- 24x7 healthcare environment
Interested in this position?
Create your free account and tailor your CV to match this job.