Sr. Service Engineer

7 days ago
Experience level:Senior
Minimum experience:5+ years
Education:Bachelor’s degree
Apply Now

Job description

Service Engineer (3P) – Service Engineering

About the Role

We are seeking a Service Engineer to support the design, implementation, automation, assessment, and operation of secure, scalable, resilient, and highly available cloud platforms. This role requires strong cloud engineering expertise, infrastructure automation skills, platform reliability knowledge, and a passion for operational excellence, modernization, and continuous improvement.

Key Responsibilities

  • Design, implement, and manage cloud infrastructure on Azure (preferred), AWS, or GCP.
  • Build and maintain Infrastructure as Code (IaC) solutions using tools such as Terraform, ARM/Bicep, or CloudFormation.
  • Automate infrastructure provisioning, configuration management, and deployment processes.
  • Support cloud-native applications, container platforms, and Kubernetes environments.
  • Implement and maintain monitoring, alerting, logging, and observability solutions.
  • Manage cloud networking, security, identity, storage, backup, and disaster recovery capabilities.
  • Collaborate with development, platform, security, architecture, and operations teams to deliver reliable cloud solutions.
  • Optimize cloud performance, availability, scalability, reliability, and cost efficiency.
  • Troubleshoot infrastructure, platform, and production issues and drive root cause analysis.
  • Support CI/CD implementation and DevOps best practices.
  • Ensure compliance with security, governance, and operational standards.
  • Contribute to cloud modernization, automation, platform engineering, and transformation initiatives.
  • Assess application infrastructure foundations, including compute, storage, networking, access management, dependencies, backup, recovery, and disaster recovery readiness.
  • Evaluate platform resilience, redundancy, reliability, and high-availability capabilities against business continuity and recovery objectives.
  • Identify infrastructure modernization, standardization, technical debt reduction, and risk mitigation opportunities to improve platform stability and operational supportability.
  • Recommend and implement enhancements that improve service continuity, operational excellence, incident reduction, and system recoverability.
  • Conduct cloud platform health assessments and architecture reviews, documenting findings, recommendations, risks, and remediation plans.
  • Participate in capacity planning, performance optimization, disaster recovery testing, and lifecycle management of cloud infrastructure and platform services.

Qualifications

  • 5–7 years of experience in Cloud Engineering, Infrastructure Engineering, Platform Engineering, Site Reliability Engineering, or related roles.
  • Strong hands-on experience with Microsoft Azure services, including Compute, Networking, Storage, Identity Management, Monitoring, Security, and Governance.
  • Experience with AWS and/or GCP is desirable.
  • Experience with Infrastructure as Code (Terraform preferred) and configuration management tools.
  • Experience with Kubernetes, Docker, AKS, and containerized platforms.
  • Strong understanding of cloud networking, load balancing, DNS, VPNs, firewalls, and security best practices.
  • Experience implementing and supporting CI/CD pipelines using Azure DevOps, GitHub Actions, Jenkins, or similar tools.
  • Knowledge of observability and monitoring tools such as Azure Monitor, Grafana, Datadog, Dynatrace, Splunk, OpenTelemetry, or similar platforms.
  • Experience with scripting and automation using PowerShell, Python, Bash, or similar languages.
  • Understanding of cloud security, governance, compliance, identity, and access management.
  • Experience performing cloud platform assessments, operational readiness reviews, and infrastructure health evaluations.
  • Knowledge of Site Reliability Engineering (SRE) principles, including availability, reliability, incident management, and service performance.
  • Understanding of disaster recovery planning, backup strategies, business continuity, and recovery testing methodologies.
  • Familiarity with FinOps practices, cloud cost governance, and resource optimization strategies.
  • Strong troubleshooting, analytical, and problem-solving skills.
  • Excellent communication and collaboration skills within cross-functional teams.
  • Azure certifications such as AZ-104, AZ-305, AZ-400, or equivalent are preferred.
  • Bachelor's degree in Computer Science, Information Technology, Engineering, or equivalent experience.

Preferred Experience

  • Experience supporting mission-critical, highly available enterprise workloads.
  • Experience working within regulated environments with strong security, compliance, and governance requirements.
  • Exposure to Platform Engineering, Cloud Center of Excellence (CCoE), or Cloud Adoption Framework (CAF) initiatives.
  • Experience supporting large-scale cloud migrations, modernization efforts, and enterprise transformation programs.

Skills mentioned


Healthcare, Mental Health

Providence is a not-for-profit Catholic health system headquartered in Renton, Washington, and one of the largest health systems in the United States. It operates around 50 hospitals and 1,000+ clinics across seven western states, alongside senior services and health plans.

Apply for this job

Use the application link supplied with this listing to apply to Providence. Check the destination before entering personal information.

Apply Now