Site Reliability Engineer III (SRE) - Guidewire Cloud Platform (Application)
Описание должности
At Guidewire, we deliver the software that Property and Casualty (P&C) insurance companies rely on to protect their customers during crises, natural disasters, accidents, and cyber risks. Our core applications enable insurers to sell and underwrite policies, settle claims, and bill their customers. We also offer a suite of innovative products for data management, digital portals, and predictive analytics.
Hundreds of insurers worldwide use Guidewire's products, running on our cutting-edge Guidewire Cloud Platform, to handle billions of dollars in business. We are dedicated to providing the tools and technology that help insurers protect and support their customers when they need it most.
We are seeking a Site Reliability Engineer III who is eager to contribute to the transformation of the insurance industry with our leading cloud platform. As a member of the SRE-Application team, you'll play a critical role in ensuring the reliability, performance, and scalability of applications running on our Guidewire Cloud Platform. This position offers a unique opportunity to apply your skills in automation, software engineering, and operational discipline to support our cloud-based solutions.
Why Guidewire
This is an opportunity to join a mission-driven company and make a real impact in the lives of people facing challenges. You'll work with cutting-edge technology, collaborate with talented peers, and grow your skills in a culture that values innovation, teamwork, and work-life balance. We offer competitive compensation, comprehensive benefits, and opportunities for career development.
If you're an SRE who combines deep technical expertise with a passion for problem-solving and a commitment to reliability, we'd love to hear from you. Join us in building the software that helps insurers care for their customers when they need it most.
This position requires participation in mandatory on-call rotations to ensure the availability and reliability of our services. This includes responding to incidents and alerts outside of regular business hours, on weekends, and during holidays, as per the established on-call schedule. Candidates must be willing and able to fulfill this critical responsibility.
Required
Skills
- Software engineering background with experience in Python, Go, or Java, following best practices (SOLID, DRY, KISS) and writing clean, testable code
- Experience with designing and implementing SLI's, SLO's, and Error Budgets
- Familiarity with application performance monitoring (APM) and telemetry tools to maintain expected service levels for applications
- Experience troubleshooting and debugging distributed systems on cloud infrastructure
- Experience with CICD pipelines within K8S and legacy ecosystems
- Experience creating monitors, dashboards, and synthetic transactions in monitoring tools like Datadog
- Experience deploying and managing scalable infrastructure within AWS and Kubernetes ecosystems using Terraform and other cloud-native approaches
- Experience with infrastructure configuration management using tools such as GitOps, Puppet, or Ansible
- Good understanding of cloud networking, security, and vulnerability management, with the ability to programmatically remediate infrastructure issues
Preferred
Skills
- SRE Certification in one or more categories
- AWS Certification in one or more categories
- Experience with SQL, database administration, data pipelines, performance tuning, and schema design
- Familiarity with pipelining tools such as Team City, Bitbucket Pipelines, Jenkins, or GitHub Actions
- Exposure to open-source distributed data processing frameworks such as Hadoop, Apache Spark, AWS RedShift, etc.
- Experience with distributed systems, including microservices and event-driven architectures
Упомянутые навыки
Откликнуться на вакансию
Use the application link supplied with this listing to apply to Guidewire Software. Check the destination before entering personal information.
