Senior Site Reliability Engineer — Platform Engineering

Guidewire Software Solutions India Private LimitedDublin, LeinsterOn-siteFull-timeSenior, 5–8 yearsListed 5 hours ago

Apply now

About this role

Summary

At Guidewire, we build the technology that helps Property and Casualty (P&C) insurance companies take care of their customers when it matters most. Our products support the core insurance lifecycle—from selling and underwriting policies to settling claims and billing customers—alongside capabilities for data management, digital portals, and predictive analytics.

These products run on the Guidewire Cloud Platform, serving hundreds of insurance providers worldwide and processing billions of dollars of business. We are a market leader, a Top Cloud Employer on Glassdoor, and a company guided by integrity, rationality, and collegiality.

We are looking for a Senior Site Reliability Engineer who wants to build software, not just operate it.

The Platform team is transforming from an operations-focused function into a modern engineering organization. We are investing in automation, developer tooling, platform capabilities, and resilient systems that make it easier to build, deploy, operate, and scale Guidewire products. This is an opportunity to help shape that transformation while working on systems that matter at global scale.

As a Senior SRE, you will design and develop software that improves the reliability, availability, performance, and operability of Guidewire’s flagship cloud platform and InsuranceSuite products. You will write production code, build internal platforms and automation, improve observability, and partner closely with product engineers to solve complex reliability challenges.

You will work across AWS, Kubernetes, Linux, databases, identity platforms, and distributed systems. The ideal candidate combines strong software engineering instincts with deep operational experience and enjoys turning recurring operational problems into reliable, maintainable software.

If your instinct is to automate, simplify, and improve the system—and if you would rather write code than click through a GUI—we would love to hear from you.

Job Description

What you’ll do

- Design, develop, test, and operate software and platform capabilities for resilient, multi-tenant SaaS systems.
- Apply software engineering and SRE practices to shared infrastructure and customer-centric application environments.
- Build automation and developer tooling that improve deployment, operations, reliability, and the engineering experience.
- Develop and maintain AWS infrastructure and services using Infrastructure as Code and GitOps practices.
- Engineer and improve containerized platforms based on Docker, Helm, Kubernetes, and Amazon EKS, including networking, CNI, and ingress.
- Contribute features, bug fixes, reliability improvements, and architectural enhancements to core infrastructure systems and, where appropriate, the product itself.
- Design and operate reliable SAML/OAuth-based authentication and single sign-on platforms.
- Build observability capabilities across logging, metrics, tracing, alerting, dashboards, and application performance monitoring.
- Help evolve the platform toward self-healing operations through automation, intelligent signals, and well-designed control loops.
- Improve incident management by helping teams identify, mitigate, and learn from reliability risks and production issues.
- Support 24x7x365 follow-the-sun operations of critical production systems through automation, documentation, and effective engineering practices.
- Create clear system documentation, runbooks, and training materials that enable the wider engineering organization.
- Collaborate with product and engineering teams to define non-functional requirements for availability, performance, observability, security, and maintainability.
- Mentor teammates and raise the standard for software quality, operational excellence, and reliability engineering.

What you bring

Engineering and technical experience

- Bachelor’s degree in Computer Science or a related field, or equivalent practical experience.
- Strong software development and task automation skills in Bash, Python, Go, or similar languages.
- A track record of building, testing, deploying, and maintaining production software—not only configuring systems.
- Deep experience with Linux systems and production engineering.
- Extensive hands-on experience designing and automating solutions on Amazon Web Services (AWS).
- Strong experience with Kubernetes and containerized production environments, including Docker, Helm, Amazon EKS, CNI, and ingress networking.
- Experience with Infrastructure as Code tools such as Terraform, Terragrunt, or Terraspace.
- Experience with DevOps/GitOps tooling such as Git, Bitbucket, Flux CD, or TeamCity, including promotion and release controls.
- Experience supporting production web applications built with Java, Apache, and/or Tomcat.
- Experience operating complex, microservice-based systems at scale.
- Strong understanding of observability and incident response, with tools such as Datadog, CloudWatch, and PagerDuty.
- Experience with relational databases such as Aurora PostgreSQL and/or Oracle RDS.
- Strong understanding of Single Sign-On, SAML, and OAuth; hands-on experience with Okta is a plus.
- Solid understanding of X.509 certificates, public-key infrastructure, and core encryption concepts.
- Familiarity with event-store and stream-processing technologies such as Kafka or AWS SQS.
- Experience with application development, JSON, web technologies, user interfaces, and distributed system architecture.
- Familiarity with Open Application Model platforms such as KubeVela or Crossplane is a plus.
- Practical experience using AI and data-driven insights to improve engineering productivity, reliability, and decision-making.
- Working knowledge of agile software development practices, including Scrum and Kanban.

How you work

- You prefer writing code over clicking through a GUI.
- You are curious, pragmatic, and comfortable learning unfamiliar technologies quickly.
- You enjoy teaching, mentoring, and working across organizational boundaries.
- You bring strong troubleshooting and analytical skills and can make sound decisions under pressure.
- You take ownership, follow through, and consistently meet your commitments.
- You communicate clearly with both technical and non-technical audiences.
- You are a positive, collaborative teammate who helps turn ambiguous problems into executable plans.
- You champion a culture of reliability through practices such as blameless postmortems, SLOs, automation, and continuous learning.
- You are excited to help shape an engineering team that treats operations as a software problem and reliability as a shared product responsibility.

Other requirements

- Ability to read, write, and speak English.
- Participation in a rotating on-call schedule for weekend production emergencies and operational support.
- Occasional travel—less than 5%—to other Guidewire offices for training and team meetings.

Why join us?

You will help transform a critical platform team while working on software that operates at meaningful global scale. You will have the opportunity to influence architecture, build engineering foundations, and make the systems used by Guidewire and its customers more reliable every day.

At Guidewire, we foster a culture of curiosity, innovation, and responsible use of AI. We empower our teams to experiment, learn, and use emerging technologies and data-driven insights to improve productivity and outcomes.

#LI-AS3

About Guidewire

Guidewire is the platform P&C insurers trust to engage, innovate, and grow efficiently. We combine digital, core, analytics, and AI to deliver our platform as a cloud service. More than 540+ insurers in 40 countries, from new ventures to the largest and most complex in the world, run on Guidewire.

As a partner to our customers, we continually evolve to enable their success. We are proud of our unparalleled implementation track record with 1600+ successful projects, supported by the largest R&D team and partner ecosystem in the industry. Our Marketplace provides hundreds of applications that accelerate integration, localization, and innovation.

For more information, please visit www.guidewire.com and follow us on Twitter: @Guidewire_PandC.

Guidewire Software, Inc. is proud to be an equal opportunity and affirmative action employer. We are committed to an inclusive workplace, and believe that a diversity of perspectives, abilities, and cultures is a key to our success. Qualified applicants will receive consideration without regard to race, color, ancestry, religion, sex, national origin, citizenship, marital status, age, sexual orientation, gender identity, gender expression, veteran status, or disability. All offers are contingent upon passing a criminal history and other background checks where it's applicable to the position.

Guidewire is proud to be a disability inclusive employer. We are committed to providing individualized accommodations and support to candidates for any part of our hiring process. If you think you may need disability-related accommodations, please email EmployAbility at [email protected], our disability partner, or schedule a call with an EmployAbility expert (calendar booking slot) for a confidential conversation.