Site Reliability Engineer III
Dublin, County Dublin, Ireland · Full Time
Be the first to apply
- Experience
- Any
- Salary
- —
- Openings
- 1
- Posted
- 2 ঘন্টা আগে
- Work mode
- In office
- Education
- Bachelor's Degree
- Resume
- Required to apply
Where you'll work
Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.
Job description
Overview
Guidewire Software develops core software applications for Property and Casualty insurance companies, assisting them with policy sales, underwriting, claims settlement, and billing. Their offerings also extend to data management, digital portals, and predictive analytics, hosted on Guidewire's Cloud Platform, serving global insurance providers handling billions in business.
Recognized as a Top Cloud Employer and industry leader, Guidewire fosters a workplace culture grounded in integrity, rationality, and collegiality. The company invites professionals passionate about collaborative product delivery and operational support to join their team and make meaningful contributions.
Role Summary
The Site Reliability Engineer (SRE) will contribute to automating processes ensuring efficient system performance within Guidewire's Platform team. This team focuses on software that enhances the reliability of production systems supporting millions of transactions for numerous customers. The role emphasizes maintaining stability, developing operational tooling, and ensuring availability for multi-tenant SaaS environments. Collaboration with product developers is essential to meet functional and non-functional system requirements, such as availability, performance, observability, and maintainability.
Key Responsibilities
- Implement and promote reliability and automation for SaaS microservices and multi-tenant infrastructure.
- Manage and enhance AWS deployments through automation of operational tasks.
- Contribute to the evolution of core infrastructure by adding features, resolving defects, and improving system reliability.
- Develop and maintain a complex single sign-on authentication platform leveraging SAML and OAuth protocols.
- Create and sustain observability tools, including metrics and dashboards, to support global infrastructure monitoring.
- Advance incident management processes by identifying risks, mitigating issues, and fostering a self-healing environment.
- Produce system documentation and training resources to empower team members.
- Collaborate with engineering teams to provide insights and contribute code enhancing product quality.
- Embrace innovation including responsible AI use and data-driven enhancements to productivity and outcomes.
Required Qualifications and Skills
- Bachelor’s degree in Computer Science or related discipline.
- Proficient software engineering and automation skills using Bash, Python, and/or Go.
- Experience with agile methodologies such as Scrum and Kanban.
- Strong knowledge of Linux operating systems.
- Expertise in AWS cloud automation and live production environment support (Java/Apache/Tomcat).
- Experience with Infrastructure as Code tools like Terraform, Terragrunt, or Terraspace.
- Familiarity with DevOps/GitOps tools including Git, Bitbucket, Flux CD, and TeamCity.
- Hands-on experience with containerization technologies: Docker, Helm, Kubernetes/EKS, CNI, and Ingress networking.
- Understanding of authentication protocols and single sign-on systems: SAML, OAuth (experience with Okta is a plus).
- Knowledge of observability platforms such as Datadog, AWS CloudWatch, and PagerDuty.
- Experience with streaming/event processing systems (Kafka, AWS SQS).
- Working knowledge of relational databases like Aurora Postgres and Oracle RDS.
- Exposure to application development fundamentals including web UI design and JSON.
- Familiarity with Open Application Models such as KubeVela or Crossplane is advantageous.
Personal Attributes
- Preference for writing maintainable code over performing manual GUI tasks.
- Aptitude for mentoring and knowledge sharing.
- Exceptional analytical and problem-solving skills with a structured, process-oriented mindset.
- Strong communication skills capable of conveying complex technical information effectively.
- Advocate for reliability culture, including blameless postmortems and SLO monitoring.
- Demonstrated openness to leveraging AI and data-driven methods to drive continual improvement.
Company Commitment
Guidewire stands as a trusted platform for over 540 insurers globally, combining digital, core systems, analytics, and AI to offer cloud-based solutions. The company boasts over 1600 successful project implementations, supported by the industry’s largest R&D team and partner network, and offers a marketplace with hundreds of apps to accelerate innovation.
Guidewire is dedicated to equal opportunity and diversity, considering all qualified applicants without regard to protected characteristics and contingent on background checks where applicable.
Minimum education
Bachelor's Degree