Senior Site Reliability Engineer
Cork, County Cork, Ireland · Full Time
Be the first to apply
- Experience
- 8+ yrs
- Salary
- —
- Openings
- 1
- Posted
- 8 saat önce
- Work mode
- In office
- Education
- Bachelor's degree in Computer Science or Engineering
- Resume
- Required to apply
Where you'll work
Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.
Job description
About OpenText
OpenText is a worldwide leader in information management known for fostering innovation, creativity, and teamwork. Joining our team means collaborating with top global enterprises to tackle complex challenges and contribute to projects that are shaping the future of digital transformation. AI is central to our mission, driving innovation and aiding digital knowledge workers. We're seeking talented individuals who complement AI to advance the future of information management.
Role Overview
As a Lead Site Reliability Engineer in our Cork Cloud Service team, you will play a key role in ensuring our cloud infrastructure is reliable, scalable, and performs optimally. You'll partner closely with development and operations teams to architect, implement, and sustain resilient and high-efficiency systems.
Key Responsibilities
- Architect and maintain secure, high-performing, and cost-effective AWS infrastructure.
- Manage infrastructure provisioning and lifecycle using Terraform and Terragrunt across various environments.
- Monitor and optimize Windows and Linux systems to maintain service performance and compliance with operational standards.
- Work alongside development teams to integrate and streamline CI/CD pipelines for microservices deployment.
- Develop and operate CI/CD workflows leveraging tools such as Octopus Deploy, Jenkins, GitLab CI, and GitHub Actions.
- Lead critical root cause analysis to resolve infrastructure and application performance issues promptly.
- Take part in on-call rotations and lead incident response activities ensuring quick recovery and comprehensive postmortem reviews.
- Contribute to AI-powered observability, alerting, and self-healing initiatives, including anomaly detection and automation for cloud infrastructure and application management.
Qualifications and Experience
- Bachelor’s degree in Computer Science, Engineering, or a related discipline.
- At least 8 years of professional experience in Site Reliability Engineering, DevOps, or infrastructure roles supporting high-availability production systems.
- In-depth knowledge and hands-on experience with Kubernetes and large-scale container orchestration.
- Solid track record designing and administering CI/CD pipelines using Octopus Deploy, Jenkins, GitLab, or GitHub Actions.
- Excellent interpersonal and communication skills with proven ability to collaborate across development, product management, and operations teams.
- Experience mentoring junior engineers and contributing to architectural strategy decisions.
- Ability to articulate reliability strategies effectively to senior leadership and manage cross-functional projects.
- Familiarity with AI/ML-driven monitoring solutions, anomaly detection methods, and automated remediation technologies.
- Keen interest to support internal efforts adopting AI for cloud operations enhancement.
Additional Information
At OpenText, we strive beyond traditional corporate boundaries to foster a truly inclusive global community grounded in trust, high standards, and a strong sense of accountability. Our Employment Equity and Diversity Policy ensures a workplace welcoming to individuals regardless of race, gender, sexual orientation, age, disability, veteran status, religion, or other protected categories. Reasonable accommodations during hiring are available upon request by contacting us.
Level
Senior
Minimum education
Bachelor's Degree