Senior Site Reliability Engineer
Cork, County Cork, Ireland · Full Time
Be the first to apply
- Experience
- 8+ yrs
- Salary
- —
- Openings
- 1
- Posted
- 5 days ago
- Work mode
- In office
- Education
- Bachelor's degree in Computer Science, Engineering or related
- Resume
- Required to apply
Where you'll work
Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.
Job description
About OpenText
OpenText stands as a global frontrunner in information management, where core values of innovation, creativity, and teamwork define the corporate culture. Joining OpenText means working alongside some of the world's foremost companies, tackling challenging and impactful projects aimed at driving digital transformation.
AI is central to OpenText's mission, powering advancements that transform work and empower digital knowledge workers. The organization seeks talents who complement AI capabilities to help shape the future of information management.
Role Overview
As a Lead Site Reliability Engineer within OpenText's Cloud Service team based in Cork, you will play a pivotal role in guaranteeing the reliability, scalability, and robust performance of the cloud infrastructure. This position involves close collaboration with development and operations teams to architect, execute, and sustain resilient and efficient systems.
Key Responsibilities
- Architect, maintain, and optimize AWS infrastructure emphasizing security, performance, and cost-efficiency.
- Utilize Terraform and Terragrunt for infrastructure as code deployment and management across diverse environments.
- Oversee Windows and Linux system monitoring to ensure compliance with operational standards alongside high reliability and performance.
- Partner with development teams to integrate CI/CD pipelines facilitating seamless deployment of microservices.
- Create and maintain CI/CD workflows leveraging tools such as Octopus Deploy, Jenkins, GitLab CI, and GitHub Actions.
- Lead root cause analysis and resolution efforts related to infrastructure and application performance challenges.
- Participate in on-call rotations and manage incident responses to ensure rapid recovery and thorough postmortem reviews for critical systems.
- Drive exploration and implementation of AI-powered observability, alerting, and self-healing mechanisms including anomaly detection and automated remediation in cloud systems.
Qualifications and Experience
- Bachelor’s degree in Computer Science, Engineering, or equivalent discipline.
- At least 8 years of professional experience in Site Reliability Engineering, DevOps, or related infrastructure roles supporting live production environments.
- Comprehensive knowledge and hands-on expertise with Kubernetes and managing container orchestration at scale.
- Demonstrated experience in designing and managing CI/CD automation pipelines using tools like Octopus Deploy, Jenkins, GitLab, or GitHub Actions.
- Excellent interpersonal and communication skills with a history of effective collaboration across product, development, and operations teams.
- Experience mentoring junior engineers and influencing system architecture decisions.
- Ability to present reliability strategies to senior leadership and spearhead initiatives that span multiple teams.
- Familiarity with AI and machine learning techniques applied to monitoring, anomaly detection, and automated remediation systems.
- Keen interest in advancing internal AI-driven approaches for cloud infrastructure operations.
Additional Information
OpenText fosters a global community rooted in trust and accountability. The company actively embraces diversity, inclusion, and equal employment opportunities irrespective of cultural or demographic background. Any candidate requiring accommodation during the recruitment process due to disability can request assistance.
Level
Senior
Minimum education
Bachelor's Degree