- Experience
- 4+ yrs
- Salary
- USD 130,000 – USD 200,000 / year
- Openings
- 1
- Posted
- 3 days ago
- Work mode
- Work from home
- Eligibility
- Open to all qualified applicants regardless of race, color, ethnicity, nationality, gender, gender identity or expression, sexual orientation, age, religion, disability, marital status, or any other legally protected characteristic.
- Resume
- Required to apply
Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.
Job description
Role Summary
We are seeking a Site Reliability Engineer to ensure our production systems remain reliable, observable, and performant. This position blends software engineering with operations, focusing on automating mundane tasks and treating reliability as a primary engineering challenge.
Key Responsibilities
- Design and manage systems to achieve high availability and optimal performance.
- Develop and maintain observability tools, including logging, metrics, and tracing systems.
- Establish and monitor Service Level Objectives (SLOs), Service Level Indicators (SLIs), and manage error budgets.
- Lead incident responses and conduct thorough post-mortem analyses.
- Automate manual operational workloads by creating tooling and enhancing platforms.
- Collaborate with application teams to ensure readiness for production deployment.
Qualifications and Skills
- Minimum of 4 years experience in Site Reliability Engineering, DevOps, or infrastructure engineering.
- Proficient in scripting and software development, preferably using Python, Go, or equivalent languages.
- Extensive experience working with cloud environments such as AWS, GCP, or Azure.
- Hands-on familiarity with Kubernetes, Terraform, and observability platforms.
- Proven leadership in managing incident response in live production settings.
- Strong grasp of distributed systems architecture and operation.
Personal Attributes
- Inquisitive nature with a drive to analyze systems and deliver impactful improvements.
- Excellent written communication skills with ability to clearly explain technical decisions.
- A test-and-learn approach characterized by rapid shipping, data-driven evaluation, and iteration.
- Comfort with asynchronous collaboration across multiple time zones.
Compensation and Benefits
- Competitive salary reflecting experience, skills, and geographic location, with an indicative range between $130,000 and $200,000 USD equivalent.
- Performance-driven bonus scheme.
- Annual stipend dedicated to learning and professional development.
- Health and wellness benefits subject to local variations.
- Opportunity to work remotely with flexible scheduling.
- Engagement on impactful projects at large scale.
Equal Opportunity Commitment
This opportunity is open to all qualified candidates without discrimination based on race, color, ethnicity, nationality, gender identity or expression, sexual orientation, age, religion, disability, marital status, or other protected statuses. Employment decisions will be made solely on merit, qualifications, and ability.
Industry
Staffing & Recruitment