- Experience
- 5+ yrs
- Salary
- —
- Openings
- 1
- Posted
- 1 day ago
- Work mode
- Work from home
- Resume
- Required to apply
Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.
Job description
About Camunda
Camunda delivers an enterprise platform for agentic orchestration, helping organizations coordinate AI agents, people, and systems throughout complex end-to-end business processes. The platform emphasizes governance, auditability, and human oversight, allowing companies to transition AI from pilots to production safely and at scale. Trusted by over 700 customers including major US banks, Camunda enables faster operational gains and enhanced customer experiences. Currently fully remote and globally active, Camunda is transforming into an AI-first company using agentic AI to automate intelligent processes while enhancing human contributions across all teams. The company is recognized by GP Bullhound's Top 100 Next Unicorn list, 2025 Great Place to Work certification, Gartner's 2025 Magic Quadrant for Business Orchestration and Automation Technologies, and ranked 3rd among the most flexible companies in 2026.
Role Overview
As a Senior Software Engineer focusing on backend reliability at Camunda, you will lead the design and execution of automated reliability testing and chaos engineering for Camunda 8. This role prioritizes curiosity, experimentation, and continuous enhancement through breaking systems intentionally in controlled environments to bolster platform stability before impacting customers. You'll influence product direction, collaborate with knowledgeable colleagues sharing core company values, and contribute significantly to delivering resilient systems and exceptional user experiences even in failure scenarios.
Key Responsibilities
- Design and run automated chaos experiments and reliability tests simulating real-world conditions on Camunda's platform.
- Diagnose, root-cause, and debug potential failures or performance degradation in Java-based products.
- Enhance current infrastructure related to load testing, chaos engineering, and observability to maintain high system reliability and performance.
- Utilize technologies including Java, Go, Kubernetes, Prometheus, and Grafana for development and operational tasks.
- Introduce innovative testing tools and methods, such as fault injection tools like zbchaos, to significantly improve operational workflows and customer outcomes.
- Collaborate closely with QA and cross-functional engineering teams, sharing insights through internal blogs and shaping team roadmaps with experimental outcomes.
- Promote autonomous and pragmatic software design approaches, tackling challenging problems while learning new technologies and principles.
- Adopt a user-centered perspective by employing Camunda’s products firsthand to advocate for reliability improvements.
Success Metrics
- Within three months, develop and deliver a rapid, reproducible load testing approach that reduces engineering feedback loops and integrates seamlessly into development lifecycles while remaining extensible for engineers.
- Create a comprehensive, automated load testing framework for Camunda Optimize, based on performance analyses and in conjunction with senior engineers and field teams.
- Communicate performance results with core engineering teams to enable retesting and document findings publicly to assist users in performance planning.
Required Qualifications
- Experience of over 5 years in backend software engineering using Java.
- A strong desire and ability to leverage Camunda’s products and contribute reliability solutions from an end-user viewpoint.
- Keen interest in experimenting, adopting cutting-edge technologies, and driving automated reliability and chaos testing within distributed systems.
- Passionate about enhancing system performance and fault-tolerance in production environments.
- Self-driven and pragmatic problem solver capable of guiding peers and integrating solutions broadly.
Preferred Skills
- Practical experience programming in Go or familiarity with multi-language system environments.
- Background in site reliability engineering with a strong software engineering foundation in distributed systems.
- Knowledge of Kubernetes, Helm Charts, Operators, and pro-production infrastructure management.
- Experience managing production applications, skilled in monitoring, troubleshooting, and analyzing performance.
- Familiarity with chaos engineering and automated load or performance testing practices.
Compensation and Benefits
Camunda provides a competitive and transparent salary structure, adjusted by location within standard or major global markets. Annual total cash compensation varies significantly based on geography and experience, with specific ranges for countries including the US (up to $231,900), UK, Singapore, and Canada. Equity participation is available via a Virtual Stock Option Plan.
The company emphasizes employee wellbeing and development globally, offering remote work flexibility, home office support, co-working allowances, flexible time off, and organized annual meetups. Health benefits include local healthcare access, mental wellbeing programs, and a Live Well Lifestyle Spending Account supporting diverse personal needs and valued at up to €1,000 annually starting 2027. Financial benefits include retirement plans with potential employer contributions and life/disability insurance as applicable. Professional development budgets support self-driven learning avenues.
Diversity and Hiring Practices
Camunda values inclusivity and equal opportunity, welcoming applicants from all backgrounds regardless of gender, race, ethnicity, religion, sexual orientation, disability, or other protected characteristics. The recruiting process may involve AI-supported screening and interviews. To protect candidates, the company warns against scams and clarifies official correspondence originates only from legitimate Camunda domains. The company does not engage unsolicited recruiting agencies.