Director of DevOps / Site Reliability Engineering - Equities Trading
Singapore · Full Time
Be the first to apply
- Experience
- Any
- Salary
- —
- Openings
- 1
- Posted
- 2 weeks ago
- Work mode
- In office
- Education
- Bachelor's degree in Computer Science, Engineering, or equivalent practical experience
- Resume
- Required to apply
Where you'll work
Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.
Job description
About the Role
Join a premier global financial institution as an Executive Director in DevOps / Site Reliability Engineering within the Corporate & Investment Bank Equities Trading division. Lead initiatives that enhance operational resilience, automation, and reliability for a high-speed, latency-sensitive trading platform spanning multiple regions worldwide.
Key Responsibilities
- Develop and implement the DevOps/SRE strategy and roadmap tailored to the Electronic Market Making platform, focusing on reliability and productivity metrics.
- Manage governance for major DevOps/SRE projects, ensuring documented objectives, timelines, dependencies, and comprehensive lifecycle updates.
- Coordinate cross-regional collaboration among application engineering, infrastructure, network, and platform teams with clear roles and responsibilities.
- Establish and enforce production excellence standards, including incident oversight, root cause analysis, and systematic remediation to minimize operational failures.
- Continuously advance resilience protocols such as automated failover, disaster recovery, and readiness controls.
- Oversee service level objectives (SLOs), service level agreements (SLAs), and observability frameworks covering monitoring, alerts, and key risk indicators to maintain uptime and performance transparency.
- Lead continuous integration/continuous deployment (CI/CD), release engineering, and multi-region automation to boost delivery efficiency and confidence.
- Empower developers through internal platforms, self-service tools, and automation aimed at expanding engineering capacity while preserving control standards.
- Implement a shared-responsibility operational model across platform/product and line-of-business engineering teams to uphold stringent security, compliance, and production standards.
- Recruit, mentor, and manage DevOps/SRE professionals and oversee operational risks with well-defined escalation and mitigation protocols.
Essential Qualifications and Experience
- Bachelor’s degree in Computer Science, Engineering, or comparable practical expertise.
- Substantial experience in DevOps, SRE, Production or Platform Engineering with proven leadership in managing teams or cross-functional projects yielding measurable operational improvements.
- Background supporting electronic trading or other low-latency environments, including market data handling, exchange connectivity, and order routing.
- Proven track record of accountability for production quality, incident management, root cause analysis, reliability enhancements, and operational resilience.
- Expertise in CI/CD and release engineering at scale, utilizing tools like Jenkins or GitLab CI, with knowledge of multi-region deployment architectures and governance.
- Strong Linux system expertise (particularly RHEL-based), including performance tuning, capacity planning, and diagnostic skills.
- Advanced scripting and automation capabilities using Python and shell for reliable and maintainable tooling as well as self-service platform development.
- Experience with infrastructure automation and configuration management technologies such as Ansible, Puppet, or Salt, along with infrastructure-as-code methodologies.
- Solid understanding of networking fundamentals (TCP/IP, DNS, load balancing) and effective collaboration with network teams supporting market data and exchanges.
- Exceptional communication skills suited to executive reporting with data-driven, transparent updates regarding risks and project statuses.
Preferred Skills and Expertise
- Proficiency in observability and operational analytics tools like Prometheus, Grafana, Splunk, or ELK, focusing on actionable alerting and service health monitoring.
- Experience deploying containerization and orchestration technologies such as Docker and Kubernetes within regulated, production-level environments.
- Familiarity with message-driven middleware or publish-subscribe systems including AMPS or Kafka.
- Knowledge of equities, options, and futures market structures.
- Experience working in regulated financial institutions with robust production controls and audit compliance requirements.
Company Overview
J.P. Morgan is an industry-leading global financial services provider committed to delivering strategic advice and comprehensive products to corporations, governments, high-net-worth individuals, and institutional investors worldwide. This commitment is supported by fostering an inclusive environment that values diversity in all its forms and offers equal opportunities across its workforce. The firm prioritizes accommodating religious practices and supporting employees’ mental and physical health needs.
Team Context
The Commercial & Investment Bank segment of J.P. Morgan drives global leadership in banking, markets, securities services, and payments. Serving corporate, governmental, and institutional clients in over 100 countries, this division excels at providing strategic advice, risk management, liquidity solutions, and capital raising services across the globe.
Minimum education
Bachelor's Degree