HPC Systems Engineer
Sydney, New South Wales, Australia · Part Time
Be the first to apply
- Experience
- 5+ yrs
- Salary
- —
- Openings
- 1
- Posted
- 3 days ago
- Work mode
- In office
- Resume
- Required to apply
Where you'll work
Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.
Job description
About Jump Trading Group
Jump Trading Group is deeply committed to cutting-edge research, drawing on exceptional talent in Mathematics, Physics, and Computer Science to push scientific boundaries and apply advanced research to global financial markets. Our culture thrives on continuous innovation, embracing creativity, intellectual integrity, and a fierce competitive drive. We foster teamwork and unlock individual potential through collaboration and mutual respect. Our research fuels superior risk-adjusted returns and supports technologies that transform industries, from funding startups to partnering with leading research institutions worldwide.
Role Overview
Our High Performance Computing (HPC) team in Sydney seeks a highly capable Systems Engineer to support our substantial computing infrastructure. Our environment poses distinct challenges, requiring seamless integration of compute resources, scheduling systems, networks, and large-scale storage to empower data pipelines and quantitative research. We're looking for a hands-on professional with deep expertise in Linux system management and a solid software development background, ready to maintain customized HPC systems at scale.
Key Responsibilities
- Design, deploy, maintain, and support HPC compute and storage platforms to ensure optimal performance and reliability
- Develop and manage systems for monitoring performance and diagnosing faults across compute, storage, and network components
- Create tools for compiling, packaging, installing, and upgrading software and OS components across multiple systems
- Collaborate across teams to author and test codebases in diverse programming languages, maintaining and enhancing system functionality
- Produce and refine documentation for systems and user support
- Engage in coordinated maintenance activities, including off-hours operations such as evenings and weekends
- Contribute to global infrastructure projects supporting our HPC ecosystem
- Work closely with research teams to optimize their HPC resource usage
- Develop and maintain tools essential for managing production computing environments
- Provide operational support on a rotating schedule and as required to maintain system health
- Manage vendor relationships, including domestic and international travel to meet current and potential partners
- Follow strict cybersecurity and IT compliance protocols, utilizing only approved hardware and software
- Perform additional duties as assigned
Required Skills and Experience
- At least five years of professional experience in HPC environments, including familiarity with parallel filesystems like Lustre or GPFS, and batch scheduling systems such as Slurm or Grid Engine; experience with high-performance network interconnects is advantageous
- Five years or more administering Linux systems with strong command over system details and performance considerations
- Proficient programming and scripting skills in at least one language such as Go, Python, or C
- Proven ability to architect, build, and maintain complex, interdependent distributed systems
- Experienced in profiling and debugging software stacks using debuggers and performance profilers
- Hands-on knowledge of configuration management tools such as SaltStack, Ansible, or Puppet
- A strong analytical mindset focused on root cause analysis for persistent issues
- Dependable availability and consistency in operational support roles
Additional Information
If you are currently a student or recent graduate, separate campus positions offering internships and full-time roles are available.