- Experience
- 5+ yrs
- Salary
- —
- Openings
- 1
- Posted
- 3 days ago
- Work mode
- In office
- Education
- Bachelor's degree in Computer Science, Engineering, Information Systems or related field
- Resume
- Required to apply
Where you'll work
Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.
Job description
About the Company
The client is an established organisation that heavily invests in cutting-edge cloud, data, and artificial intelligence technologies to drive digital transformation and innovation across various business divisions.
Role Overview
This position involves designing, developing, and maintaining large-scale batch and real-time data pipelines on contemporary cloud platforms. You will create robust data integration solutions leveraging APIs, databases, file systems, and event-driven architectures. Additionally, the role includes developing AI-powered knowledge repositories and retrieval-based solutions to support emerging generative AI applications.
Responsibilities
- Design, develop, and maintain scalable batch and streaming data pipelines using modern cloud technologies.
- Create and support data integration workflows involving APIs, databases, file systems, and event-driven systems.
- Develop and manage knowledge base systems and retrieval-based AI solutions to enable generative AI use cases.
- Provide technical leadership regarding platform architecture, uphold coding standards, streamline deployment practices, and ensure operational excellence.
- Enhance platform stability by implementing monitoring solutions, troubleshooting issues, managing releases, and driving continuous process improvements.
Candidate Profile
- Possess a bachelor's degree in Computer Science, Engineering, Information Systems, or related fields.
- Have a minimum of five years' experience in data engineering or delivering cloud-based data platforms.
- Demonstrate solid expertise in Databricks, Apache Spark (including PySpark), and SQL; familiarity with Python, Kafka, Delta Lake, and Microsoft Fabric will be advantageous.
- Show proven experience in designing and supporting production-grade batch and streaming pipelines, including orchestration, monitoring, troubleshooting, and optimizing performance.
- Experienced with implementing Knowledge Base and Retrieval-Augmented Generation (RAG) solutions, covering embeddings, vectorization, retrieval techniques, and generative AI applications.
- Have a strong understanding of CI/CD pipelines, data governance, cloud security, access controls, and platform operations, alongside effective stakeholder communication abilities.
Benefits and Opportunities
- Competitive compensation package paired with promising career progression opportunities.
- Chance to contribute to advanced cloud, data, and AI transformation initiatives.
- Exposure to enterprise-level platforms, varied stakeholder groups, and impactful business projects.
Minimum education
Bachelor's Degree