ADB, PySpark and Kafka Developer
Hyderabad, Telangana, India · Full Time
Be the first to apply
- Experience
- 7–12 yrs
- Salary
- —
- Openings
- 1
- Posted
- 1 day ago
- Work mode
- In office
- Education
- Any graduate
- Eligibility
- Open to applicants holding any graduate degree.
- Resume
- Required to apply
Where you'll work
Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.
Job description
About Hexaware Technologies
Hexaware Technologies is a rapidly growing provider of automation-driven next-generation IT, BPO and consulting services. The company is powered by strong strategic initiatives, passionate global teams, and a culture emphasizing innovation and automation. Hexaware's digital solutions help clients achieve operational excellence and superior customer experiences. Their approach focuses on leading clients to gain competitive edges through customer intimacy. Currently, Hexaware is transforming how customers experience services by leveraging industry-leading delivery models based on the principles "Automate Everything™", "Cloudify Everything™", and "Transform Customer Experiences™". At the core of Hexaware’s offerings is Bottom-Up Disruption, a crowdsourcing initiative fostering innovation and simplifying complexities to accelerate client business growth. Hexaware employs a globally diverse workforce exceeding 25,000 professionals dedicated to customer success. The company reported over USD 1 billion in global revenue in 2019.
Role Summary and Responsibilities
- Design, develop, and optimize scalable ETL/ELT data pipelines using Databricks platforms.
- Create production-ready data transformation processes utilizing Python, PySpark, and SQL languages.
- Construct Bronze, Silver, and Gold data layers in observance of the Delta Lake medallion architecture methodology.
- Apply both batch and streaming ingestion mechanisms where applicable.
- Leverage Databricks Workflows and Unity Catalog for pipeline orchestration and data governance.
- Integrate Databricks solutions with Azure data services such as ADLS Gen2 and Synapse Analytics.
- Implement rigorous data quality assessments, fault handling, monitoring, and tune the performance of data processes.
- Develop automated deployment pipelines using Continuous Integration and Continuous Delivery (CI/CD) practices and adhere to software engineering best practices.
- Provide troubleshooting and maintenance support for production data pipelines ensuring their reliability and scalability.
Mandatory Skills
- Strong, practical experience with Databricks platform, Python/PySpark scripting, and SQL queries.
- Familiarity with Delta Lake and usage of Databricks Workflows for data pipeline management.
- Expertise in ETL/ELT development and performance tuning techniques.
- Experience working with ADLS Gen2 storage and the broader Azure data ecosystem.
- Knowledge of Unity Catalog for data governance and access control.
- Proficiency in Git version control system and CI/CD pipeline design and execution.
Application Eligibility
Any graduate can apply for this position.
Minimum education
Bachelor's Degree