Hexaware Technologies

ADB, PySpark and Kafka Developer

Hexaware Technologies

Hyderabad, Telangana, India · Full Time

Be the first to apply

Experience
7–12 yrs
Salary
—
Openings
1
Posted
1 day ago
Work mode
In office
Education
Any graduate
Eligibility
Open to applicants holding any graduate degree.
Resume
Required to apply

Where you'll work

Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.

Job description

About Hexaware Technologies

Hexaware Technologies is a rapidly growing provider of automation-driven next-generation IT, BPO and consulting services. The company is powered by strong strategic initiatives, passionate global teams, and a culture emphasizing innovation and automation. Hexaware's digital solutions help clients achieve operational excellence and superior customer experiences. Their approach focuses on leading clients to gain competitive edges through customer intimacy. Currently, Hexaware is transforming how customers experience services by leveraging industry-leading delivery models based on the principles "Automate Everything™", "Cloudify Everything™", and "Transform Customer Experiences™". At the core of Hexaware’s offerings is Bottom-Up Disruption, a crowdsourcing initiative fostering innovation and simplifying complexities to accelerate client business growth. Hexaware employs a globally diverse workforce exceeding 25,000 professionals dedicated to customer success. The company reported over USD 1 billion in global revenue in 2019.

Role Summary and Responsibilities

  • Design, develop, and optimize scalable ETL/ELT data pipelines using Databricks platforms.
  • Create production-ready data transformation processes utilizing Python, PySpark, and SQL languages.
  • Construct Bronze, Silver, and Gold data layers in observance of the Delta Lake medallion architecture methodology.
  • Apply both batch and streaming ingestion mechanisms where applicable.
  • Leverage Databricks Workflows and Unity Catalog for pipeline orchestration and data governance.
  • Integrate Databricks solutions with Azure data services such as ADLS Gen2 and Synapse Analytics.
  • Implement rigorous data quality assessments, fault handling, monitoring, and tune the performance of data processes.
  • Develop automated deployment pipelines using Continuous Integration and Continuous Delivery (CI/CD) practices and adhere to software engineering best practices.
  • Provide troubleshooting and maintenance support for production data pipelines ensuring their reliability and scalability.

Mandatory Skills

  • Strong, practical experience with Databricks platform, Python/PySpark scripting, and SQL queries.
  • Familiarity with Delta Lake and usage of Databricks Workflows for data pipeline management.
  • Expertise in ETL/ELT development and performance tuning techniques.
  • Experience working with ADLS Gen2 storage and the broader Azure data ecosystem.
  • Knowledge of Unity Catalog for data governance and access control.
  • Proficiency in Git version control system and CI/CD pipeline design and execution.

Application Eligibility

Any graduate can apply for this position.

Minimum education

Bachelor's Degree

Tools & software

Git required Databricks Workflows required Databricks required
🤖
Online · instant AI help
Broxer