Infosys

Big Data Engineer

Infosys

Noida, Uttar Pradesh, India · Full Time

Be the first to apply

Experience
5–15 yrs
Salary
Openings
1
Posted
3 days ago
Work mode
In office
Education
B.Tech / B.E.
Eligibility
Candidates with a bachelor's degree in Engineering or Technology (B.Tech/B.E.) in any specialization are eligible to apply.
Resume
Required to apply

Where you'll work

Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.

Job description

About the Role

We seek seasoned Big Data Engineers with 5 to 15 years of expertise to architect, develop, and enhance large-scale data processing systems and analytics platforms. The ideal professional will have comprehensive experience in Big Data technologies, distributed computing setups, cloud environments, and robust data engineering methodologies to drive enterprise-wide data initiatives.

Key Responsibilities

  • Design, develop, and support scalable Big Data architectures to handle extensive structured and unstructured datasets.
  • Construct and fine-tune data pipelines, ETL/ELT workflows, and real-time streaming applications.
  • Establish data ingestion frameworks integrating diverse sources such as APIs, databases, file systems, and streaming platforms.
  • Implement both batch and streaming data processing using Big Data technology stacks.
  • Collaborate closely with Data Architects, Scientists, Business Analysts, and other stakeholders to gather and refine business needs.
  • Maintain high standards of data quality, integrity, governance, and security throughout the data lifecycle.
  • Enhance data processing efficiency and resolve production-level challenges.
  • Actively participate in architecture and code reviews along with technical brainstorming sessions.
  • Assist in cloud migration strategies and modernization projects.
  • Document technical processes and operational guidelines comprehensively.

Required Expertise

  • Proven command over Big Data ecosystems including Hadoop, Apache Spark (Core, SQL, Streaming), Hive, HDFS, and YARN.
  • Development experience with distributed data processing solutions.
  • Strong background in ETL/ELT workflows, data modeling, and data warehousing principles.
  • Experience managing large-scale data transformation and migration endeavors.
  • Knowledge of data governance and quality assurance frameworks.
  • Programming skills in Python, Scala, Java, and SQL with automation and scripting capabilities.
  • Hands-on experience with streaming platforms such as Apache Kafka, Spark Streaming, preferred Apache Flink, and Apache NiFi.
  • Familiarity with cloud service providers including AWS, Microsoft Azure, and Google Cloud Platform.
  • Database proficiency with SQL Server, Oracle, PostgreSQL, MySQL, MongoDB, and preferably Cassandra.
  • Practical knowledge of tools like Databricks, Snowflake, and Airflow.

Additional Information

Preferred Notice Period: Immediate to 30 days.

Minimum education

Bachelor's Degree

Tools & software

Java AWS Python required Apache Spark required Apache Kafka required Scala required
🤖
Online · instant AI help
Broxer