Job Summary
Seeking a Hadoop Developer with 8 years of experience in big data engineering to design, build, and optimize scalable data pipelines on Hadoop ecosystem technologies. The role involves batch/stream processing, data integration, performance tuning, and close collaboration with business teams.
Key Responsibilities
1. Design and develop robust ETL/data pipelines using Hadoop ecosystem tools.
2. Build and optimize large-scale data processing jobs using Spark and/or Hive.
3. Develop workflows using Airflow.
4. Write complex SQL/HiveQL for data transformation and reporting needs.
5. Ingest structured and unstructured data from multiple sources (Kafka, Sqoop, APIs, files).
6. Ensure data quality, lineage, and governance standards are followed.
7. Monitor, troubleshoot, and tune jobs for performance and scalability.
8. Create technical documentation and follow coding best practices.
Required Skills
1. 8 years of hands-on experience in Hadoop/big data development.
2. Strong knowledge of HDFS, Hive, Spark,
3. String knowledge in Python/Scala
4. Strong SQL and data modelling fundamentals.
5. Experience with workflow orchestration tools (Airflow).
6. Good understanding of Linux, shell scripting, and distributed systems.
8. Familiarity with version control (Git) and Agile delivery practices.
Tech Mahindra represents the connected world, offering innovative and customer-centric information technology experiences, enabling Enterprises, Associates and the Society to Rise . We are a USD 4.9 billion company with 121,840+ professionals across 90 countries, helping over 935 global customers including Fortune 500 companies. Our convergent, digital, design experiences, innovation platforms and reusable assets connect across a number of technologies to deliver tangible business value and experiences to our stakeholders. Tech Mahindra is the highest ranked Non-U.S. company in the Forbes Global Digital 100 list (2018) and in the Forbes Fab 50 companies in Asia (2018).
By continuing you agree to our Terms & Privacy Policy.