Job description Location Mumbai Role Overview : As a Big Data Engineer, you'll design and build robust data pipelines on Cloudera using Spark (Scala/PySpark) for ingestion, transformation, and processing of high-volume data from banking systems. Key Responsibilities : Build scalable batch and real-time ETL pipelines using Spark and Hive Integrate structured and unstructured data sources Perform performance tuning and code optimization Support orchestration and job scheduling (NiFi, Airflow) Required education Bachelor's Degree Preferred education Master's Degree Required technical and professional expertise Skills Required : Proficiency in PySpark/Scala with Hive/Impala Experience with data partitioning, bucketing, and optimization Familiarity with Kafka, Iceberg, NiFi is a must Knowledge of banking or financial datasets is a plus Role: Data Engineer Industry Type: IT Services & Consulting Department: Engineering - Software & QA Employment Type: Full Time, Permanent Role Category: Software Development Education UG: Any Graduate PG: Any Postgraduate Key Skills Skills highlighted with ‘‘ are preferred keyskills hivescalaimpalaapache nifikafka clouderapysparkapache pigsqlapachejavaunix shell scriptingsparkflumelinuxmysqlhadoopbig datahbasepythonoracleooziemicrosoft azurenosqlmapreducesqoopawsunix
Job Summary
Company
IBM
Location
Mumbai
Salary Range
Not disclosed
Experience
1-4 Yrs
Verification
Government Verified Recruiter
TrueCV Commitment Pledge
Zero Ghosting Guarantee
Every recruiter on TrueCV commits to review applications, respect interview time, and authenticate joining via our Day 1 Handshake.