
at Citi
Bulge Bracket Investment BanksPosted 12 days ago
No clicks
**Senior Big Data Engineer**: Design, build, and optimize large-scale data pipelines using PySpark, Hadoop ecosystem (Hive, HDFS, Sqoop, Spark, Impala), and streaming data platforms. Expertise is sought in data processing, transformation, analysis, and ensuring high availability and low-latency data delivery. 4-7 years of relevant experience required. Hybrid role based in Pune.
- Compensation
- Not specified
- City
- Not specified
- Country
- India
Currency: Not specified
Full Job Description
Senior Big Data Engineer
Job Req Id:
26979589
Location(s):
Pune, Maharashtra, India
Job Type:
Hybrid
Posted:
Jul. 22, 2026
Discover your future at Citi
Working at Citi is far more than just a job. A career with us means joining a team of approximately 219,000 dedicated people from around the globe. At Citi, youll have the opportunity to grow your career, give back to your community and make a real impact.
Job Overview
Citi is looking for a Senior Big Data Engineer to design, build, and optimize large-scale data pipelines and distributed data systems that power critical business intelligence across the organization. Based in Pune and operating in a hybrid model, you will work within a high-performing engineering team where your expertise in PySpark, the Hadoop ecosystem, and streaming data platforms will directly shape the reliability and performance of Citi's data infrastructure.
Responsibilities
- Build and maintain scalable data pipelines using PySpark within a Big Data environment to process and transform large volumes of structured and unstructured data.
- Design and develop solutions across the Hadoop ecosystem including Hive, HDFS, Sqoop, Spark, Impala, and Scala to enable efficient data ingestion, processing, and storage.
- Develop and manage real-time and batch data workflows using streaming data platforms, ensuring high availability and low-latency data delivery.
- Write complex SQL queries to extract, validate, and analyze data across distributed systems, supporting data-driven decision-making.
- Design and implement data models and data architecture patterns aligned with data warehouse principles, ensuring scalability, accuracy, and consistency.
- Automate pipeline scheduling and orchestration using shell scripting and Autosys, reducing manual intervention and improving operational reliability.
- Independently identify, assess, and resolve technical risks and data issues in a timely manner, maintaining system integrity across the data platform.
Required Qualifications & Skills
- 4 -7 years of relevant experience.
- Hands-on expertise in PySpark and Big Data processing, with the ability to build and optimize distributed data workflows at scale.
- Practical knowledge of the Hadoop ecosystem, including Hive, HDFS, Sqoop, Spark, Impala, and Scala, applied in a production environment.
- Proficiency in complex SQL query development for data analysis, transformation, and validation across large datasets.
- Solid understanding of distributed systems architecture and how data flows across interconnected processing layers.
- Demonstrated knowledge of data modelling and data design, with familiarity in data warehouse concepts and dimensional modelling techniques.
- Competence in shell scripting and job scheduling using Autosys or equivalent workflow automation tools.
- Strong analytical and problem-solving ability, with a track record of working independently to diagnose and resolve complex data engineering challenges.
- Clear and effective communication skills, with the ability to articulate technical concepts to both technical and non-technical audiences.
------------------------------------------------------
Job Family Group:
Technology------------------------------------------------------
Job Family:
Applications Development------------------------------------------------------
Time Type:
Full time------------------------------------------------------
Most Relevant Skills
Please see the requirements listed above.------------------------------------------------------
Other Relevant Skills
For complementary skills, please see above and/or contact the recruiter.------------------------------------------------------
Citi is an equal opportunity employer, and qualified candidates will receive consideration without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, disability, status as a protected veteran, or any other characteristic protected by law.
If you are a person with a disability and need a reasonable accommodation to use our search tools and/or apply for a career opportunity review Accessibility at Citi( opens in new window).
View Citis EEO Policy Statement( opens in new window) and the Know Your Rights( opens in new window) poster.




