Capco Logo

Capco

Data Engineer (Databrick + Pyspark)

Sorry, this job was removed at 01:48 p.m. (IST) on Monday, Jul 06, 2026
Remote or Hybrid
Hiring Remotely in India
Remote or Hybrid
Hiring Remotely in India

Similar Jobs at Capco

9 Hours Ago
Remote or Hybrid
India
Mid level
Mid level
Fintech • Professional Services • Consulting • Energy • Financial Services • Cybersecurity • Generative AI
The Senior Data Analyst will define source data, create data mapping, report inconsistencies, assist with data migration, and manage stakeholder expectations while delivering insights in financial services.
Top Skills: CmdHiveNote++PuttyPysparkPythonSQL
9 Hours Ago
Remote or Hybrid
India
Senior level
Senior level
Fintech • Professional Services • Consulting • Energy • Financial Services • Cybersecurity • Generative AI
The Business Analyst will support the Trading Wind Down programme by managing project milestones, collaborating with stakeholders, tracking activities, and generating reports.
Top Skills: AlteryxConfluenceJIRAMS OfficeMicrosoft ProjectsPower BI
9 Hours Ago
Remote or Hybrid
India
Mid level
Mid level
Fintech • Professional Services • Consulting • Energy • Financial Services • Cybersecurity • Generative AI
The Data Analyst role involves defining source data, mapping data sets, reporting inconsistencies, assisting with data migrations, and managing stakeholder expectations. A focus on optimal solutions and strong communication skills in a banking context is essential.
Top Skills: HivePysparkPythonSQL

Job Title: Data Engineer (PySpark / Databricks)

Experience: 5–9 Years Location: Pune (Hybrid – Capco Office)

Job Summary

We are looking for a skilled Data Engineer with strong expertise in PySpark, Databricks, and modern data engineering practices. The ideal candidate will have hands-on experience in building scalable data pipelines, working with large datasets, and leveraging cloud-based data platforms.

Key Responsibilities Design, develop, and maintain scalable ETL/ELT data pipelines Work extensively with PySpark and Apache Spark for large-scale data processing Build and manage workflows using Apache Airflow Develop and optimize data solutions on Databricks (Jobs, Delta Lake) Work with cloud-based data lakes (S3 or equivalent) Write efficient and complex SQL queries for data transformation and analysis Run and manage Spark workloads on EMR Serverless or other managed Spark platforms Ensure data quality, reliability, and performance optimization of pipelines Must Have Skills Strong hands-on experience with PySpark and Apache Spark internals Experience with Databricks (Jobs, Delta Lake) Proficiency in Apache Airflow for workflow orchestration Solid experience building ETL/ELT pipelines at scale Strong SQL skills and experience with Data Warehouse (DWH) systems Experience running Spark workloads on EMR Serverless or managed Spark platforms Hands-on experience with cloud data lakes (S3 or equivalent) Good to Have Skills Experience with Delta Lake / Apache Iceberg Exposure to streaming frameworks (Spark Structured Streaming, Kafka) Familiarity with CI/CD pipelines for data engineering workflows Knowledge of data governance, cataloging, and lineage tools

Capco Mumbai, Maharashtra, IND Office

Capco Technologies PVT Ltd, C/O Wipro Limited, 4 LandsEnd, B.J.Road, Opp Sister's Bungalow, Bandstand, Bandra West, Mumbai, 4000, Mumbai, India, 400050

What you need to know about the Mumbai Tech Scene

From haggling for the best price at Chor Bazaar to the bustle of Crawford Market, the energy of Mumbai's traditional markets is a key part of the city's charm. And while these markets will always have their place, the city also boasts a thriving e-commerce scene, ranking among the largest in the region. Driven by online sales in everything from snacks to licensed sports merchandise to children's apparel, the local industry is worth billions, with companies actively recruiting to meet the demands of continued growth.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account