Designs and implements data pipelines and processing workflows, optimizes data storage and processing performance, ensures data accuracy, provides technical support, and collaborates with data scientists, analysts, and stakeholders to resolve data issues and meet data requirements.
Data Engineers design and build data systems and pipelines. Responsibilities include developing data processing workflows, optimizing data storage, and ensuring data accuracy. You will collaborate with data scientists and analysts to meet data requirements and resolve data issues. Strong experience in data engineering and problem-solving skills are required.
ResponsibilitiesDesign and implement data pipelines, Optimize data processing and storage, Ensure data solutions meet performance standards, Provide technical support, Collaborate with stakeholders.
QualificationsBachelor's/Master's in Engineering 2-5 years
Similar Jobs
Cloud • Software
Build and maintain batch and streaming data pipelines across AWS, Azure, and GCP using Databricks, Spark, PySpark, SQL, and Delta Lake. Implement medallion architecture, CDC, SCD, data quality, governance, security, metadata, and lineage controls. Configure storage, orchestration, CI/CD, monitoring, troubleshooting, and performance optimization. Collaborate with architects, data scientists, engineers, analysts, and product teams while documenting pipelines, transformations, testing, and operational procedures.
Top Skills:
Amazon S3Apache AirflowSparkAWSAzure Data FactoryAzure StorageCi/CdDatabricksDatabricks LakeflowDatabricks WorkflowsDelta LakeDelta Live TablesGitGoogle Cloud PlatformGoogle Cloud StorageAzureMicrosoft Azure Dp-203Microsoft PurviewPower BIPysparkSQLUnity Catalog
Artificial Intelligence • Computer Vision • Software
Design, build, and maintain scalable data pipelines and ETL/ELT workflows using Azure Data Factory, Databricks, PySpark, Python, and SQL. Transform large datasets, optimize data ingestion and queries, monitor pipeline performance, resolve data quality issues, support data modeling, document engineering standards, and contribute to cloud data infrastructure improvements. The role also involves collaboration with architects and senior engineers and requires familiarity with Azure Synapse, Delta Lake, governance, and security.
Top Skills:
Azure Data FactoryAzure Synapse AnalyticsCi/CdDatabricksDelta LakeDelta SharingLakehouse FederationAzurePandasPysparkPythonSparkSQLSQL ServerTerraform
Artificial Intelligence • Cloud • Information Technology • Machine Learning • Natural Language Processing • Consulting
Principal Data Engineer responsible for developing and supporting Databricks pipelines, integrating heterogeneous healthcare data sources, modeling unified data products, and ensuring data quality and availability. The role partners with business stakeholders, product teams, data engineers, and clinical leaders to gather requirements, create datasets, support ad hoc analytics, enable machine learning and AI use cases, and assist with acquisition integrations. Extensive healthcare data, cloud, SQL, PySpark, and Python experience is required.
Top Skills:
AdtSparkAWSDatabricksEhrFhirGCPHealthcare Claims DataAzurePysparkPythonSQL
What you need to know about the Mumbai Tech Scene
From haggling for the best price at Chor Bazaar to the bustle of Crawford Market, the energy of Mumbai's traditional markets is a key part of the city's charm. And while these markets will always have their place, the city also boasts a thriving e-commerce scene, ranking among the largest in the region. Driven by online sales in everything from snacks to licensed sports merchandise to children's apparel, the local industry is worth billions, with companies actively recruiting to meet the demands of continued growth.


