Maximum of 25 job preferences reached.
Top Data Engineer Jobs in Kolkata
Information Technology
Build and optimize scalable data pipelines, ETL/ELT processes, data lakes, and real-time and batch processing solutions. Use Databricks, Spark, PySpark, cloud platforms, Python, SQL, orchestration tools, and streaming technologies. Ensure data quality, security, governance, availability, and performance while collaborating with architects, data scientists, analysts, and engineers.
Top Skills:
Amazon RedshiftAmazon S3Apache AirflowApache KafkaSparkAvroAWSAzure Blob StorageBigQueryCi/CdDatabricks Unified Data Analytics PlatformDbtDevOpsGoogle Cloud PlatformInformaticaJSONMatillionAzureParquetPysparkPythonSQLTalend
Information Technology
Design, develop, and maintain data pipelines and ETL processes for data generation, integration, migration, and deployment. Ensure data quality, optimize pipeline performance, and support scalable data solutions across systems. Collaborate with cross-functional teams, document workflows, contribute to technical problem-solving, and guide junior team members while serving as a data engineering subject matter expert.
Top Skills:
Cloud PlatformsData EngineeringData PipelinesData WarehousingDatabase Management SystemsDistributed Computing FrameworksETL
Information Technology
Design, develop, and maintain data solutions, pipelines, data models, and ETL processes. Ensure data quality, reliability, and efficient data migration across systems. Collaborate with multiple teams, provide technical solutions, make team decisions, and serve as a subject matter expert. Monitor and troubleshoot pipeline performance while supporting analytical and operational data needs.
Top Skills:
Cloud Data MigrationData EngineeringData GovernanceData IntegrationData ModelingData Quality FrameworksData WarehousingETL
An Hour AgoSaved
Artificial Intelligence • Cloud • Information Technology • Machine Learning • Natural Language Processing • Consulting
Build and support scalable healthcare data platforms, unified data models, datasets, and ad hoc data access solutions. Extract and integrate heterogeneous healthcare data, improve data quality, support acquisitions, and model data for BI, software applications, machine learning, and AI. Collaborate with product, engineering, business, and clinical stakeholders while following software development lifecycle best practices.
Top Skills:
AdtAWSBi ToolsDatabricksEhrFhirGoogle Cloud PlatformAzurePysparkPythonSparkSQL
Information Technology • Software • Consulting • Cybersecurity
Design, build, and maintain scalable Azure-based data pipelines and systems (Databricks, Data Lake, Synapse, ADF). Collaborate with cross-functional teams, ensure data quality and integrity, optimize performance, monitor and troubleshoot systems, and stay current with industry trends to improve solutions.
Top Skills:
SparkAzure Data Factory (Adf)Azure DatabricksData LakePythonSQLSynapse
Big Data • Cloud • Logistics • Machine Learning • Retail
Designs and optimizes scalable data pipelines and architectures supporting analytics and operational needs. The role develops ETL, streaming, data quality, governance, and cloud-based workflows using modern big data technologies. Responsibilities include translating business requirements into technical solutions, monitoring data systems, improving reliability and performance, and providing technical guidance on data modeling, storage, and metadata management.
Top Skills:
Apache HiveApache KafkaSparkCi/CdData ModelingData WarehousingETLGCPMetadata ManagementPython
eCommerce • Marketing Tech • Design • SEO
Build and maintain scalable ETL/ELT pipelines, cloud data infrastructure, automated workflows, and cross-platform integrations. Manage relational databases and cloud warehouses, optimize data access, and route clean data to AI systems, dashboards, and reporting tools. Implement data quality checks, deduplication, monitoring, governance, and security controls. Collaborate with AI, backend, and marketing teams while managing multiple client integration projects remotely.
Top Skills:
Ai ApisAirflowAWSAzureBigQueryEltETLGa4GCPGoogle AdsJSONMetaN8NPostgresPythonRedshiftRest ApisSalesforceScalaServicenowSnowflakeSQLWebhooksXML
Artificial Intelligence • Information Technology • Software • Cybersecurity
Designs and develops scalable Microsoft Fabric and Azure data platforms, including ETL/ELT pipelines, APIs, integrations, data warehouses, lakehouses, and analytical data models. The role supports Power BI reporting, Salesforce integration, data quality, monitoring, CI/CD, performance optimization, production troubleshooting, and collaboration with business, CRM, reporting, and engineering teams.
Top Skills:
Azure Data FactoryAzure Data Lake StorageAzure Synapse AnalyticsCi/CdDaxGitAzureMicrosoft FabricPower BIPysparkPythonRest ApisSalesforce CRMSQL
Software
Design and maintain batch and real-time data pipelines, data lakes, warehouses, marts, and reusable dbt models using Spark, Kafka, Maxwell, S3, Trino, BigQuery, and related technologies. Optimize distributed processing, data models, queries, storage, reliability, and costs. Establish data quality, observability, governance, automation, and CI/CD practices while partnering with Product, Engineering, Analytics, Data Science, and business stakeholders.
Top Skills:
Amazon S3Apache KafkaSparkCi/CdDbtGitGoogle BigqueryMaxwellMetabasePythonScalaSQLTrino
Blockchain • Fintech • Software • Cryptocurrency • Metaverse
Build and operate scalable financial market data pipelines for Binance stock products and AI applications. Responsibilities include sourcing, ingesting, modeling, processing, storing, and serving real-time and historical data; designing multi-market data frameworks; optimizing Flink pipelines; and establishing quality, reconciliation, monitoring, lineage, and recovery systems. The role also evaluates vendors and data sources, defines data contracts, and collaborates with trading, product, AI, legal, and compliance teams.
Top Skills:
Apache DorisApache FlinkApache HbaseApache KafkaSparkCi/CdClickhouseElasticsearchJavaPythonScalaSQL
Blockchain • Fintech • Software • Cryptocurrency • Metaverse
Build and operate scalable financial market data infrastructure, including source evaluation, ingestion, cleansing, modeling, real-time and batch processing, quality governance, reconciliation, monitoring, and data services. Develop Flink-based pipelines supporting trading, research, and AI applications. Partner with product, trading, compliance, procurement, and engineering teams to define data contracts, source strategies, and consistent financial data semantics across markets.
Top Skills:
Anomaly DetectionApache FlinkApache KafkaSparkCi/CdClickhouseData GovernanceData LineageData ModelingDistributed StorageDorisElasticsearchHbaseJavaKnowledge GraphsMetadata ManagementPythonRagScalaSQLTask Orchestration
Software
Build, monitor, optimize, and maintain cloud-based ETL/ELT pipelines supporting business intelligence and reporting. Develop high-performance SQL transformations, manage data warehouse schemas, tune query and pipeline performance, and optimize cloud costs. Collaborate with analysts, data scientists, actuaries, and architects to translate insurance and SaaS requirements into scalable technical solutions. Participate in code reviews, documentation, troubleshooting, and data integrity initiatives.
Top Skills:
BigQueryCloud ComposerDataflowDdlEltETLGCPGoogle Cloud StorageIamSnowflake SchemaSQLStar Schema
New
Track Smarter, Apply Better.
Ditch the spreadsheets. Organize your job search with our freeApplication Tracker.
Use For Free
Artificial Intelligence • Information Technology • Software • Cybersecurity
Design, build, and support production-grade AWS data pipelines that transform operational data into secure, high-quality, AI-ready datasets. Responsibilities include distributed data processing, Parquet curation, privacy-preserving transformations, orchestration, data quality monitoring, schema management, CI/CD, infrastructure as code, metadata and lineage management, troubleshooting, and reliable backfills.
Top Skills:
AirflowApache HudiApache IcebergSparkAWSAws DmsAws Step FunctionsCi/CdCloudFormationDagsterDebeziumDelta LakeGitParquetPythonSQLTerraform
eCommerce • Information Technology • Consulting
Design, develop, and maintain scalable Microsoft Fabric and Azure data solutions. Build Lakehouse, Data Warehouse, OneLake, ETL/ELT pipelines, data transformations, and integrations using Python, PySpark, SQL, and Fabric tools. Implement medallion architecture, optimize data models and workloads, establish data quality monitoring, and support CI/CD deployments. Collaborate with analytics and business teams to deliver enterprise-scale data platforms.
Top Skills:
AdlsAzure Data FactoryAzure DevopsAzure OpenaiAzure SqlAzure SynapseCi/CdData PipelinesDataflows Gen2DaxDelta LakeDirect LakeFabric Data FactoryFabric Data WarehouseFabric NotebooksGitLakehouseMicrosoft FabricOnelakePower BIPysparkPythonRest ApisSemantic ModelsSQLT-Sql
Cloud • Software
Build and maintain batch and streaming data pipelines across AWS, Azure, and GCP using Databricks, Spark, PySpark, SQL, and Delta Lake. Implement medallion architecture, CDC, SCD, data quality, governance, security, metadata, and lineage controls. Configure storage, orchestration, CI/CD, monitoring, troubleshooting, and performance optimization. Collaborate with architects, data scientists, engineers, analysts, and product teams while documenting pipelines, transformations, testing, and operational procedures.
Top Skills:
Amazon S3Apache AirflowSparkAWSAzure Data FactoryAzure StorageCi/CdDatabricksDatabricks LakeflowDatabricks WorkflowsDelta LakeDelta Live TablesGitGoogle Cloud PlatformGoogle Cloud StorageAzureMicrosoft Azure Dp-203Microsoft PurviewPower BIPysparkSQLUnity Catalog
Information Technology
Designs and supports enterprise data platform blueprints using Snowflake, collaborating with integration and data architects to align systems and data models. The role also involves designing and delivering Generative AI and Agentic AI solutions, applying prompt engineering and AI evaluation frameworks to automate processes and generate business value.
Top Skills:
Agentic AiAi Evaluation FrameworksGenerative AiPrompt EngineeringSnowflake Data Warehouse
Information Technology
Leads enterprise Collibra data governance implementations, including platform configuration, business glossaries, catalogs, lineage, policies, workflows, permissions, and stewardship models. Integrates Collibra with other systems through APIs and connectors, supports data quality and lineage initiatives, manages milestones and stakeholders, and mentors junior team members. Requires strong Collibra expertise, cloud and SQL knowledge, scripting ability, and experience leading enterprise deployments.
Top Skills:
AgileAWSAzureBpmnBusiness GlossaryCollibra Data GovernanceCollibra Data Intelligence CloudData CatalogGCPJobserverLineage HarvesterPolicy ManagerPythonRest ApisScrumSQLWorkflow Designer
Artificial Intelligence • Computer Vision • Software
Design, build, and maintain scalable data pipelines and ETL/ELT workflows using Azure Data Factory, Databricks, PySpark, Python, and SQL. Transform large datasets, optimize data ingestion and queries, monitor pipeline performance, resolve data quality issues, support data modeling, document engineering standards, and contribute to cloud data infrastructure improvements. The role also involves collaboration with architects and senior engineers and requires familiarity with Azure Synapse, Delta Lake, governance, and security.
Top Skills:
Azure Data FactoryAzure Synapse AnalyticsCi/CdDatabricksDelta LakeDelta SharingLakehouse FederationAzurePandasPysparkPythonSparkSQLSQL ServerTerraform
16 Days AgoSaved
Artificial Intelligence • Cloud • Information Technology • Machine Learning • Natural Language Processing • Consulting
Principal Data Engineer responsible for developing and supporting Databricks pipelines, integrating heterogeneous healthcare data sources, modeling unified data products, and ensuring data quality and availability. The role partners with business stakeholders, product teams, data engineers, and clinical leaders to gather requirements, create datasets, support ad hoc analytics, enable machine learning and AI use cases, and assist with acquisition integrations. Extensive healthcare data, cloud, SQL, PySpark, and Python experience is required.
Top Skills:
AdtSparkAWSDatabricksEhrFhirGCPHealthcare Claims DataAzurePysparkPythonSQL
Artificial Intelligence • Cloud • Software
Build and maintain scalable data products using Snowflake, dbt, and Airflow. Responsibilities include dimensional data modeling, SQL transformations, pipeline orchestration, data quality testing, exploratory analysis, SLA monitoring, documentation, and stakeholder requirement gathering. Partner with analytics, product, engineering, and business teams to deliver reliable, production-ready datasets while following software engineering, CI/CD, testing, and deployment practices.
Top Skills:
Apache AirflowDbtPythonSnowflakeSQL
Cloud • Consulting
Design, build, and optimize end-to-end data solutions using Microsoft Fabric, Azure Synapse, Databricks, Azure Data Factory, and Power BI. Develop scalable real-time and batch pipelines, data models, semantic and dimensional models, dashboards, reports, and self-service analytics. Apply medallion architecture, DataOps, monitoring, security, governance, quality controls, and performance tuning. Integrate Azure Machine Learning and streaming technologies when needed while delivering enterprise data and BI solutions for clients.
Top Skills:
Azure Data FactoryAzure Machine LearningAzure Synapse AnalyticsDatabricksDataopsDaxEvent HubsKafkaMicrosoft FabricPower BIPower Query MPythonScalaSpark SqlSQLT-Sql
Greentech • Energy • Solar • Renewable Energy
Designs and maintains SAP-integrated ETL/ELT pipelines connecting ECC, S/4HANA, BW/4HANA, and cloud data platforms. Responsibilities include SAP data extraction, transformation, real-time integrations, dimensional modeling, database optimization, data quality monitoring, alerting, cost optimization, and documentation. The role partners with functional consultants, BI teams, FP&A stakeholders, and business leaders while developing cloud skills across AWS services such as Glue, Lambda, Airflow, and Redshift.
Top Skills:
Amazon RedshiftApache AirflowAws Ec2Aws GlueAws LambdaBapiIdocsNoSQLOdpPysparkPythonRfcSap AbapSap Analytics CloudSap BodsSap BtpSap Bw ExtractorsSap Bw/4HanaSap Cds ViewsSap Data IntelligenceSap Data ServicesSap DatasphereSap EccSap Integration SuiteSap RiseSap S/4HanaSap SltSparkSQL
Fitness • Healthtech • Payments • Software
Lead data engineering design and delivery for CRM data platforms supporting member aggregation, segmentation, audiences, campaigns, reporting, and Connect use cases. Build and optimize SQL workflows across MySQL and Redshift, operate Iceberg-based lakehouse models, own pipeline and infrastructure reliability, and mentor engineers. Partner with Product and engineering teams using Agile, Lean, DevOps, CI/CD, and modern data quality practices.
Top Skills:
AgileAmazon RedshiftApache IcebergAWSCi/CdDbtDevOpsEltETLEvent-Driven ReplicationGoJavaKafkaLeanMySQLSQLSqsTerraformTrino
eCommerce • Other • Retail
Design, build, and maintain enterprise ETL/ELT pipelines in Microsoft Fabric; ingest and transform data from diverse sources into dimensional star schemas; optimize performance and cost; implement data quality checks and monitoring; document architecture and pipelines; collaborate with analysts, stakeholders, and global engineering teams to deliver Power BI-ready analytics datasets.
Top Skills:
Advertising PlatformsAmazon Seller CentralAmazon Vendor CentralAPIsAzure Data FactoryBusiness CentralDataflow Gen2Fabric LakehouseFabric WarehouseJSONMicrosoft Dynamics NavMicrosoft FabricPower BIPythonQuickbooksRestSalesforce Commerce CloudShopifySparkSQLWebhooks
Marketing Tech
As a Senior Data Engineer, you will develop and maintain data pipelines for healthcare data, ensuring data quality, reliability, and adherence to enterprise standards while collaborating with architects and stakeholders.
Top Skills:
AirflowConfluenceDagsterDbtJIRAMatillionPythonSnowflakeSQL
Let Your Resume Do The Work
Upload your resume to be matched with jobs you're a great fit for.
Success! We'll use this to further personalize your experience.
Top Kolkata Companies Hiring Data Engineers
See AllPopular Job Searches
Tech Jobs & Startup Jobs in Kolkata
Remote Jobs in Kolkata
Accounts Executive Jobs in Kolkata
Android Developer Jobs in Kolkata
Artificial Intelligence Jobs in Kolkata
Backend Jobs in Kolkata
Business Analyst Jobs in Kolkata
Content Writing Jobs in Kolkata
Cyber Security Jobs in Kolkata
Data Analyst Jobs in Kolkata
Data Engineer Jobs in Kolkata
Data Science Jobs in Kolkata
Design Engineer Jobs in Kolkata
DevOps Engineer Jobs in Kolkata
Digital Marketing Jobs in Kolkata
Engineering Jobs in Kolkata
Finance Jobs in Kolkata
Front End Developer Jobs in Kolkata
Graphic Designer Jobs in Kolkata
HR Jobs in Kolkata
IT Jobs in Kolkata
IT Support Jobs in Kolkata
Java Developer Jobs in Kolkata
Legal Jobs in Kolkata
Machine Learning Jobs in Kolkata
Marketing Jobs in Kolkata
NET Jobs in Kolkata
Network Engineer Jobs in Kolkata
Operations Manager Jobs in Kolkata
Product Manager Jobs in Kolkata
QA Jobs in Kolkata
React JS Jobs in Kolkata
Sales Jobs in Kolkata
Sales Manager Jobs in Kolkata
SEO Jobs in Kolkata
Software Engineer Jobs in Kolkata
Tech Support Jobs in Kolkata
UX Designer Jobs in Kolkata
Web Developer Jobs in Kolkata
All Filters
Total selected ()
No Results
No Results


























