Top Data Engineer Jobs in Bengaluru

3 Days AgoSaved
In-Office
Bengaluru, KA
Senior level
Senior level
Artificial Intelligence • eCommerce • Machine Learning • Software
Design, build, and maintain scalable batch and real-time data pipelines, ETL/ELT workflows, data models, warehouses, and lakes. Ensure data quality, reliability, security, and performance while integrating APIs, databases, cloud platforms, and third-party systems. Collaborate with data science, machine learning, engineering, and product teams to provide datasets for analytics and machine learning. Monitor pipelines, troubleshoot infrastructure issues, implement testing and validation frameworks, and contribute to data platform architecture.
Top Skills: Apache AirflowSparkAWSAzureCi/CdDagsterData LakesData WarehousesDistributed SystemsEtl/EltGCPGitPysparkPythonRelational DatabasesSQL
4 Days AgoSaved
In-Office
Bengaluru, KA
Junior
Junior
Artificial Intelligence • Fintech • Software • Financial Services
Build and manage scalable data pipelines in Databricks, transform raw data with SQL and dbt, and produce reliable data assets. Collaborate across teams to deliver insights for products, AI, and business decisions. Support machine learning research by validating model performance, optimize data processing and reporting, and work with AWS services, Airflow, visualization tools, and databases.
Top Skills: AirflowAWSCloudfrontDatabricksDbtEc2EksHexLambdaMongoDBNode.jsNoSQLPostgresPower BIPythonRReactRedisS3SQLSqsTableau
Senior level
HR Tech • Information Technology • Professional Services
Build scalable telemetry and usage-data pipelines using AWS Glue, PySpark, Python, SQL, and Snowflake. Transform raw data into rated, validated, invoice-ready outputs for SAP billing and usage-based billing systems. Implement rating logic, data quality checks, reconciliation, audit trails, exception handling, monitoring, and CI/CD deployments. Optimize data models and pipelines while supporting batch and event-driven ingestion and collaborating with engineering, finance, billing, and operations teams.
Top Skills: Aws CodepipelineAws GlueCi/CdGitGithub ActionsGitlab CiJenkinsPysparkPythonSAPSnowflakeSQL
7 Days AgoSaved
In-Office
Bengaluru, KA
Mid level
Mid level
Financial Services
Build and optimize high-throughput market and execution-data pipelines using Python, R, and PostgreSQL. Orchestrate scheduled workflows, maintain 24/7 data availability, monitor pipeline health, implement data-quality checks, and develop REST APIs for analytics. Collaborate with quants and product managers to turn transaction-cost research into production features and client analyses while applying statistical methods to order-flow data.
Top Skills: AIAirflowC++DagsterDjangoFastapiMlNumpyPandasPolarsPostgresPrefectPythonRRest Apis
YesterdaySaved
Remote
Bengaluru, KA
Mid level
Mid level
Artificial Intelligence • Computer Vision • Software
Design, build, and maintain scalable data pipelines and ETL/ELT workflows using Azure Data Factory, Databricks, PySpark, Python, and SQL. Transform large datasets, optimize data ingestion and queries, monitor pipeline performance, resolve data quality issues, support data modeling, document engineering standards, and contribute to cloud data infrastructure improvements. The role also involves collaboration with architects and senior engineers and requires familiarity with Azure Synapse, Delta Lake, governance, and security.
Top Skills: Azure Data FactoryAzure Synapse AnalyticsCi/CdDatabricksDelta LakeDelta SharingLakehouse FederationAzurePandasPysparkPythonSparkSQLSQL ServerTerraform
Expert/Leader
Artificial Intelligence • Cloud • Information Technology • Machine Learning • Natural Language Processing • Consulting
Principal Data Engineer responsible for developing and supporting Databricks pipelines, integrating heterogeneous healthcare data sources, modeling unified data products, and ensuring data quality and availability. The role partners with business stakeholders, product teams, data engineers, and clinical leaders to gather requirements, create datasets, support ad hoc analytics, enable machine learning and AI use cases, and assist with acquisition integrations. Extensive healthcare data, cloud, SQL, PySpark, and Python experience is required.
Top Skills: AdtSparkAWSDatabricksEhrFhirGCPHealthcare Claims DataAzurePysparkPythonSQL
12 Days AgoSaved
In-Office
Bengaluru, KA
Mid level
Mid level
eCommerce • Business Intelligence
Build and maintain scalable data pipelines, BigQuery warehouses, Dataform/dbt transformations, data services, APIs, and data quality monitoring in GCP. Support machine learning workflows and develop an AI-driven insights layer. Collaborate with analysts, data scientists, and stakeholders to deliver reliable data solutions for Easyship’s SaaS and eCommerce shipping platform. This is a full-time onsite role in Bangalore aligned with UK working hours.
Top Skills: Apache AirflowAPIsBigQueryCloud ComposerCloud FunctionsDataformDbtGoogle Cloud Platform (Gcp)KubeflowMlflowPub/SubSQLVertex Ai
14 Days AgoSaved
In-Office
Bengaluru, KA
Senior level
Senior level
Artificial Intelligence • Fintech • Software • Financial Services
Lead migration from Matillion ETL to AWS-native data platforms. Design and optimize scalable pipelines using AWS Glue and Apache Airflow, build reusable ETL/ELT components, and work with Snowflake for data warehousing and performance tuning. Implement data quality, monitoring, security, governance, and compliance practices. Collaborate with architects and stakeholders, support CI/CD and DevOps processes, conduct code reviews, mentor engineers, and improve processing performance, cost, and scalability.
Top Skills: Apache AirflowAWSAws GlueAws LambdaAws Step FunctionsCi/CdDevOpsGitMatillion EtlPysparkPythonShell ScriptingSnowflakeSnowpipeSnowsqlSparkSQL
15 Days AgoSaved
In-Office
Bengaluru, KA
Expert/Leader
Expert/Leader
Professional Services • Consulting • Financial Services
Designs, builds, and optimizes scalable data pipelines, ETL processes, data models, APIs, and cloud-based data infrastructure. The role integrates structured and unstructured data, manages databases, warehouses, lakehouses, and distributed systems, and ensures data quality, security, governance, and performance. Responsibilities also include troubleshooting, technical documentation, architectural contribution, cross-functional collaboration, and mentoring team members.
Top Skills: AWSAzureAzure Data FactoryData LakesDatabricksETLGCPHadoopJavaKafkaLakehousesNoSQLPythonRest ApisScalaSnowflakeSparkSQLSsisStored Procedures
16 Days AgoSaved
In-Office
Bengaluru, KA
Senior level
Senior level
Information Technology • Consulting • Financial Services
Design, develop, optimize, and maintain scalable ETL pipelines using PySpark and AWS Glue. Build data ingestion and transformation workflows, orchestrate processes with AWS Step Functions, and develop Python-based AWS Lambda functions. Support AWS data lake architectures, analytical use cases, monitoring, reliability, and performance optimization. The role also involves SQL and may include Java microservices, REST APIs, and backend integrations.
Top Skills: AWSAws Data LakesAws GlueAws LambdaAws Step FunctionsETLJavaPysparkPythonRest ApisSQL
Reposted 16 Days AgoSaved
In-Office
Bengaluru, KA
Mid level
Mid level
Greentech
Design, build, and maintain scalable data pipelines, ETL processes, and data models to support analytics and product teams. Ensure data quality and reliability, instrument data workflows, and collaborate with US and global engineering teams to operationalize data infrastructure for SPAN's electrification and home energy solutions.
Reposted 15 Days AgoSaved
Remote
Bengaluru, KA
Senior level
Senior level
Analytics
Build, optimize, and maintain scalable AWS-based data pipelines and data lakes (Snowflake/Databricks). Implement streaming ingestion (Kinesis/Kafka), CDC, automated data quality checks, and ML-ready feature stores. Collaborate with Product, ML, and Analytics teams, own pipeline QA, observability, and participate in on-call rotations to resolve production data issues.
Top Skills: Apache IcebergAWSAws BedrockAws GlueAws KinesisAws S3Aws SnsAws SqsAws Step FunctionsDatabricksDatabricks AiDatabricks WorkflowsDebeziumDelta LakeFeature StoreFivetranKafkaParquetPysparkPythonSnowflakeSnowflake CortexSQL
New

Track Smarter, Apply Better.

Ditch the spreadsheets. Organize your job search with our freeApplication Tracker.

Use For Free
Application Tracker Preview
4 Days AgoSaved
In-Office
Bengaluru, KA
Senior level
Senior level
Information Technology • Software
Design, develop, and optimize scalable ETL pipelines and distributed data applications for batch and streaming workloads. Manage workflows with Apache Airflow, operate data platforms on AWS and Amazon S3, ensure data quality and performance, troubleshoot production issues, and collaborate with engineering, product, and analytics teams.
Top Skills: Amazon S3Apache AirflowApache KafkaSparkAWSDatabricksPythonScalaSQL
Reposted 18 Days AgoSaved
Remote
Bengaluru, KA
Senior level
Senior level
Energy • Renewable Energy
Design, build, and maintain production data pipelines and integrations across applications, databases, and cloud services. Support data architecture, ingestion, embedding workflows for AI, data quality, monitoring, and platform reliability. Collaborate with Full Stack and AI teams to deploy scalable cloud-based data infrastructure and establish data engineering standards and best practices.
Top Skills: AirflowAws CdkAws Ec2Aws EventbridgeAws GlueAws IamAws LambdaAws RdsAws S3Aws Step FunctionsBigQueryCi/CdContainerizationDbtDockerEmbedding PipelinesEvent-Driven ArchitecturesGitGitPandasPolarsPostgisPostgresPythonRag ArchitecturesRedshiftSnowflakeSQLTerraformVector Databases
Reposted 28 Days AgoSaved
In-Office
Bengaluru, KA
Mid level
Mid level
Information Technology • Consulting
As a Data Engineer, you'll optimize data workflows, design robust data solutions, and collaborate with teams to support data-driven initiatives.
Top Skills: Apache AirflowSparkAws RedshiftAzure Synapse AnalyticsGoogle Bigquery
7 Days AgoSaved
In-Office
Bengaluru, KA
Senior level
Senior level
Software
Build and own Zinier’s AWS-based medallion data platform, including canonical data models, CDC pipelines, ETL transformations, dimensional schemas, tenant isolation, governance, lineage, and self-service BI. Migrate divergent client schemas into a unified model, enable QuickSight reporting, and establish clean historical data foundations for AI/ML. Collaborate with Product, Solutions, Engineering, Customer Success, and AI teams while balancing scalable architecture with hands-on implementation and troubleshooting.
Top Skills: 3NfAmazon AthenaAmazon QuicksightAmazon RdsAmazon RedshiftAmazon S3AWSAws Database Migration Service (Dms)Aws GlueAws Glue Data CatalogCdcEltETLGoJavaMedallion ArchitecturePythonRedshift ServerlessStar Schemas
Reposted One Month AgoSaved
In-Office
Bengaluru, KA
Mid level
Mid level
Information Technology • Software • Analytics • Biotech
The AWS Data Engineer will design and implement data pipelines using AWS services, develop ETL processes, and optimize data storage solutions.
Top Skills: Aws DynamodbAws EmrAws GlueAws RdsAws RedshiftAws S3PysparkPythonSQL
Reposted One Month AgoSaved
In-Office
Bengaluru, KA
Mid level
Mid level
Artificial Intelligence • eCommerce • Sales • Software
Build and maintain scalable ETL/ELT pipelines, data warehouses/lakes, and real-time analytics infrastructure. Ensure data quality, governance, and deliver high-quality datasets for product, analytics, and customer-facing dashboards.
Top Skills: AirflowSparkAWSAzureBigQueryDbtDockerGCPKafkaKubernetesLookerPower BIPythonRedshiftSnowflakeSQLTableau
Reposted One Month AgoSaved
In-Office
Bengaluru, KA
Senior level
Senior level
Financial Services
Build and scale end-to-end ETL/ELT pipelines (batch and near-real-time), design ingestion systems, optimize cloud data warehouses, enforce data quality and monitoring, define data models/contracts, enable ML/AI workflows, and operate data infrastructure using Python, PostgreSQL, MongoDB, and AWS.
Top Skills: AirflowAWSData WarehouseDbtDockerEltETLFeature StoreKafkaKubernetesLakehouseMongoDBPlaywrightPostgresPythonSeleniumSparkSQL
Reposted 25 Days AgoSaved
Remote
Bengaluru, KA
Mid level
Mid level
eCommerce • Healthtech • Retail
Owner of company-wide reporting and analytics for Growth, Marketing, Product, Finance and Operations. Build and maintain dashboards, automated reports and data pipelines (dbt), ensure data quality and governance, run growth and product analyses (CAC, ROAS, LTV, retention), support A/B testing and ad-hoc decision support, leverage AI for automation, and coordinate VAs on data tasks.
Top Skills: Ai ToolsBigQueryDatabricksDbtETLExcelGoogle SheetsLookerMetabasePower BIRedshiftSnowflakeSQLTableau
12 Days AgoSaved
In-Office
Bengaluru, KA
Senior level
Senior level
Artificial Intelligence • Cloud • Machine Learning • Consulting
Lead enterprise data engineering initiatives, define data architecture, and build scalable cloud-native platforms and Lakehouse solutions. Lead engineering teams, cloud migrations, governance, quality, metadata, security, and operational excellence efforts. Review technical designs and code, translate business requirements into solutions, and champion CI/CD, Infrastructure as Code, automation, and DataOps. Mentor engineers and drive modernization across the software development lifecycle.
Top Skills: Apache KafkaSparkAzure Data FactoryAzure Data Lake StorageAzure Synapse AnalyticsCi/CdDatabricksDataopsDevOpsInfrastructure As CodeLakehouse ArchitectureAzurePythonScalaSnowflakeTerraform
26 Days AgoSaved
Remote
Bengaluru, KA
Junior
Junior
Artificial Intelligence • Insurance • Software • Automation
Build and operate ingestion and transformation pipelines to turn messy insurance data into clean datasets. Design bronze/silver/fact layers, model and tune Postgres and analytical stores, define KPIs and reconciliation checks, orchestrate observable data jobs, and ensure end-to-end data quality for analytics and AI products.
Top Skills: AirflowClickhouseDagsterPostgresPythonSQLTemporal
Reposted One Month AgoSaved
In-Office
Bengaluru, KA
Mid level
Mid level
Information Technology • Software • Consulting
Design, build, test, and deploy ETL/ELT data pipelines and data‑warehouse solutions for AML/compliance domains. Produce technical specifications, coordinate stakeholders, automate data ingestion controls/tests, implement DevOps practices, and support releases while ensuring data quality, lineage, and regulatory requirements are met.
Top Skills: Apache AirflowApache AtlasSparkAwr/AshAws EmrAzure Data FactoryAzure SynapseBashCollibraCx_OracleDbtDockerDuckdbExternal TablesFunctions)Gcp DataprocGitGithub ActionsGrafanaJenkinsKafkaKubernetesMxObject StorageOci (Autonomous DbOci Resource ManagerOracleOracle Data PumpOracledbPandasPl/SqlPolarsPowershellPrefectPrometheusPyarrowPysparkPytestPythonRabbitMQScalaSql*LoaderSqlalchemySwiftTerraformUnittest
Reposted One Month AgoSaved
Remote
Bengaluru, KA
Mid level
Mid level
Software
Design, implement, and maintain data pipelines using Elasticsearch and data warehousing solutions, integrate data sources, and ensure data accuracy while collaborating with cross-functional teams.
Top Skills: DatabricksElasticsearchElk StackFlaskGitJavaLinuxMongoDBMySQLPandasPostgresPysparkPythonRedisSparkSQL
14 Days AgoSaved
In-Office
Bengaluru, KA
Senior level
Senior level
Software
Own the architecture, development, and operation of scalable data pipelines, distributed data systems, databases, and data products. Define data quality and reliability standards, optimize data models and SQL workloads, build reusable infrastructure, and establish testing, observability, and deployment practices. Partner with engineering, AI, data science, and product teams to deliver reliable data solutions, influence platform strategy, and mentor engineers.
Top Skills: BigQueryC++FlinkGoJavaJavaScriptKafkaNosql DatabasesPrestoPythonRelational DatabasesSnowflakeSparkSQLTrino
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account