Sr.Data Engineering

Bosch

bengaluru, , IndiaPosted 27 days ago
Bosch logo

Skill Required

EngineeringQuery OptimizationGitHub ActionsETLObservabilityCD pipelinesSQL ServerR ProgrammingDatabricksKubernetesdesigningsecuritybuildingSparkPrometheusTestNGAzurePythondesignGitKafkaVaultCI/CDHelmAPIsFulltime

Key highlights

  • 6–8 years of relevant experience required
  • BE, MCA, or M Tech educational qualification
  • Strong focus on Azure data ecosystem (Databricks, Data Factory, SQL Server, Key Vault)
  • Hands-on OpenShift and HELM experience for deployment
  • CI/CD automation using GitHub Actions
  • Monitoring with Grafana for data quality and pipeline health

Role overview

Bosch Global Software Technologies Private Limited, a 100% owned subsidiary of Robert Bosch GmbH, is seeking a data engineering professional with 6–8 years of relevant experience. The role involves designing and optimizing scalable ETL/ELT data pipelines using Python and PySpark within the Azure ecosystem (Databricks, Data Factory, SQL Server, Key Vault), managing cloud infrastructure on OpenShift, and driving CI/CD automation. The candidate will collaborate with data scientists, analysts, and engineering teams to deliver robust data solutions, ensure data quality, and maintain performance across the data platform.

Responsibilities

  • Design, build, and optimize robust, scalable, and efficient ETL/ELT data pipelines using Python and PySpark, primarily within Azure Databricks and Azure Data Factory.
  • Develop and manage processes for ingesting data from various sources (e.g., transactional databases, APIs, streaming sources) and transforming it into clean, usable formats for downstream consumption.
  • Implement comprehensive unit and integration test coverage for data pipelines.
  • Establish and maintain monitoring, alerting, and dashboarding solutions (e.g., Grafana) for data quality, pipeline health, and performance.
  • Contribute to the setup, configuration, and maintenance of data-related infrastructure on OpenShift, ensuring deployment readiness and leveraging tools like HELM for application packaging and deployment.
  • Drive CI/CD best practices using GitHub Actions, ensuring automated testing (unit tests), build, and deployment processes for data solutions to environments like OpenShift.
  • Develop and optimize complex SQL queries for data extraction, transformation, and loading.
  • Apply strong data modeling principles for efficient data storage and retrieval in SQL Server and other data stores.
  • Utilize a broad range of Azure data and analytics services, including Azure Data Factory, Azure Databricks, Azure SQL Server, Azure Key Vault, and others to build comprehensive data solutions.
  • Proactively identify and resolve performance bottlenecks in data pipelines and databases through query optimization, indexing strategies, and efficient data processing techniques.
  • Work closely with data scientists, analysts, and other engineering teams to understand data requirements.
  • Create clear and concise documentation for data pipelines, architecture, and processes.
  • Be flexible to adapt to project requirements needs.

Requirements

  • 6–8 years of relevant experience.
  • BE, MCA, or M Tech.
  • Strong proficiency in Python and PySpark for large-scale data processing and ETL development.
  • Expertise in SQL for complex querying, data manipulation, and schema design.
  • Demonstrable experience in designing, building, and maintaining robust ETL/ELT data pipelines.
  • Hands-on experience with Azure Databricks.
  • Proficiency with core Azure Analytics Services including Azure Data Factory, Azure SQL Server, and Azure Key Vault.
  • Experience implementing CI/CD pipelines from GitHub (including GitHub Actions) for automated testing (unit tests), build, and deployment processes.
  • Familiarity and practical experience with OpenShift (setup, deployment-ready configurations, and management).
  • Experience with HELM for deploying applications on Kubernetes/OpenShift.
  • Experience in setting up and configuring Grafana for dashboards to monitor data quality and pipeline health.

Nice to have

  • Proven experience in SQL optimization and performance tuning.
  • Bachelor's or Master's degree in Computer Science, Engineering, Data Science, or a related quantitative field.
  • Relevant Azure certifications (e.g., Azure Data Engineer Associate).
  • Experience with real-time data processing frameworks (e.g., Kafka, Azure Event Hubs).
  • Understanding of data governance, data security, and compliance best practices.

Additional details

  • Bosch Global Software Technologies Private Limited is a 100% owned subsidiary of Robert Bosch GmbH, one of the world's leading global supplier of technology and services, offering end-to-end Engineering, IT and Business Solutions.
  • With over 27,000+ associates, it's the largest software development center of Bosch, outside Germany, indicating that it is the Technology Powerhouse of Bosch in India with a global footprint and presence in the US, Europe and the Asia Pacific region.
Apply now