Sr.Data Engineering
Bosch
bengaluru, , IndiaPosted 27 days ago
Skill Required
EngineeringQuery OptimizationGitHub ActionsETLObservabilityCD pipelinesSQL ServerR ProgrammingDatabricksKubernetesdesigningsecuritybuildingSparkPrometheusTestNGAzurePythondesignGitKafkaVaultCI/CDHelmAPIsFulltime
Key highlights
- 6–8 years of relevant experience required
- BE, MCA, or M Tech educational qualification
- Strong focus on Azure data ecosystem (Databricks, Data Factory, SQL Server, Key Vault)
- Hands-on OpenShift and HELM experience for deployment
- CI/CD automation using GitHub Actions
- Monitoring with Grafana for data quality and pipeline health
Role overview
Bosch Global Software Technologies Private Limited, a 100% owned subsidiary of Robert Bosch GmbH, is seeking a data engineering professional with 6–8 years of relevant experience. The role involves designing and optimizing scalable ETL/ELT data pipelines using Python and PySpark within the Azure ecosystem (Databricks, Data Factory, SQL Server, Key Vault), managing cloud infrastructure on OpenShift, and driving CI/CD automation. The candidate will collaborate with data scientists, analysts, and engineering teams to deliver robust data solutions, ensure data quality, and maintain performance across the data platform.
Responsibilities
- Design, build, and optimize robust, scalable, and efficient ETL/ELT data pipelines using Python and PySpark, primarily within Azure Databricks and Azure Data Factory.
- Develop and manage processes for ingesting data from various sources (e.g., transactional databases, APIs, streaming sources) and transforming it into clean, usable formats for downstream consumption.
- Implement comprehensive unit and integration test coverage for data pipelines.
- Establish and maintain monitoring, alerting, and dashboarding solutions (e.g., Grafana) for data quality, pipeline health, and performance.
- Contribute to the setup, configuration, and maintenance of data-related infrastructure on OpenShift, ensuring deployment readiness and leveraging tools like HELM for application packaging and deployment.
- Drive CI/CD best practices using GitHub Actions, ensuring automated testing (unit tests), build, and deployment processes for data solutions to environments like OpenShift.
- Develop and optimize complex SQL queries for data extraction, transformation, and loading.
- Apply strong data modeling principles for efficient data storage and retrieval in SQL Server and other data stores.
- Utilize a broad range of Azure data and analytics services, including Azure Data Factory, Azure Databricks, Azure SQL Server, Azure Key Vault, and others to build comprehensive data solutions.
- Proactively identify and resolve performance bottlenecks in data pipelines and databases through query optimization, indexing strategies, and efficient data processing techniques.
- Work closely with data scientists, analysts, and other engineering teams to understand data requirements.
- Create clear and concise documentation for data pipelines, architecture, and processes.
- Be flexible to adapt to project requirements needs.
Requirements
- 6–8 years of relevant experience.
- BE, MCA, or M Tech.
- Strong proficiency in Python and PySpark for large-scale data processing and ETL development.
- Expertise in SQL for complex querying, data manipulation, and schema design.
- Demonstrable experience in designing, building, and maintaining robust ETL/ELT data pipelines.
- Hands-on experience with Azure Databricks.
- Proficiency with core Azure Analytics Services including Azure Data Factory, Azure SQL Server, and Azure Key Vault.
- Experience implementing CI/CD pipelines from GitHub (including GitHub Actions) for automated testing (unit tests), build, and deployment processes.
- Familiarity and practical experience with OpenShift (setup, deployment-ready configurations, and management).
- Experience with HELM for deploying applications on Kubernetes/OpenShift.
- Experience in setting up and configuring Grafana for dashboards to monitor data quality and pipeline health.
Nice to have
- Proven experience in SQL optimization and performance tuning.
- Bachelor's or Master's degree in Computer Science, Engineering, Data Science, or a related quantitative field.
- Relevant Azure certifications (e.g., Azure Data Engineer Associate).
- Experience with real-time data processing frameworks (e.g., Kafka, Azure Event Hubs).
- Understanding of data governance, data security, and compliance best practices.
Additional details
- Bosch Global Software Technologies Private Limited is a 100% owned subsidiary of Robert Bosch GmbH, one of the world's leading global supplier of technology and services, offering end-to-end Engineering, IT and Business Solutions.
- With over 27,000+ associates, it's the largest software development center of Bosch, outside Germany, indicating that it is the Technology Powerhouse of Bosch in India with a global footprint and presence in the US, Europe and the Asia Pacific region.