Databricks Engineer
Blend360
IndiaremotePosted 23 days ago
Skill Required
Databricks-EngineerData-EngineeringBig-Data-EngineeringData-Pipeline-EngineerCloud-Data-EngineeringDatabricks-Data-EngineerDatabricks-EngineeringSenior-Databricks-EngineerDatabricks-DeveloperData-Engineering-(Databricks)GCP-Databricks-EngineerDatabricks-ArchitectDatabricks-SpecialistAzure-Databricks-DeveloperDatabricksQuery Optimizationdata engineeringGitHub ActionsCloud SecurityCD pipelinesAzureEngineeringObservabilityanalyticssecuritySparkKafkaPythondesignDevOpsGitCI/CDCloudETLSQLandCICDAIFulltime
Key highlights
- Requires 4+ years of Data Engineering experience
- Must have strong hands‑on Azure Databricks expertise
- Proficiency in Python and advanced SQL required
- Experience with Delta Lake and Azure Data Lake Storage (ADLS Gen2)
- Exposure to CI/CD pipelines and streaming data preferred
- Collaboration with data scientists/analysts for ML use cases
Role overview
We are looking for an experienced Azure Databricks Engineer with strong hands‑on expertise in Python, SQL, and Apache Spark to design, build, and optimize scalable data pipelines and analytics solutions on the Azure cloud platform.
Responsibilities
- Design, develop, and maintain scalable data pipelines using Azure Databricks
- Implement ETL/ELT workflows using PySpark, Spark SQL, and Python
- Optimize Spark jobs for performance, cost, and scalability
- Work with structured and semi‑structured data (Parquet, Delta, JSON, CSV)
- Build and manage Delta Lake tables (ACID, time travel, schema evolution)
- Integrate Databricks with Azure Data Lake Storage (ADLS Gen2)
- Develop complex queries and transformations using SQL
- Collaborate with data scientists, analysts, and stakeholders to support analytics and ML use cases
- Ensure data quality, validation, and monitoring
- Follow best practices for security, access control, and governance in Azure
Requirements
- 4+ years of experience in Data Engineering
- Strong hands‑on experience with Azure Databricks
- Proficiency in Python for data processing
- Strong knowledge of SQL (joins, window functions, performance tuning)
- Hands‑on experience with Apache Spark / PySpark
- Experience working with Delta Lake
- Knowledge of Azure Data Lake Storage (ADLS Gen2)
- Understanding of distributed computing concepts
- Experience with Git version control
- Experience with Azure Data Factory
- Basic understanding of data modeling
- Familiarity with cloud security and RBAC in Azure
Nice to have
- Exposure to CI/CD pipelines (Azure DevOps, GitHub Actions)
- Exposure to streaming data (Spark Structured Streaming, Event Hub, Kafka)
Additional details
- Blend is a premier AI services provider, committed to co‑creating meaningful impact for its clients through the power of data science, AI, technology, and people. With a mission to fuel bold visions, Blend tackles significant challenges by seamlessly aligning human expertise with artificial intelligence. The company is dedicated to unlocking value and fostering innovation for its clients by harnessing world‑class people and data‑driven strategy. We believe that the power of people and AI can have a meaningful impact on your world, creating more fulfilling work and projects for our people and clients. For more information, visit .
- Originally posted on Himalayas