Posted today · be early
Senior Data Scientist
Proximity Works
IndiaremotePosted today
Skill Required
Data-ScienceMachine-Learning-EngineeringData-ScientistApplied-Machine-LearningData-EngineeringSenior-Data-ScienceSenior-Staff-Data-ScientistSenior-Machine-Learning-Data-ScientistSenior-Data-Scientist-JobsSenior-Data-Science-SpecialistSenior-Data-Scientist-LeadSenior-Staff-Data-ScienceData ScientistQuery OptimizationMachine LearningScikit-learnEngineeringTensorFlowDatabricksautomationFirewallanalyticsMLflowPyTorchbuildingAirflowetc.)PythonPandasHadoopdesignSparkCloudDesign PatternsjQuerySQLAWSandContract
Key highlights
- Minimum 5 years of experience
- Best-in-class salary
- Expertise in deploying ML models into production
- Hands-on experience with Databricks, Airflow, and AWS EMR
Role overview
We’re seeking a highly skilled, execution-focused Senior Data Scientist with a minimum of 5 years of experience. This role demands hands-on expertise in building, deploying, and optimizing machine learning models at scale, while working with big data technologies and modern cloud platforms. You will be responsible for driving data-driven solutions from experimentation to production, leveraging advanced tools and frameworks across Python, SQL, Spark, and AWS. The role requires strong technical depth, problem-solving ability, and ownership in delivering business impact through data science.
Responsibilities
- Design, build, and deploy scalable machine learning models into production systems.
- Develop advanced analytics and predictive models using Python, SQL, and popular ML/DL frameworks (Pandas, Scikit-learn, TensorFlow, PyTorch).
- Leverage Databricks, Apache Spark, and Hadoop for large-scale data processing and model training.
- Implement workflows and pipelines using Airflow and AWS EMR for automation and orchestration.
- Collaborate with engineering teams to integrate models into cloud-based applications on AWS.
- Optimize query performance, storage usage, and data pipelines for efficiency.
- Conduct end-to-end experiments, including data preprocessing, feature engineering, model training, validation, and deployment.
- Drive initiatives independently with high ownership and accountability.
- Stay up to date with industry best practices in machine learning, big data, and cloud-native deployments.
Requirements
- Minimum 5 years of experience in Data Science or Applied Machine Learning.
- Strong proficiency in Python, SQL, and ML libraries (Pandas, Scikit-learn, TensorFlow, PyTorch).
- Proven expertise in deploying ML models into production systems.
- Experience with big data platforms (Hadoop, Spark) and distributed data processing.
- Hands-on experience with Databricks, Airflow, and AWS EMR.
- Strong knowledge of AWS cloud services (S3, Lambda, SageMaker, EC2, etc.).
- Solid understanding of query optimization, storage systems, and data pipelines.
- Excellent problem-solving skills, with the ability to design scalable solutions.
- Strong communication and collaboration skills to work in cross-functional teams.
Benefits
- Best-in-class salary: We hire strong talent and compensate accordingly.
- Proximity Talks: Meet and learn from designers, engineers, product leaders, and AI practitioners.
- Continuous learning: Work with a world-class team and stay close to the latest in AI, engineering, and product development.
- High-impact work: Build AI-first systems and products used at scale by global clients.
Additional details
- Proximity is the trusted technology, design, and consulting partner for some of the biggest Sports, Media, and Entertainment companies in the world.
- We’re headquartered in San Francisco and have offices in Palo Alto, Dubai, Mumbai, and Bangalore.
- Since 2019, Proximity has built high-impact, scalable products used by millions of users every day.
- Today, we are a global team of engineers, designers, product managers, and experts solving complex problems and building cutting-edge technology at scale.
- Originally posted on Himalayas