Posted today · be early
Data Engineer - Global Team (India)
YipitData
IndiaremotePosted today
Skill Required
Data-EngineeringData-EngineerData-Pipeline-EngineeringAnalytics-EngineeringETL-EngineeringData-Engineering-TeamData-Engineer-JobsData-Engineering-PositionsData Engineerdata engineeringETLEngineeringDatabricksanalyticalSnowflakeanalyticsbuildingAirflowwrittenSparkCloudDesign PatternsdbtSQLandAIFulltime
Key highlights
- Required experience: 4-7 years of data engineering experience
- Notable requirement: Proficiency in PySpark, Delta, and Databricks
- Key benefit: Learning reimbursement and parental leave
- Location: Remote based in India
Role overview
YipitData is a leading market research and analytics firm for the disruptive economy, utilizing proprietary technology to analyze billions of alternative data points. We are seeking a highly skilled Data Engineer to join our dynamic Data Engineering team as the first hire for a strategic pipeline, with the potential to build and lead the team as responsibilities expand. This high-impact, high-visibility role serves as the gold standard for all other YipitData analyst teams, building and maintaining the core pipelines and tooling that power the company's products. The role is a remote opportunity based in India.
Responsibilities
- Report directly to a Manager of Data Engineering, who will provide significant, hands-on training on cutting-edge data tools and techniques.
- Build and maintain end-to-end data pipelines.
- Help with setting best practices for our data modeling and pipeline builds.
- Create AI-ready analytical datasets designed with the structure, metadata, documentation, and business context needed for effective use by AI agents and insight-driven applications.
- Become an expert at solving complex data pipeline issues using PySpark and SQL.
- Collaborate with stakeholders to incorporate business logic into our central pipelines.
- Deeply learn Databricks, Spark, AI, and other ETL toolings developed internally.
Requirements
- 4-7 years of data engineering experience.
- Solid understanding of Spark.Pyspark, and SQL.
- Data pipeline experience.
- Bachelor’s or Master’s degree in Computer Science, STEM, or a related technical discipline.
- 4+ years of experience as a Data Engineer or in other technical functions.
- Excited about solving data challenges and learning new skills.
- Great understanding of working with data or building data pipelines.
- Comfortable working with large-scale datasets using PySpark, Delta, and Databricks.
- Understand business needs and the rationale behind data transformations to ensure alignment with organizational goals and data strategy.
- Eager to constantly learn new technologies.
- Self-starter who enjoys working collaboratively with stakeholders.
- Exceptional verbal and written communication skills.
Nice to have
- Experience with Airflow, dbt, Snowflake, or equivalent.
Benefits
- Comprehensive benefits, perks, and a competitive salary.
- Vacation time.
- Parental leave.
- Team events.
- Learning reimbursement.
- Environment focused on ownership, respect, and trust where growth is determined by impact rather than tenure, unnecessary facetime, or office politics.
Additional details
- YipitData recently raised $475M from The Carlyle Group at a valuation of over $1B.
- Offices are located in the US (NYC, Austin, Miami, Mountain View), APAC (Hong Kong, Shanghai, Beijing, Guangzhou, Singapore), and India.
- Recognized by Inc. as a Best Workplace for three consecutive years.
- Remote opportunity based in India.
- During training and onboarding, several hours of overlap with US working hours are expected.
- Afterward, standard IST working hours are permitted, with the exception of 1-2 days per week for meetings with the US team.
- Committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, marital status, disability, gender, gender identity or expression, or veteran status.
- Originally posted on Himalayas