Data Scientist L1 | IITs, NITs, IIITs, or those with relevant Master’s and Ph.D. degrees. - 6 Positions
Core
Build and optimize scalable data infrastructure, ETL/ELT pipelines, and feature stores to support AI/ML model development and data-driven insights.
Role type
Data Engineer (AI/ML support)
Builds
Production data pipelines, feature stores, and data governance systems
Domain
Artificial Intelligence / Machine Learning / Data Engineering
Deliverable
infrastructure
Required skills
Python, SQL, data preprocessing, feature engineering, ETL/ELT pipeline design, cloud platforms (AWS/Azure/GCP), big data technologies (Spark/Hadoop), data warehousing, data governance
Preferred skills
Experience with Apache Airflow, Talend, data security and privacy best practices
Technologies
Python, SQL, Pandas, NumPy, Scikit-learn, AWS, Azure, Google Cloud, Spark, Hadoop, Databricks, MySQL, PostgreSQL, Apache Airflow, Talend
Responsibilities
Design and maintain efficient ETL/ELT pipelines for data ingestion and transformation; Build reusable feature stores and feature transformation pipelines; Prepare datasets for model training, validation, and testing; Develop dashboards and reports to translate business problems into data-driven solutions; Implement data cataloging, metadata management, and version control practices; Monitor and optimize data processes for performance, scalability, and reliability.
Seniority
Individual Contributor