Must-Have Skills
β’ 7+ years of experience in Python Development & Data Engineering
β’ Strong hands-on experience in Offline Evaluation Metrics (nDCG, MRR, MAP@K)
β’ Expertise in PySpark and distributed data processing
β’ Experience in building scalable ETL/data pipelines
β’ Pipeline automation and workflow orchestration experience
β’ Strong SQL and data transformation skills
β’ Experience with large-scale datasets and performance tuning
β’ Understanding of ML/Recommender System evaluation workflows
β’ Experience with Git, CI/CD, and production-grade deployments
β’ Good problem-solving and debugging skills
Good-to-Have Skills
β’ Experience with Airflow, Luigi, or Prefect
β’ Exposure to AWS/GCP/Azure data services
β’ Knowledge of ML Ops / model evaluation frameworks
β’ Experience with Databricks or Hadoop ecosystem
β’ Familiarity with Docker/Kubernetes
β’ Exposure to recommendation systems or search ranking platforms
β’ Knowledge of monitoring/logging tools
β’ Agile/Scrum development experience