Mandatory Experience:
• Developer should have Strong background in Software development with experience in ingest, transform and store data from large datasets using Pyspark in Azure Databricks with strong knowledge on distributed computing concepts
• Must have 2 years’ experience in designing and developing ETL Pipelines in Pyspark in Azure Databricks with strong python scripting exposure like list comprehensions, Dictionary variables etc.
• Must have 2 year’s of exposure and good proficiency in data warehousing concepts.
• Experience in Ingestion and ETL release pipelines(Databricks) creation and deployments
• Knowledge of Azure cloud computing platform with at least 2 years’ experience in Azure Synapse, ADLS and HDI.
• At least a year of experience with good proficiency in Delta table and delta file operations like merge, Insert override, Partition overrides etc.
Benefits:
• Exposure to new processes and technologies.
• Competitive salary at par with the best in the industry.
• Flexible and employee friendly environment