Job Description
-Must Have-Databricks-Advanced PySpark, Spark SQL, Delta Lake, Unity Catalog, Databricks Workflows/Jobs, performance tuning
-Must Have-Azure Data Factory-Pipeline development, orchestration, parameterisation, monitoring, error handling, incremental loading.
-Must Have-SQL-Strong analytical SQL, data modelling, window functions, CTEs, performance optimisation
-Must Have-Python-Production-quality Python, type hints, modular code, reusable libraries, testing
-Must Have-Software Engineering Practices-Git, pull requests, code reviews, structured development practices
-Must Have-CI/CD-Azure DevOps pipelines, deployment automation, environment promotion
-Must Have-Data Pipeline Testing-Unit testing, integration testing, data quality validation
-Must Have-Documentation & Maintainability-Well-structured code, clear naming conventions, reusable components, comprehensive documentation
-Strongly Preferred-GitHub Copilot / AI-Assisted Development-Practical experience using AI coding assistants to accelerate delivery while maintaining quality.
-Strongly Preferred-AI-Maintainable Coding Practices-Writing code that can be easily understood and extended by AI agents and future developers.
-Strongly Preferred-Databricks AI Capabilities-Awareness of AI Functions, Genie, AI/BI Dashboards, Mosaic AI or similar capabilities.
-Strongly Preferred-DataOps / Infrastructure-as-Code-Terraform, Bicep, Databricks Asset Bundles (DABs), deployment as code.
-Strongly Preferred-Prompt Engineering for Development-Ability to specify requirements effectively for AI-assisted coding workflows.
-Preferred-Docker-Containerising Python applications and services
-Preferred-Kubernetes Literacy-Basic understanding of Kubernetes concepts and deployment patterns
-Preferred-DevOps Practices-Ability to deploy and support own solutions with minimal platform team involvement
-Preferred-Power BI-Azure-connected reporting, gateway configuration, scheduled refresh
-Nice to Have-Basic Geoscience Domain Knowledge Drillhole, geochemistry, geoscience data structures and workflows.
-Nice to Have-Workflow Orchestration-Argo Workflows orchestration.
-Nice to Have-Data Governance & Metadata-Data cataloguing, lineage and metadata management
Required Technical/ Functional Competencies
Domain/ Industry Knowledge:
Requirement Gathering and Analysis:
Product/ Technology Knowledge:
Architecture tools and frameworks:
Architecture concepts and principles:
Analytics Solution Design:
Tools & Platform Knowledge:
Required Behavioral Competencies
Accountability:
Collaboration:
Agility:
Customer Focus:
Communication:
Drives Results:
Resolves Conflict:
Certifications
Mandatory At YASH, you are empowered to create a career that will take you to where you want to go while working in an inclusive team environment. We leverage career-oriented skilling models and optimize our collective intelligence aided with technology for continuous learning, unlearning, and relearning at a rapid pace and scale. Our Hyperlearning workplace is grounded upon four principlesModule Lead - Azure Job • Pune, MH, IN