Position: Python Data Engineer
Location: Indore / Pune
Experience: 6–10 Years
Work Mode: Onsite during training; Hybrid thereafter
About the Role
We are seeking an experienced Python Data Engineer to design and build scalable, efficient, and reliable data pipelines and applications. The ideal candidate should have strong proficiency in Python, PySpark, SQL, and modern cloud data technologies (Databricks, ADF) with a focus on performance, scalability, and clean code practices.
Key Responsibilities
- Design, develop, and maintain data pipelines and applications using Python and PySpark.
- Build end-to-end data solutions ensuring performance, modularity, and reusability.
- Write and optimize complex SQL queries for large-scale data processing and validation.
- Implement data workflow orchestration using Databricks and Azure Data Factory (ADF).
- Refactor and modernize legacy codebases to improve reliability and maintainability.
- Collaborate with data engineers, analysts, and business teams to translate requirements into technical solutions.
- Drive best practices in code review, performance tuning, and deployment automation.
- Stay up to date with emerging technologies in data engineering and contribute to process improvements.
Required Skills
- 8+ years of experience in data or software engineering.
- Strong proficiency in Python, PySpark, and SQL.
- Expertise in Databricks and Azure Data Factory (ADF).
- In-depth understanding of data structures, design patterns, and scalable architectures.
- Experience with cloud platforms (Azure, AWS, or GCP).
- Strong debugging, performance optimization, and automation skills.
- Excellent problem-solving and analytical thinking abilities.
Stay updated with our latest job opportunities and company news by following us on LinkedIn: :https://www.linkedin.com/company/sourcebae
Skills Required
Pyspark, Sql, Databricks, Python