Lead Data Engineer, GCP Scala.
This mandate is run by SYMPHONI HR, a Mumbai-based executive search firm, est. 2003. Applications are email-verified and reach the search team running the role — confidentially, always.
About the organisation.
Our client is a prominent computer software company known for its innovative solutions and commitment to technological advancement. They operate on a global scale, delivering cutting-edge products and services to a diverse client base. The company fosters a culture of continuous learning and collaboration, encouraging employees to push the boundaries of what's possible in the tech industry.
About the role.
This role is for a Lead Data Engineer with a strong background in Google Cloud Platform (GCP) and Scala. You will be instrumental in architecting and implementing robust data solutions, focusing on end-to-end data pipeline development. The position requires deep expertise in cloud data technologies and the ability to drive complex data initiatives within a dynamic environment.
What you will do.
- Architect and implement solutions on Google Cloud Platform using various GCP components
- Create end-to-end data pipelines using Apache Beam, Google Dataflow, or Apache Spark
- Identify downstream implications of data loads and migrations, including data quality and regulatory aspects
- Implement data pipelines to automate ingestion, transformation, and augmentation of data sources
- Provide best practices for data pipeline operations and maintenance
- Work in a rapidly changing business environment to enable simplified user access to massive data
- Build scalable data solutions to support evolving business needs
- Perform advanced SQL writing and data mining for complex datasets
What you bring.
- 12–15 years of experience in IT or professional services, with a focus on IT delivery or large-scale IT analytics projects
- Expertise in Google Cloud Platform (GCP) data engineering, including BigQuery, Cloud Composer/Python, Cloud Functions, Dataproc+PySpark, Dataflow+Pub/Sub, and Scala
- Strong experience with Apache Beam, Google Dataflow, or Apache Spark for data pipeline creation
- Expert knowledge in SQL development and advanced SQL writing
- Proficiency in building data integration and preparation tools using cloud technologies (e.g., Snaplogic, Google Dataflow, Cloud Dataprep, Python)
- Experience programming in Python, Java, or similar languages
- Expertise in at least two of the following: Relational Databases, Analytical Databases, NoSQL databases
- Certified in Google Professional Data Engineer/Solution Architect is a significant advantage
Other open roles.
Data Engineer, Looker & Big Data — Bangalore · Work From Office
AWS Data Engineer — Hyderabad · Hybrid
Lead Research Analyst — Hyderabad · Hybrid
