GCP Data Engineer.
This mandate is run by SYMPHONI HR, a Mumbai-based executive search firm, est. 2003. Applications are email-verified and reach the search team running the role — confidentially, always.
About the organisation.
Our client is a leading global computer software company known for its innovative solutions and commitment to technological excellence. They operate at the forefront of data-driven innovation, serving a diverse range of industries. The company fosters a collaborative and dynamic work environment, encouraging continuous learning and professional growth.
About the role.
We are seeking an experienced GCP Data Engineer to join a dynamic team. This role involves designing, developing, and maintaining robust, scalable data pipelines on Google Cloud Platform. The ideal candidate will leverage strong expertise in Python and PySpark to ensure data quality, performance, and reliability across various data platforms.
What you will do.
- Design, build, and optimize scalable data pipelines and ETL processes using GCP services like BigQuery, Dataflow, Dataproc, and Cloud Storage
- Develop high-quality, efficient, and well-documented code in Python and PySpark for data processing, manipulation, and analytics
- Utilize PySpark for distributed data processing and analysis of large datasets to support analytical and business needs
- Orchestrate data workflows using tools like Cloud Composer (Apache Airflow) and automate processes using Python and shell scripting
- Implement data modeling, data warehousing concepts, and data governance practices to ensure data integrity, security, and quality
- Monitor, troubleshoot, and optimize data processing jobs and queries for performance, reliability, and cost efficiency
- Work closely with data scientists, analysts, and cross-functional teams to gather requirements and translate them into technical solutions
What you bring.
- Bachelor's or Master's degree in Computer Science, Engineering, or a related field
- 8–12 years of proven experience in a data engineering role with a strong focus on GCP
- 2+ years of hands-on experience with GCP data services including BigQuery, Cloud Storage, Dataflow, Dataproc, and Pub/Sub
- Strong proficiency in Python and SQL is essential, with demonstrated expertise in PySpark for big data processing
- Solid understanding of data modeling, ETL/ELT processes, and data warehousing principles
- Excellent problem-solving, analytical, and communication skills
- Experience with version control systems like Git
Other open roles.
Databricks Data Engineer — bangalore · Hybrid
FRTB Model Validation Analyst — Mumbai · Hybrid
Full Stack Engineer - Java & Angular — Bengaluru · Hybrid
