Data Engineer GCP – Python / Spark – Climate Risk & Credit
Context: This project aims to implement a solution for ingesting, aggregating, and exposing data from multiple sources to generate the business reports necessary for credit risk monitoring. The solution is being developed on Google Cloud Platform (GCP) by an experienced team, on a high-stakes, high-profile project. Experience in the field of credit risk is preferred.
Main Responsibilities: The consultant will be involved in:- Designing and developing data solutions on Google Cloud Platform; Developing and integrating Python/Spark processes; Implementing and evolving data ingestion and processing pipelines; Ingesting, aggregating, and exposing data from multiple sources; Developing high-volume processing; Collecting, storing, and leveraging data streams; Monitoring and integrating data of various types; Ensuring data quality and security ; Maintaining and monitoring the operational status of developed applications; Maintaining associated data infrastructures ; Documenting and mapping different data sources; Contributing to the industrialization of processes and models; Proposing improvements to existing data infrastructures and solutions. The role of Data Engineer also includes structuring databases, managing metadata and repositories, industrializing machine learning models, and maintaining them in operational mode.
- Python: excellent command; Spark: excellent command; Development and integration of Data solutions on GCP.
- Big
- Dataproc; Cloud Composer; Cloud Functions; Cloud Run Jobs; Google Cloud Storage (GCS).
Dev
Ops – good proficiency expected:- Terraform; Jenkins; XL Deploy (XLD). These skills are explicitly presented as essential given the strategic nature and visibility of the project.
- Experience in the field of credit risk; Knowledge of banking environments; Experience with high-volume data projects; Understanding of data quality, security and governance issues.
We are looking for a senior GCP Data Engineer with 3 to 6 years of experience and proven expertise in Python/Spark development on Google Cloud Platform. The candidate must demonstrate significant operational experience with Big
Query, Dataproc, Composer, Cloud Functions, Cloud Run Jobs, and GCS. A strong understanding of the Dev
Ops pipeline, particularly Terraform, Jenkins, and XL Deploy, is also expected. Beyond technical skills, the consultant must be autonomous, possess excellent communication skills, and be a strong team player. Experience in credit risk management would be a significant advantage.
Application: CV + cover letter + copies of diplomas to be sent to
#J-18808-Ljbffr