Data Engineer Hadoop / PySpark - Datalake @ > 3 years XP @ Bank @ Paris - Freelance (M/F)
Daily Rate (TJM): 400-420 € excl. tax depending on profile
Location: IDF Desired Start Date: 28/09/2026 Estimated End Date: 26/02/2027 Seniority: > 3 years of experience Sector: Bank / Financial Services
CONTEXT & CHALLENGES
As part of a mission within the IT department of a major player in the banking and financial services sector, we are looking for a Data Engineer Hadoop / PySpark to work on the CIO Office Datalake.
The IT department concerned provides, in particular, the applications necessary for Technology & Operations activities as well as Retail Banking and Insurance activities.
The Data Engineer will join a Data environment requiring close collaboration with application production and DevOps teams, notably to implement technical evolutions of the platform.
The main challenges involve the integration of new data, its preparation and transformation, as well as the development and evolution of processing pipelines within a Hadoop environment.
DETAILED MISSIONS
The consultant will notably be responsible for:
- Implementing new data ingestions within the Datalake.
- Performing Data Preparation and data transformation processing.
- Developing new data pipelines in PySpark.
- Maintaining and evolving existing Data pipelines and processing.
- Performing Shell Scripting under Unix/Linux.
- Working in collaboration with the DevOps team on platform evolution requests.
- Participating in maintaining the operational conditions of Data processing.
- Contributing to the technical evolutions of the platform in connection with the application production teams.
- Analyzing and resolving technical issues encountered in data processing and flows.
- Writing technical specifications.
- Documenting the developments and technical solutions implemented.
- Participating in the team's work within an Agile / Scrum environment.
- Collaborating regularly with a significant part of the IT team based in Porto, requiring professional use of English both written and spoken.
TECHNICAL STACK
Indispensable skills:
- Hadoop
- Python
- PySpark
- Hive
- SQL
- Shell Scripting Unix/Linux
Environment & tools:
- Git
- Jenkins
- Jira
- Agile / Scrum
Appreciated / differentiating skills:
- Indexima
- Alteryx
- Altair
- GCP
- BigQuery
- Python libraries oriented towards APIs
SEARCHED PROFILE
We are looking for a Data Engineer ideally having 3 to 6 years of experience, with a very good mastery of the Hadoop / PySpark ecosystem.
The profile must be particularly operational in the development and processing of large volumes of data, with a genuine appetite for Data issues and development.
A very good mastery of Hadoop, Python/PySpark, Hive, SQL and Shell Scripting Unix/Linux is indispensable.
The consultant must also be able to intervene both in the creation of new processing and in the maintenance and evolution of existing pipelines, in an environment involving Data, DevOps and application production teams.
The following qualities are particularly expected:
- Autonomy
- Dynamism
- Proactivity
- Technical curiosity
- Strong analytical capacity
- Ability to solve technical problems
- Excellent interpersonal skills
- Ability to work effectively in a team
- Ease in an international environment
Professional English, both written and spoken, is indispensable, given the regular exchanges with the IT teams located in Porto.