Jobs and Careers
BL

Data Engineer

Blockchain.com
FranceRemotefull_timeVerifiedPosted 30 Oct 2023

About the role

Blockchain.com is the world's leading software platform for digital assets. Offering the largest production blockchain platform in the world, we share the passion to code, create, and ultimately build an open, accessible and fair financial future, one piece of software. 

​We are looking for a talented Data Engineer to join our Data Science team and work from our office in Paris. The group is part of a larger DS team, informing all product decisions and creating models and infrastructure to improve efficiency, growth, and security. To do this, we use data from various sources and of varying quality. Our automated ETL processes serve both the broader company (in the form of clean, simplified tables of aggregated statistics and dashboards) and the Data Science team itself (cleaning and processing data for analysis and modeling purposes, ensuring reproducibility). 

We are looking for someone with experience in designing, building, and maintaining a scalable and robust Data Infra that makes data easily accessible to the Data Science team and the broader audience via different tools. As a data engineer, you will be involved in all aspects of the data infrastructure, from understanding current bottlenecks and requirements to ensuring the quality and availability of data. You will collaborate closely with data scientists, platform, and front-end engineers, defining requirements and designing new data processes for both streaming and batch processing of data, as well as maintaining and improving existing ones. We are looking for someone passionate about high-quality data who understands their impact in solving real-life problems. Being proactive in identifying issues, digging deep into their source, and developing solutions, are at the heart of this role.  

JUNIOR: 

What you will do

  • Maintain and evolve the current data lake infrastructure and look to evolve it for new requirements
  • Maintain and extend our core data infrastructure and existing data pipelines and ETLs 
  • Provide best practices and frameworks for data testing and validation and ensure reliability and accuracy of data
  • Design, develop and implement data visualization and analytics tools and data products.

What you will need

  • Bachelor’s degree in Computer Science, Applied Mathematics, Engineering or any other technology-related field
  • Previous experience working in a data engineering project or role
  • Fluency in Python
  • Previous experience with ETL pipelines and data processing
  • Good knowledge of SQL and no-SQL databases
  • Good knowledge of coding principles, including Oriented Object Programming 
  • Experience with Git

Nice to have

  • Experience with Airflow, Google Composer or Kubernetes Engine
  • Experience working with Google Cloud Platform
  • Experiences with other programming languages, like Java, Kotlin or Scala
  • Experience with Spark or other Big Data frameworks
  • Experience with distributed and real-time technologies (Kafka, etc..)
  • 1-2 years commercial experience in a related role 

MID:

What you will do

  • Maintain and evolve the current data infrastructure and look to evolve it for new requirements
  • Maintain and extend our core data infrastructure and existing data pipelines and ETLs 
  • Provide best practices and frameworks for data testing and validation and ensure reliability and accuracy of data
  • Design, develop and implement data visualization and analytics tools and data products.

What you will need

  • Bachelor’s degree in Computer Science, Applied Mathematics, Engineering or any other technology-related field
  • Previous experience working in a data engineering role
  • Fluency in Python
  • Previous experience with ETL pipelines
  • Experience working with Google Cloud Platform
  • In-depth knowledge of SQL and no-SQL databases
  • In-depth knowledge of coding principles, including Oriented Object Programming 
  • Experience with Git

Nice to have

  • Experience with code optimisation, parallel processing
  • Experience with Airflow, Google Composer or Kubernetes Engine
  • Experiences with other programming languages, like Java, Kotlin or Scala
  • Experience with Spark or other Big Data frameworks
  • Experience with distributed and real-time technologies (Kafka, etc..)
  • 2-5 years commercial experience in a related role 

SENIOR:

What you will do

  • Maintain and evolve the current d

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Blockchain.com

View company profile →