Data & AI Engineer - Sr Data Engineer
CapgeminiAbout the role
Description
About the job you’re considering
Name of the position: Sr .Data Engineer(Python & Spark)
Reports to: Team Lead/Delivery Manager
Department/Project: Engineering
As a Senior Data Engineer, you will build distributed data processing solution and highly loaded database solutions for various businesses cases including reporting, product analytics, marketing optimization and financial reporting. Contribute as part of self-organized team of experienced data engineers working in a challenging, innovative environment for our client, creating the foundation for decision-making at a company dealing with billions of events per day.
Investigate, create, and implement the solutions for existing technical challenges. Provide guidance, instruction, direction, leadership to a development team with the purpose of achieving project goals.
Your role
- Obtains tasks from the project lead or Team Lead (TL), prepares functional and design specifications, approves them with all stakeholders.
- Ensures that assigned area/areas are delivered within set deadlines and required quality objectives.
- Provides estimations, agrees task duration with the manager and contributes to project plan of assigned area.
- Analyzes scope of alternative solutions and makes decision about area implementation based on his/her experience and technical expertise.
- Leads functional and architectural design of assigned areas. Makes sure design decisions on the project meet architectural and design requirements.
- Addresses area-level risks, provides and implements mitigation plan.
- Reports about area readiness/quality, and raises red flags in crisis situations which are beyond his/her AOR.
- Responsible for resolving crisis situations within his/her AOR.
- Design, develop and implement large scale, high volume, high performance data models and pipelines for Data Lake and Data Warehouse.
- Develop and implement data quality checks, conduct QA and implement monitoring routines Improve the reliability and scalability of our ETL processes
- Initiates and conducts code reviews, creates code standards, conventions and guidelines.
- Suggests technical and functional improvements to add value to the product;
- Constantly improves his/her professional level.
- Collaborates with other teams.
Your skills and experience
- University degree in Computer Related Sciences or similar.
- 5+ years experience working in data engineering, business intelligence, or a similar role Proficiency in programming languages Python.
- Expert in Database fundamentals, Advanced proficiency in Complex SQL and distributed computing
- 5+ years of experience in ETL orchestration and workflow management tools like Airflow, Flink, Oozie and Azkaban using AWS/GCP
- Experience in Spark, Snowflake & Databricks.
- Must have ability to debug spark jobs.
- Experience working with Snowflake, Redshift, PostgreSQL and/or other DBMS platforms.
- Excellent communication skills and experience working with technical and non-technical teams.
- Able to clear hacker rank code test.
- Experience in AWS (EC2/S2/IAM).
- Experience working with technical and non-technical teams Knowledge of reporting tools such as Tableau, Superset and Looker.
- Strong python coder with expert data migration experience .
Life at Capgem
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s