Staff, Data Engineer
WalmartAbout the role
Position Summary...
What you'll do...
What you'll do...
Come join our International Data Lake organization @Walmart where you’ll have the opportunity to manage International Supply Chain data enablement for International Markets at Walmart scale and make sure the consistent data is being enabled foe various data use-cases.
About Team:
The international lake organization is tasked with building and maintaining the data lake(s) for all of the international markets. This will enable both our in-market and home office associates to make more informed decisions driven by accurate, up to date and easily accessible data.
What you'll do:
- Design, Build and Enhance high performance reusable framework for data pipeline
- Develop and deploy cutting edge solutions at scale, impacting millions of customers worldwide to drive value from data at Walmart Scale
- Support Intl Supply chain Data applications ensuring Data Quality and Governance standards.
- Interact with Walmart engineering teams across geographies to leverage expertise and contribute to the technical community.
- Ensure data ingested and processed is accurate and of high quality by implementing data quality checks, data validation, and data cleaning processes
- Analyze, translate business requirements into strategies, initiatives, and projects and aligns them to business strategy and objectives, and drives the execution of deliverables.
- Defines and identifies the most suitable sources for required data that is fit for purpose, referring to external sources as required.
- Builds the infrastructure required for optimal transformation and integration from a wide variety of data sources using appropriate data integration technologies.
- Uses modern tools, techniques, and architectures to automate the most-common, repeatable and tedious data preparation and integration tasks partially or completely.
- Provides guidance to business stakeholders, peers, and junior associates on the implementation of data governance practices.
- Collaborates with team members in solving business problems effectively.
What you'll bring:
- Bachelor's/master’s degree in computer science or a related field
- With 8+ years' experience in development of big data technologies/data pipelines
- Proficiency in managing and manipulating huge datasets in the order of terabytes (TB) is essential.
- Expertise in big data technologies like Hadoop, Apache Spark, Apache Hive, or similar frameworks on the cloud (GCP preferred, AWS, Azure etc.) to build batch data pipelines with strong focus on optimization, SLA adherence and fault tolerance.
- Proven track record coding with programming language like, Scala and Python.
- Expertise in building idempotent workflows using orchestrators like Airflow.
- Expertise in writing SQL to analyze, optimize, profile data preferably in Big Query or SPARK SQL
- Hands on working experience in any messaging platform like Kafka is preferred.
- Strong data modeling skills are necessary for designing a schema that can accommodate the evolution of data sources and facilitate seamless data joins across various datasets.
- Ability to work directly with stakeholders to understand data requirements and translate that to pipeline development / data solution work.
- Strong analytical and problem-solving skills are c
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s