Senior, Data Engineer
WalmartAbout the role
Position Summary...
What you'll do...
What you'll do...
Come join our International Data Lake organization @Walmart where you'll have the opportunity to manage International Supply Chain data enablement for International Markets at Walmart scale and make sure the consistent data is being enabled for various data use-cases.
About Team:
The international data organization is tasked with building and maintaining the data lake(s) for all the international markets. This will enable both our in-market and home office associates to make more informed decisions driven by accurate, up to date and easily accessible data.
What you'll do:
- Design, Build and Enhance high performance reusable framework for data pipeline
- Develop and deploy cutting edge solutions at scale, impacting millions of customers worldwide to drive value from data at Walmart Scale
- Support Intl Supply Chain Data applications ensuring Data Quality and Governance standards.
- Interact with Walmart engineering teams across geographies to leverage expertise and contribute to the technical community.
- Ensure data ingested and processed is accurate and of high quality by implementing data quality checks, data validation, and data cleaning processes
- Analyze,translatebusiness requirements into strategies, initiatives, and projects and aligns them to business strategy and objectives, and drives the execution of deliverables.
- Defines and identifies the most suitable sources for required data that is fit for purpose, referring to external sources as required.
- Builds the infrastructure required for optimal transformation and integration from a wide variety of data sources using appropriate data integration technologies.
- Uses modern tools, techniques, and architectures to automate the most-common, repeatable, and tedious data preparation and integration tasks partially or completely.
- Collaborates with team members in solving business problems effectively.
What you'll bring:
- Bachelors/masters degree in computer science or a related field
- With5+ yearsexperience in development of big data technologies/data pipelines
- Proficiency in managing and manipulating huge datasets in the order of terabytes (TB) is essential.
- Proficiency in big data technologies like Hadoop, Apache Spark, Apache Hive, or similar frameworks on the cloud (GCP preferred, Azure etc.) to build batch data pipelines with strong focus on optimization, SLA adherence and fault tolerance.
- Proven track record codingwith programming language like, Scala and Python.
- Proficiency in building idempotent workflows using orchestrators like Airflow.
- Proficiency in writing SQL to analyze, optimize, profile data preferably in Big Query or SPARK SQL
- Hands on working experience in any messaging platform like Kafka is preferred.
- Strong data modeling skills are necessary for designing a schema that can accommodate the evolution of data sources and facilitate seamless data joins across various datasets.
- Ability to work directly with stakeholders to understand data requirements and translate that to pipeline development / data solution work.
- Analytical and problem-solving skills are crucial for identifying and resolving issues that may arise during the data integration and schema evolution process.
- Ability to move at a r
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s