Senior Data Engineer
WalmartAbout the role
This notice is being provided as a result of the filing of an Application for Permanent Alien Labor Certification. Any person may provide documentary evidence bearing on the application to the Certifying Officer of the Department of Labor: U.S. Department of Labor, Employment and Training Administration, Office of Foreign Labor Certification, 200 Constitution Avenue, NW, Room N-5311, Washington, DC 20210
What you'll do...
Position: Senior Data Engineer
Job Location: 702 SW 8th Street, Bentonville, AR 72716
Duties: Identifies possible options to address the business problems within one's discipline through analytics, big data analytics, and automation. Supports the development of business cases and recommendations. Owns delivery of project activity and tasks assigned by others. Supports process updates and changes. Solves business issues. Supports the documentation of data governance processes. Supports the implementation of data governance practices. Understands, articulates, and applies principles of the defined strategy to routine business problems that involve a single function. Extracts data from identified databases. Creates data pipelines and transform data to a structure that is relevant to the problem by selecting appropriate techniques. Develops knowledge of current data science and analytics trends. Supports the understanding of the priority order of requirements and service level agreements. Helps identify the most suitable source for data that is fit for purpose. Performs initial data quality checks on extracted data. Analyzes complex data elements, systems, data flows, dependencies, and relationships to contribute to conceptual, physical, and logical data models. Develops the Logical Data Model and Physical Data Models including data warehouse and data mart designs. Defines relational tables, primary and foreign keys, and stored procedures to create a data model structure. Evaluates existing data models and physical databases for variances and discrepancies. Develops efficient data flows. Analyzes data-related system integration challenges and proposes appropriate solutions. Creates training documentation and trains end-users on data modeling. Oversees the tasks of less experienced programmers and stipulates system troubleshooting supports. Writes code to develop the required solution and application features by determining the appropriate programming language and leveraging business, technical, and data requirements. Creates test cases to review and validate the proposed solution design. Creates proofs of concept. Tests the code using the appropriate testing approach. Deploys software to production servers. Contributes code documentation, maintains playbooks, and provides timely progress updates. Demonstrates up-to-date expertise and applies this to the development, execution, and improvement of action plans by providing expert advice and guidance to others in the application of information and best practices; supporting and aligning efforts to meet customer and business needs; and building commitment for perspectives and rationales.
Minimum education and experience required: Master’s degree or the equivalent in Computer Science or related field plus 1 year of experience in software engineering or a related field; OR Bachelor’s degree or the equivalent in Computer Science or related field plus 3 years of experience in software engineering or a related field.
Skills Required: Must have experience with: Developing ETL data pipelines in HADOOP and Google Cloud Platform GCP and loading final data in HDFS Google Cloud Storage GCS and BIGQUERY; Importing data from Relational Database Management Systems (ORACLE, MySQL, DB2, TERADATA) Source to Target layer of HDFS and GCP Using SQOOP; Transforming raw data into clean data with MAPREDUCE, HIVE, SQL, and SPARK Framework; SPARK SCALA RDD's and Dataframes scripts for table creation, data loading, unit test cases, data transformation and achieve data quality; Developing User Defined Functions (UDF's) using JAVA, PYTHON Code to extend HIVE functionality and developing UNIX scripts for creating, dropping tables, and splitting file into Header, Detail and Trailer Files; Developing physical and logical data models using ERWIN Data Modeler; Optimizing Load Performance for ETL Jobs and performing various scenarios like file watcher and automated validations for data quality with SHELL Scripts with SQL and HIVE Queries; Scheduling Data Pipeline Workflows in production with AIRFLO
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s