Senior Back-end Engineer - Data/ML
YahooAbout the role
A Lot About You:
The Senior Software Dev Engineer will analyze, program, debug and modify software enhancements and/or new products. You will have the opportunity to participate in the development of data platform designs with our team of Big Data engineers, and work in an agile scrum driven environment to deliver new and innovative products. You will design applications, write code, develop and test and debug, and document work and results.
Responsibilities:
Perform all phases of software engineering including requirements analysis, application design, and code development & testing
Design and implement reusable frameworks, libraries and components, product features in collaboration with business and IT stakeholders (25%)
Develop and implement anomaly detection models using machine learning and statistical techniques (10%)
Collaborate with cross-functional data engineers to improve data quality, ETL pipelines, and logging systems (10%)
Define alerting mechanisms for automatic detection and response to anomalies. Conduct root cause analysis on detected anomalies to prevent future occurrences (10%)
Interpret and visualize anomalies using dashboards (e.g., Looker Core, Looker Studio) (10%)
Ingest data from various structured and unstructured data sources into Hadoop and other distributed Big Data systems (5%)
Support the sustainment and delivery of an automated ETL pipeline and Validate data that is extracted from sources like HDFS, databases, and other repositories using scripts and other automated capabilities, logs, and queries (5%)
Enrich and transform extracted data, as required. Monitor and report the data flow through the ETL process (5%)
Perform data extractions, data purges, or data fixes in accordance with current internal procedures and policies (5%)
Track development and operational support via user stories and decomposed technical tasks in a provided issue tracking software, including GIT, Maven, and JIRA (5%)
Troubleshooting production support issues post-deployment and come up with solutions as required (5%)
Mentor junior engineers within the team for development and delivery (5%)
Qualifications:
Bachelor's degree or Master's degree in computer science or related discipline; or, equivalent experience
Five years of related industry experience
Experience in back-end programming, like Java, Python, and OOAD and ETL Tools
Experience in Data analysis, data investigation and implementing data quality
Strong Knowledge of GCP especially logging, alerting and metrics.
Experience of working with large scale databases like Bigquery
Knowledge and experience of Unix (Linux) Platforms and Shell Scripting
Experience in writing Pig Latin scripts, MapReduce jobs, HiveQL etc.
Experience with Machine Learning frameworks
Strong knowledge of anomaly detection algorithms
Strong knowledge of visualization tools. (Looker Core, Looker Studio)
Good knowledge of database structures, theories, principles, and practices
Familiarity with data loading tools like Flume, Sqoop
Knowledge of workflow/schedulers like Oozie, Airflow
Analytical and problem solving skills, applied to Big Data domain
Proven understanding with Hadoop, HBase, Hive, Pig
Writing high-performance, reliable and maintainable code
Expertise in version control tools like GIT
Good aptitude in multi-threading and concurrency concepts
Effective analytical, troubleshooting and problem-solving skills
Strong customer focus, ownership, urgency and drive
#LI-AC1
The material job duties and responsibilities of this role include those listed a
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s