Jobs and Careers
LI

Software Reliability Engineer - Warehouse Management Systems

Lineage
United StatesRemotefull_timeVerifiedPosted 29 May 2026

About the role

The Software Reliability Engineer (SRE) will play a critical role in ensuring that our Warehouse Management Software (WMS) runs seamlessly across both automated and manual facilities. This role focuses on investigating, diagnosing, and resolving operational software issues that impact warehouse performance—freeing developers to focus on new features and ensuring WMS never disrupts day-to-day operations.
Please note: We are unable to sponsor work authorization now or in the future for this role. 

Roles and Responsibilities 

1. Operational Issue Investigation and Quick Resolution 

  • Monitor and respond to operational issues affecting WMS functions (e.g., receiving, shipping, inventory). 

  • Analyze system logs, error reports, and transaction flows to identify anomalies or failures. 

  • Work closely with Level 1 support and warehouse operation teams to understand incident symptoms and timelines. 

  • Execute quick resolutions by using extended user rights, database interventions, or WMS configuration changes. 

2. Code-Level Debugging 

  • Debug application code, workflows, customizations, and interfaces to identify bugs or performance bottlenecks. 

  • Collaborate with WMS QA team to reproduce issues in test environments and trace through application workflows to isolate root causes. 

  • Collaborate with Product/Development teams to propose, implement, and test code fixes. 

3. Real-Time System Monitoring 

  • Use tools like Datadog or internal diagnostics to monitor WMS behavior. 

  • Proactively set up or refine alerts for failure patterns (e.g., inventory mismatches, interface timeouts, RF disconnects). 

  • Improve observability by suggesting/implement better logging practices and metric coverage. 

4. Interface Troubleshooting 

  • Investigate communication failures between WMS and other Products (e.g., LinOS, Link, EDI, Easymetrics). 

  • Troubleshoot integration issues between the WMS and external systems (e.g., DevOps, DCOps). 

  • Provide software-side support during integration testing, mainly remote and on-site by occasion. 

5. Incident Management & Escalation 

  • Participate in on-call rotations or site support shifts for time-sensitive incidents. 

  • Coordinate with operations, IT, and engineering during critical events to ensure fast resolution. 

  • Document incidents thoroughly, including root causes, fixes, and follow-up actions. 

6. Post-Incident Review & Continuous Improvement 

  • Contribute to postmortem analysis for high-impact incidents. 

  • Recommend and implement configuration changes or process improvements to prevent repeated issues. 

  • Update or create playbooks and troubleshooting guides for known WMS issues. 

7. Internal Tooling and Automation 

  • Develop scripts o

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Lineage

View company profile →