Robotics System Reliability Specialist
Sonwil · West Seneca, NY, US · United States · On-site
Pay: USD 55,203 – 69,992 a year
Posted Sep 4, 2026
Sign up free: we match you to jobs like this, tailor your application and fill the form. 2 free applications every day.
*** This position REQUIRES work at heights of up to 60+ feet, please read the full job description before applying***
Job purpose
The Robotic System Reliability Specialist is a member of the FourSquare Solutions development team within Sonwil Technology Services Group, focused on improving the reliability and performance of robotic systems used to move pallets throughout the solution. This is a hands-on, field-based technical role working directly with live robotic systems and devices, often at significant heights. The Specialist observes systems in operation, investigates and reproduces failures, performs root cause analysis, and works alongside the development team to develop, implement, and validate improvements across the controls, electrical, mechanical, and operational aspects of the solution. The position plays a critical role in systematically eliminating failure modes and reducing the need for operational intervention, advancing the solution toward continuous, highly autonomous operation.
Duties and responsibilities
Work hands-on with live robotic systems and devices, including equipment operating at significant heights, to observe system behavior and understand performance under actual operating conditions.
Investigate robotic system failures and abnormal behaviors by observing, reproducing, and troubleshooting issues to identify root causes across controls, electrical, mechanical, and operational components of the solution.
Collect, document, and analyze failure information, including system behavior, operating conditions, fault data, and other relevant observations necessary to support effective root cause analysis.
Analyze failures collectively to identify common characteristics, recurring failure modes, systemic issues, and emerging trends rather than treating failures solely as individual events.
Review system reliability and failure KPI data to identify trends, quantify the operational impact of failure modes, and help prioritize improvement efforts that provide the greatest benefit to overall system stability.
Clearly communicate root cause findings, failure trends, and technical observations to the development team, translating field observations and system data into actionable development priorities.
Participate in the continuous integration and continuous operation (CICO) development process by providing feedback from live system operation, collaborating on corrective actions and product improvements, and helping move changes through implementation and validation.
Participate in the implementation of system improvements, including PLC controls and configuration changes as knowledge, training, and system access permit.
Test and validate corrective actions and system changes on live equipment to confirm intended behavior, identify unintended effects, and determine whether the targeted failure modes have been effectively addressed.
Evaluate system performance following implemented changes to determine their effect…