Symbotic · 채용 중 118건
Staff Senior Reliability Engineer
Staff Senior Reliability Engineer
데브옵스 엔지니어정규직시니어 · 8년 이상
Symbotic은 복잡한 운영 환경의 신뢰성 엔지니어링을 총괄할 전문가를 찾습니다. RCA(근본 원인 분석) 및 ITIL 문제 관리 분야에서 5년 이상의 경력을 포함해 총 8년 이상의 실무 경험이 필수입니다. 대규모 분산 시스템의 장애를 분석하고 고객 및 경영진과 소통하며 기술적 개선을 주도합니다. 로봇 공학이나 자동화 환경 경험자를 우대합니다.
USA Remote & Travel
Time type: full time Posted on: July 22, 2026 Job requisition id: R7618
You’ll be the Senior Reliability Engineer who owns complex production investigations end to end — from root cause, through customer communication, through verified corrective action. You’ll be the face of our RCA process to customers and senior leadership, the analyst who turns incident trends into strategic improvements, and a mentor who raises the quality bar for the whole team.
Lead high-impact RCA investigations for complex production incidents spanning software, infrastructure, industrial controls, and production SOP execution
Chair structured, blameless RCA reviews aligned with ITIL Problem Management — fact-based analysis, clear ownership, timely resolution
Serve as the customer-facing technical lead for RCA discussions, updates, and formal deliverables within defined SLA timelines
Analyze logs, telemetry, and incident trends across many issues to find recurring failure patterns — then influence teams across the organization to eliminate them
Present findings, risks, and recommendations to senior internal leadership and customer stakeholders, backed by data you own end to end
Drive Continuous Service Improvement initiatives that measurably reduce repeat incidents and investigation toil through automation, tooling, and reporting
Mentor teammates and elevate RCA quality standards as a senior individual contributor
Minimum 8 years supporting complex, business-critical production environments, with a career centered on reliability
Minimum 5 years leading technical RCA, post-incident reviews, or ITIL-aligned Problem Management across software, infrastructure, systems, or industrial technology domains
Strong hands-on troubleshooting and data analysis across large-scale distributed systems, on-prem infrastructure, custom software, logs, telemetry, and incident datasets
A track record of regular, ongoing customer interaction — you can tell us who you worked with, at what level, and how often — and of earning trust with executives, technical and non-technical alike
Proven ability to run multiple high-priority investigations in parallel while influencing cross-functional teams, without direct authority, to close actions on time
Bachelor’s degree in a technical field, or equivalent practical experience
Hands-on experience with warehouse automation, robotics, industrial controls, or large-scale physical production environments (strongly preferred)
Kubernetes, VMware, Linux, Grafana, Prometheus, Zabbix, GitLab, and AI-assisted troubleshooting and analysis
Power BI, Tableau, reporting automation, or dashboard development for RCA analysis and leadership visibility
Depth in ITSM/ITIL practice — Incident Management, Problem Management, RCA, and Continuous Service Improvement
#LI-PA2
#LI-Remote
The base range for this position in the posted location is $120,000.00 - $165,000.00 however, base pay offered may vary depending on job-related knowledge, skills, and experience. The compensation package includes medical, dental, vision, disability, 401K, PTO and/or other benefits.
Corporate Functions & IT | USA Wilmington, MA - HQ
Corporate Functions & IT | USA Remote & Travel
Corporate Functions & IT | USA Wilmington, MA - HQ
Corporate Functions & IT | USA Wilmington, MA - HQ