Staff Reliability Engineer
$154,000–$154,000 year
On-siteSecaucus, New Jersey, United States
Job Summary
Design for Reliability principles to ensure cloud hardware meets specified use-conditions and stresses. Act as internal consultant on reliability matters, interfacing with program management, vendors, and design engineering to support software/script development needs. Create or revise reliability engineering guidelines to improve product field performance through design enhancements. Use performance evaluation and prediction principles to improve reliability and maintainability of Cloud Infrastructure servers. Identify, collect, analyze, and manage data to minimize failures and improve product performance. Develop scripts representing expected environment and operational conditions. Collaborate with cross-functional teams to apply Design for Reliability principles and ensure products meet customer expectations.
Required Qualifications
- Minimum B.S. in Electrical Engineering, Computer with Science/Engineering, or Software development
- 2+ years of relevant work experience
- Knowledge of computer systems/hardware structure, as well as switch/network interfaces
- Knowledge and/or experience with programming languages like Python or Unix (Bash and/or PowerShell)
- Knowledge of statistical & probability techniques and reliability modeling
- Ability to communicate, collaborate and lead cross-functionally to resolve issues, including those with customers
- Fundamental knowledge of Computer Architecture, Server architecture at the block level, and Hardware/Firmware/OS interactions
- Working knowledge of PCBA (printed circuit board assembly) design, fabrication, and validation testing
- Experience using tools such as ReliaSoft & JMP statistical software packages
- Working knowledge of electronic components/devices and their failure modes & failure mechanism
- Knowledge of industry standards, IPC, JEDEC, Telcordia, and MIL-STD
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.