Senior Hardware Reliability Engineer

CoreWeave
Summary
Join CoreWeave, a leading AI hyperscaler, as a highly skilled Engineer on our Hardware Provisioning team. You will play a vital role in designing, developing, and optimizing our server hardware infrastructure. Collaborate with cross-functional teams and vendors to deliver high-performance hardware solutions. This position requires proficiency in Ansible/Python, in-depth server hardware knowledge, and a strong passion for automation. CoreWeave offers a competitive salary ranging from $160,000 to $220,000 annually, along with comprehensive benefits including 100% employer-paid medical, dental, and vision insurance, life insurance, disability insurance, 401k matching, flexible PTO, and more. We operate as a hybrid workplace, offering flexibility between in-office and remote work, with onboarding in person within the first month for remote employees.
Requirements
- Proficiency in ansible/python and experience with programmatically interacting with server BMCs, using IPMI or Redfish
- In-depth knowledge of server hardware, components, and management technologies
- Proven ability to stay updated with the latest industry technologies and trends
- Previous experience collaborating with hardware vendors
- Strong passion for automation, with a commitment to automating processes comprehensively
- Excellent documentation skills and attention to detail
- Strong analytical and problem-solving abilities
Responsibilities
- Develop and maintain hardware/firmware management services
- Automate all aspects of the server hardware lifecycle
- Serve as the senior point of contact for hardware escalation and troubleshooting
- Collaborate with cross-functional teams to define hardware requirements, specifications, and system architecture
- Create and maintain accurate documentation of hardware designs, specifications, test procedures, and results
- Analyze and optimize the performance of hardware systems, identify bottlenecks, and propose improvements for enhanced efficiency
- Establish processes for internal hardware testing, deployment, and performance optimization
Benefits
- Medical, dental, and vision insurance - 100% paid for by CoreWeave
- Company-paid Life Insurance
- Voluntary supplemental life insurance
- Short and long-term disability insurance
- Flexible Spending Account
- Health Savings Account
- Tuition Reimbursement
- Mental Wellness Benefits through Spring Health
- Family-Forming support provided by Carrot
- Paid Parental Leave
- Flexible, full-service childcare support with Kinside
- 401(k) with a generous employer match
- Flexible PTO
- Catered lunch each day in our office and data center locations
- A casual work environment
- A work culture focused on innovative disruption
- Hybrid work environment with flexibility between in-office and remote work