$ cat jobs/staff-engineer-hardware-reliability-linkedin3-74030afa9388.json
Staff Engineer, Hardware Reliability
LinkedIn is the world's largest professional network, built to create economic opportunity for every member of the global workforce. Our products help people make powerful connections, discover exciting opportunities, build necessary skills, and gain valuable insights every day. We're also committed to providing transformational opportunities for our own employees by investing in their growth. We aspire to create a culture that's built on trust, care, inclusion, and fun – where everyone can succeed. Join us to transform the way the world works. At LinkedIn, our approach to flexible work is centered on trust and optimized for culture, connection, clarity, and the evolving needs of our business. The work location of this role is hybrid, meaning it will be performed both from home and from a LinkedIn office on select days, as determined by the business needs of the team. This role will be based in Sunnyvale, CA. We are looking for a highly skilled, self-motivated Staff Engineer to join our Hardware Capacity Engineering (HCE) team and help us scale and sustain the infrastructure that powers LinkedIn. HCE qualifies, integrates, and operates the full range of hardware in our on-prem data centers , such as general-purpose compute, GPU/accelerator, storage, and networking platforms, across a large-scale, multi-vendor, multi-generation fleet. This role spans both bringing new platforms into production and keeping our existing fleet healthy, performant, and reliable, backed by the software, firmware automation, and fleet-health systems the team builds. In this role, you will identify requirements and the best-suited hardware platform or solution, integrate that solution into our on-prem data center environment, and help operate and continuously improve the existing fleet at scale. You will build software and automation that make the fleet observable, performant, and reliable, and partner closely with SRE, software engineering, AI/ML, and hardware vendors. Responsibilities Col
Similar remote roles
DevOps Engineer - AI Model Evaluator
mercor · Worldwide · mid
Site Reliability Engineer (SRE)
GetEpic.com · Worldwide · mid
DualEntry | Full-time | REMOTE
DualEntry · Worldwide · mid
Engineering Manager - Site Reliability *EU/UK remote* (m/f/d)
Pliant · Worldwide · mid
Senior II Security Engineer - Platform
Preply · UK · senior
Senior Platform Software Engineer
Preply · UK · senior
Senior Site Reliability Engineer (SRE)
branch · US · senior
Engineering Manager (Site Reliability)
moniepoint · APAC · mid