Logo for STN Incorporated

Hardware Engineer

Role overview

Qualifications

  • 5+ years in hardware engineering, systems engineering, or data center engineering
  • Deep knowledge of x86 server architecture, GPU systems, and modern storage
  • Hands-on experience with NVIDIA HGX, DGX, or hyperscale-class systems
  • Strong Linux fundamentals and scripting skills (Python, Bash)

Responsibilities

  • Monitor GPU and server health including thermal, error rates, and component failures
  • Drive the RMA process with vendors (NVIDIA, Supermicro, HPE, and others) end-to-end
  • Manage firmware, BIOS, and BMC upgrade campaigns across the fleet
  • Develop hardware burn-in and acceptance test procedures, including NCCL and stress tests

Key facts

  • Remote from: California (USA)
  • Full time
  • Senior (5-10 years)
  • Hardware Engineer
  • English

Hard skills

Other skills

  • Problem Solving
  • Collaboration

About the company

STN Incorporated logo

STN Incorporated

IT Services & IT Consulting

At STN, we don’t just deliver technology, we build the foundation that modern organizations run on. From enterprise IT to AI infrastructure, we design, operate, and support systems that are reliable, secure, and built for performance. But what sets us apart isn’t just our stack it’s how we show up. We don’t believe in one-size-fits-all. We don’t drop in hardware and disappear. We work side by side with our customers to understand what they actually need and build solutions that fit, flex, and scale as they grow.Whether you're running business-critical systems, deploying AI models, or training large-scale workloads on NVIDIA GPUs, we’re here to make sure your infrastructure isn’t just running, it’s working for you. Our team brings deep technical expertise, hands-on support, and a people-first mindset to everything we do. Because we believe technology should unlock potential, not create more problems. STN exists to help teams thrive in complex environments with custom engineering, real partnership, and a clear plan forward.

Company details

IndustryIT Services & IT Consulting
Company size11 - 50

Your match analysis

See how your profile stacks up against this role.

We compared the job requirements to your profile to show where you're strong and where you fall short.

Job description

Hardware Engineer

Infrastructure operations · shared across sites

Reports to: Director, Hardware Engineering

Location: Pleasanton, CA (hybrid) or assigned site; travel up to 25%

Department: Infrastructure & DC Operations / Systems Engineering

Position summary

The Hardware Engineer owns hardware lifecycle for GPU and supporting infrastructure assets, including fleet health monitoring, RMA workflows, firmware management, and long-range capacity planning. The role is the technical owner of the physical compute platform.

Key responsibilities

  • Monitor GPU and server health including thermal, error rates, and component failures

  • Drive the RMA process with vendors (NVIDIA, Supermicro, HPE, and others) end-to-end

  • Manage firmware, BIOS, and BMC upgrade campaigns across the fleet

  • Develop hardware burn-in and acceptance test procedures, including NCCL and stress tests

  • Investigate hardware failures and produce vendor-grade root cause analyses

  • Maintain hardware inventory, asset records, and CMDB accuracy

  • Drive capacity planning across compute, storage, and networking

  • Coordinate with Procurement on spare parts strategy and stocking levels

  • Author hardware engineering runbooks and operational procedures

  • Support new platform bring-up, qualification, and reference architecture validation

Required qualifications

  • 5+ years in hardware engineering, systems engineering, or data center engineering

  • Deep knowledge of x86 server architecture, GPU systems, and modern storage

  • Hands-on experience with NVIDIA HGX, DGX, or hyperscale-class systems

  • Strong Linux fundamentals and scripting skills (Python, Bash)

  • Bachelor's degree in computer science, electrical engineering, or related field

Preferred qualifications

  • Experience with NVIDIA Mission Control, Base Command Manager, or Bright Cluster Manager

  • Familiarity with IPMI, Redfish, and vendor management interfaces

  • Knowledge of liquid cooling and high-density power architectures

  • Experience operating fleets of 1,000+ GPUs

Apply once. Then go straight to the hiring manager.

After you apply, unlock the direct contact details of the people who actually make the call. A quick follow-up makes you 5x more likely to land an interview.

MR

Marcus Rivera

Chief Revenue Officer

m.rivera@company.com
linkedin.com/in/marcusrivera
Unlocked after you apply
·

Hardware Engineer Related jobs

Other jobs at STN Incorporated

Premium

Reach out to the hiring manager directly.

Gain access to the contact details of the hiring managers who actually decide, and reach out to network with them directly. That, plus more when you upgrade:

  • Full match report with fit score and gaps
  • Career diagnostics on how recruiters read you
  • Curated company matches and warm intros
  • 48h early access to new roles

Cancel anytime.