Principal Systems Software Engineer, LPU Job at NVIDIA Gruppe, Santa Clara, CA

R01lRkN1Q3hDeTFqWGlnT0pBcldWWXF5Wmc9PQ==
  • NVIDIA Gruppe
  • Santa Clara, CA

Job Description

We are now looking for a Principal Software Engineer for LPX System Software! NVIDIA’s LPX System Software team builds the foundational software that turns a novel deterministic compute architecture into a platform that compiler teams and data center operators can rely on. We shift complexity out of silicon and into software: the hardware abstraction layers, core system libraries, drivers, and runtime components that workloads enter the platform through. We build this stack in Rust. For system software living at the boundary between hardware and everything above it, we treat memory safety, explicit ownership, and long-lived API stability as the baseline rather than the goal — the foundation that lets us spend our judgment on the hard problems instead of on classes of bugs that should not exist.

What you’ll be doing:

  • Shape the architecture of the hardware abstraction layers and core system libraries, and own the API contracts for the components you lead.
  • Design and implement drivers, runtimes, and data movement and aggregation pipelines that execute workloads on novel silicon.
  • Build runtime interfaces for launching, monitoring, and managing workloads at production scale.
  • Drive triage of the most difficult sequencing, initialization, and cross‑component runtime failures, and produce root‑cause analyses that change how the system is built.
  • Lead new platform bring‑up and NPI for new boards and silicon, in tight partnership with hardware engineering, compiler teams, and data center operations.
  • Multiply the team — establish the agent‑assisted engineering practices, reusable abstractions, diagnostics, and documentation that let everyone move faster without destabilizing the platform.
  • Communicate architecture and design tradeoffs clearly, in writing and in diagrams, to audiences ranging from individual engineers to executive staff.

What we need to see:

  • MS in CS, CE, EE, or a related STEM field, or equivalent experience, and 12+ years building production system software.
  • Deep systems‑programming expertise, with Rust as your language of choice for low-level work. You have shipped production Rust at the hardware or kernel boundary — drivers, firmware, runtimes, or similar — and you can articulate from experience where Rust earns its keep in system software and where it costs you.
  • A track record of designing and evolving libraries and APIs meant to be supported for years, including ABI and compatibility discipline.
  • Fluency in large, multi‑repository codebases with layered dependencies.
  • Demonstrated leadership driving triage of difficult reliability issues to clear, written root‑cause analysis.
  • Low‑level platform experience: firmware and boot flows, RTOS, BMCs/MCUs, RISC‑V, or closely related system software.
  • Linux driver or kernel‑adjacent experience (for example, VFIO or similar subsystems).
  • Hardware bring‑up and system triage experience: fault analysis, diagnostics, and validation in lab environments.
  • An established habit: building with AI coding agents — not as a novelty, but as a way you already ship and raise leverage. You can speak to how you design work to be agent‑amenable and where you keep humans in the loop.

Ways to stand out from the crowd:

  • Experience having built Rust system software at the scale of a hyperscaler or a Rust‑native hardware company.
  • Distributed systems experience: gRPC and RPC frameworks, coordination and telemetry patterns, MPI. Inference systems and token serving experience (vLLM or similar serving and runtime stacks) a huge plus.
  • Experience shipping and supporting customer‑facing SDKs, including documentation and ABI compatibility practices.
  • Production readiness and delivery depth: CI/CD and release workflows, monitoring and alerting practices, Kubernetes, and data center operational workflows.

Compensation

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 272,000 USD - 431,250 USD. You will also be eligible for equity and benefits.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

#J-18808-Ljbffr

Job Tags

Shift work

Similar Jobs

University Of California Irvine

Regulatory Affairs Coordinator Job at University Of California Irvine

 ...University of California, Irvine is hiring a Regulatory Affairs Coordinator. The role manages regulatory compliance for clinical trials...  ....Ability to work across multiple research sites including UCI Medical Center and community clinics.Compensation ranges from $35.77... 

Wheelhouse Credit Union

VP/ Chief Lending Officer Job at Wheelhouse Credit Union

 ...Job Description Job Description Description: Join Our Leadership Team! Wheelhouse Credit Union is seeking a Chief Lending Officer (CLO). This role is a key member of the executive leadership team and is responsible for overseeing all lending operations, including... 

New Paradigm Staffing

Remote Medical Records Indexing Clerk Entry Level Job at New Paradigm Staffing

 ...healthcare staffing agency is looking for an Entry-Level Medical Indexing Clerk to assist in organizing patient information remotely. The role involves categorizing...  ...systems. This is a great opportunity to start a career in health information management.#J-18808-Ljbffr

CLHG-Minden LLC

Mental Health Technician Job at CLHG-Minden LLC

 ...Job Description Job Description Join Minden Medical Center as a PRN Mental Health Technician and immerse yourself in a dynamic and impactful work environment located in Minden, LA. This onsite role offers you the opportunity to engage directly with patients, contributing... 

HAN Staffing

AWS and Datadog engineer Job at HAN Staffing

Job ID: 20252770Reference Number: 23-01244Title: AWS and Datadog engineerLocation: Iselin, NJ, 08830Posted Date: 2023-08-02Company: HAN StaffingAWS and Datadog engineer