Back to job search
NVIDIA logo

Senior Security Engineer, Infrastructure Security Engineering - DGX Cloud

  • Canada
  • Remote
  • Posted Sep 5, 2026
  • 1 position

$170,000–$275,000 / year

Opens an external site

Sign in to save this job
Employment type
Full-time
Experience level
Lead · 10+ years
Minimum education
Bachelor’s degree
Posting language
English
Working hours
40 hours per week

Job summary

You will architect and build foundational security primitives and automated guardrails to protect large-scale AI infrastructure. The role involves designing production-grade security services and integrating them into the CI/CD pipeline to ensure platform integrity.

Job details

NVIDIA DGX Cloud is the AI supercomputing-as-a-service substrate designed to power the next generation of AI and industrial-scale breakthroughs. As a Security Engineer within our Infrastructure Security Engineering organization, you will not just help "secure" our platform—you will architect and build the foundational security primitives that protect massive-scale GPU clusters. You will design automated, resilient security systems that help ensure the integrity of our omni-cloud and on-premise AI infrastructure. What You Will Be Doing: * Security Engineering: Design, build, and integrate production-grade security services. You will focus on the engineering of security products—transforming third-party and open-source tools into seamless, API-driven components of the DGX Cloud security stack. * Automated Policy Enforcement: Shift security "left" by developing Infrastructure as Code and Policy as Code to automate security enforcement and compliance at the speed of cloud-scale deployment. * Orchestration Security & Guardrails: Architect and implement the security control plane. You will engineer automated guardrails, controllers, and runtime security policies that validate and enforce the integrity of tenant boundaries. * Security-as-a-Service Approach: Designing and operating security services as a scalable platform. Building "self-service" security primitives (e.g., Identity-as-a-Service, automated secrets management, and real-time scanning APIs) that allow developer teams to move fast. * Security Tooling & Lifecycle: Develop internal security frameworks and automated response systems. Responsible for the full software development lifecycle (SDLC) of the security tools, including testing, deployment, and maintenance. * Threat Modeling & System Design: Conduct deep-dive threat models on complex distributed systems and the DGX Cloud stack, identifying architectural gaps in security and engineering the solutions to close them. * Multi-Functional Collaboration: Partner with DGX Cloud platform teams, broader NVIDIA security teams, and product engineering to understand their needs and build paved paths that seamlessly embed security into the CI/CD pipeline and the hardware lifecycle. What We Need to See: We are looking for high-caliber engineers with deep spikes of expertise in a few of these areas and the intellectual curiosity to dive into the rest. If your experience aligns with the core of this role—building resilient security systems—and you can show us how, we want to hear from you! * Infrastructure Engineering: Experience (typically 8+ years) in SRE, Software Engineering, and Infrastructure Security. You focus on building systemic solutions rather than performing manual operations or "tool administration." * Production-Grade Coding: A strong software engineering background with the ability to write clean, maintainable, and well-tested code. You should be comfortable building and maintaining production service at scale. * Distributed Systems Expertise: Understanding of cloud-native architecture, container orchestration (Kubernetes), and the security challenges inherent in high-throughput, low-latency environments. * Platformizing Security: Transform complex security requirements into consumable internal services. You will focus on the "Developer Experience" of security, ensuring that our infrastructure security controls are delivered as robust, API-first platforms that integrate seamlessly with NVIDIA’s internal engineering workflows. * Security Product Integration: Proven track record of taking complex security products (AuthN/AuthZ, Vaulting, Scanning, IDS) and integrating them into an automated infrastructure via APIs and custom glue-code. * Linux Internals: Strong hands-on experience with Linux systems security, including kernel-level primitives (eBPF, AppArmor, or SELinux). * Foundation: Bachelor’s degree in Computer Science, Engineering, or a related technical field (or equivalent experience). Ways To Stand Out from the Crowd: * HPC/AI Security: Experience securing high-performance computing environments, RDMA-based networks, or GPU-specific security challenges. * Cloud-Native Identity: Expertise in workload identity frameworks (e.g., SPIFFE/SPIRE) and hardware-root-of-trust (TPM/HSM) integration. * Open Source Impact: Notable contributions to security-focused open-source projects or a track record of engineering-focused security research. How have you represented and helped advance the industry? Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 170,000 CAD - 220,000 CAD for Level 4, and 225,000 CAD - 275,000 CAD for Level 5. You will also be eligible for equity and benefits [https://www.nvidia.com/en-us/benefits/]. Applications for this job will be accepted at least until September 18, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.

What you’ll do

You will architect and build foundational security primitives and automated guardrails to protect large-scale AI infrastructure. The role involves designing production-grade security services and integrating them into the CI/CD pipeline to ensure platform integrity.

Requirements

Candidates must have over 8 years of experience in SRE, software engineering, and infrastructure security with strong coding skills. Expertise in distributed systems, Linux security, and cloud-native orchestration is essential for this role.

Benefits

  • Equity
  • Benefits

Listed skills

  • Kubernetes · Preferred

Other relevant skills

Identified from the job description. Confirm important requirements above.

  • Infrastructure Security
  • Software Engineering
  • SRE
  • Kubernetes
  • Cloud-native Architecture
  • Linux Internals
  • eBPF
  • AppArmor
  • SELinux
  • Threat Modeling
  • API Development
  • Infrastructure as Code
  • Policy as Code
  • Distributed Systems
  • Security Automation
  • Identity Management
  • Security Tools
  • AI Security
  • Intellectual Curiosity
  • Cloud-Native Architecture
  • Cloud-Native Computing
  • CI/CD
  • Infrastructure Automation
  • Workflow Management
  • Resilience
  • Infrastructure as Code (IaC)
  • Research
  • Application Programming Interface (API)
  • Artificial Intelligence
  • Automation
  • Management
  • Cloud Security
  • Computer Science
  • Security Controls
  • Linux
  • Software Development Life Cycle
  • Scalability
  • Security Engineering
  • Operations
  • Policy Enforcement
  • Product Engineering
  • Systems Development Life Cycle
  • Remote Direct Memory Access
  • Security Policies
  • Security Requirements Analysis
  • Self Service Technologies
  • Supercomputing
  • Systems Design
  • Tooling
  • Security Systems

Job areas

  • Technology
  • Security & Safety
  • Software
  • Engineering
  • Data & Analytics
  • Infrastructure Security Engineer
  • Cyber Security Engineer
  • Database and Network Professionals Not Elsewhere Classified
  • Information Security Engineers
  • Computer Occupations, All Other

More jobs from NVIDIA

See all jobs from NVIDIA