NVIDIA Logo

NVIDIA

Senior Security Engineer, Infrastructure Security Engineering - DGX Cloud

Posted Yesterday
Be an Early Applicant
Remote
Hiring Remotely in Canada
Senior level
Remote
Hiring Remotely in Canada
Senior level
Design and build foundational security services for NVIDIA DGX Cloud’s large-scale GPU infrastructure. Develop automated Infrastructure-as-Code and Policy-as-Code enforcement, orchestration guardrails, identity and secrets services, scanning APIs, and security response systems. Integrate security products into API-first platforms and CI/CD workflows, conduct threat modeling, and secure distributed cloud-native systems. Collaborate with infrastructure, security, and product engineering teams across omni-cloud and on-premise environments.
The summary above was generated by AI

NVIDIA DGX Cloud is the AI supercomputing-as-a-service substrate designed to power the next generation of AI and industrial-scale breakthroughs. As a Security Engineer within our Infrastructure Security Engineering organization, you will not just help "secure" our platform—you will architect and build the foundational security primitives that protect massive-scale GPU clusters. You will design automated, resilient security systems that help ensure the integrity of our omni-cloud and on-premise AI infrastructure.

What You Will Be Doing: 

  • Security Engineering: Design, build, and integrate production-grade security services. You will focus on the engineering of security products—transforming third-party and open-source tools into seamless, API-driven components of the DGX Cloud security stack.

  • Automated Policy Enforcement: Shift security "left" by developing Infrastructure as Code and Policy as Code to automate security enforcement and compliance at the speed of cloud-scale deployment.

  • Orchestration Security & Guardrails: Architect and implement the security control plane. You will engineer automated guardrails, controllers, and runtime security policies that validate and enforce the integrity of tenant boundaries.

  • Security-as-a-Service Approach: Designing and operating security services as a scalable platform. Building "self-service" security primitives (e.g., Identity-as-a-Service, automated secrets management, and real-time scanning APIs) that allow developer teams to move fast.

  • Security Tooling & Lifecycle: Develop internal security frameworks and automated response systems. Responsible for the full software development lifecycle (SDLC) of the security tools, including testing, deployment, and maintenance.

  • Threat Modeling & System Design: Conduct deep-dive threat models on complex distributed systems and the DGX Cloud stack, identifying architectural gaps in security and engineering the solutions to close them.

  • Multi-Functional Collaboration: Partner with DGX Cloud platform teams, broader NVIDIA security teams, and product engineering to understand their needs and build paved paths that seamlessly embed security into the CI/CD pipeline and the hardware lifecycle.

What We Need to See: 

We are looking for high-caliber engineers with deep spikes of expertise in a few of these areas and the intellectual curiosity to dive into the rest. If your experience aligns with the core of this role—building resilient security systems—and you can show us how, we want to hear from you!

  • Infrastructure Engineering: Experience (typically 8+ years) in SRE, Software Engineering, and Infrastructure Security. You focus on building systemic solutions rather than performing manual operations or "tool administration."

  • Production-Grade Coding: A strong software engineering background with the ability to write clean, maintainable, and well-tested code. You should be comfortable building and maintaining production service at scale.

  • Distributed Systems Expertise: Understanding of cloud-native architecture, container orchestration (Kubernetes), and the security challenges inherent in high-throughput, low-latency environments.

  • Platformizing Security: Transform complex security requirements into consumable internal services. You will focus on the "Developer Experience" of security, ensuring that our infrastructure security controls are delivered as robust, API-first platforms that integrate seamlessly with NVIDIA’s internal engineering workflows.

  • Security Product Integration: Proven track record of taking complex security products (AuthN/AuthZ, Vaulting, Scanning, IDS) and integrating them into an automated infrastructure via APIs and custom glue-code.

  • Linux Internals: Strong hands-on experience with Linux systems security, including kernel-level primitives (eBPF, AppArmor, or SELinux).

  • Foundation: Bachelor’s degree in Computer Science, Engineering, or a related technical field (or equivalent experience).

Ways To Stand Out from the Crowd:

  • HPC/AI Security: Experience securing high-performance computing environments, RDMA-based networks, or GPU-specific security challenges.

  • Cloud-Native Identity: Expertise in workload identity frameworks (e.g., SPIFFE/SPIRE) and hardware-root-of-trust (TPM/HSM) integration.

  • Open Source Impact: Notable contributions to security-focused open-source projects or a track record of engineering-focused security research. How have you represented and helped advance the industry?

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 170,000 CAD - 220,000 CAD for Level 4, and 225,000 CAD - 275,000 CAD for Level 5.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August 26, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA Toronto, Ontario, CAN Office

Toronto, Ontario, Canada

Similar Jobs

An Hour Ago
Remote
British Columbia, BC, CAN
Junior
Junior
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Provides technical support for production radiology systems affecting patient care. Responsibilities include responding to support calls, managing a 20–30 case ticket backlog, documenting incidents, troubleshooting Windows, network, server, hardware, storage, and Oracle database issues, and coordinating with internal teams to resolve customer cases. The role is Canada-based and includes weekday shifts beginning between 5:00 and 9:30 a.m. PST, plus an on-call rotation.
Top Skills: Active DirectoryCC#C++Enterprise HardwareLanOracle DatabasesPacsPowershellPythonRisStorage SystemsTcp/IpWanWindowsWindows Command LineWindows ServerWindows Workstations
An Hour Ago
In-Office or Remote
CA
Senior level
Senior level
Blockchain • eCommerce • Fintech • Payments • Software • Financial Services • Cryptocurrency
Lead the technical design and delivery of banking features at Cash App, collaborating across teams, driving architecture improvements, and mentoring engineers.
Top Skills: AWSDatadogDynamoDBGrpcGuiceHibernateHTTPJavaJettyJSONJunitKafkaKotlinMySQLOkhttpPrometheusProtocol BuffersSignalfx
An Hour Ago
In-Office or Remote
CA
Senior level
Senior level
Blockchain • eCommerce • Fintech • Payments • Software • Financial Services • Cryptocurrency
Manage a portfolio of mid-market sellers, expanding use cases and technical solutions, negotiating contracts, and providing exceptional client service.
Top Skills: Technical Solutions

What you need to know about the Toronto Tech Scene

Although home to some of the biggest names in tech, including Google, Microsoft and Amazon, Toronto has established itself as one of the largest startup ecosystems in the world. And with over 2,000 startups — more than 30 percent of the country's total startups — Toronto continues to attract new businesses. Be it helping entrepreneurs manage their finances, simplifying business operations by automating payroll or assisting pharmaceutical companies in launching new drugs, the city's tech scene is just getting started.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account