Provide Tier 3 technical support for InfiniBand and Ethernet fabrics in large-scale AI, HPC, and storage environments. Manage complex fabric installations, troubleshoot end-to-end networking issues, and collaborate with engineering teams to resolve customer problems.
NVIDIA
156 Remote Job Openings at NVIDIA
You will build and deploy sophisticated AI-powered tools to support the operation and optimization of the global GeForce NOW service. This involves developing LLM-based systems to automate root cause analysis and managing large-scale data pipelines for operational intelligence.
Provide onsite technical support for large-scale NVIDIA AI datacenter deployments and resolve complex hardware issues. Establish strong technical relationships with customer architects and provide training to sales teams and partners.
Senior Software Engineer, Core Infrastructure Services - DGX Cloud
NVIDIA
·
Full Time
·
a day ago
NVIDIA
Build and operate core infrastructure services that power global AI infrastructure and DGX Cloud deployments. Develop scalable cloud-native platform services and automate infrastructure provisioning and lifecycle management.
You will profile, benchmark, and analyze AI and HPC workloads on GPU and CPU clusters to identify performance bottlenecks. Additionally, you will collaborate with cross-functional teams to develop diagnostic tools and provide actionable performance insights.
You will be responsible for defining and implementing chip pad rings, substrate interconnect schemes, and leading the package layout design process. You will collaborate with cross-functional teams including circuit, signal integrity, and system design engineers to ensure robust electrical package performance.
Drive growth in the Cisco EMEA channel business through partner development, enablement, and the execution of strategic business initiatives. Collaborate with Cisco channel teams and partners to accelerate pipeline creation and improve market execution for NVIDIA platforms.
The Research Scientist will build machine learning models and datasets to connect early drug discovery to clinical applications, including drug perturbation prediction and patient stratification. They will also design benchmarks for agentic systems, publish original research, and collaborate with internal and external teams to transition research into products.
You will identify and cultivate strategic relationships across the AI ecosystem to drive the adoption of NVIDIA's Vision AI and VLM capabilities. Additionally, you will provide technical leadership to partners and developers while guiding go-to-market strategies for intelligent environments.
You will design and implement accelerated computing data center solutions to support AI inference and physical AI workflows. Additionally, you will lead technical sales activities and collaborate with customers to optimize infrastructure performance and hybrid cloud deployments.
Develop and execute a technical strategy to drive the adoption of NVIDIA GPU platforms and AI solutions among startups and software partners in Southern Europe. Act as a technical mentor and evangelist to engage developer communities and influence technical leaders across key industries.
The engineer will perform hardware and software verification and validation to ensure product quality and stability. They will also troubleshoot testing procedures, optimize production jigs, and provide detailed activity reports.
You will serve as a technical advisor and champion for the EMEA AI developer ecosystem to drive the adoption of NVIDIA technologies. This involves collaborating with engineering and product teams to accelerate workloads, provide technical enablement resources, and influence product roadmaps based on partner feedback.
Drive the integration of NVIDIA's software libraries, models, and frameworks across strategic partners like Google Cloud and DeepMind. Lead cross-functional teams to accelerate critical workloads and influence partner product roadmaps through technical advocacy.
Senior State and Local Government Business Development Manager
NVIDIA
·
Full Time
·
3 days ago
NVIDIA
You will lead strategic partnerships with Global Public Sector ISVs to accelerate the adoption and integration of NVIDIA AI technologies. This role involves driving account strategy, fostering executive-level relationships, and developing co-sell motions to scale revenue across government and public-sector ecosystems.
You will serve as a technical advisor and champion for the EMEA AI developer ecosystem to drive the adoption of NVIDIA technologies. This involves collaborating with internal teams and partners to accelerate workloads, provide technical enablement, and influence product roadmaps based on field insights.
You will support the creation, implementation, and verification of AI factories by running and debugging AI/LLM workloads on Linux-based GPU clusters. Additionally, you will collaborate with cross-functional teams to optimize performance, scalability, and observability for customer-ready AI environments.
You will lead strategic customer initiatives by partnering with engineering and product teams to accelerate AI workload deployment within the Amazon ecosystem. Additionally, you will serve as a trusted business advisor to influence partnership direction and drive measurable business outcomes through deep technical and market expertise.
Senior Software Engineer, Infrastructure Automation and Distributed Systems
NVIDIA
·
Full Time
·
5 days ago
NVIDIA
Design, build, and maintain large-scale infrastructure services while managing the software lifecycle to meet business goals. Participate in observability strategies, incident response, and the elimination of operational toil through automation.
The role involves bridging cutting-edge AI research with production-grade infrastructure by identifying, prototyping, and integrating agentic system methods. You will also design evaluation harnesses and secure workflows to ensure agent reliability and safety in enterprise environments.
Analyze and improve the performance of HPC applications by driving compiler optimizations and library enhancements. Collaborate with compiler development and application engineering teams to deliver effective solutions for CPU and GPU platforms.
The role involves designing and optimizing large-scale AI infrastructure while collaborating with customers to maximize GPU utilization and workload throughput. You will also profile distributed training and inference workloads to identify bottlenecks and lead technical engagements for high-performance AI systems.
Develop and maintain Compute Sanitizer tools for GPUs across various operating systems including Linux, Windows, and embedded RTOS. Collaborate with compiler, architecture, and driver teams to design, implement, and verify new features while mentoring junior engineers.
Develop and manage automation software for NVIDIA's GPU Cloud and SuperPod deployments to enable efficient network design and lifecycle management. Collaborate with infrastructure experts to build scalable, zero-touch deployment solutions and integrate with various service APIs.
You will design and develop a massively distributed scalable platform to identify, diagnose, and remediate non-performant GPU assets. Additionally, you will collaborate across teams to ensure production AI clusters run reliably and consistently with maximum performance.
Drive revenue growth and achieve quota attainment for NVIDIA's networking portfolio across Enterprise accounts. Lead customer-facing discussions to position networking solutions within broader AI and data center initiatives while collaborating with technical teams.
Analyze high-performance computing applications to identify performance characteristics and optimization opportunities. Provide technical guidance to compiler and application engineering teams to improve GPU acceleration.
You will develop and execute a strategy for the healthcare robotics developer ecosystem, focusing on spatial intelligence and simulation workflows. Additionally, you will engage with MedTech partners and developers to provide technical guidance, create educational content, and represent the developer voice within NVIDIA.
The role involves partnering with internal customers to ensure successful adoption of DGX Cloud, removing technical blockers, and optimizing workload performance. You will translate complex business and technical requirements into scalable infrastructure recommendations while driving product improvements through customer insights.
The engineer will build, deploy, and optimize both on-premises and cloud-based data center infrastructure while managing network connectivity and routing. They will also utilize automation tools to streamline operations and lead root cause analysis for complex infrastructure faults.
Lead and coach a team of Storage Production Engineers to design, deploy, and maintain large-scale, high-performance storage systems. Partner with engineering and AI/ML teams to optimize data pipelines, ensure system reliability, and drive operational excellence.
You will develop and implement business plans to drive partner engagement and accelerate the adoption of NVIDIA's AI computing platform. Additionally, you will maintain executive-level relationships with key energy companies and the broader ecosystem to foster successful AI practices.
Drive NVIDIA platform adoption by nurturing strategic partnerships with CAE developers and ISVs while providing technical expertise. Inform product strategy by representing the developer community and leading technical collaborations across internal teams.
Drive NVIDIA platform adoption by nurturing strategic partnerships with CAE developers and ISVs. Inform product strategy by representing the developer community and identifying high-impact technical solutions.
You will design and operate scalable vulnerability management workflows while leveraging AI-assisted automation to enhance detection and remediation. The role involves partnering with cross-functional engineering teams to drive risk-based prioritization and embed security practices into the software development lifecycle.
You will scout, recruit, and support AI startups in the public sector, energy, and CAE industries to join the NVIDIA Inception program. Additionally, you will collaborate with cross-functional teams to drive go-to-market initiatives and accelerate the adoption of NVIDIA solutions.
Lead a team of software and production engineers in building and operating DGX Cloud infrastructure across various environments. Drive execution in cluster operations, Kubernetes operability, automation, and incident response while mentoring technical leaders.
The Solution Architect will partner with government clients to translate mission goals into AI-enabled strategies and deploy technical solutions. They will also mentor partners and customers while building resources to support the adoption of NVIDIA's AI software and hardware technologies.
You will architect and deploy AI-powered solutions to optimize semiconductor design and validation toolchains. This involves collaborating with cross-functional teams to integrate AI workflows that enhance efficiency and scalability across the production lifecycle.
Provide hands-on technical mentorship to customers on the NVIDIA GenAI stack and guide the development of Agentic AI workflows. Build demonstrations and POCs while partnering with engineering and sales teams to secure design wins.
Design and operate scalable backend services and data platform integrations for customer data, personalization, and AI-enabled marketing workflows. Partner with cross-functional teams to modernize identity systems, consent workflows, and improve platform reliability through observability.
Senior Sales Account Manager, Smart Spaces and Local Government
NVIDIA
·
Full Time
·
10 days ago
NVIDIA
Lead strategic relationships with cities and public-sector organizations to drive the adoption of NVIDIA's AI platforms and Digital Twin technology. Focus on crafting revenue growth and developing AI factories to modernize public services and urban infrastructure.
Support the RAPIDS project by managing CI/CD pipelines, container build processes, and infrastructure for data science libraries. Collaborate with engineering teams to implement DevOps best practices and ensure high-quality software releases.
Lead the automation, design, and operation of Linux platforms supporting physical clients and large-scale AI agent compute platforms. Develop automated workflows for provisioning, patching, and security hardening across hybrid-cloud environments.
Coordinate incident response, maintenance, and reporting across NVIDIA's datacenter portfolio to ensure reliability and minimize service disruptions. Lead root cause analysis and develop reliability standards, automation, and health scores to predict and isolate points of failure.
The role involves managing international immigration processes for employees and candidates across EMEA, APAC, and Canada. Key duties include partnering with global vendors, guiding stakeholders on policies, and identifying process improvements to enhance the employee experience.
Lead the design, implementation, and validation of high-performance data center networks for strategic customers using the PDI framework. Provide technical expertise in Ethernet routing and automation to optimize GPU-enabled AI applications and server infrastructure.
Design and implement innovations for managing GPU-based AI servers, focusing on BMC firmware and OOB management. Collaborate with global teams, hardware engineers, and industry partners to deliver high-end enterprise server platforms.
The role involves productizing new GPU boards for datacenter architectures and optimizing production lines for HGX GPU Accelerated Server Platforms. Responsibilities include developing diagnostic tests, ensuring DFx compliance, and collaborating with contract manufacturers to optimize assembly and yield.
The role involves developing and deploying scalable AI/ML solutions on cloud-based GPU platforms for strategic customers. You will build custom PoCs, conduct technical meetings, and partner with sales teams to drive NVIDIA technology adoption.
Serve as a technical advisor to security companies to drive the adoption of NVIDIA's software stack for AI and Agentic environments. Collaborate with internal engineering and product teams to influence roadmaps based on developer feedback and ecosystem trends.
Lead the architecture for cloud-networking, security, and orchestration solutions for DPUs and NICs. Develop end-to-end solutions from the application level to hardware and create reliable architecture specifications.
Own the mechanical aspects of production test fixtures, jigs, and enclosures while overseeing maintenance and calibration across global contract manufacturing sites. Support NPI mechanical readiness and provide technical guidance to ensure production quality and manufacturing excellence.
Analyze HPC applications to improve performance by driving compiler optimizations and library enhancements. Evaluate the impact of NVHPC components and port applications across different platforms and programming models.
Support the Enterprise EMEA Sales team by managing sales operations, order execution, and customer queries. Act as a liaison between Sales, Operations, and Finance to ensure efficient processes and accurate reporting.
Senior Enterprise Sales Account Manager, Media and Entertainment
NVIDIA
·
Full Time
·
16 days ago
NVIDIA
The role involves connecting with Media and Entertainment customers to identify business needs and quantify market opportunities. The manager will orchestrate resources to implement partnership strategies and drive the adoption of NVIDIA technology solutions.
Solutions Architect β Power and Utilities, AI Grid, Power Systems
NVIDIA
·
Full Time
·
16 days ago
NVIDIA
Develop and prototype GPU-accelerated AI solutions for grid modernization, including forecasting, resilience, and load orchestration. Act as a technical advisor to partners and customers to integrate NVIDIA technology into power and utility architectures.
Responsible for product engineering of NVIDIA adaptors, switches, and interconnect products. This includes establishing and training CM engineering teams and collaborating with cross-functional teams to resolve bugs and improve product quality.
Design and deploy GPU-accelerated AI solutions and industrial digital twins for energy operations across power generation, grid, and oil & gas sectors. Act as a technical advisor to integrate NVIDIA platforms into control rooms and field environments to solve last-mile deployment challenges.
Develop and harden a distributed systems platform providing secure, sandboxed runtimes for autonomous AI agents. This includes implementing network security, inference routing, and control plane systems while ensuring production observability.
Lead NVIDIA's engagement with the French automotive ecosystem by building C-level relationships across OEMs and Tier-1 suppliers. Develop multi-year account strategies to drive the adoption of NVIDIA's AI hardware and software platforms.
Drive factory execution across manufacturing partners to ensure production output aligns with customer demand and AVC commitments. Manage build plans, resolve material shortages, and support manufacturing expansion initiatives in Mexico.
Create and deliver technical content, tutorials, and workshops to educate the developer community on the CUDA platform. Collaborate with product and engineering teams to provide developer-centric insights and inform the product roadmap.
Manage strategic relationships with Ford and Rivian to drive design wins and revenue growth in automated driving and cockpit applications. Collaborate with internal stakeholders and external R&D teams to position NVIDIA's AI and deep learning solutions within the automotive ecosystem.
Drive the deployment and integration of next-generation AI networking platforms and fabrics at strategic customer data centers. Partner with customers to guide architecture decisions and translate technical requirements into product feedback for engineering teams.
Maintain the reliability, availability, and performance of the GeForce NOW cloud gaming platform across cloud and datacenter environments. Drive observability initiatives and build automation tools to eliminate operational toil and improve service SLOs.
The manager will lead a team to operationalize NVIDIA technologies, bridging the gap between prototypes and production-ready deployments. Responsibilities include managing deployments, resolving obstacles, and collaborating with engineering teams.
Design and optimize GPU-accelerated software for deep learning inference, focusing on LLM and Generative AI models. Contribute to open-source frameworks and libraries like vLLM and SGLang to improve model serving pipelines across NVIDIA accelerators.
Drive the deployment of end-to-end AI hardware and software solutions in customer data centers. Act as a technical advisor to strategic customers, guiding network design and providing feedback for the product roadmap.
Lead technical engagements to accelerate customer AI workloads and reduce infrastructure costs. Develop proof-of-concepts and debug software for NVIDIA and open-source AI frameworks.
Design and evaluate routing policies for LLM traffic and build agentic benchmarks to measure algorithm quality. Collaborate with engineering teams to integrate software across the NVIDIA accelerated serving stack and contribute to open-source repositories.
Lead the definition, positioning, and go-to-market strategy for AI physics products, models, and frameworks. Collaborate with engineering and marketing teams to translate technical capabilities into customer value and partner enablement assets.
Support SLT NPI manufacturing activities from product bring-up through mass production. Develop automated test systems, thermal solutions, and leverage AI technologies to improve testing efficiency and yield.
The engineer will perform technical analysis and debugging of returned networking products at contract manufacturer locations. They are responsible for identifying failure trends and escalating complex systemic issues to global engineering and design teams.
Define and drive structural test plans for baseboards and systems to detect SMT, PCB, and component issues early. Collaborate with cross-functional teams and suppliers to develop automated test infrastructure and hardware solutions.
Provide onsite technical engagement and support for large-scale NVIDIA AI datacenter deployments. Act as a technical specialist for GPU and networking products to support sales account managers and establish relationships with customer architects.
Senior Platform Safety Analysis Engineer, HALOs Platforms - Autonomous Vehicles
NVIDIA
·
Full Time
·
a month ago
NVIDIA
Lead the development of safety architecture and requirements for AI-powered autonomous driving platforms. Perform complex safety analyses using qualitative and quantitative methods to ensure hardware and software fault metric compliance.
Lead the design and development of end-to-end reference system stacks for 5G/6G baseband systems. Optimize CPU, GPU, and NIC sub-systems to ensure low-latency and maximum throughput while collaborating with partners and customers.
Responsible for production systems enabling large scalable GPU clusters for AI workloads, including asset provisioning and lifecycle management. Focuses on implementing monitoring, health management, and incident response to ensure maximum performance and reliability.
Architect and scale high-performance distributed AI infrastructure using NVIDIA GPU supercomputers for diverse customers. Provide technical leadership and on-site support throughout the product lifecycle to ensure successful deployment and customer satisfaction.
Build and maintain automated infrastructure across bare-metal, virtualized, and containerized environments. Drive continuous improvements in CI/CD pipelines and provide automation support for development and verification teams.
Provide onsite technical engagement and support for large-scale NVIDIA AI datacenter deployments. Act as a technical specialist for GPU and networking products to support sales and establish relationships with customer architects.
Architect and build scalable RL post-training infrastructure that spans from single GPU experimentation to production across thousands of nodes. Collaborate with researchers to optimize deep learning frameworks and improve distributed runtimes like Ray and Monarch.
Serve as the primary post-sale relationship owner for U.S. federal agencies to drive the adoption and value realization of NVIDIA's AI portfolio. Develop tailored success plans and lead executive business reviews to align technology deployments with mission-critical outcomes.
Build and maintain a Kubernetes-native control plane to gather, aggregate, and normalize topology data for GPU-accelerated infrastructure. Interface with NVIDIA hardware to optimize GPU-to-GPU communication for large-scale workloads across multiple cloud providers.
Define and drive NVIDIA's JAX strategy to ensure peak performance across heterogeneous supercomputing platforms. Lead and mentor a high-performing engineering organization while coordinating contributions across the JAX ecosystem and partnering with external open-source projects.
Senior Systems Software Engineer, Accelerated Kubernetes Performance and Scale - DGX Cloud
NVIDIA
·
Full Time
·
a month ago
NVIDIA
Lead performance and scalability analysis for the Kubernetes-based accelerated runtime stack to optimize AI infrastructure. Design architectural changes for the Kubernetes control plane and contribute to open-source projects to enable hyperscale AI workloads.
Define high-level SoC subsystem architecture for LPU products and convert requirements into detailed architectural specifications for uncore and I/O. Collaborate with IP and software teams to build functional models and drive tradeoffs in bandwidth, power, and latency.
Drive the deployment of end-to-end AI hardware and software technology solutions at strategic customer data centers. Act as a technical advisor to guide network design, debug performance issues, and provide feedback for the product roadmap.
Senior Developer Relations Manager β Cloud Provider AI Factory
NVIDIA
·
Full Time
·
a month ago
NVIDIA
Drive the adoption of NVIDIA's AI and computing platforms by serving as a technical advisor to cloud provider and hosting ecosystems. You will integrate the NVIDIA software stack into partner products and collaborate with internal engineering teams to inform product roadmaps.
Drive end-to-end performance and scale characterization for the NVIDIA DGX Cloud software stack, focusing on Kubernetes and GPU components. Design monitoring tools and collaborate with the open-source community to optimize AI infrastructure and reduce cost per token.
Provide comprehensive technical support and debugging for AI hardware and software products on the DGX Platform. Collaborate with Engineering and Marketing teams to improve product requirements and support methodologies.
Senior Systems Software Engineer, Kubernetes Scale - DGX Cloud
NVIDIA
·
Full Time
·
a month ago
NVIDIA
Drive end-to-end performance and scale characterization for the NVIDIA DGX Cloud software stack, focusing on Kubernetes and GPU components. Design monitoring tools and collaborate with the open-source community to optimize AI infrastructure and reduce cost per token.
Partner with research universities to co-create innovative HPC and AI solutions using NVIDIA's accelerated computing platform. Architect ground-breaking workflows in computational physics and optimize AI training and inference workloads for scientific applications.
Senior Solutions Architect, AI Factory Observability and Visualization - NVIS
NVIDIA
·
Full Time
·
a month ago
NVIDIA
Develop full-spectrum visibility and observability for HPC systems and AI factories to ensure optimal performance. This includes running validation tools, building telemetry surfaces, and collaborating across teams to transform complex data into actionable insights.
Partner with university researchers to advance foundation models, multimodal AI, and agentic systems using NVIDIA's accelerated computing platforms. Provide technical guidance on GPU-accelerated training, inference studies, and the development of research prototypes.
Deploy, manage, and maintain large-scale AI infrastructure for customers as a domain expert. Collaborate with internal teams to provide feedback, document workarounds, and implement AI Factory projects.
Research and implement model architecture improvements to enhance video generation fidelity and human-centric quality for world foundation models. Translate research results into production-grade checkpoints and benchmarks to improve physical AI and simulation.
The role focuses on owning the mechanical aspects of production test fixtures, jigs, and enclosures across global contract manufacturing sites. It involves managing maintenance, calibration, and NPI mechanical readiness to ensure high-volume manufacturing excellence.
Lead the GeForce business by managing AIC/OEM partners and distributors to drive sales growth and market dominance. Develop end-user demand generation programs and collaborate with gaming ecosystem partners to increase brand preference.
Lead the growth of the Federal Physical AI ecosystem by driving the adoption of NVIDIA's AI and computing platforms through strategic partnerships. Collaborate cross-functionally to develop go-to-market resources and provide actionable field insights to influence internal product roadmaps.
Senior Software Engineer β Accelerated Quantum Chemistry and cuEST
NVIDIA
·
Full Time
·
2 months ago
NVIDIA
Architect, implement, and optimize production-grade GPU solutions for the cuEST quantum chemistry library. Collaborate with the broader community and internal teams to drive the adoption and development of GPU-accelerated electronic structure computations.
Partner with engineering and sales teams to secure design wins and optimize ML/DL models for financial services clients. Develop technical collateral and proof-of-concepts to accelerate high-performance computing workloads in capital markets.
The role focuses on enabling AI startups and enterprise customers to accelerate their applications using NVIDIA SDKs and technologies. This includes creating high-quality technical content, tutorials, and training materials to drive adoption across the developer ecosystem.
The role focuses on building a pipeline of new opportunities for NVIDIA's Nordic sales teams through lead management and demand generation. This includes identifying prospects via phone, email, and LinkedIn, and collaborating with marketing and sales teams to convert leads.
Develop energy-specific applications and platforms using NVIDIA technology to serve the Oil and Gas, Renewables, and Power Utilities sectors. Act as a technical advisor to partners and customers, bridging the gap between industry teams and technical implementation.
Design, deploy, and operate large-scale storage and data platforms on Kubernetes using automation and infrastructure-as-code. Develop telemetry and observability tools to ensure system health and participate in sustainable incident response and on-call rotations.
Profile and optimize end-to-end neural reconstruction and Gaussian Splatting workflows to improve speed, scalability, and reliability. Translate Python and PyTorch bottlenecks into efficient CUDA/C++ implementations while ensuring reconstruction quality is preserved.
Collaborate with internal and external teams to define system requirements for fault-tolerant quantum computing. Develop novel approaches for real-time quantum error correction and calibration across various qubit modalities.
The role involves acting as a technical consultant for ISV developers to foster the adoption of NVIDIA's AI and computing platforms. You will collaborate cross-functionally to identify growth opportunities, guide partner onboarding, and influence product roadmaps based on field feedback.
Senior Solutions Architect, Cloud Infrastructure and DevOps - NVIS
NVIDIA
·
Full Time
·
2 months ago
NVIDIA
Maintain large-scale HPC/AI clusters and develop automation tooling for deployment, monitoring, and resource consumption. Collaborate with customers and internal teams to analyze and implement large-scale networking projects.
Senior Solutions Architect, Generative AI - AI Models and Systems at NVAITC
NVIDIA
·
Full Time
·
2 months ago
NVIDIA
Collaborate with university research labs to identify and execute high-impact Generative AI projects. Act as a strategic bridge between academic partners and NVIDIA's engineering teams to drive the adoption of NVIDIA software platforms.
Senior Solutions Architect, Physical AI and Robotics at NVAITC
NVIDIA
·
Full Time
·
2 months ago
NVIDIA
Collaborate with university PIs on high-impact Physical AI and Robotics research projects while championing the adoption of NVIDIA software platforms. Act as a strategic bridge between academic partners and NVIDIA's internal engineering teams to drive world-class research and institutional agreements.
Design and deploy Agentic AI applications to automate telecommunications network operations using generative models and RAG pipelines. Provide technical guidance to strategic partners and translate integration challenges into reference architectures for the NVIDIA accelerated computing stack.
Profile and optimize end-to-end neural reconstruction and Gaussian Splatting workflows to improve speed, scalability, and reliability. Translate Python and PyTorch bottlenecks into efficient CUDA/C++ implementations while ensuring reconstruction quality is preserved.
Senior Solutions Architect, Infiniband and Networking Ethernet - NVIS
NVIDIA
·
Full Time
·
2 months ago
NVIDIA
Build and support large-scale AI/HPC infrastructure for customers, focusing on performance, reliability, and real-time monitoring. Collaborate with internal teams to refine services and implement large-scale networking projects.
Lead research in AI for quantum algorithm discovery and drive technical collaborations with supercomputing centers and QPU builders. Develop innovative quantum-classical applications and publish impactful research to drive NVIDIA's quantum product adoption.
Manage NVIDIA Interconnect products by ensuring flawless engineering implementation, maintenance, and yield management. Coordinate with manufacturing partners and cross-functional teams to resolve product failures and scale capabilities.
Senior Solutions Architect, Simulations - Clinical Sciences and Autonomous Lab
NVIDIA
·
Full Time
·
3 months ago
NVIDIA
Drive innovation in healthcare and life sciences by designing and optimizing GPU-accelerated AI software for clinical sciences and autonomous labs. Partner with pharmaceutical companies to implement patient modeling, robotic systems, and biomedical agentic AI.
Lead the research, design, and implementation of security architectures for next-generation NVIDIA Networking products. Collaborate with cross-functional teams and external partners to develop hardware security primitives and trusted platforms.
Senior Technical Program Manager, Pre-Silicon Software Enablement and Workload Studies
NVIDIA
·
Full Time
·
3 months ago
NVIDIA
Drive NVIDIA's software left-shift program to ensure software teams have necessary infrastructure to begin development early in the silicon lifecycle. Lead cross-functional alignment between architecture, modeling, and software teams to resolve dependencies and improve pre-silicon results.
Conduct in-depth performance characterization and analysis on large multi-GPU and multi-node clusters. Triage and root-cause performance issues while building tools to visualize and analyze performance data.
The Senior Product Engineer will manage board product lifecycles, including yield management, manufacturing process optimization, and failure resolution. They will collaborate with cross-functional teams to ensure high-quality product execution across global contract manufacturing sites.
The Senior Supplier Quality Engineer will lead factory quality onsite activities, including NPI and mass production, while ensuring compliance with NVIDIA quality standards. They will also facilitate root cause analysis, manage supplier performance metrics, and drive continuous improvement initiatives across the supply chain.
You will develop, deploy, and validate AI factory environments by running and debugging complex AI/LLM workloads on GPU clusters. Additionally, you will build automation and observability tools to optimize performance, latency, and scalability for distributed training.
You will lead technical engagement efforts with defense partners to integrate NVIDIA's accelerated computing stack into autonomous aerial platforms and uncrewed systems. This involves architecting perception and planning pipelines, providing reference designs, and guiding product roadmaps for edge AI technologies.
Lead and mentor technical teams while driving AI research and strategic collaborations with academia and industry. Develop and implement NVIDIA technology-related tutorials, workshops, and demos to foster accelerated AI adoption.
The Solutions Architect will serve as a technical advisor to drive the design, integration, and deployment of large-scale AI and GPU infrastructure for strategic partners. They will collaborate with cross-functional teams to deliver technical content, conduct workshops, and ensure successful implementation of NVIDIA hardware and software solutions.
Manage end-to-end production infrastructure supply chain from NPI to mass production delivery. Coordinate capacity management, risk assessment, and maintenance activities across multiple production sites.
You will design and deploy sophisticated Agentic AI systems for top-tier retail and enterprise clients using NVIDIA's core technology stack. This role involves building reference architectures, optimizing inference performance, and enabling partner engineering teams through technical workshops and documentation.
The Solutions Architect will engage with customers and partners to deliver high-value technical solutions leveraging NVIDIA's AI, HPC, and networking platforms. They will act as a trusted technical advisor, creating documentation and educational content while collaborating with internal teams to drive customer success.
You will plan and establish processes, define test requirements, and optimize production lines to successfully launch new GPU boards for datacenter architectures. Additionally, you will collaborate with cross-functional teams and contract manufacturers to ensure cost and quality metrics are met while resolving yield and test problems.
The role involves guiding partners in adopting end-to-end Agentic AI solutions and collaborating with customers and partners to deploy AI solutions at scale. Solution Architects will also assist with demos, proof-of-concepts, and knowledge sharing.
The Project Technical Delivery Manager will oversee the complete project lifecycle from initiation to close, ensuring technical requirements are met within budget. They will facilitate technical architecture decisions, manage complex installations, and maintain effective relationships with stakeholders and customers.
You will serve as a technical SME to design developer tools, APIs, and workflows for chemistry and materials science. You will also collaborate across research and engineering teams to shape the NVIDIA ALCHEMI software stack and represent the company at scientific conferences.
The researcher will identify hardware vulnerabilities on SoC and GPU designs and develop advanced security tools and techniques. They will also guide the integration of security mitigations and conduct research into side-channel, fault, and physical attacks.
Senior Software Engineer - Accelerated Kubernetes Runtime Team
NVIDIA
·
Full Time
·
4 months ago
NVIDIA
Design and implement automation systems to orchestrate the lifecycle of runtime components across thousands of Kubernetes clusters. Develop Kubernetes controllers, operators, and CRDs to manage the installation, upgrade, and validation of accelerated compute components.
Primary responsibilities include building and operating AI/HPC infrastructure for new and existing customers. The role involves supporting operational and reliability aspects of large-scale AI clusters.
Maintain large scale computational and AI infrastructure, focusing on monitoring, logging, and workload orchestration. Serve as a key technical resource, developing and documenting standard methodologies and operational guidelines.
This role involves serving as a trusted technical advisor and champion for the EMEA AI Natives developer ecosystem, driving adoption of NVIDIA technologies by demonstrating groundbreaking solutions and accelerating critical workloads. The manager will also guide partners and startups through integration, track ecosystem growth, and collaborate cross-functionally to optimize adoption strategies.
This role involves researching and developing techniques to optimize key Cloud and HPC CPU workloads specifically on NVIDIA's CPU, requiring in-depth analysis for current and future generations. Responsibilities also include engaging with the developer community, guiding framework developers, and contributing directly to their software stack or developing reference code.
The engineer will develop and implement CUDA Core Libraries in C++ and/or Python, focusing on parallel algorithms and idiomatic language bindings for core CUDA functionality. Responsibilities also include composing, optimizing, and evolving GPU algorithms and APIs, owning features end-to-end, and improving the overall developer experience.
The Senior DFT Engineer will define and implement SCAN, MBIST, and JTAG debug structures, driving post-silicon testing plans and creating ATPG and MBIST test vectors. They will also build DFT timing constraints, partner with physical design teams, and work with the post-silicon team to bring up test patterns on actual silicon.
Primary responsibilities involve deploying, managing, and maintaining AI/HPC infrastructure in Linux-based environments for customers, acting as the domain expert during planning and implementation phases. This role also requires creating handover documentation, performing knowledge transfers, and providing feedback to internal teams regarding bugs and improvements.
The role involves developing a highly optimized inference framework that runs on the worldβs largest supercomputers and data centers, focusing on performance and scalability in AI networking acceleration.
This role focuses on redefining AI hardware development methodology by inventing next-wave techniques, pioneering AI-driven automation for sophisticated ASIC conception, exploration, and closure. Responsibilities include taking a comprehensive view of the ASIC lifecycle to identify bottlenecks where automation and AI can improve convergence and turnaround time.
The engineer will take a comprehensive view of the ASIC development lifecycle to identify bottlenecks where automation and AI can improve predictability and turnaround time. They will also establish quantitative metrics to measure efficiency and serve as a technical catalyst by sharing best practices and mentoring engineers on emerging AI-enabled techniques.
EMEA Sales Senior Account Manager, Smart Spaces and Local Government
NVIDIA
·
Full Time
·
6 months ago
NVIDIA
This role involves owning strategic relationships with leading cities and public-sector organizations to position NVIDIA's platforms and AI Factory strategy for modernizing public services and competitiveness across EMEA. Key activities include crafting revenue growth for Smart Cities solutions, building a robust pipeline focused on AI deployment, and developing long-term relationships with senior city leaders.
As a Senior Formal Verification Engineer, you will verify ASICs using formal verification tools and define the verification scope to ensure correctness. You will collaborate with various teams to resolve design issues and improve verification methodologies.
As a key member of the ASIC Verification team, you will verify the design and implementation of the inference accelerator. You will collaborate with architects, designers, and verification teams to ensure the correctness of the design.
As a key member of the Design team, you will implement, document and deliver high performance, area and power efficient RTL. You will collaborate with various teams to analyze architectural trade-offs and deliver fully verified designs.
As a Senior Formal Verification Engineer, you will verify ASICs using formal verification tools and define the verification scope to ensure correctness. You will collaborate with various teams to improve methodologies and deliver high-quality results on schedule.
Analyze Deep Learning models and investigate TensorRT stability and performance issues. Work with an internationally distributed team for CUDA and TensorRT development.
Lead architecture for cloud-networking and security solutions while designing state-of-the-art system architecture for DPUs & NICs technologies. Collaborate with global teams to innovate and develop proof of concept prototypes into full-fledged products.
Develop automation for deploying Kubernetes clusters for streaming media use cases and monitor and manage these clusters. Collaborate with other NVIDIA R&D teams globally in a fast-paced environment.
Senior Systems Software Security Engineer β Data Center Systems
NVIDIA
·
Full Time
·
9 months ago
NVIDIA
You will focus on securing NVIDIAβs Data Center Systems by delivering necessary security features and engaging with teams to drive implementation. Your role will involve designing and developing optimized security solutions following industry standards.