19 open positions available
Design and develop scalable automation and orchestration systems for cloud infrastructure workflows. | Senior-level experience with programming, cloud infrastructure, container orchestration, and distributed systems. | We are looking for a Senior Software Engineer to join our DGX Cloud team and build the foundational systems that drive NVIDIA's high-performance GPU infrastructure. You will play a critical role in designing scalable automation solutions, integrating diverse systems, and enabling seamless workflows across global cloud operations. What You'll Be Doing: • Design and develop APIs to orchestrate and integrate operational workflows. • Build state management and workflow automation systems that streamline infrastructure lifecycle processes. • Collaborate across teams to codify business processes into scalable, self-measuring systems. • Develop extensible, schema-driven platforms for reducing manual toil and ensuring operational consistency. • Drive integrations with container orchestration tools like Kubernetes and observability systems such as Prometheus, OpenTelemetry, Grafana. • Optimize the reliability and efficiency of cloud operations through automated workflows and telemetry systems. • Lead and ship impactful technical projects, ensuring quality and scalability at every stage. What we need to see: • 8+ years of industry experience with a Bachelor's or Master's degree (or equivalent experience), or 2+ years with a PhD. • Expertise in designing, building, and operating services in a high reliability environment. • Proficiency in programming languages such as Go, Java, or Python. • Strong understanding of cloud infrastructure (AWS, GCP, Azure) and container technologies like Docker and Kubernetes. • Experience with high-scale distributed systems, including architectural patterns for APIs and data pipelines. • Outstanding communication and collaboration skills, with a focus on solving complex operational challenges. • A passion for automating manual processes and driving system efficiency. Ways to Stand Out from the Crowd: • A track record of designing workflow orchestration systems for large-scale infrastructure. • Proven experience in reducing operational inefficiencies through automation and integration. • Strong debugging and problem-solving skills in distributed environments. • Prior experience or strong familiarity with the operational aspects of the NVIDIA AI/ML software stack (e.g., CUDA, cuDNN, containerization) Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until July 27, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.
Lead and manage the AI Factory Architecture Review Board process end to end, coordinating multiple teams and ensuring deployment readiness. | 5+ years program management in large-scale infrastructure or AI deployments, with strong coordination and communication skills. | In this role, you will lead the AI Factory Architecture Review Board process that helps turn complex customer infrastructure designs into validated, deployment-ready solutions! What you will be doing: • Own and lead the AI Factory ARB process end to end, from initial intake through operations and pre-deployment handoff. • Lead architecture governance, assessments of deployment preparedness, and the validation and sign-off activities for complex AI Factory deployments. • Coordinate account teams, solution architects, NVIDIA infrastructure specialists, engineering, networking, storage, security, facilities, operations, OEM partners, and NVIDIA Cloud Partner teams. • Manage ARB schedules, agendas, decision logs, action owners, risk tracking, blocking issues, and executive reporting. • Set and maintain deployment gates, design checkpoints, exception handling, and sign-off processes aligned to NVIDIA Reference Architectures, Reference Designs, and deployment standards. • Review ARB submissions, classify requests by review path, prepare teams for timely decisions, and maintain clear documentation in shared workspaces. • Coordinate remediation and exception paths for designs that do not meet Reference Architecture or Reference Design criteria, including approval tracking and follow-through on blocking issues. • Improve ARB templates, tools, documentation, reporting, automation, and AI-assisted workflow agents; partner with global teams to scale ARB governance across regions and partners. What we need to see: • Bachelor's degree in engineering, computer science, business, operations, or a related field, or equivalent experience. • 5+ years of program or project management experience in large-scale infrastructure, cloud, networking, data center, hardware, or AI Factory deployment environments. • Experience coordinating architecture reviews, governance boards, technical approval processes, or deployment readiness reviews. • Working understanding of AI infrastructure, GPU clusters, networking, storage, data center operations, deployment readiness, and technical validation processes. • Experience using Jira or similar tools for agile execution, risk management, dependency tracking, and deployment lifecycle management. • Ability to lead multi-functional programs across technical, commercial, operations, and partner teams through clear ownership and follow-through. • Clear written and verbal communication skills, including experience preparing executive updates and translating technical risks into business impact. Ways to stand out from the crowd: • Experience with GPU-accelerated computing, AI infrastructure, HPC, DGX SuperPOD, OEM, NVIDIA Cloud Partner, or enterprise data center environments. • Familiarity with NVIDIA Reference Architecture and Reference Design compliance, customer infrastructure validation, or pre-deployment handoff processes. • Experience with Salesforce and Wrike for pipeline tracking, data quality, dashboards, reporting, or workflow automation. • PMP, PgMP, Agile, Scrum, or equivalent program management certification or experience improving global governance processes. • Use of reporting, analytics, automation, or AI tools such as Sheets, Copilot, Claude, Gemini, or Perplexity to improve program transparency and execution. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 168,000 USD - 258,750 USD. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until June 15, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.
Lead and scale logistics and service fulfillment programs ensuring operational excellence and continuous improvement. | 8+ years in program or operations management in logistics or service operations, strong data-driven decision making, vendor management, and experience with ERP and logistics systems. | We are looking for an experienced Program Manager to lead and scale service fulfillment and Reverse logistics operations across the U.S. You will design, implement, and continuously improve end-to-end processes for installation, field service fulfillment, parts logistics, returns, and depot repair to meet aggressive SLAs, cost targets, and customer experience goals. This position is based in Dallas, Texas. What you'll be doing: • Lead cross-functional programs to operationalize new products, services, and service offerings, ensuring readiness across fulfillment, logistics, field service, and support. • Design and optimize end-to-end fulfillment flows: order intake, kitting, staging, shipping, installation scheduling, returns, repair, and reverse logistics. • Manage relationships with 3PLs, carriers, depot/repair partners, warehousing teams, and field service providers; negotiate contracts and measure performance. • Define program benchmarks (OTD, SLA compliance, RMA turnaround, inventory turns, shipping cost per unit, fill rate, first-time fix rate) and build dashboards for leadership. • Lead continuous improvement initiatives (Lean, Six Sigma tools) to reduce cost, cycle time, and failures while increasing capacity and quality. • Develop capacity planning and inventory strategies based on projected demand to support product launches and seasonal demand spikes. • Lead change management: rollout of new systems (WMS, TMS, FSM, ERP integrations), SOPs, training, and governance. • Own risk management and business continuity planning for logistics and field operations (disruption mitigation, contingency carriers, spare pools). • Coordinate cross-functional collaborators (product, engineering, sales, finance, customer success) to align service SLAs with commercial commitments. What we need to see: • 8+ years of program or operations management experience focused on fulfillment, logistics, field service, or service operations in a high-tech or hardware environment. • Bachelor's degree or equivalent experience in supply chain, Engineering, Business, or related; MBA or relevant certification (CSCP, Six Sigma) preferred. • Proven track record managing complex, cross-functional programs with external partners (3PLs, carriers, repair depots). • Strong data-driven decision-making: experience with benchmark definition, dashboards, and regular operational reviews. • Hands-on experience with WMS/TMS, FSM platforms, and ERP integration projects. • Demonstrated ability to optimize inventory, forecast demand, and manage reverse logistics/RMA processes. • Excellent collaborator management, negotiation, and vendor management skills. • Comfortable working in fast-paced, ambiguous environments with tight timelines. • Willingness to travel (up to 25%) and work across Texas time zones as needed. Ways to Stand out from the crowd: • Strong project management (PMP, Agile) and process improvement capabilities. • Analytical proficiency with SQL, Excel modeling, and BI tools (Tableau, Power BI). • Excellent written/verbal communication, able to present to executive audiences. • Customer-centric approach with a focus on operational excellence and continuous improvement. • Experience scaling operations for rapid growth and new product launches. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 136,000 USD - 218,500 USD for Level 4, and 176,000 USD - 276,000 USD for Level 5. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until June 12, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.
Develop and maintain CUDA Core Libraries with focus on performance and developer experience. | 8+ years development experience with strong C++ and Python skills, systems-level programming, and parallel computing expertise. | We are hiring a full-time Software Engineer to work on the CUDA Core Libraries that power GPU computing for both C++ and Python developers. This includes projects such as CCCL (Thrust, CUB, libcudacxx), cuda-python, and numba-cuda. You will join the team building the foundational libraries, algorithms, and language/runtime infrastructure that make CUDA a speed-of-light experience for developers across deep learning, scientific computing, and data analytics! What you'll be doing: • Develop and implement CUDA Core Libraries in C++ and/or Python, including parallel algorithms and idiomatic language bindings for core CUDA functionality. • Compose, optimize, and evolve GPU algorithms and APIs, from high-level interfaces down to low-level performance tuning involving memory, parallelism, and synchronization. • Own features end-to-end: develop, implementation, testing, benchmarking, documentation, and long-term maintenance. • Improve developer experience across the stack: CI, tests, benchmarks, packaging, examples, and docs. • Collaborate with senior CUDA engineers in design reviews, code reviews, and open-source-style workflows. • Engage with real users through issues, performance investigations, and API feedback. What we need to see: • BS, MS, or PhD in Computer Science, Computer Engineering, or a related field or equivalent experience. • Minimum of 8+ years of related development experience • Strong programming skills in C++, Python, or both, with proven interest in systems-level software (performance, memory, concurrency, API design). • Solid understanding of modern C++ (templates, generics, standard library) and/or Python library development and packaging. • Practical experience with parallel or heterogeneous programming (CUDA, OpenMP, GPU-accelerated Python, or similar). • Experience contributing to production software or open-source libraries, including testing, profiling, and code review. • Ability to work independently, scope problems, and drive projects to completion. • Clear written communication for technical design and documentation. • Comfort navigating large, multi-language codebases (C++, Python, CMake, Pixi, CI systems). Ways to stand out from the crowd: • Strong understanding of CPU/GPU architecture and how hardware details affect performance. • Hands-on experience with CUDA C++, CUDA Python, PyTorch, JAX, Numba, CuPy, or similar GPU-accelerated stacks. • Familiarity with Thrust, CUB, libcudacxx, or other modern C++/GPU libraries. • Experience with compiler infrastructure or tooling (LLVM, Clang tooling, MLIR). • Demonstrated interest in developer tools, library design, and making other developers faster. If you care deeply about performance, enjoy working at the C++/Python boundary, and want to shape the core CUDA libraries relied on by thousands of developers, this role is a direct fit. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until June 5, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.
Develop and lead testing infrastructure and automation for autonomous vehicle simulation platforms. | Senior-level experience in software testing, simulation environments, and strong programming skills in Python/C++. | The Automotive Vehicles team is searching for a creative and experienced Senior Software Engineer in Test to help us bring NVIDIA's autonomous vehicle solution out to the world. You will participate in a focused effort to develop and productize ground-breaking solutions that will redefine the world of transportation and the growing field of self-driving cars. You will work with hardworking and dedicated multi-functional engineering development teams across various vehicle subsystems to integrate their work into our AV SW platform, while achieving or exceeding all meaningful NVIDIA and automotive standards & guidelines. You'll find the work is exciting, fun, and relevant. We have deadlines, customers, and competition. What you will be doing: • Build novel design solutions and drive execution to improve the use of end to end Simulation at scale to evaluate Autonomous Vehicle performance through test architecture and tool development. • Lead efforts to streamline and automate the development of simulation tests from all aspects of creation, run, and analysis infrastructure. • Build reliable and scalable infrastructure and operations that can make it easy to scale and interact with simulation tests. • Analyze complex technical issues and independently drive resolution. • Guide technical aspects of development of test scenario creation and usability in Simulation. • Work closely with developers and operational teams responsible for Simulation test creation to understand their needs and improve scalability and stability of the platform. • Collaborate across many teams, bridging needs of Autonomous Vehicle development and innovative end-to-end Simulation capabilities, applying your skills to understand and improve the modules that simulate the real world driving. • Mentor and guide junior engineers, promoting a culture of technical excellence and collaborative problem solving. What we need to see: • PhD with 4+ years, MS with 6+ years, or BS (or equivalent experience) with 8+ years of relevant experience in Computer Science, Computer Engineering, or a related technical field. • Experience building testing, infrastructure and test automation. • Passion for developing software that makes others more productive. • Experience working on Advanced Driver Assistance Systems (ADAS), Autonomous Driving, Replay testing, or Simulation environments. • Substantial experience with Python/C++ in a software driven environment. • Strong leadership and interpersonal skills, with the ability to drive alignment across large organizations. Ways to stand out from the crowd: • Experience with scaling the use of autonomous vehicles simulation frameworks. • Background with Bazel, Docker, Jenkins. • Experience with or knowledge of AI/ML systems, or a background in working with AI-native systems • Experience in large-scale simulation for hardware or software validation of any robotics solution. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until May 30, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.
Design and implement profiling tools features for cloud and cluster environments, collaborating across teams and product lifecycle. | Senior software engineer with 5+ years experience, fluent in C++ and Python, with Kubernetes and distributed systems knowledge. | As part of the growing Developer Tools team, you will work on tools like Nsight Operator, Nsight Cloud, Nsight Systems, and more. Our work is based on years of expertise collecting detailed trace and sampling data, which we now apply to the ever growing in complexity cloud and cluster profiling use cases. The position includes elements of systems design, frontend and backend development, and working across the full product feature lifecycle: from prototypes to user demos. What you will be doing: • Join the Developer Tools team as a senior software engineer to work on profiling tools within the growing Nsight family. • Design and implement product features that would help make it possible and easy to collect, analyze, and visualize performance profiling data in cluster and cloud environments. • Communicate across multiple teams to collect and understand the requirements, user needs, and expectations. Understand how the underlying hardware and software works, and use that knowledge to deliver valuable features to the users. • Collaborate with team members across multiple time zones in a dynamic, high-energy work environment. • Interact with internal and external users, help them get the maximum value out of our products, and deliver their feedback to the product team. What we need to see: • Excellent problem solving, collaborative, and interpersonal skills. Experience working in distributed teams is welcome. • Ability to work across the full stack, including the server side code in Python and frontend code in JavaScript. • Fluency in C++ and Python. • Track record of working with Kubernetes and in distributed environments. • Strong understanding of algorithms and computer architecture. • BS or MS in EE, CE, CS, Systems Engineering (or equivalent experience) • 5 years of experience in a related software position. Ways to stand out from the crowd: • Experience with GPUs, CUDA, HPC, clusters, networking, and performance optimization in cloud environments. • Participation in deployment and maintenance of microservices. • Experience building web APIs, familiarity with GraphQL. • Proficiency with Datadog, ClickHouse, and Grafana for observability, analytics, and dashboarding. • Experience working in programming languages like Go and Rust. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until May 24, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.
Analyze large-scale workloads and infrastructure signals to find improvement opportunities and build visualizations and ML tools. | 5+ years analyzing complex datasets, strong Python and JavaScript skills, experience with telemetry stacks, and collaboration skills. | NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. Join a team that analyzes large-scale datacenter workloads on GPU-accelerated clusters. You will turn telemetry and workload data into clear findings and visuals. You will partner with OS, container, GPU, and systems engineers. When useful, you will apply machine learning and deep learning techniques for categorization and forecasting. These will be coordinated into tools the team actually uses. What you’ll be doing: * Analyze large-scale workloads and infrastructure signals to find application and platform improvement opportunities. * Work with high-dimensional data: spot trends, tie changes to known events, summarize conclusions, and communicate results to engineers and leadership. * Partner with the team to clarify questions, scope analyses, and document methods so others can extend your work. * Build and maintain practical visualizations and lightweight implementations (e.g. ML/DL for classification/prediction) inside existing software workflows. What we need to see: * 5+ years analyzing complex datasets, debugging data issues, and communicating trends clearly. * BS or MS in Engineering, Mathematics, Physics, Computer Science, or equivalent experience. * Strong Python and JavaScript; * Comfortable being responsible for an analysis end-to-end. * Hands-on use of telemetry / observability stacks (e.g. Grafana, Elasticsearch, Splunk). * Shown grasp of core ML concepts; quick learner; strong analytical and problem-solving skills. * Collaboration and communication. Ways to stand out from the crowd: * TensorFlow or PyTorch * Linux and HPC / large-scale or performance-sensitive environments * Experience visualizing high-dimensional problems * Diligent, action-biased analysis style NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative and autonomous, we want to hear from you! NVIDIA also offers a comprehensive benefits package. We provide health care coverage, dental and vision, 401(K), including company matching and after tax contributions, Employee Stock Purchase Program (ESPP), Employee Assistance Program (EAP), company paid holidays, paid sick leave, vacation leave, professional time off, life and disability protection. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until May 22, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes. NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Solve customer issues and optimize performance for NCCL and CUDA libraries in datacenter GPU workloads. | Expertise in CUDA, NCCL, C/C++ programming, parallel programming models, performance profiling, and datacenter system architecture with 8+ years experience. | NVIDIA is seeking a Senior Software Engineer, NCCL and CUDA specialization to join our Cloud Service Provider (CSP)Engagements team, focusing on ML software stack functionality and performance for datacenter products such as GB300 and Vera Rubin. This role involves working with the customers to understand the functional and performance issues in the libraries layer for deployment at scale. The role combines deep technical expertise in workloads, NCCL and CUDA libraries, frameworks, and system software interaction to solve the customer issues and bring the next generation of innovation. What you will be doing: * Engage with our CSPs to root cause functional and performance issues in NCCL and CUDA libraries. * Analyze and improve multi-GPU workloads performance through profiling, benchmarking, and tuning. * Understand and solve NCCL and NVSHMEM data movement issues in multi-node clusters. * Understand and solve CUDA porting issues for customer workloads. * Apply datacenter-specific scheduling and topologies for optimal performance * Debug and resolve complex issues related to GPU computation, memory, and transports. * Collaborate with customers to understand their workload integration specific challenges to NCCL and CUDIA libraries and suggest tailored solutions aligned with the NVIDIA ecosystem. * Collaborate with AE, FAE, and solution architects to deliver integrated customer solutions and technical documentation. * Collaborate with internal teams to help customers use the latest advancements in CUDA and in NCCL. What we need to see: * Experience with parallel programming models and with communication libraries (MPI, NCCL, NVSHMEM) run time. * Experience with performance optimization and profiling tools (e.g., Nsight, nvprof) * Excellent C/C++ programming and debugging skills, with experience in CUDA development. * Good exposure to PCIe and NVLINK. * Deep understanding of operating systems and data-center system architecture. * Knowledge of high-performance networking like InfiniBand, and RoCE. * Proficient understanding of compute, networking and cloud deployment, specifically on bare-metal and VMs. * BS or MS in Computer Engineering, Computer Science, or related field (or equivalent experience). * Familiarity with containers, cloud provisioning and scheduling tools such as Docker, Kubernetes, SLURM, and Ansible. * 8+ years of system software validation experience. * Ability to communicate effectively and collaborate with partner and customer teams. Ways to stand out from the crowd: * Strong software architecture experience. * Experience with deep learning workloads training and inferencing * Experience conducting performance benchmarking and developing tooling on HPC clusters NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High-Performance Computing and Visualization. The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our products and services. We have some of the most forward-thinking and hard-working people on the planet working for us. If you're creative, hard-working and self-motivated, we want to hear from you! Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until May 18, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes. NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Architect and support large-scale NVIDIA data center AI infrastructure deployments and act as a technical bridge between product and customers. | Requires 4+ years in system design or technical support in high-tech electronics, strong problem-solving, communication skills, and preferably direct data center and NVIDIA hardware experience. | Do you flourish with taking a strategic product from launch to go to market at scale across the world’s largest customers? NVIDIA is seeking a hands-on, action-oriented Senior Solutions Architect to join our team, focused on the technical execution across engineering, product, and sales teams and support to achieve Hyperscaler-scale deployment for our groundbreaking data center products. This role requires a strong passion for system design and a successful history of engaging with customers on a deep technical level, preferably with direct datacenter knowledge and experience debugging datacenter platforms. You will serve as a critical bridge between product strategy and large-scale customer deployment. What you’ll be doing: * Help architect and scale high-performance, distributed AI infrastructure on-prem or in the cloud built with the latest NVIDIA GPU supercomputers for new and existing customers. * Act as a technical specialist on GPU and networking products, directly supporting sales account managers to secure design wins. * Provide technical and onsite support to solve complex hardware and software problems, with a focus on deep learning inference. * Collect, maintain, and analyze complex deployment data and logs to assess product health, identify technical challenges, and guide the customer to resolution of roadblocks. * Actively establish and cultivate technical relationships with engineers and architects at key customer accounts. * Identify customer architectures and key product requirements in the CSP/OEM AI market to efficiently implement NVIDIA's solutions. * Develop technical solutions including hardware & software demos and example system designs. * Offer technical and sales training to direct sales teams and channel partners. What we need to see: * BS/MS/PhD in Electrical/Computer Engineering, Computer Science, Physics, or other Engineering fields or equivalent experience. * 4+ years of work-related experience in the high-tech electronics industry, particularly in system design or technical customer support roles. * Technical competence to roll up your sleeves, look at logs and deployment data, and guide a customer to the resolution of technical issues * Experienced in analytical and problem-solving abilities. * Capable of excelling in a dynamic, constantly evolving environment. * Excellent written and oral communication skills in English, with the ability to collaborate effectively with management and engineering teams. Ways to stand out from the crowd: * Experience working in a Hyperscaler or Cloud Service Provider (CSP) environment or extensive experience engaging with them as a strategic partner. * Hands-on experience with NVIDIA hardware (e.g., ARM CPU, H100, GB200) and software libraries, with an understanding of performance tuning and error diagnostics. * Established track record of driving a product from the pilot phase to high-volume, at-scale deployment in a data center environment. * Practical knowledge of NVIDIA systems technology such as NCCL, DCGM, and UFM. * Knowledge of Embedded Linux Systems, APIs, and similar embedded OS. NVIDIA is widely considered to be one of the technology world's most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative and autonomous, we want to hear from you! Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until May 15, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes. NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Maintain and troubleshoot datacenter cluster infrastructure and manage software and firmware updates. | 5+ years experience with cluster deployment, datacenter hardware management, networking, and scripting for automation. | NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High-Performance Computing and Visualization. The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our products and services. Our work opens up new universes to explore, enables amazing creativity and discovery, and powers what were once science fiction inventions from artificial intelligence to autonomous cars. NVIDIA is looking for phenomenal people like you to help us accelerate the next wave of artificial intelligence. Join our team of innovative engineers who develop and maintain software facilitating GPU communication, driving groundbreaking solutions in High Performance Computing and Deep Learning. We are seeking a highly motivated EngOps Engineer (5+ years of experience) to join our advanced infrastructure software team. In this role, you will be responsible for maintaining high-performance, rack-scale management solutions for datacenter environments. You will work directly with our Infrastructure Service software development team to support deployment and debug of our hardware and Infrastructure Manager. What you’ll be doing: * Take ownership of daily cluster failures and issues, troubleshooting them promptly to maintain optimal cluster availability and performance. * Manage updates to the site controller management nodes. * Manage the rollout and rollback of cluster software and firmware updates, ensuring smooth transitions and minimal disruptions. What we need to see: * BS or MS in Computer Science, Computer Engineering, Electrical Engineering, or a related field, or equivalent experience. * 5+ years of hands-on experience in deploying and administrating clusters, servers, switches, and related infrastructure. * Experience with deployment and configuration of operating systems, computer networks, and high-performance applications. * Proven ability to work effectively with developers and test engineers across different teams and time zones. * Experience deploying services in Kubernetes. * Datacenter or computer architecture experience is required—you should understand server, rack, and network topologies, as well as hardware/firmware/software interactions. * Background with hardware management protocols (Redfish, IPMI, BMC) and firmware update automation. * Experience configuring and debugging complex data center networks. * Experience developing scripts to automate recovery actions for management controllers and datacenter systems. Ways to stand out from the crowd: * Direct experience with industry standard alerting tools and emergency response practices. Experience with observability tools such as Grafana. * Hands-on experience with GPU-focused hardware and software, such as DGX systems and Compute Clusters. * Proficiency in designing large scale networking technologies and the associated challenges. Experience with OpenStack and Foreman NVIDIA is often recognized as one of the technology industry's most esteemed employers. We have some of the brightest and most driven individuals in the world working with us. If you are a self-motivated and imaginative individual, we encourage you to apply! Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until May 1, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes. NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Lead performance engineering and optimization of large-scale GPU infrastructure for AI workloads. | Requires deep expertise in GPU computing, CUDA, HPC systems, Linux, and strong programming in C/C++/Python with 8+ years experience. | NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. We are looking for a dedicated engineer for the Senior Systems Software Engineer role, focusing on GPU Performance at Scale. At NVIDIA, this role is uniquely positioned to drive innovation in AI and GPU computing. You will contribute to world-class computing hardware and software, fueling groundbreaking advancements in artificial intelligence. You will provide insights on large-scale system composition and tuning mechanisms for high-performance compute runs. Collaborate with researchers, developers, and customers to craft improved workflows and develop new, leading solutions. Engage with HPC, OS, CPU, GPU compute, and systems specialists to architect, build, and optimize large-scale performance platforms. What you'll be doing: * Lead the implementation of performance practices in large-scale GPU infrastructure, delivering powerful tools, methodologies, and flows to validate and improve multiple datacenter products concurrently. * Align next-generation AI workloads with next-generation datacenter builds for NVIDIA GPUs, CPUs, and networking hardware. Engage early with HW/FW/SW/platform internal and customer teams. * Develop engineering solutions that provide continuous insights into the performance of AI workloads in evolving environments, generating swift insights into improvements and regressions. * Decompose high-complexity performance or stability issues into minimal reproduction cases, working towards identifying the root cause. * Participate in collaborations with various SW and FW teams (BMC/SBIOS/OS/drivers, etc.) to develop outstanding methods and tools. Analyze, debug, and resolve critical firmware and software issues to achieve the highest AI workload performance at scale. What we need to see: * Proven understanding of accelerated computing software stacks (CUDA). * Experience with modern cloud and container-based enterprise computing architectures, with Slurm preferred. * Strong programming and scripting experience in C/C++/Python/Bash. * Deep expertise in systems architecture and the impact of various components on performance. * Experience with container technology and Linux-based OSes, with Docker preferred. * Experience supporting high-performance computing or deep learning in engineering or academic research communities. * Strong teamwork and communication skills, coupled with results-focused analytical abilities. * BS in Engineering, Mathematics, Physics, or Computer Science (or equivalent experience); MS or PhD desirable with 8+ years of applicable experience. Ways to Stand Out From the Crowd * End-to-end GPU performance engineering from the profiler to systems analysis. * Linux systems programming and optimization experience. * Exposure to virtualization techniques and cloud platform solutions. * Experience with scheduling and resource management systems. * Experience with large-scale HPC environments. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until April 30, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes. NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Engage with partners to troubleshoot NCCL issues, conduct performance analysis, develop tools, and provide training for HPC applications. | Requires 5+ years experience in HPC or AI with strong C/C++ skills, parallel programming, networking knowledge, and Linux expertise. | NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High Performance Computing and Visualization. The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our products and services. Our work opens up new universes to explore, enables amazing creativity and discovery, and powers what were once science fiction inventions from artificial intelligence to autonomous cars. We are the GPU Communications Libraries and Networking team at NVIDIA. We deliver communication runtimes like NCCL and NVSHMEM for Deep Learning and HPC applications. We are looking for a motivated Partner Enablement Engineer to guide our key partners and customers with NCCL. Most DL/HPC applications run on large clusters with high-speed networking (Infiniband, RoCE, Ethernet). This is an outstanding opportunity to get an end to end understanding of the AI networking stack. Are you ready for to contribute to the development of innovative technologies and help realize NVIDIA's vision? What you will be doing: * Engage with our partners and customers to root cause functional and performance issues reported with NCCL * Conduct performance characterization and analysis of NCCL and DL applications on groundbreaking GPU clusters * Develop tools and automation to isolate issues on new systems and platforms, including cloud platforms (Azure, AWS, GCP, etc.) * Guide our customers and support teams on HPC knowledge and standard methodologies for running applications on multi-node clusters * Document and conduct trainings/webinars for NCCL * Engage with internal teams in different time zones on networking, GPUs, storage, infrastructure and support. What we need to see: * B.S./M.S. degree in CS/CE or equivalent experience with 5+ years of relevant experience. Experience with parallel programming and at least one communication runtime (MPI, NCCL, UCX, NVSHMEM) * Excellent C/C++ programming skills, including debugging, profiling, code optimization, performance analysis, and test design * Experience working with engineering or academic research community supporting HPC or AI * Practical experience with high performance networking: Infiniband/RoCE/Ethernet networks, RDMA, topologies, congestion control * Expert in Linux fundamentals and a scripting language, preferably Python * Familiar with containers, cloud provisioning and scheduling tools (Docker, Docker Swarm, Kubernetes, SLURM, Ansible) * Adaptability and passion to learn new areas and tools * Flexibility to work and communicate effectively across different teams and timezones Ways to stand out from the crowd: * Experience conducting performance benchmarking and developing infrastructure on HPC clusters. Prior system administration experience, esp for large clusters. Experience debugging network configuration issues in large scale deployments * Familiarity with CUDA programming and/or GPUs. Good understanding of Machine Learning concepts and experience with Deep Learning Frameworks such PyTorch, TensorFlow * Deep understanding of technology and passionate about what you do Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until April 28, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes. NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Manage cross-functional collaboration to launch products and accelerate partner readiness in channel NPI programs. | Over 12 years in channel partner programs or NPI in high-tech global companies with strong project management and communication skills. | NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we're tapping into the unlimited potential of AI to define the next era of computing – an era in which our GPU acts as the brains of computers, agents, robots, and self-driving cars that can understand the world. Doing what's never been done before takes vision, innovation, and the world's best talent. As an NVIDIAN, you'll be immersed in a diverse, encouraging environment where everyone is inspired to do their best work. As a key member of the NVIDIA Partner Network (NPN) Organization, we are looking for a Senior Program Manager in Santa Clara, California who will drive NPI channel initiatives through cross-functional collaboration across the organization to successfully launch products in the channel and accelerate partner readiness! What you'll be doing: * Liaise between the extended product launch teams as the primary NPN team contact. * Develop and manage the master channel NPI launch plan, tracking all breakthroughs, deliverables, and dependencies for channel readiness. * Facilitate regular cross-functional NPI status meetings, driving clarity on ownership, decisions, and next steps. * Identify and mitigate risks to the channel launch schedule, quality, and partner experience. * Validate operational readiness, including pricing, quoting, ordering (SKUs), and fulfillment processes are in place with fulfillment partners, including distributors. * Review and contribute to NPI Plans of Record (POR) and Business Requirements Documents (BRDs). * Project manage and help build channel readiness materials including Frequently Asked Questions (FAQs) and other partner reference documents. * Monitor post-launch performance, gather feedback from Channel Partners and Internal Stakeholders, and lead continuous improvement efforts for future NPIs. What we need to see: * 12+ years of established capability in channel partner programs, operations, or NPI in a high technology global enterprise company. * Bachelor’s degree in business, marketing, or related field (or equivalent experience). * Proven success in leading or actively participating in channel go-to-market product launch strategy development and implementation. * Demonstrated expertise with complex channel route-to-market strategies and possess strong knowledge of channel rebates, pricing, and discount structures. * Extensive knowledge of channel partner ecosystems, business models, and market dynamics. * Have a track record of leading and closing multiple priority projects. * Excellent communication and interpersonal skills, proficient in establishing relationships across all levels. * Hands-on experience using AI-enabled tools (i.e. for planning, reporting, content creation, analysis, or automation) and an eagerness to experiment with new technologies. * Solid experience using project tracking tools. Ways to stand out from the crowd: * Strong ownership approach: you anticipate issues, drive decisions, and follow through without needing constant direction. * A passion for winning, a solid aptitude for business strategy, and strong collaboration skills. * Comfort to work in a fast-paced environment, leading multiple high-priority initiatives simultaneously, where strong project management and organizational skills are crucial. * A strategic problem-solver, and the ability to drive solutions that improve partner performance and satisfaction. * Passion for learning: you actively seek out new tools, methods, and guidelines, and share them with the broader team. NVIDIA is widely considered to be one of the world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If this opportunity sounds like you, we want to hear from you! Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 176,000 USD - 276,000 USD. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until April 27, 2026. This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes. NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Build scalable CI/CD frameworks, visualization tools, and AI agents to improve research productivity and GPU utilization for robotics research. | 8+ years in large-scale MLOps and AI infrastructure, strong full-stack development skills, knowledge of GPU technologies, and experience with modern web technologies. | We are seeking a Senior Software Engineer to join a new team building the foundational infrastructure for Robotics Research. This new team will work very closely with NVIDIA's Generalist Embodied Agent Research (GEAR) group. The near term focus is Project GR00T, NVIDIA's moonshot initiative at building foundation models and full-stack technology for humanoid robots. This particular position will focus on ML productivity tooling - we are looking for people who have high agency and love creating tools to make researchers more productive and GPUs more efficiently used. You will work with an amazing and collaborative research team that consistently produces influential works on multimodal foundation models, large-scale robot learning, embodied AI, and physics simulation. Your contributions will have a significant impact on our research projects and product roadmaps. What you'll be doing: • Build highly scalable, robust, and efficient CI/CD frameworks. The workload is data intensive and requires CPU/GPU heterogeneous computation. • Build world-class visualization tools for analyzing and optimizing for all our datasets and compute jobs (across 10s of thousands of GPUs) • Develop and apply AI agents to significantly improve programming efficiency within the team, and decrease the human effort in fixing job failures • Overall, collaborate with researchers to gather requirements, understand tooling / visualization / automation needs, and deliver full-stack solutions that move the needle with speed of light. What we need to see: • Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent experience • 8+ years of full-time industry experience in large-scale MLOps and AI infrastructure. • Strong experience in full-stack software development, with a focus on building CI/CD or visualization tools. • Proficient in both front-end and back-end programming with Python, JavaScript, SQL, or similar. • Familiar with modern web front/back end technologies like React, Node.js. • Knowledge of GPU technologies like CUDA and NCCL • Bonus: experience with PyTorch, Ray, Kubernetes Ways to stand out from the crowd: • Master's or PhD's degree in Computer Science, Robotics, Engineering, or a related field • Demonstrated Tech Lead experience, coordinating a team of engineers and driving projects from conception to deployment • Strong experience at building and operating large-scale tooling infrastructure in production • Strong background and curiosity in frontier AI research NVIDIA is widely considered to be one of the technology world's most desirable employers. We have some of the most forward-thinking and productive people in the world. Please join us and be part of the forefront of developing general-purpose robots and large-scale foundation models! Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until October 5, 2025. NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Lead a team to design and develop scalable cloud services and APIs for GPU telemetry ingestion and operational automation across global cloud operations. | 12+ years experience, expertise in scalable REST APIs with PostgreSQL, proficiency in Go/Java/Python, modern JS frameworks, cloud and container tech, distributed systems, and strong communication skills. | Today, we're tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what's never been done before takes vision, innovation, and the world's best talent. As an NVIDIAN, you'll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. NVIDIA is widely recognized as one of the most desirable employers, with some of the most talented people in the world working for us. If you're passionate about building scalable, efficient systems to power cloud operations, we invite you to join our team. We are looking for a Lead Software Engineer to join our DGX Cloud team and build the foundational systems that drive NVIDIA's high-performance GPU infrastructure. You will play a technical lead role in designing scalable cloud services that integrate with diverse systems including GPU telemetry in datacenters, and enabling operational automation across global cloud operations. What You'll Be Doing: • Act as technical lead for a team of software engineers designing cloud services backed by databases and data warehouses. • Design and develop RESTful APIs to ingest telemetry from AI datacenters. • Build scalable cloud services for high-volume ingestion, processing, and storage of large datasets. • Build and manage data pipelines for online and offline data storage. • Collaborate across teams to codify business processes into scalable, self-measuring systems. • Optimize the reliability and efficiency of cloud services and operations. • Lead and ship impactful technical projects, ensuring quality and scalability at every stage. What We Need To See: • At least 12+ years of industry experience with a Bachelor's or Master's degree (or equivalent experience); PhD degree preferred. • Expertise in building scalable REST APIs backed by PostgreSQL-compatible data stores. • Proficiency in programming languages such as Go, Java, or Python. • Familiarity with modern JavaScript frameworks (e.g., React, Angular, Next.js). • Expertise in cloud infrastructure (AWS, GCP, Azure, etc) and container technologies like Docker and Kubernetes. • Expertise with high-scale distributed systems, including architectural patterns for APIs and data pipelines. • Outstanding communication and collaboration skills, with a focus on solving complex operational challenges. • A passion for delivering scalable and efficient cloud services. • Familiarity with Linux operating systems. Ways to Stand Out from the Crowd: • A track record of leading engineers to successful delivery and operations of high-performance cloud services at Internet scale. • Experience operating NVIDIA datacenter GPUs. • Strong debugging and problem-solving skills in distributed environments. NVIDIA is committed to creating an environment where diverse perspectives drive innovation. As part of the DGX Cloud team, you'll work on cutting-edge technology that powers the future of AI and cloud computing. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 224,000 USD - 356,500 USD for Level 5, and 272,000 USD - 425,500 USD for Level 6. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until October 5, 2025. NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Collaborate with global research teams to deploy and optimize generative AI software for scientific discovery and provide feedback to engineering teams. | 5+ years experience with AI/ML models using Python or C++, deep knowledge of GenAI architectures, and business-level English plus Japanese communication skills. | Generative AI is revolutionizing ranging from supercomputing, higher education, manufacturing, semiconductors, energy storage, climate science and agriculture. The advent of GenAI models and tools such as ClimaX, GenSLM, MatterGen, etc are some exemplars of adapting GenAI for scientific research and discoveries. We are now looking for a senior Application engineer to work with leading science and research institutes in Japan, promoting the adoption GPU-accelerated computing solutions-including machine learning, deep learning, and especially generative AI for scientific discovery and research. Do you enjoy working with top researchers who are coming up with innovative model architectures and workflows for GenAI for Science ? Do you find it exciting to help collaborators solve problems in training, scaling and deploying GenAI models at scale? If converting a small proof of concept into an inspiring solution for the larger community is exciting, then we encourage you to apply! What you'll be doing: • Work alongside research teams worldwide adopting NV's GenAI SW for scientific data/models • Debug, profile and recommend optimizations • Provide customer feedback back into engineering teams on improving of NV's GenAI SW for science research and discovery. • Participate in Gen AI hands-on hackathons/workshops that drive adoption of NV's GenAI SW stack in the scientific community What we need to see: • Excellent verbal, written communication, and technical presentation skills in Japanese. Business level English communication is also a requirement. • Degree or equivalent experience in Computer Science, Applied Mathematics, or related engineering field (Ph.D. or Masters preferred) • 5+ years hands-on experience developing, optimizing and running AI/ML models or LLMs using popular frameworks (PyTorch, or similar), and languages (C++, Python) • Deep understanding of Gen AI model architectures (GPT-x, Llama, MoE, etc) • Ability to work independently and as part of a globally distributed team. Ways to stand out from the crowd: • Expertise in deploying large-scale training and inferencing pipeline • Hands-on experience with customizing AI models for multi-modal scientific data • Contributions to cornerstone research papers or open-source projects involving GenAI/AI for science applications • Experience with complex workflows integrating simulation, AI models/GenAI models and running them on large scale HPC systems With competitive salaries and a generous benefits package, we are widely considered to be one of the technology world's most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us and, due to unprecedented growth, our exclusive engineering teams are rapidly growing. If you're a creative and autonomous engineer with a real passion for technology, we want to hear from you. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 124,000 USD - 195,500 USD for Level 3, and 144,000 USD - 230,000 USD for Level 4. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until September 20, 2025. NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Develop methodologies and tooling for testing sufficiency, analyze naturalistic driving data, automate workflows, and create dashboards to track testing coverage. | 8+ years experience, strong Python and SQL skills, data/statistical analysis experience, excellent problem solving and communication skills. | NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It's an outstanding legacy of innovation that's powered by great technology-and outstanding people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what's never been done before takes vision, innovation, and the world's best talent. As an NVIDIAN, you'll be immersed in a diverse, encouraging environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. We are thrilled to present an outstanding opportunity to join NVIDIA as a Senior Systems Software Engineer. This role is perfect for an ambitious individual who is passionate about driving world-class safety and security solutions within our engineering teams through the use data analytics. You will have the chance to work on transformative projects in a collaborative and encouraging environment, helping to craft the future of computing. What you'll be doing: • Developing methodologies and tooling to demonstrate the sufficiency of on-road and offline testing. • Performing analyses measuring exposure rates of objects, events and environments in naturalistic driving data. • Working jointly with test engineers, developers, and other team members to identify and close coverage gaps. • Automating workflows and reports to support a growing number of product releases and deployments. • Creating dashboards and visualization tools to track testing coverage and identify trends. • Develop and share standard methodologies in analytics and data driven validation with colleagues. What we need to see: • BS/MS or equivalent experience in engineering, science, or a related technical field • 8+ years of experience • Strong background in python and SQL or similar languages while working with large datasets • Previous experience in data or statistical analysis • Excellent problem solving skills and attention to detail • Strong communication skills with the ability to convey data meaningfully to drive decisions Ways to stand out from the crowd: • Have a track record of building scalable data pipelines and analytics platforms. • A background of 2+ years in autonomous vehicle validation or development. Join us at NVIDIA and make a lasting impact on the world! Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until September 20, 2025. NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law. #deeplearning
Research and develop GPU-accelerated high performance database and ETL applications, optimize data intensive workloads, and influence next-generation hardware and software design. | Masters or PhD in Computer Science or related field, 5+ years relevant experience, strong C/C++ and parallel programming skills, deep understanding of GPU/CPU architectures, and domain expertise in high performance databases and ETL. | NVIDIA is currently seeking a Senior Developer Technology Engineer for High-Performance Databases! Would you enjoy researching new algorithms and memory management techniques to accelerate databases on modern computer architectures? Do you likeinvestigating hardware and system bottlenecks, and optimizing performance of data intensive applications? Are you excited about the opportunity to work on the leading edge of technology with both visibility and impact to the success of a leader like NVIDIA? If so, the Developer Technology Team invites you to consider this opportunity. What you will be doing: • In this role, you will research and develop techniques to GPU-accelerate high performance database and ETL applications. • Work directly with other technical experts in their fields (industry and academia) to perform in-depth analysis and optimization of complex data intensive workloads to ensure the best possible performance of current GPU architectures. • Influence the design of next-generation hardware architectures, software, and programming models in collaboration with research, hardware, system software, libraries, and tools teams at NVIDIA What we need to see: • A Masters or PhD in Computer Science, Computer Engineering, or related computationally focused science degree (or equivalent experience). • At least 5+ years of relevant work or research experience. • Programming fluency in C/C++ with a deep understanding of algorithms and software design. • Hands-on experience with low-level parallel programming, e.g., CUDA, OpenACC, OpenMP, MPI, pthreads, TBB, etc. • In-depth expertise with CPU/GPU architecture fundamentals, especially memory subsystem. • Domain expertise in high performance databases, ETL and data analytics • Good communication and organization skills, with a logical approach to problem solving, and prioritization skills. Ways to stand out from the crowd: • Experience optimizing the performance of distributed database systems and frameworks (e.g. production database or Spark). • Background with compression, storage systems, networking, and distributed computer architectures. Data Analytics is one of the rapidly growing fields in GPU accelerated computing. Data preprocessing and data engineering are traditionally CPU based and are becoming the bottleneck for Machine Learning (ML) and Deep Learning (DL) applications, as performance of the frameworks and core ML/DL libraries has been highly optimized leveraging GPUs. Many of today's applications have complex data analytics pipelines that can benefit from optimizations in memory management, compression, parallel algorithms like sort, search, join, aggregation, groupby, scaling up to multi GPU systems, and scaling out to many nodes. Take a look at some of the open-source projects that our Devtech team have worked on: NVIDIA nvcomp, NVIDIA Distributed join, NVIDIA cuCollections. #LI-Hybrid Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 148,000 USD - 235,750 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until September 6, 2025. NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Lead and govern integrated mechanical, electrical, and automation system designs for AI-enabled data centers, collaborating across multiple engineering and construction teams to ensure innovation and execution. | Master's or Bachelor's in CS, EE, or related, 15+ years as senior technical contributor in mechanical/electrical/automation systems, expertise in AI infrastructure, and strong multi-disciplinary collaboration skills. | We are looking for a Data Center Distinguished Engineer. NVIDIA's data centers host ground-breaking products across high-performance computing to machine learning applications for autonomous vehicles and healthcare. At the heart of our data centers is the ability to engineer integrated system designs in close coupling to NVIDIA's industry-leading GPU products. We are seeking a Data Center Distinguished Engineer to accelerate next-generation solutions. The Distinguished Engineer will be responsible for shaping and governing the multi-horizon integration of mechanical, electrical, and automation systems within AI-enabled data centers. This role orchestrates alignment of people, processes, and technology across teams and partners - including Architecture, Engineering, and Construction and MEP organizations - to enable innovation, ensure readiness, and drive flawless execution at scale. What you'll be doing: • Provide technical leadership in coordination with product requirements and configurations of mechanical and electrical system designs. • Be responsible for the strategic development and oversight of harmonized mechanical, electrical, and automation systems for AI-enabled data centers. • Collaborate with AEC and MEP partners to align development priorities, technical standards, and integrated engineering solutions. • Advance the creation of engineering environments to support development, testing, and validation across partner ecosystems. • Provide strategic direction and integration oversight for energy generation initiatives dedicated to scaling data centers, ensuring alignment with infrastructure and transmission requirements. • Collaborate with Technical Program Management (TPM) programs to ensure initiatives are completed on time, within scope, and aligned across multiple engineering subject areas. • Provide deep technical guidance to partner engineering teams to ensure compatibility, scalability, and future-readiness. • Facilitate teamwork across R&D, peer engineering groups, and partner organizations to accelerate innovation and integration. • Participate in site selection reviews to ensure infrastructure readiness, technical feasibility, and alignment with strategic goals. • Anticipate and incorporate new technologies, market shifts, and partner capabilities into long-term strategy. What we need to see: • Master's or Bachelor's degree in Computer Science or Electrical Engineering or CE or equivalent experience. • 15+ years of experience as a senior technical individual contributor (Distinguished Engineer, Principal Engineer, or equivalent) in mechanical, electrical, or automation systems. • Minimum of years of experience in partner-driven engineering development, particularly with AEC, MEP, and energy infrastructure stakeholders. • Expertise in AI applications for infrastructure and data center ecosystems. • Outstanding ability to align multi-disciplinary systems across diverse collaborator groups. • Excellent communication, facilitation, and critical thinking skills. With competitive salaries and a generous benefits package, we are widely considered to be one of the technology world's most desirable employers, and we have some of the most forward-thinking and hardworking people in the world working for us. Due to outstanding growth, our best-in-class teams are rapidly growing so if you're creative and autonomous with a real passion for technology, we want to hear from you. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 308,000 USD - 471,500 USD. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until August 29, 2025. NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Create tailored applications specifically for Nvidia Corporation with our AI-powered resume builder
Get Started for Free