Latest Autonomous Systems Jobs

NVIDIA logo

Systems Software Engineer, Kubernetes Scale - DGX Cloud

This role involves driving performance and scalability of the NVIDIA DGX Cloud software stack, focusing on Kubernetes and NVIDIA components like GPU Operator, DCGM, and NIM. The engineer will diagnose complex distributed systems issues, build automated testing and monitoring tools, and contribute to open-source communities to optimize AI infrastructure at scale.

NVIDIA Germany PLN 176,250 – PLN 383,500 pa
Remote Permanent
NVIDIA logo

Systems Software Engineer, Kubernetes Scale - DGX Cloud

This role involves deep performance and scale analysis of the NVIDIA DGX Cloud software stack, focusing on Kubernetes and NVIDIA components like GPU Operator, DCGM, and NIM. The engineer will diagnose distributed systems issues, build automated testing frameworks, and contribute to open-source communities to optimize large-scale AI infrastructure. Work includes continuous performance testing, root cause analysis, and collaboration with AI teams and cloud platforms.

NVIDIA PLN 176,250 – PLN 383,500 pa
Remote Permanent
NVIDIA logo

Systems Software Engineer, Kubernetes Scale - DGX Cloud

This role involves scaling and optimizing AI infrastructure on Kubernetes for NVIDIA's DGX Cloud, focusing on performance, reliability, and cost efficiency at massive scale. The engineer will diagnose complex distributed systems issues, develop automated testing frameworks, and collaborate with AI teams and open-source communities. Work includes deep performance analysis across the full stack—from orchestration down to GPU hardware—using tools like Kubernetes, DCGM, and NIM.

NVIDIA PLN 176,250 – PLN 383,500 pa
Remote Permanent
NVIDIA logo

Systems Software Engineer, Kubernetes Scale - DGX Cloud

This role involves driving performance and scale characterization for NVIDIA's DGX Cloud software stack, focusing on Kubernetes and NVIDIA components like GPU Operator, DCGM, and NIM. The engineer will diagnose complex distributed systems issues, build automated testing and monitoring tools, and contribute to open-source communities to optimize large-scale AI infrastructure. Work includes deep performance analysis from orchestration down to hardware level, with a focus on reducing cost per token in AI workloads.

NVIDIA PLN 176,250 – PLN 383,500 pa
Remote Permanent
NVIDIA logo

Systems Software Engineer, Kubernetes Scale - DGX Cloud

This role involves driving performance and scale characterization of NVIDIA's DGX Cloud software stack, focusing on Kubernetes and NVIDIA components like GPU Operator and DCGM. The engineer will deep-dive into distributed systems issues, build automated testing and monitoring tools, and collaborate with AI teams and open-source communities to optimize large-scale AI infrastructure. Work includes root cause analysis, CI/CD integration, and contributing to upstream projects like Kubernetes and CNCF.

NVIDIA PLN 176,250 – PLN 383,500 pa
Remote Permanent
NVIDIA logo

Systems Software Engineer, Kubernetes Scale - DGX Cloud

This role involves driving performance and scalability of NVIDIA's DGX Cloud software stack, focusing on Kubernetes and NVIDIA components like GPU Operator, DCGM, and NIM. The engineer will diagnose complex distributed systems issues, build automated testing frameworks, and collaborate with AI teams and open-source communities to optimize large-scale AI infrastructure.

NVIDIA PLN 176,250 – PLN 383,500 pa
Remote Permanent
NVIDIA logo

Senior Software Engineer, RL Post-Training Frameworks

This role involves designing and building scalable reinforcement learning (RL) post-training infrastructure that supports the full lifecycle of RL workflows, from experimentation to production at massive scale. The engineer will optimize distributed training-inference-rollout loops across heterogeneous hardware, contribute to open-source RL frameworks, and collaborate with research and hardware teams to advance AI capabilities. Work includes improving system resilience, performance, and integration with next-generation technologies.

NVIDIA Germany
Remote Permanent
NVIDIA logo

Senior Software Engineer, RL Post-Training Frameworks

This role involves designing and building scalable reinforcement learning (RL) post-training infrastructure that supports the full lifecycle of training, inference, and rollout across heterogeneous hardware. You'll contribute to open-source RL frameworks, optimize distributed systems for performance and fault tolerance, and collaborate with AI researchers, infrastructure teams, and hardware engineers to enable next-generation AI capabilities. The work spans deep integration with PyTorch, Kubernetes, and distributed runtimes like Ray and Monarch, focusing on real-world challenges in large-scale RL deployment.

Remote Permanent
NVIDIA logo

Senior Software Engineer, RL Post-Training Frameworks

This role involves designing and building scalable reinforcement learning (RL) post-training infrastructure that supports the full lifecycle of training-inference-rollout loops across heterogeneous hardware. The engineer will contribute to open-source RL frameworks, optimize distributed systems for performance and fault tolerance, and collaborate with research and hardware teams to advance AI capabilities. Work spans deep system-level tuning, cross-team advocacy for RL-specific needs, and integration of CPU, GPU, and LPU-accelerated workloads in large-scale environments.

Remote Permanent
NVIDIA logo

Senior Software QA Test Development Engineer

This role involves designing and implementing automated test frameworks for distributed systems managing thousands of NVIDIA resources. The engineer will develop test plans, automate validation for UIs, APIs, and performance, integrate testing into CI/CD pipelines, and leverage AI/LLMs to enhance testing efficiency. A strong focus is placed on improving QA workflows through AI-driven tools and ensuring robust, scalable software quality across large-scale infrastructure.

NVIDIA Reading, United Kingdom
Hybrid Permanent
NVIDIA logo

Senior HPC Performance Engineer

Analyze HPC applications across distributed and heterogeneous systems to identify performance bottlenecks and optimization opportunities. Focus on GPU acceleration of scientific workloads and provide technical insights to compiler and application engineering teams. Work with large-scale simulations in domains like climatology, fluid dynamics, and defense.

Remote Permanent
NVIDIA logo

Senior HPC Performance Engineer

Analyze HPC applications across diverse architectures to identify performance bottlenecks and optimization opportunities. Support compiler and application engineering teams by providing deep technical insights. Work with large-scale, distributed-memory systems featuring CPUs, GPUs, and many-core processors to accelerate scientific computing workloads.

NVIDIA Germany
Remote Permanent
NVIDIA logo

Senior HPC Performance Engineer

Analyze HPC applications across heterogeneous systems to identify performance bottlenecks and optimization opportunities. Provide technical insights to compiler and application engineering teams. Work with large-scale parallel applications using GPUs, CPUs, and many-core processors, focusing on performance tuning and compiler feedback.

Remote Permanent
NVIDIA logo

Senior CPU Compiler Performance Engineer

We are looking to hire a senior CPU Compiler performance Engineer for an exciting and fun role at NVIDIA. We craft outstanding compilers that realise the potential of NVIDIA's CPUs designed for the world's largest AI and computing markets: https://www.nvidia.com/en-in/data-center/grace-cpu/....

NVIDIA Cambridge, United Kingdom
NVIDIA logo

Senior HPC Performance Engineer

As a Senior HPC Performance Engineer, you will conduct in-depth performance analysis on large multi-GPU and multi-node clusters, evaluate proof-of-concepts, and triage performance issues. You will collaborate with a dynamic team across multiple time zones to advance the state of the art in GPU communication libraries and HPC applications.

NVIDIA £221,250 – £507,000 pa
Remote Permanent