Latest Data Science Jobs

Staff / Senior Software Engineer (Agentic Search) - Index

Design and operate large-scale indexing systems and data pipelines for an AI-native search platform, processing petabyte-scale datasets at high throughput. Focus on building efficient, reliable data structures and storage strategies that support low-latency, observable search APIs for AI agents. Collaborate with ML and runtime teams to ensure data freshness, correctness, and performance at scale.

Nebius United Kingdom

Technical Program Manager - Compute Systems Engineering

This role involves managing complex technical programs across hardware and software engineering teams to scale and improve Nebius's high-performance compute infrastructure. The Technical Program Manager will coordinate system-level initiatives for GPU and CPU platforms, ensuring reliability, performance, and scalability across global data centers. Key responsibilities include driving cross-functional execution, mitigating risks, and enhancing development and operational processes in a fast-moving AI cloud environment.

Nebius United Kingdom

Senior Engineer, Storage Services

This role involves designing, building, and operating storage services across block, file, parallel, and object storage technologies within a high-performance GPU cloud platform. The engineer will ensure storage systems are reliable, scalable, and tightly integrated with compute and networking infrastructure, while driving automation, observability, and incident resolution. The position emphasizes deep technical collaboration, operational excellence, and continuous improvement in a fast-moving AI-focused environment.

NScale United Kingdom
Remote Permanent

Senior Software Engineer (Capacity and Quota Management)

Design and develop core systems for GPU capacity reservation and quota management in a distributed cloud environment. Build scalable APIs and algorithms that allocate and manage compute resources across Nebius services. Work on reliability, consistency, and efficiency in high-stakes infrastructure impacting AI workloads globally.

Nebius United Kingdom

Senior Backend Software Engineer (Cloud Monetization Platform)

Design and build scalable backend systems for a cloud monetization platform handling billing, payments, and enterprise contracts. Work on distributed workflows with strong consistency requirements using Python and microservices. Help shape a core revenue-critical system in a high-growth AI cloud environment.

Nebius United Kingdom

Senior Backend Engineer

Develop and maintain scalable cloud infrastructure services including managed databases, Kubernetes, billing, and monitoring systems. Work on complex backend challenges in Golang within a distributed, high-performance environment. Collaborate with a global engineering team focused on solving hard problems in compute, storage, and AI infrastructure.

Nebius United Kingdom

Senior Software Engineer (Serverless)

Design and build core components of a GPU-native serverless platform for AI workloads, including the control plane, scheduler, runtime, and APIs. Solve hard distributed systems problems like cold-start latency, GPU scheduling under contention, and multi-tenancy. Operate the service with SRE practices, lead design and code reviews, and collaborate directly with customers and cross-functional teams to shape the technical roadmap.

Nebius United Kingdom

Senior Site Reliability Engineer — Token Factory (Inference Platform)

Design and maintain telemetry pipelines for metrics, logs, and traces at scale, while optimizing Kubernetes and Terraform configurations for GPU-intensive inference workloads. Focus on building self-healing, observable systems that ensure high reliability and performance across a distributed AI infrastructure. Collaborate with engineering teams to harden request routing, autoscaling, and incident response mechanisms for large-scale model deployment.

Nebius United Kingdom

Senior ML Engineer (Token Factory)

Design and optimize large-scale machine learning systems for training and deploying foundation models across text, vision, and multimodal domains. Focus on fine-tuning, inference optimization, and low-precision methodologies using JAX and distributed deep learning frameworks. Work closely with hardware and infrastructure teams to solve performance bottlenecks at scale.

Nebius United Kingdom

Senior ML Engineer (Token Factory)

The role involves optimizing large-scale LLM inference and fine-tuning performance across tens of thousands of GPUs, focusing on reducing latency and cost-per-token. You'll work on cutting-edge problems like speculative decoding, low-precision pipelines, and inference engine improvements, with deep involvement in transformer architectures and GPU compute efficiency. The position offers technical leadership opportunities within a globally distributed engineering team pushing the boundaries of AI infrastructure.

Nebius United Kingdom

Senior Software Engineer (Storage Virtualization Team)

Design and implement high-performance, distributed storage systems with sub-millisecond latency and massive throughput. Focus on storage virtualization using virtio-blk, virtio-fs, and QEMU, while ensuring fault tolerance and self-healing capabilities. Work closely on I/O performance, replication, and low-level systems integration across VM and storage layers.

Nebius United Kingdom