科技職缺
職缺列表
GPU Cluster Architect
Architect end-to-end GPU cluster design for a next-generation AI cloud platform, making decisions across compute, networking, and storage that power LLM training and inference at the scale of tens of thousands of GPUs. Core work spans InfiniBand/RoCE interconnect topology, GPU architecture (NVIDIA/AMD), performance modeling, and automation tooling.
Group Product Manager - Platform Experience
Group Product Manager leading the Platform Experience group for Nebius AI Cloud, managing a team of PMs and owning the unified developer and customer experience across APIs, CLI, Terraform, SDKs, documentation, AI assistance, notifications, quotas, and the resource model.
Incident Response Lead
Lead and build Nebius' global incident response function across cloud, platform, and endpoint environments, managing a follow-the-sun team of responders while personally driving hands-on forensic investigations, containment, and recovery for high-severity incidents.
Instructional Designer, Nebius Academy
Designs and develops courses on Nebius AI Cloud for an international online learning platform, creating lessons, practice exercises, video content, and assessments for B2C and B2B audiences on technical topics like software engineering and data science.
IT Infrastructure Engineer (East London)
Support IT operations across multiple global data centers at this AI cloud infrastructure company, troubleshooting complex firmware and hardware issues on servers, acting as a subject matter expert and point of escalation, and collaborating with R&D to improve hardware designs.
IT infrastructure engineer (RMA & Diag)
IT infrastructure engineer at Nebius troubleshooting complex server hardware/firmware issues in data centers, managing RMA warranty replacements with vendors, and acting as escalation point for L1/L2 technicians. Core stack: server hardware, Linux/Unix, firmware, diagnostics.
IT infrastructure engineer (RMA & Diag)
Troubleshoots complex data center server hardware and firmware issues, develops workarounds, handles vendor RMAs, and supports escalated L1/L2 cases. The role also improves hardware-support processes, documentation, and training using Linux/Unix, monitoring, data analysis, and network troubleshooting skills.
IT Infrastructure Engineer (RMA & Diag)
On-site IT infrastructure engineer in Newport, UK, troubleshooting complex data center server hardware and firmware issues, managing RMA warranty replacements with vendors, and acting as an escalation point for L1/L2 technicians at an AI cloud provider.
Key Customers Solutions Architect
As a trusted technical advisor for key customers of Nebius's AI cloud platform, you'll help clients design, deploy, and scale GPU-powered AI/ML workloads (hundreds to thousands of GPUs), troubleshoot complex issues, and partner with sales and product teams to drive growth.
ML Infrastructure Engineer
ML Infrastructure Engineer at Nebius benchmarking and profiling GPU platforms for deep learning/AI workloads, optimizing training and inference performance, and building internal tooling and dashboards. Core stack spans CUDA, ROCm, NCCL, PyTorch, JAX, Megatron-LM, TensorRT-LLM, Docker, and Kubernetes.
Offensive Security Lead
Lead and build Nebius' internal red team, designing adversarial simulations against their AI cloud platform — covering GPU infrastructure, Kubernetes, inference stacks, and multi-tenant isolation — and translating findings into security improvements via penetration testing and purple-team exercises.
Partner Solutions Architect
Partner Solutions Architect serving as the technical interface between Nebius AI Cloud and strategic technology partners (data platforms, MLOps vendors, AI frameworks, ISVs). Designs integrations, builds reference implementations, and drives joint customer success across GPU-accelerated AI/ML infrastructure.