網站可靠性工程:3,092 個職缺
瀏覽與 網站可靠性工程 相關的開放職缺。
Senior Site Reliability Engineer
At SolarWinds, we’re a people-first company. Our purpose is to enrich the lives of the people we serve—including our employees, customers, shareholders, partners, and communities. Join us in our mission to help…
Senior Site Reliability Engineer (In-Office Required)
About Tavily We're building the infrastructure layer for agentic web interaction at scale. Our API is designed from the ground up to power Retrieval-Augmented Generation (RAG) and real-time reasoning in AI systems.…
Senior Site Reliability Engineer — Token Factory (Inference Platform)
Senior SRE on Nebius Token Factory, owning reliability, performance, and observability for a GPU-accelerated AI inference platform running tens of thousands of GPUs. Core stack: Kubernetes, Prometheus, Grafana, Terraform, Python/Bash, with GPU serving tools (vLLM, Triton, Ray).
Site Reliability Engineer
SRE on the Hardware Infrastructure team ensuring fault-tolerance, scale, and uninterrupted operations for AI cloud services. Day-to-day work involves monitoring data-center and IT equipment, asset tracking, and CI/CD using Linux, Python, and Bash.
Site Reliability Engineer in Hardware Infrastructure
SRE on Nebius's Hardware Automation team building internal platforms and tooling for large-scale data center infrastructure. Day-to-day work centers on ensuring fault-tolerance, scaling services, implementing CI/CD, and troubleshooting hardware, software, and networking issues using Linux, Python, and Bash.
Site Reliability Engineer (SRE) AI Infrastructure (Early Career)
Early-career SRE assisting with day-to-day network infrastructure operations, deploying approved changes, and executing small SRE projects at an AI cloud company. Core stack includes Python/Go/C++, Linux, Kubernetes, Terraform, and low-level networking (eBPF, DPDK, TCP/IP).
Senior Site Reliability Engineer (SRE) | Feeld
GT was founded in 2019 by a former Apple, Nest, and Google executive. GT’s mission is to connect the world’s best talent with product careers offered by high-growth companies in the UK, USA, Canada, Germany, and the…
SRE-инженер
Наша команда занимается эксплуатацией и развитием внутренних инфраструктур, систем и продуктов, включающей в себя набор инструментов по управлению геораспределенной экосистемой компании, а также интегрированные с…
Director, Site Reliability Engineering
At Klaviyo, we value the unique backgrounds, experiences and perspectives each Klaviyo (we call ourselves Klaviyos) brings to our workplace each and every day. We believe everyone deserves a fair shot at success and…
Senior Site Reliability Engineer (AWS, Docker, Kubernetes)
Join a Digital Health Engineering project focused on building a highly automated multi-cloud provisioning platform. You will own reliability, scalability, automation and production infrastructure as a senior technical…
ESPECIALISTA DE SRE I - Remoto
Na Inmetrics , a inovação e a excelência operam lado a lado em um ambiente de trabalho colaborativo, saudável e dinâmico. Nossa cultura valoriza: 📚 Aprendizado constante 🗣️ Transparência na comunicação 🔄…
Senior Site Reliability Engineer
Do you want to change the world? At Cabify, that’s what we’re doing. We aim to make cities better places to live by improving mobility for the people living in them, connecting riders to drivers, providing mobility…
Site Reliability Engineer (SRE) On-Prem
Every nation has data. Few can protect it. Fewer still can act on it. Dream is the sovereign AI and national cyber-defense company for governments. We help nations secure their most critical systems, connect fragmented…
Business Site Reliability Engineer
About Us Established in 2018, Bybit is one of the world’s leading cryptocurrency exchanges and digital financial platforms, serving over 80 million users across more than 200 countries and regions. Powered by…