Software Engineer Infrastructure

You will own and improve the systems behind a real-time conversational product. You will operate GPU inference infrastructure across providers and regions, expand clusters on EKS, improve routing and scheduling, raise uptime, support security and SOC2 work, and improve deployment pipelines. You will also investigate and fix issues across infrastructure, backend services, and product code.

Responsibilities

  • Own GPU inference deployments serving live conversations across multiple providers and regions
  • Tune GPU generations and reduce cold-start and model load times
  • Expand GPU capacity by onboarding providers and regions
  • Stand up clusters on EKS
  • Build routing, scheduling, and throughput for fast weight loading
  • Improve infrastructure uptime
  • Support security and SOC2 work
  • Identify, fix, or escalate infrastructure and software problems
  • Improve deployment pipelines

Requirements

  • Hands-on experience deploying and optimizing GPU inference workloads
  • Experience building reliable systems on GPU cloud providers
  • Deep Kubernetes and EKS knowledge, including routing and scheduling
  • Experience designing workload placement across a fleet
  • Experience writing services for workload placement
  • Deep AWS experience
  • Senior-level ownership experience, including setting technical direction and delivering ambiguous work
  • Ability to explain complex ideas clearly to engineers and non-engineers

See also

要針對這個職缺調整履歷嗎?

目前無法檢查您與這個職缺的符合程度;請先將履歷加入個人檔案,下次即可查看。

A new version of freehire is available