Sr DevOps/Platform Engineer

Storage Strategies, Inc. (SSI) is seeking a Senior DevOps / Platform Engineer in San Diego, CA to lead the container platform, CI/CD pipelines, and observability capabilities in AWS GovCloud. The successful candidate will enable large Navy development teams to deliver secure capabilities rapidly using Kubernetes, GitOps, and modern DevSecOps practices. Must have an active Secret clearance. Position is hybrid work.


DUTIES AND RESPONSIBILITIES:

  • Design, build, and operate the CI/CD infrastructure serving the program's engineering team and customer base
  • Configure, deploy, and operate containerized applications using Kubernetes on AWS EKS
  • Using Helm and Kustomize, design and manage Kubernetes resources including Deployments, StatefulSets, ingresses, storage classes, PodDisruptionBudgets, HorizontalPodAutoscalers, and related resources to meet availability and performance objectives.
  • Implement and administer Kubernetes RBAC (Roles, RoleBindings), EKS Pod Identities / IRSA, and network policies to enforce least privilege and workload isolation.
  • Define and enforce platform guardrails using ValidatingAdmissionPolicies for resource paradigm enforcement.
  • Utilize GitOps based workflows using ArgoCD for environment configuration and application deployments, targeting near zero configuration drift.
  • Deploy, configure, and optimize GitLab runners (Docker, Fleeting autoscaler, and Kubernetes runners) to support large scale CI/CD job execution.
  • Configure AWS Controllers for Kubernetes (ACK) and IAM policies to safely provision AWS resources via Kubernetes and GitOps workflows while preserving cluster isolation.
  • Integrate and maintain DevSecOps tools (e.g., Artifactory, SonarQube, Fortify) within CI/CD pipelines.
  • Implement quality and security gates in projects’ merge request settings.
  • Develop and configure monitoring agents (e.g., Telegraf) and observability pipelines using Prometheus, InfluxDB, CloudWatch, OpenTelemetry, and related tooling.
  • Build and maintain Grafana dashboards and alert rules for system health, application availability, performance metrics, certificate expiration, and capacity management.
  • Use CloudWatch metrics and Log Insights to troubleshoot incidents, perform root cause analysis, and continuously improve platform reliability.
  • Author and maintain runbooks for application and platform upgrades, migrations, and incident response.
  • Mentor development teams on Kubernetes, GitOps, CI/CD, and secure deployment patterns.

See also

要針對這個職缺調整履歷嗎?

目前無法檢查您與這個職缺的符合程度;請先將履歷加入個人檔案,下次即可查看。

A new version of freehire is available