Senior Platform Engineer
Operate and improve Chronicle’s AWS and EKS infrastructure, manage Kubernetes workloads and cloud systems, strengthen reliability and security, support blockchain infrastructure, and collaborate across engineering teams.
Responsibilities
- Operate and improve AWS and EKS infrastructure
- Manage Kubernetes workloads, networking, IAM, storage, and compute through Terraform, Helm, and GitOps
- Improve CI/CD safety, traceability, rollback practices, and operational automation
- Support platform upgrades, migrations, capacity planning, and architecture changes
- Maintain observability across metrics, logs, traces, and alerts
- Improve service health checks, alert quality, reliability targets, and recovery procedures
- Provide scheduled production on-call coverage
- Reduce recurring incidents and technical or human single points of failure
- Operate oracle infrastructure, RPC nodes, relayers, signing services, and data workloads
- Support messaging systems such as Kafka and NATS
- Investigate failures across cloud services, applications, external dependencies, and blockchain networks
- Apply access control, secrets management, change management, and infrastructure security practices
- Maintain records for production changes, privileged access, and recovery procedures
- Support customer, security, and audit-related technical reviews
- Maintain architecture documents, runbooks, handovers, and operational procedures
- Distribute critical knowledge through documentation, pairing, walkthroughs, and reviews
- Collaborate with DevOps, on-chain, off-chain, and full-stack engineers
- Own infrastructure work from investigation through deployment and follow-up
- Review changes for reliability, security, maintainability, performance, and cost
- Shape platform priorities around production risks and recurring engineering needs
Requirements
- Strong experience in Platform Engineering, DevOps, SRE, or cloud infrastructure
- Hands-on production experience with AWS, Kubernetes, Linux, networking, and IAM
- Practical knowledge of Terraform, Infrastructure as Code, GitOps, and CI/CD
- Experience operating observability systems and participating in on-call, incident response, and recovery work
- Ability to automate operational tasks using Go, Python, Bash, or a comparable language
- Understanding of distributed systems, security, and event-driven architectures
- Clear written communication and disciplined documentation and handover practices
- Ability to work autonomously across a globally distributed engineering team
- Based in the Central European time zone and able to work primarily within CET/CEST business hours
Benefits
- A clear path to broader technical ownership
- Direct influence over production architecture and operational practices
- Exposure to decentralized and institutional financial infrastructure
- Experienced, globally distributed engineering team
- Competitive compensation package
- Equity participation
- Flexible payment options in USD, EUR, and USDC
- 6 weeks vacation
- Remote-first environment
- Flexible working schedule
- Annual company offsites