Senior DevOps / Infrastructure Engineer
Skills
Computer ScienceDevopsDomain Name SystemGrafanaLarge Language ModelsLinuxOperations
What the job involves
The main requirements, responsibilities and hiring steps.
Requirements
- 5+ years in DevOps SRE or Infrastructure Engineering operating production systems at scale
- Strong Linux systemd networking and shell fundamentals
- Comfort debugging live systems over SSH
- Deep hands-on experience with Ansible and Terraform
- Experience with observability stacks such as Prometheus Grafana and Loki
- Hands-on fluency with AI-assisted engineering and coding agents
- Experience designing automation with safe guardrails
- Programming and scripting experience in Python or bash
- Experience with Kubernetes and GitOps is a plus
- Experience building AI agent tooling or MCP servers is a plus
- Experience serving inference is a plus
- Experience with blockchain clients or node operations is a plus
- Bachelor's degree in Computer Science Engineering or related field is a plus
Nice to have
- Calm incident response
- Methodical debugging
- Low ego
- Collaborative mindset
- High-quality output
- AI-native
Day to day
- Operate a globally distributed blockchain node fleet across mainnet and testnet, keeping validators full nodes archive nodes and indexers healthy through upgrades recovery and incident response
- Own infrastructure as code and platform automation using Ansible Terraform Atlantis Kubernetes and Flux while maintaining reliable cloud DNS and GitOps workflows
- Build observability alerting and release automation with Prometheus Grafana and Loki and develop safe AI-driven operational tooling with guardrails human oversight and low toil
- Harden systems manage secrets and codify operational knowledge into reusable runbooks tooling and agent workflows that support both engineers and autonomous agents
