Senior DevOps / Infrastructure Engineer

Skills
Computer ScienceDevopsDomain Name SystemGrafanaLarge Language ModelsLinuxOperations
Role

What the job involves

The main requirements, responsibilities and hiring steps.

Requirements

  • 5+ years in DevOps SRE or Infrastructure Engineering operating production systems at scale
  • Strong Linux systemd networking and shell fundamentals
  • Comfort debugging live systems over SSH
  • Deep hands-on experience with Ansible and Terraform
  • Experience with observability stacks such as Prometheus Grafana and Loki
  • Hands-on fluency with AI-assisted engineering and coding agents
  • Experience designing automation with safe guardrails
  • Programming and scripting experience in Python or bash
  • Experience with Kubernetes and GitOps is a plus
  • Experience building AI agent tooling or MCP servers is a plus
  • Experience serving inference is a plus
  • Experience with blockchain clients or node operations is a plus
  • Bachelor's degree in Computer Science Engineering or related field is a plus

Nice to have

  • Calm incident response
  • Methodical debugging
  • Low ego
  • Collaborative mindset
  • High-quality output
  • AI-native

Day to day

  • Operate a globally distributed blockchain node fleet across mainnet and testnet, keeping validators full nodes archive nodes and indexers healthy through upgrades recovery and incident response
  • Own infrastructure as code and platform automation using Ansible Terraform Atlantis Kubernetes and Flux while maintaining reliable cloud DNS and GitOps workflows
  • Build observability alerting and release automation with Prometheus Grafana and Loki and develop safe AI-driven operational tooling with guardrails human oversight and low toil
  • Harden systems manage secrets and codify operational knowledge into reusable runbooks tooling and agent workflows that support both engineers and autonomous agents