How Remoteville checks and expires listings
Cloud DevOps Engineer
Skills
Microsoft AzureAmazon Web ServicesGood Clinical PracticeGoogle Cloud PlatformKubernetesLinux System AdministrationNetwork Infrastructure
What the job involves
The main requirements, responsibilities and hiring steps.
Requirements
- Passion for working on deeply technical projects and wanting to work on distributed systems, concurrency & parallelism, and correctness
- 5+ years of in-depth knowledges in managing the lifecycle and operations of Kubernetes clusters in the public cloud
- Experience in shaping, managing, and deploying workloads in Kubernetes
- Strong understanding of Go
- Expert knowledge of at least one: AWS, Azure, GCP including experience with IAC both system and network infrastructure
- Understanding of stream processing
- Experience and comfortable working with a 100% distributed engineering team, collaborating on GitHub, in the open and a self-starter
- Passion to grow a culture of written tradition
- Excellent communication and interpersonal skills, capable of engaging with stakeholders at all levels of the organization
- Working knowledge of building Kubernetes operators for storage systems and stateful workloads
- Experience building a SaaS platform
- Operated and used streaming platforms either as a user or provider
Nice to have
- Working knowledge of building Kubernetes operators
- Experience building a SaaS platform
- Experience with streaming platforms
Day to day
- We are continuing to invest and grow our Go-based Kubernetes team at Redpanda
- In this role, you will focus on developing and maintaining our global Kubernetes fleet
- You will collaborate closely with our Cloud and Core engineering teams to build robust and scalable systems with an emphasis on stability and safe upgrades with zero downtime
- Build and maintain the Fleet Management tool stack that manages our global footprint of Kubernetes clusters
- Contribute to a Kubernetes operator to manage Redpanda clusters in a zero-downtime fashion
- Participate in an on-call escalation chain, help drive down problems that lead to on-call escalations, and help drive improving on-call quality of life.
