How Remoteville checks and expires listings

Site Reliability Engineer

Skills
.NET FrameworkCloud ComputingJavaMicrosoft AzureAmazon Web ServicesEngineeringGoogle Cloud Platform
Role

What the job involves

The main requirements, responsibilities and hiring steps.

Requirements

  • Experience in infrastructure as code tooling such as Terraform or ARM Bicep
  • Hands-on experience with core CI/CD tooling such as Azure DevOps GitHub Actions or GitLab
  • Experience with monitoring tools such as Datadog Splunk New Relic Azure Monitor or AWS CloudWatch
  • Demonstrable experience in multiple core technologies such as Dotnet Java AI Data Engineering or Golang
  • Strong troubleshooting skills with the ability to identify systemic issues from incidents and failures
  • Experience implementing fixes features automation and service request resolution improvements
  • Experience leading incident resolution post mortems documentation and mitigation planning
  • Experience with service requests change management and stakeholder management
  • Experience proactively mitigating security risks across code infrastructure and dependencies
  • Experience working with cloud platforms such as AWS Azure or GCP

Nice to have

  • Client-facing
  • Analytical
  • Proactive
  • Collaborative
  • Detail-oriented
  • Process-driven

Day to day

  • Act as a technical escalation point for unresolved data platform issues and support the stability of critical SRE services
  • Monitor maintain and troubleshoot databases data warehouses and related infrastructure to keep platforms reliable and performant
  • Collaborate with data engineering teams and external suppliers to improve data flow automate support tasks reduce toil and strengthen incident response