How Remoteville checks and expires listings
Site Reliability Engineer
Skills
.NET FrameworkCloud ComputingJavaMicrosoft AzureAmazon Web ServicesEngineeringGoogle Cloud Platform
What the job involves
The main requirements, responsibilities and hiring steps.
Requirements
- Experience in infrastructure as code tooling such as Terraform or ARM Bicep
- Hands-on experience with core CI/CD tooling such as Azure DevOps GitHub Actions or GitLab
- Experience with monitoring tools such as Datadog Splunk New Relic Azure Monitor or AWS CloudWatch
- Demonstrable experience in multiple core technologies such as Dotnet Java AI Data Engineering or Golang
- Strong troubleshooting skills with the ability to identify systemic issues from incidents and failures
- Experience implementing fixes features automation and service request resolution improvements
- Experience leading incident resolution post mortems documentation and mitigation planning
- Experience with service requests change management and stakeholder management
- Experience proactively mitigating security risks across code infrastructure and dependencies
- Experience working with cloud platforms such as AWS Azure or GCP
Nice to have
- Client-facing
- Analytical
- Proactive
- Collaborative
- Detail-oriented
- Process-driven
Day to day
- Act as a technical escalation point for unresolved data platform issues and support the stability of critical SRE services
- Monitor maintain and troubleshoot databases data warehouses and related infrastructure to keep platforms reliable and performant
- Collaborate with data engineering teams and external suppliers to improve data flow automate support tasks reduce toil and strengthen incident response
