How Remoteville checks and expires listings
Reliability Engineer
Skills
Computer ScienceDevopsSoftware As A ServiceContinuous ImprovementDatabasesOptimizationReadiness
What the job involves
The main requirements, responsibilities and hiring steps.
Requirements
- Bachelors degree in Computer Science Information Technology or related field or relevant experience
- 8+ years experience in DevOps Site Reliability Engineering Cloud Engineering or similar roles
- Strong hands-on experience with Microsoft Azure including Azure SQL Azure Functions Azure App Services and Azure Containers
- Ability to read and interpret telemetry logs metrics and resource usage data
- Experience working with production systems requiring high availability and reliability
- Comfort owning work end to end from issue identification to improvement execution
- Experience adjusting pipelines hosting configurations and deployment processes
- Solid understanding of cloud cost drivers and usage optimization
- Strong problem-solving skills and ability to collaborate across engineering and support teams
- Ability to read and interpret application code for troubleshooting and root cause analysis
Nice to have
- Proactive
- Analytical
- Collaborative
- Continuous improvement mindset
Day to day
- Monitor and improve the health availability performance and cost efficiency of Azure-based production systems
- Use telemetry logs metrics and resource usage data to identify bottlenecks reliability risks and performance issues
- Investigate complex production incidents perform root cause analysis and implement practical improvements across deployments systems and operations
