What the job involves
The main requirements, responsibilities and hiring steps.
Requirements
- 4+ years of professional software engineering experience
- Experience building and maintaining large-scale software platforms
- Strong experience developing ETL and data ingestion pipelines
- Strong knowledge of Apache Spark
- Understanding of distributed data processing systems
- Experience with data modeling concepts and software design patterns
- Strong collaboration and communication skills
- Scala development experience
- Databricks platform experience
- Experience with cloud infrastructure AWS preferred
- Terraform or Infrastructure as Code experience
- Strong Python development skills
- Experience with DevOps and cloud platform management
- Experience using AI-assisted software development tools
- Experience working with large-scale data platforms
Nice to have
- Collaborative
- Product-minded
- Curious
- Ownership-oriented
- Detail-oriented
Day to day
- Design, build, and maintain core data platform software for large-scale distributed systems
- Develop and optimize ETL pipelines that ingest transform and normalize massive datasets
- Build high-performance data processing applications using Apache Spark Databricks and Scala
- Collaborate with software engineers data scientists and business stakeholders on high-impact technical initiatives
- Improve platform reliability scalability and developer productivity
- Participate in architecture discussions code reviews and continuous improvement initiatives
- Utilize AI-assisted development tools such as Claude Code and GitHub Copilot to improve engineering efficiency
