Data Engineer

Skills
PysparkML/AI ModelsSemantic MatchingData NormalizationData PipelinesData PlatformsMastering
Role

What the job involves

The main requirements, responsibilities and hiring steps.

Requirements

  • Strong data engineering experience
  • Experience leveraging ML/AI models for semantic matching and data normalization
  • Ability to improve mastering accuracy across ingestion merging and roll up layers
  • Hands-on experience with large-scale data pipelines
  • Hands-on experience with production-grade data platforms
  • Pyspark experience
  • Work authorization in the USA without sponsorship

Nice to have

  • Detail oriented
  • Analytical mindset
  • Collaborative

Day to day

  • Work on mastering and semantic matching capabilities for large-scale data platforms.
  • Improve mastering accuracy across ingestion merging and roll up layers.
  • Build and support production-grade data pipelines using Pyspark and related technologies.