Job Description
Job Description
Daily responsibilities
Develop Spark/ BQ data pipelines in GCP DataProc Cluster
Develop data models and schema to support business requirements
Ensure data quality and reliability by developing automated checks in place
Optimize performance and ensure adherence to SLAs
Document and share learnings with the rest of the team.
Candidate will be working on and migrating workloads from Hadoop OnPrem cluster to GCP.
Team is geographically spread based mainly out of Charlotte, NC and Bangalore, India.
Candidate should be able to communicate in a concise and clear manner to keep the team
appraised of day-to-day progress.
Required Skills
GCP/ Big query Experience ? 1 year minimum
Scala/ Python - 3+ years
SQL ? 3+ years
Strong SQL, one or more programming language Scala/ Python
Preferred Skills
Exposure to building semantic layer in LookML
GCP Certification
Master?s Degree In related field
