Senior Cloud Data Engineer (GCP Data Fabric)
Senior Cloud Data Engineer (GCP Data Fabric)
POSITION OVERVIEWWe are seeking a Senior Cloud Data Engineer to serve as the technical lead for a large-scale enterprise data migration and cloud modernization initiative. This individual will be responsible for migrating more than 100TB of business data from Cloudera Hadoop environments and AWS storage into Google Cloud Platform (GCP), helping build a modern Data Fabric centered around BigQuery.
The ideal candidate will have deep expertise in Cloudera, Hadoop ecosystems, cloud data engineering, and large-scale data migrations, along with experience building modern, scalable data pipelines using Medallion Architecture.
MUST-HAVE SKILLS Candidates should have most or all of the following:6+ years of Cloud Data Engineering experienceAdvanced experience with Cloudera (HDFS/Hive)Experience migrating large-scale enterprise data platformsStrong Google Cloud Platform (GCP) experienceExperience with BigQueryExperience building modern cloud data pipelinesExperience working with Hadoop ecosystems (HDFS, Hive)Experience with AWS data storage services (S3 and/or FSx)Experience designing and implementing Medallion Architecture (Bronze, Silver, Gold)Strong SQL and data engineering fundamentalsExperience optimizing large-scale data storage and processingExcellent troubleshooting and problem-solving skills
PREFERRED SKILLSData Fabric implementation experienceServerless data architectureData lake modernizationData governance and securityData quality and validation frameworksCI/CD for data engineeringInfrastructure as CodePython or Spark development experienceEnterprise cloud migration experience
CORE RESPONSIBILITIESCloud Data MigrationLead migration of over 100TB of enterprise data from Cloudera Hadoop environments into Google Cloud PlatformMigrate data from HDFS, Hive, AWS S3, and FSx into BigQueryValidate data integrity throughout migration activitiesOptimize storage, performance, and scalabilityData EngineeringDesign and develop modern cloud-native data pipelinesBuild scalable ingestion and transformation frameworksDevelop optimized BigQuery storage modelsSupport enterprise Data Fabric implementationPlatform ArchitectureDesign and implement Medallion Architecture (Bronze, Silver, Gold)Develop secure and scalable data movement frameworksImprove platform reliability, performance, and operational efficiencySupport cloud-native data engineering best practices