Role OverviewJoin Capgemini and help build a hybrid data platform that bridges an on-premises big data cluster with Google Cloud Platform. This role is central to establishing a repeatable production-grade pattern for multitenant data platform architecture.
What You Will Do
Configure and provision the on-premises Cloudera cluster environment, design and implement Apache Iceberg table structures, integrate the Iceberg REST Catalog with Apache Gravitino, and establish GCP as a production tenant on the existing on-prem Cloudera cluster.
Why It Might Be a Fit
This role requires 5 years of relevant experience in data platform engineering, hands-on experience with Google Cloud Platform, and experience with on-premises Cloudera big data clusters, Apache Iceberg table format, and Iceberg REST Catalog IRC.
Requirements
- 5 years of relevant experience in data platform engineering
- Hands-on experience with Google Cloud Platform
- Experience with on-premises Cloudera big data clusters
- Experience with Apache Iceberg table format
- Experience with Iceberg REST Catalog IRC
- Experience with Apache Gravitino for metadata catalog integration
Benefits
- Paid time off based on employee grade (A-F)
- Company paid holidays
- Personal Days
- Sick Leave
- Medical, dental, and vision coverage
- Retirement savings plans (e.g., 401(k) in the U.S., RRSP in Canada)
- Life and disability insurance
- Employee assistance programs
]]>