Role OverviewAs a Cloud Site Reliability Engineer (SRE) on the Data Management and Analytics Platform (DMAP) team, you will play a critical role in driving analytics throughout the organization to improve products, engage with customers, create efficiencies, and unlock new business opportunities through data-driven insights.
What You Will Do
You will focus on ensuring the availability, performance, and scalability of critical data pipelines and analytics infrastructure, designing, building, and operating highly available, scalable, and resilient cloud infrastructure, and improving observability across data pipelines and platforms.
Why It Might Be a Fit
You will need to have strong proficiency in at least one programming or scripting language, experience supporting production systems with a focus on reliability, scalability, and observability, and hands-on experience operating or designing highly available distributed systems.
Requirements
- 5+ years of experience in Site Reliability Engineering, DevOps, or Cloud Infrastructure roles
- Strong proficiency in at least one programming or scripting language (Python, and/or Go)
- Experience supporting production systems with a focus on reliability, scalability, and observability
- Hands-on experience operating or designing highly available distributed systems
- A Bachelor’s degree in Computer Science, Engineering, Mathematics, or a related field, or equivalent professional experience
Benefits
- Comprehensive and generous benefits plan
- Merit increases
- Incentive compensation (exempt roles only)
- Paid holidays
- Paid time off
- Medical, dental, vision
- Short and long term disability benefits
- 401(k) +match
- Life insurance
- Wellness programs
- Bonus
]]>