Staff Technical Program Manager, Site Reliability Engineering
Remote · Full Time
Be the first to apply
- Experience
- 8+ yrs
- Salary
- —
- Openings
- 1
- Posted
- há 4 dias
- Work mode
- Work from home
- Resume
- Required to apply
Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.
Job description
Overview
As a Staff Technical Program Manager (TPM) specializing in Site Reliability Engineering (SRE), you will collaborate closely with SRE leaders and practitioners to expand the critical platform that supports MongoDB's cloud services. This role involves steering the execution of programs, enhancing reliability practices in production environments, and facilitating coordination among teams spanning the US and EMEA regions. Your impact will be evident in improved product launches, clearer developmental roadmaps, enhanced reliability measurements, and an empowered SRE organization capable of scalable, predictable delivery.
This position is open for candidates based in MongoDB's Dublin or Cork offices or who prefer working remotely from within Ireland.
Responsibilities
- Define and manage program scopes, set milestones, and determine success criteria alongside SRE engineers and leaders.
- Oversee dependencies among platform teams, track progress using tools such as Jira, and ensure timely program completion.
- Lead initiatives to improve production reliability by driving change management and launch readiness processes.
- Collaborate with SRE and product teams to establish, implement, and operationalize Service Level Objectives (SLOs) and Service Level Indicators (SLIs).
- Use incident analyses, key metrics, and capacity data to prioritize tasks and enable continuous improvement cycles.
- Coordinate cross-functional alignment between SRE, Security, Compliance, Cloud platform, and other engineering divisions.
- Manage incident response activities across teams, ensuring transparent follow-ups and fostering trust as a central facilitator of complex efforts.
- Develop and implement scalable frameworks and communication strategies to enable SRE teams to operate effectively and autonomously.
Requirements
- A minimum of 8 years' experience in technical program management, engineering management, or similar roles collaborating with software engineering teams.
- Proven ability to lead large-scale, complex platform projects successfully, navigating ambiguity and organizational change.
- In-depth understanding of production change management, software development lifecycle practices, and reliability metrics such as SLOs and SLIs.
- Expertise in planning roadmaps and managing inter-team dependencies.
- Proficient in analyzing and interpreting metrics, logs, and other data to guide decision-making and communicate risk effectively.
- Exceptional communication skills with clarity and composure, proficient in engaging with engineers, cross-functional partners, and leadership.
- Collaborative attitude with a low ego, taking ownership of challenging problems from start to finish.
Preferred Qualifications
- Practical or collaborative experience with Kubernetes, cloud network architecture, or observability tools including metrics, logging, tracing, and alerting systems.
- Prior involvement with SRE teams or understanding of their workflows.
- Background in large-scale cloud infrastructure or platform engineering environments.
- Familiarity with MongoDB Atlas or similar modern cloud database platforms.
Why Join MongoDB?
This role embodies the core values of MongoDB by owning challenging, undefined problems holistically. You will partner with multiple teams—SRE, Engineering, Security, and Compliance—to co-develop scalable solutions that enhance the organization as a whole. Positioned at the heart of impactful initiatives, you will collaborate with executive stakeholders and receive support for professional development.
About MongoDB
MongoDB empowers its customers and employees to innovate rapidly, adapting to market changes with agility. Its pioneering data platform addresses the AI era’s challenges, enabling organizations to modernize applications, foster innovation, and leverage AI. MongoDB Atlas stands as the only multi-cloud, globally distributed data platform supporting AWS, Google Cloud, and Microsoft Azure.
Serving over 67,000 customers including a significant portion of Fortune 100 companies and AI-driven startups, MongoDB powers the next wave of software innovation globally.
MongoDB's culture is guided by its Leadership Commitment, which shapes decision-making, collaborative behavior, and team success. The company supports employee well-being through initiatives like affinity groups, fertility support, and generous parental leave policies, fostering an enriching and inclusive workplace.
Equal Opportunity and Accommodation
MongoDB is dedicated to providing accommodations for candidates with disabilities during application and interview phases upon request. The company is an equal opportunities employer ensuring a diverse and inclusive hiring process.