Senior Technical Program Manager Engineering Infrastructure Reliability
Location: Miami FL, Onsite
About the Role
We're looking for a Senior Technical Program Manager to serve as the cross-domain orchestration leader for our Engineering Infrastructure organization. This organization spans seven specialized functions Site Reliability Engineering (SRE), Platform Engineering, Security, Quality Engineering, Databases, Application Support, and Production Defect Resolution under a single director. This is not a traditional TPM role embedded in one team's delivery. You will sit across all seven domains, reporting directly to the Director of Engineering Infrastructure, and own the programs, governance, and coordination that make these domains operate as one unified organization rather than seven silos. You will be the connective tissue between proactive teams that build foundations (SRE, Platform, Security, Quality, Databases) and reactive teams that resolve customer and production issues (Application Support, Defect Resolution) ensuring insights flow in both directions and that cross-domain initiatives are executed with accountability and speed.
What We're Looking For
Required:
- 10+ years in technical program management, with at least 3 years managing programs that span 5+ engineering teams with competing priorities and hard dependencies.
- Deep technical fluency in infrastructure and platform engineering domains.
- Demonstrated ability to influence without authority.
- Strong stakeholder management skills comfortable operating at the Director/VP level for leadership communication and at the senior engineer level for technical coordination.
- Experience with metrics, dashboards, and data-driven decision support.
- Excellent facilitation skills running focused, decision-oriented meetings with senior technical leaders.
- Experience with organizational change guiding teams through operating model transformations, not just feature delivery.
Strongly Preferred:
- Background in SRE, platform engineering, DevOps, or infrastructure organizations.
- Experience with incident management processes and post-incident reviews.
- Familiarity with SRE methodologies (SLOs/SLIs/error budgets), DORA metrics, chaos engineering, and platform-as-a-product principles.
- Experience working across both proactive engineering functions (platform, reliability, security, quality) and reactive operational functions (support, defect resolution) or similar environments with mixed proactive/reactive team dynamics.
- Track record of simplifying processes, not adding them.