Senior Staff Technical Program Manager
3 нед. назад
USALeadOnsite
Lead and coordinate programs within 's Cloud Engineering to ensure operational visibility and cross-team collaboration for AI infrastructure.
Будет плюсом
- Time inside AWS, GCP, Azure, CoreWeave, Lambda Labs, or a similar cloud provider.
- Experience running or improving a data center deployment or build-to-turnup process.
- Familiarity with, Opsgenie, PagerDuty, or similar incident management platforms at scale.
- Background in AI/ML infrastructure.
Условия
- Competitive compensation and equity packages
- Restricted Stock Units
- Paid time off, paid holidays & leave of absence programs
- Comprehensive health, dental & vision insurance
- Employer contributions to HSA account
- Paid parental leave
- Paid life insurance, short-term and long-term disability
- Professional development & tuition reimbursement
- Mental health & wellness support
- Commuter benefits (parking & transit)
- Cell phone stipend
- 401(k) Retirement plan with company match up to 4% of salary
- Volunteer time off
- Global travel insurance & emergency assistance
- Daily meals allowance
- Additional perks & programs specific to location
- Compensation will be paid in the range of up to $230,000 - $280,000 + Bonus. Restricted Stock Units are included in all offers. Compensation to be determined by the applicant's knowledge, education, and abilities, as well as internal equity and alignment with market data.
- is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, disability, genetic information, pregnancy, citizenship, marital status, sex/gender, sexual preference/ orientation, gender identity, age, veteran status, national origin, or any other status protected by law or regulation.
Другое
- is on a mission to accelerate the abundance of energy and intelligence . As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack — from electrons to tokens — to power the world's most ambitious AI workloads. When you join , you join a team that is building the future, faster.
- We're in the midst of the greatest industrial revolution of our time. The demand for AI compute is boundless, and power is a bottleneck. We're solving that — with an energy-first approach that makes AI infrastructure better for the world and faster for the people innovating with AI.
- We're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved — people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services.
- If you want to do the most meaningful work of your career, help our customers and partners advance their AI strategies, and be part of a high-performing team that believes in each other, come build with us at .
- 's Cloud Engineering organization is 380 people today and hiring toward 550. At that scale, operational visibility, cross-team coordination, and clear communication don't happen by accident anymore. This role exists to make sure they happen on purpose.
- You'll drive the programs that improve the KPI infrastructure and give leadership real insight into operational health, along with the connective tissue between cross-functional teams for data center deployments (DCOps, Cloud Engineering, and TPM). You own the checks that ensure a quality hand-off between teams, with authority to pull in resources to keep compute capacity delivery on track. You don't just report when the process breaks; you fix it as the fleet scales.
- This isn't a role scoped to a single data center deployment or a single dashboard. You will be a strategic advisor to leadership on whether teams are actually executing operationally.
- The ideal candidate has a technical background (software engineering, TPM, or product in an infrastructure context) and enough exposure to physical data center builds, or the instinct to ramp fast on them, to hold cross-functional teams to a shared definition of done.
- Engineering Operations
- Run the programs that improve the KPI infrastructure, covering SLO dashboards, incident follow-up completion rates, on-call burden, deployment velocity, and initiative trackers. Surface where the org needs to improve and drive execution against it. Data should tell the story before anyone has to ask.
- Support the Chief of Staff on strategic projects and cross-cutting initiatives that don't have a natural single owner.
- Partner with engineering, SRE, customer success, and data engineering to keep operational data accurate and consistently reported.
- Work toward a single pane of glass that makes the health of the organization and its deployments easy to understand at a glance.
- Deployment Ownership
- Own the checks that ensure a quality hand-off between teams, with the authority to pull in resources needed to deliver compute capacity to customers.
- Be the single point of authority to pull in the right teams the moment an issue surfaces during a deployment, replacing today's back-and-forth, ticket-driven handback.
- Own the deployment process end to end and evolve it as headcount and site count grow.
- 12+ years in software engineering, technical program management, or a technical product role, close enough to production systems to know what an SLO breach actually means.
- You've operated in high-growth infrastructure environments where processes are still being built; ambiguity doesn't paralyze you.
- You're a natural coordinator who works across teams without formal authority. DCOps, Network, and engineering leads trust you because you follow through.
- You can turn messy, multi-source data into a clear picture of organizational health, and you understand, or can quickly learn, the full data center build lifecycle: from project planning through the physical-to-digital handoff.
- Scrappy, low-ego, high-drive. You build the program and tooling yourself when it doesn't exist yet, and you care more about the outcome than the credit.