About this opportunity
Ursus Inc lists this Technical Program Manager III opportunity in san francisco, California. Review the employer’s description below for duties, qualifications and application requirements.
Job description
Technical Program Manager III
Our client is a unified analytics platform that helps organizations accelerate innovation by simplifying data engineering, machine learning, and analytics at scale.
Job Description
The R&D Operations Organization is seeking a Senior Technical Program Manager (TPM) with experience in compute capacity management, infrastructure operations, forecasting, resource planning, and large-scale distributed systems.
This role focuses on capacity management across the compute platform, ensuring infrastructure resources are efficiently allocated, forecasted, and scaled to meet customer and business demand. The TPM will partner closely with Infrastructure Engineering, Platform Engineering, Product, Support, and Operations teams to drive operational excellence, capacity planning, utilization optimization, and strategic infrastructure initiatives.
This role involves supporting platform operations, handling customer escalations, and monitoring cluster health.
Additionally, you'll ensure optimal compute resource allocation aligned with product, sales, and engineering priorities while driving decisions through data analysis, reporting, forecasting, and operational insights.
The Impact You Will Have
Act as a single point of contact for escalations from sales and global support teams and help with various billing and support issues
Drive compute capacity planning initiatives to ensure platform teams have the necessary infrastructure resources to support customer demand, product growth, and operational scalability.
Improve platform efficiency and cost optimization through analysis of compute utilization, forecasting models, infrastructure allocation strategies, and capacity planning processes.
Establish and maintain effective communication with technical and non-technical stakeholders and customers, including regular project updates, status reports, and presentations.
Drive measurable improvements in compute efficiency, forecasting accuracy, infrastructure utilization, and operational scalability through process optimization and automation initiatives.
Ensure strategic alignment across Sales, Global Support, and Engineering
Lead short-term and long-term compute capacity forecasting and planning efforts across infrastructure environments.
Partner with Infrastructure Engineering teams to monitor platform capacity, utilization trends, operational risk, and service readiness.
Develop dashboards, reporting, and analytics to track compute consumption, demand forecasts, resource allocation, and operational performance.
Drive cross-functional programs involving Platform Engineering, Capacity Management, Support, Product, Finance, and Leadership stakeholders.
Investigate capacity constraints, service bottlenecks, and resource utilization trends, leveraging data-driven decision making to develop mitigation plans.
Competencies and Requirements
6+ years of professional experience with a Bachelor's degree and related experience in technical program management, distributed platforms, resource management, execution and strategic planning.
Proven track record of driving cross-functional teams to deliver complex technical projects on time and with high quality.
Excellent communication, negotiation and analytical skills, with the ability to document standard operating procedures and processes
Advanced working SQL Knowledge, Ability to build and maintain analytics to track, forecast, and visualize consumption through ad-hoc SQL, reports, and dashboards
Experience performing root cause analysis on internal and external data and processes to answer specific business questions and identify opportunities for improvement.
Self-motivated and able to work independently, as well as in a team environment.
Preferred experience supporting large-scale cloud infrastructure, distributed systems, compute platforms, capacity planning, or infrastructure operations.
Familiarity with big data technologies such as Apache Spark, Delta Lake, and MLflow is a plus.
Experience with compute capacity management, as well as financial analysis or sales/deal desk quoting, is a plus.
Demonstrated ability to quickly learn complex technical domains and adapt to evolving infrastructure, platform, and business priorities.
Must Have
Advanced Project Management
Technical Program Management
Infrastructure Operations or Platform Engineering Exposure
Capacity Planning / Forecasting Experience
Cross-functional Stakeholder Management
Data-Driven Decision Making
Advanced SQL proficiency with experience building analytics, dashboards, forecasting models, and operational reporting to support capacity planning and business decisions.
Cloud (aws/azure) knowledge
Fluency in business english
Nice To Have
Dashboarding / BI tools (Tableau, Power BI, Looker)
Python
Compute capacity management
Infrastructure scaling programs
Distributed systems knowledge
Platform operations
Cloud resource optimization
Financial modeling and capacity forecasting
Apache Spark, Delta Lake, and MLflow experience
BENEFITS SUMMARY: Individual compensation is determined by skills, qualifications, experience, and location. Compensation details listed in this posting reflect the base hourly rate or annual salary only, unless otherwise stated. In addition to base compensation, full-time roles are eligible for Medical, Dental, Vision, Commuter and 401K benefits with company matching.
Worksite address
san francisco, CA, 94199, US
Who can apply
Review the original listing for work authorization, qualifications and employer requirements.