Senior TPM – Global Reliability

Job not on LinkedIn

🔥 9 minutes ago

🇺🇸 United States – Remote

⏰ Full Time

🟠 Senior

⚒️ Technical Product Manager (TPM)

🦅 H1B Visa Sponsor

infoinfo

👻 Ghost score 10%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Salesforce

Salesforce

10,000+ employees

Founded 1999

💼 Consulting

📣 Marketing

☁️ SaaS

Consulting • Marketing • SaaS

Salesforce is a leading cloud-based software company that provides a comprehensive customer relationship management (CRM) platform. The platform connects all company data and teams on an integrated, AI-driven system to improve sales, customer service, and marketing efforts. Salesforce offers scalable solutions for businesses of all sizes, including small and medium enterprises, and provides industry-specific solutions to modernize operations, save time, and reduce costs. Salesforce also offers educational resources through Trailhead and a wide selection of apps on AppExchange to extend its CRM capabilities.

📋 Description

• Own and mature Slack’s incident management program end-to-end, from detection and triage through response, resolution, and post-incident review • Drive the strategic transition of incident response operations across the Customer Experience team and Salesforce’s Command Incident Center • Align processes, integrate tooling, hand off runbooks, and conduct cross-organization training • Establish and continuously improve incident severity frameworks, escalation paths, and communication protocols • Partner with Reliability leadership to measure and reduce customer-impacting incident volume and mean time to resolution • Run Slack’s reliability initiatives review for tracking, prioritizing, and delivering reliability improvements • Own load management and compute services programs to support traffic spikes and graceful scaling • Drive capacity planning and load-shedding strategy with infrastructure engineering teams • Track and report reliability and availability metrics, including SLOs, error budgets, and incident trends • Define reliability standards and SLO frameworks for agentic workloads • Develop incident response playbooks for AI failure scenarios • Champion reliability-as-a-feature in AI product development • Coordinate across platform, infrastructure, security, and Salesforce partner teams • Design scalable policies, processes, and operating rhythms • Build and maintain timelines, risk registers, dependency maps, and executive dashboards

🎯 Requirements

• 8+ years leading technical programs in a dynamic product or engineering organization, with progressive scope and complexity • Strong verbal and written interpersonal skills • Sufficient technical capability to communicate effectively with engineers and identify technical risks • Ability to work independently and communicate across multiple time zones • Excellent organizational and interpersonal/social skills • Experience handling activities across multiple teams • Ability to analyze large data sets and synthesize them into stories and slides • SQL experience extracting large data sets into executive-level dashboards • Consistent track record of delivering complex technical projects and programs with multifunctional teams • 3+ years of experience actively developing and managing programs within an SRE, reliability, or infrastructure organization • Proficient in AWS Cloud offerings or similar cloud services • Direct experience with incident management programs, including severity frameworks, escalation processes, and post-incident review practices • Experience with data residency, compliance, or regulated infrastructure programs spanning multiple regions or jurisdictions is a strong plus • Familiarity with AI/ML infrastructure, LLM serving platforms, or agentic systems is a plus, or demonstrated ability to rapidly develop technical fluency in emerging domains • Experience navigating large-org integrations is highly valued • A related technical degree required

🏖️ Benefits

• Time off programs • Medical insurance • Dental insurance • Vision insurance • Mental health support • Paid parental leave • Life insurance • Disability insurance • 401(k) • Employee stock purchasing program • Reasonable accommodation during the application or recruiting process

Apply Now

Similar Jobs

🕒 3 days ago

Mirantis

501 - 1000

💼 Consulting

🏥 Healthcare

📦 Logistics

Technical Product Manager owning k0rdent AI’s managed Kubernetes services. Shaping scalable cluster orchestration for GPU clouds, NeoClouds, and sovereign infrastructure.

🕒 4 days ago

Hanover

51 - 200

🎯 Recruiter

💼 Consulting

Senior Technical Product Owner shaping insurance technology products, roadmaps, and agile delivery. Translating business needs into secure, scalable digital solutions and guiding cross-functional teams.

🕒 4 days ago

The Hanover Insurance Group

5001 - 10000

🛡️ Insurance

🤝 B2B

👥 B2C

Senior Technical Product Owner shaping insurance technology products, roadmaps, and governance. Translating business needs into secure, scalable solutions across enterprise architecture and agile teams.

🇺🇸 United States – Remote

💰 $500M Post-IPO Debt - The Hanover Insurance Group on 2025-08

⏰ Full Time

🟠 Senior

⚒️ Technical Product Manager (TPM)

🕒 4 days ago

Clubessential Holdings

1001 - 5000

☁️ SaaS

🤝 B2B

🏨 Hospitality

Technical Product Owner advancing Clubessential’s SaaS products, AI capabilities, and agentic workflows. Owning backlogs, releases, and cross-functional product delivery.

🕒 4 days ago

Clubessential

51 - 200

💼 Consulting

🏨 Hospitality

📣 Marketing

Technical Product Owner advancing Clubessential SaaS products, platform integrations, and AI capabilities at Xplor Technologies. Owning strategy, backlog, releases, and cross-functional delivery.