
Selector is building an operational intelligence platform for digital infrastructure. Using an AI/ML-based analytics approach, the platform provides actionable, multi-dimensional insights to network, cloud, and application operators. It helps operations teams meet their KPIs through seamless collaboration, a search-driven conversational experience, and automated data engineering pipelines.
Our solutions are used by leading Telecom, Media, Health Care, Finance, Retail, Professional Sports, and Fortune 500 enterprise organizations around the world. Our novel approach and rapidly expanding footprint position us for continued growth as a category leader.
Title: Incident & Escalation Manager
Location: Santa Clara HQ preferred
Selector AI is building a dedicated Incident & Escalation function to strengthen how we manage critical incidents and customer escalations as we scale.
We are looking for an experienced leader who can define the vision, build the operating model, and drive execution. You will bring together existing practices across NOC, SRE, Solution Engineering, Engineering, Product, Support, and Account teams and establish a consistent, scalable approach to incident and escalation management.
This is not simply an incident coordination role. You will build the function—from vision and process through execution, training, metrics, and continuous improvement
What You Will Own
What You Bring
What Success Looks Like
You will build a predictable, scalable Incident & Escalation capability where the right people engage quickly, ownership is clear, communication is consistent, resolution is driven, and lessons become measurable improvements.
This is an opportunity to build and lead an important operational capability at a growing AI company.
165K–200K base
Perks: discretionary PTO, health insurance, 401k, bonus potential, and more.

Selector AI is the industry leading AIOps platform designed to provide instant, real-time actionable insights for managing multi-domain network and application infrastructures. By bringing together multiple sources of data into one easy to use platform, IT teams can troubleshoot network issues faster, avoid downtime, reduce MTTR and improve efficiency.