We are seeking a **Principal Software Engineer** to serve as an operations lead of our AI Operations platform within the NSIS organization, with primary responsibility for the ongoing operational success of the **Network AIOps** platform. This is a senior technical leadership role focused on post-implementation platform stewardship, ensuring Network AIOps remains effective, well-governed, and aligned with network operations needs.
This role will support **NetOps** teams and **service level owners (SLOs)** by driving operational improvements across platform administration, workflow tuning, dashboard support, signal quality, governance, production monitoring coordination, and vendor engagement. The role will also support global platform coverage in a **follow-the-sun operating model**, with close coordination across engineering partners in the **US and India**.
Key Responsibilities
- Serve as the primary internal technical owner for the Network AIOps platform
- Partner with **NetOps** and **SLOs** on intake, prioritization, and operational support needs
- Drive dashboard fixes, workflow improvements, and platform usability enhancements
- Improve signal quality, event relevance, and the effectiveness of AI-assisted operational insights
- Work with the **SelectorAI NOC** to monitor the production platform and coordinate issue response
- Partner with the vendor on production support, issue resolution, platform releases, and ongoing improvements
- Support follow-the-sun operations through effective cross-regional coordination and operational handoffs
- Establish and maintain governance for platform updates, access, configuration changes, and support practices
- Define KPIs and provide reporting on platform health, adoption, and operational value
- Communicate platform status, risks, priorities, and recommendations to both technical and business stakeholders
Required Qualifications
- Bachelor’s degree in Computer Science, Software Engineering, Information Systems, Network Engineering, or a related technical field; or equivalent practical experience
- 15+ years of experience in software engineering, platform operations, observability, incident/event management, or related operational technology domains
- 5+ years in a senior, lead, staff, principal, or equivalent technical role with ownership of a business-critical enterprise platform or service
- Experience operating and maintaining a vendor-delivered platform in production post implementation
- Experience with AIOps, observability, event intelligence, incident management, or AI-enabled operations platforms
- Experience improving signal quality, reducing operational noise, and strengthening triage and investigation workflows
- Experience with stakeholder intake, prioritization, governance, and operational support practices
- Experience working in a globally distributed support environment with cross-time-zone coordination
- Strong vendor engagement and cross-functional collaboration skills
- Strong written and verbal communication skills
- Ability to present technical concepts, platform health, risks, and recommendations clearly to non-technical business leaders
- Ability to influence decisions and drive alignment across teams without direct authority
- Strong organizational skills with the ability to manage competing priorities and drive follow-through
Preferred Qualifications
- Experience with **SelectorAI** or similar AIOps or operational intelligence platforms
- Experience supporting **NetOps, NOC, service assurance, or network operations** teams
- Background in **network engineering, network operations, network observability, or telecom operations**
- Experience supporting a **follow-the-sun operating model**
- Familiarity with **network telemetry and monitoring ecosystems**
- Experience with dashboards, operational reporting, APIs, scripting, Python, SQL, or similar platform administration and reporting tools
- Experience preparing executive-ready updates, operational reviews, or status reporting for leadership audiences
This role is needed to provide senior technical ownership of the Network AIOps platform. The position will help improve platform reliability, usability, governance, and adoption while strengthening global support coverage and operational continuity across regions. It will also serve as the primary technical point of coordination across internal stakeholders and vendor teams.