Department Manager of SRE/AIOps Platforms

Consolidated Edison Company of New York, Inc. (CECONY)New York City, New YorkOn-siteFull-timeStaff, 8–12 yearsListed 2 days ago

Apply now

About this role

As the Manager of Site Reliability Engineering (SRE) and AI IT Operations (AIOps) Platforms at Con Edison, you will lead the strategy, implementation, and continuous evolution of the companys SRE and AIOps platforms, practices, and transformation initiatives. You will be responsible for developing automation across IT Operations by leveraging Agentic AI, observability, and industry-leading reliability engineering practices, with the long-term goal of transforming the organization to an SRE operating model. This includes modernizing and automating IT Operations, the Service Desk, NOC, Infrastructure Operations, Incident and Major Incident Response, while continuously improving reliability, reducing outages and MTTR, and enhancing the employee experience. You will build, mentor, and develop high-performing teams, foster a culture of operational excellence and continuous learning, and manage budgets, vendor partnerships, and technology investments to maximize business value. Ultimately, you will empower Con Edison employees to serve customers and the people of New York City through a highly reliable, secure, performant, and intuitive technology environment.

This position does not provide employment pursuant to the terms of a STEM OPT Training Plan.

Required Education/Experience
- Bachelor's Degree in Computer Science, Information Technology, Engineering or a related field and 12 years of related work experience or
- Master's Degree in Computer Science, Information Technology, Engineering or a related field and 10 years of related work experience.

Preferred Education/Experience
- Bachelor's Degree and 10 years of related work experience working in customer communications, back office program management, billing and case management related field work. Experience working in the Clean Energy Marketplace.

Relevant Work Experience
- Extensive experience in IT Automation Implementation (former developer profile), with a proven track record of leadership, required.
- Strong knowledge of IT network, infrastructure, cloud and digital workplace, required.
- Demonstrated success in mentoring, upskilling, and incubating engineering talent, required.
- Experience with AI and AI Ops platforms required, preferably in the cloud, preferred.
- Experience with SRE model and implementation with IT teams and leading enterprise transformation, preferred.
- Experience leading 24x7 IT operations, preferred.
- Experience managing Capital and O&M budgets, purchase orders, appropriation requests, and financial governance, preferred.
- Relevant experience with cloud, SRE and AI certifications, preferred.

Skills and Abilities
- Demonstrated analytical skills
- Strong verbal communication and listening skills

Licenses and Certifications
- Driver's License Required

Physical Demands
- Sit or stand to answer a phone for the duration of the workday
- Ability to stoop, bend, reach, and kneel throughout the workday
- Ability to read small print and symbols

Additional Physical Demands
- The selected candidate will be assigned a System Emergency Assignment (i.e., an emergency response role) and will be expected to work non-business hours during emergencies, which may include nights, weekends, and holidays.
- Travel as necessary