About this role
Who We Are
Applied Materials is the global leader in materials science and engineering solutions that are at the foundation of virtually every new semiconductor chip and advanced display in the world. The equipment that we create and service is essential to advancing AI and accelerating the commercialization of next-generation semiconductor chips. Join us and push the boundaries of materials science and engineering in a company at the foundation of the electronics industry. The work we do together advances the world’s technology.
What We Offer
Location:
Bangalore,IND
You’ll benefit from a supportive work culture that encourages you to learn, develop, and grow your career as you take on challenges and drive innovative solutions for our customers. We empower our team to push the boundaries of what is possible—while learning every day in a supportive leading global company. Visit our Careers website to learn more.
At Applied Materials, we care about the health and wellbeing of our employees. We’re committed to providing programs and support that encourage personal and professional growth and care for you at work, at home, or wherever you may go. Learn more about our benefits .
Key Responsibilities
- Own and modernize the enterprise monitoring and observability program — shifting from traditional up/down alerting to predictive detection, automation, and intelligent event correlation across our toolchain (Datadog, BigPanda, SolarWinds, Zabbix, Grafana, and others).
- Manages specific IT systems or sets of systems for availability and performance; reports anomalies through predefined process. Participates in evaluation/recommendation of patches, point releases, major upgrades, and new systems purchases.
- Manages execution of support services to internal customers for monitoring/observability tooling. Adheres to service management processes to meet SLAs and customer satisfaction. Completes root cause analysis of outages and incident trends (often working with managed services partners); recommends and implements preventative and predictive actions.
- Builds automated remediation and self-healing workflows for well-understood failure patterns; integrates monitoring data with ITSM/orchestration platforms to automate ticketing, escalation, and response.
- Partners with the cybersecurity/SOC team to surface observability data as contextual signal for security investigations, helping correlate operational anomalies with security events.
- Negotiates with internal/external stakeholders to design solutions addressing complex, cross-functional monitoring needs. Identifies best-known-methods for integrated solution design and standards; keeps abreast of market technologies and performs technology evaluations as required. Works with multiple vendors for quality and competitiveness.
- Manages personnel and/or contract developers performing configuration, coding, testing (unit, integration, performance, acceptance), and QA related to monitoring tooling and automation, following established processes and guidelines.
- Plans and manages projects to ensure effective, efficient execution within established scope, timeline, budget, and quality guardrails.
- Defines and implements team goals and objectives in line with organizational strategy. Attracts, develops, and retains personnel; continuously improves team capabilities.
- Works with business leaders to develop IT enablement strategy for the monitoring/observability portfolio and establishes metrics to measure and actively manage its value.
##
What You Bring
- 6–10 years in IT operations, infrastructure engineering, or SRE, with hands-on experience across several of: Datadog, BigPanda, SolarWinds, Zabbix, Grafana (or equivalents such as Prometheus, Splunk, AppDynamics).
- Demonstrated experience maturing an observability program end-to-end — architecture, predictive/anomaly-based alerting, noise reduction — not just operating existing dashboards.
- Strong scripting/automation ability (Python, PowerShell, or similar) and experience integrating monitoring platforms via APIs.
- Experience supporting a large-scale, global, hybrid (cloud + on-prem) enterprise environment.
- People management experience or strong readiness for a first-line management role, with technical credibility to lead senior engineers.
##
Functional Knowledge
Demonstrates in-depth understanding of concepts, theories, and principles in own job family and basic knowledge of related job families.
##
Business Expertise
Applies understanding of the industry and how own area contributes to achievement of objectives.
##
Leadership
Manages a generally homogeneous team; adapts plans and priorities to meet service and/or operational challenges.
##
Problem Solving
Identifies and resolves technical, operational, and organizational problems.
##
Impact
Impacts the level of service and the team's ability to meet quality, volume, and timeliness objectives. Guided by policies and resource requirements within business unit, department, or sub-function.
##
Interpersonal Skills
Guides, influences, and persuades others internally in related areas or externally.
Position requires understanding of Applied Materials global Standards of Business Conduct and compliance with these standards at all times. This includes demonstrating the highest level of ethical conduct reflecting Applied Materials' core values.
Additional Information
Time Type:
Full time
Employee Type:
Assignee / Regular
Travel:
Yes, 10% of the Time
Relocation Eligible:
No
Applied Materials is an Equal Opportunity Employer. Qualified applicants will receive consideration for employment without regard to race, color, national origin, citizenship, ancestry, religion, creed, sex, sexual orientation, gender identity, age, disability, veteran or military status, or any other basis prohibited by law.