🎁 Before you apply, rehearse this interview. Create your free WorkMundi account and get an Interview Training on HelpsYouSpeak — no cost, no card. I want my training →
Senior IT Infrastructure Manager We are committed to continuous improvement in Enterprise Monitoring and Observability to ensure optimal efficiency, enhanced communication, and service delivery aligned with our business priorities. As the Senior IT Infrastructure Manager, you will play a pivotal role in shaping our Enterprise Monitoring and Observability strategies and processes. You will be responsible for the strategy, architecture, and operational delivery of enterprise monitoring and observability capabilities across infrastructure, applications, cloud platforms, and enduser experience. This role leads the design, implementation, and ongoing optimization of monitoring platforms to ensure service reliability, performance, and proactive issue detection across the enterprise. The role serves as the authoritative owner for observability standards, tooling, and practices, partnering closely with Infrastructure, Application Development, Cloud, Network, Security, and IT Operations teams to improve availability, reduce MTTR, and enable datadriven operational decisionmaking. What Youll Do: Define and own the enterprise monitoring and observability strategy, roadmap, and target-state architecture. Establish governance, standards, and best practices for metrics, logs, traces, alerts, and dashboards. Own enterprise monitoring and observability platforms (e.g., infrastructure, APM, digital experience, network, cloud, and log analytics). Ensure scalable, resilient, and cost-effective platform architecture. Define integration patterns with ITSM, CMDB, CI/CD, and incident management tools. Ensure monitoring solutions enable proactive detection, root cause analysis, and rapid incident resolution. Define alerting standards to minimize noise and support actionable alerts. Drive continuous improvement using operational data and service insights. Partner with application owners, infrastructure teams, cloud teams, and security leadership to align observability capabilities with business needs. Translate technical telemetry into meaningful insights for leadership and non-technical stakeholders. Ensure monitoring platforms comply with enterprise security, privacy, and regulatory requirements. Partner with Security teams to enable threat visibility and operational monitoring use cases Oversee data retention, access controls, and audit requirements. What We Seek: Minimum Qualifications Education Bachelors degree in Computer Science or a related field. Experience: 8+ years of experience in enterprise IT operations, monitoring, observability, or infrastructure engineering. 3+ years of people management or technical leadership experience. Strong knowledge of: Infrastructure, application, and cloud monitoring Observability concepts (metrics, logs, traces, SLIs/SLOs) Incident, problem, and change management processes Experience managing enterprise platforms at scale (thousands of hosts or services). Leadership and Management Skills: Proven experience leading teams in an enterprise environment with direct reports. Experience in managing customer relationships and understanding business imperatives. Ability to negotiate win-win outcomes and shape service propositions. Project, Technical and Problem-Solving Skills: Strong understanding of IT infrastructure, systems, and service management frameworks. Strong analytical skills to identify issues and implement effective solutions. Capacity to handle critical incidents and escalate. Strong project management skills, capable of handling multiple projects simultaneously in a fast-paced environment. Financial Acumen: Strong financial background with a proven ability to manage contract details and asset lifecycle. Communication Skills: Excellent written and verbal communication skills, with the ability to present complex information clearly to diverse audiences. Ability to prepare reports and dashboards, developing metrics to measure process effectiveness and efficiency. Preferred / Bonus Skills Experience with modern observability platforms (e.g., Dynatrace, Datadog, Grafana Labs, Prometheus, OpenTelemetry). Experience supporting hybrid and multi-cloud environments (AWS, Azure, GCP). Familiarity with SRE practices and reliability engineering concepts. Experience integrating monitoring with ITSM tools such as ServiceNow. Strong financial and vendor management experience. Industry certifications related to cloud, observability, or IT service management. Your success will be measured by: People and Teamwork Focus on aligning goals and priorities across functions. Enhancing efficiency and ease of doing business with IT. Accountability, Integrity, and Trust Creativity and Innovation Collaborating internally and externally to find solutions to business challenges and needs. Dedication to Excellence Focusing on experiencing and delivery of key efforts What We Offer: At Entegris, we invest in providing opportunity to our employees and promote