Job Description - Senior Enterprise Monitoring and Event Management Engineer
Manifest Solutions is currently seeking a Infrastructure Engineer Sr / Principal – Enterprise Observability & Monitoring Engineer (Dynatrace) in Columbus, OH.
Key Responsibilities
Enterprise Monitoring & Observability
Administer and support Dynatrace enterprise monitoring environments.
Design and implement monitoring solutions for business-critical applications and infrastructure.
Develop standards for monitoring, alerting, dashboards, and operational observability.
Continuously improve enterprise monitoring coverage and accuracy.
Tune alerting thresholds and event correlation to reduce false positives while ensuring timely incident detection.
Support enterprise observability initiatives across on-premises, cloud, and hybrid platforms.
Application Performance Monitoring (APM)
Deploy and administer Dynatrace OneAgent technologies.
Onboarding large scale applications
Monitor application health, service dependencies, user experience, and transaction performance.
Support Real User Monitoring (RUM) and Synthetic Monitoring implementations.
Analyze application performance issues and identify performance bottlenecks.
Provide end-to-end visibility across application ecosystems.
Infrastructure Monitoring
Monitor Windows, Linux, VMware, OpenShift, and cloud-hosted environments.
Support monitoring for databases, middleware, network infrastructure, storage systems, domain services, load balancers, and enterprise applications.
Assist infrastructure teams with capacity planning and performance optimization.
Proactively identify monitoring gaps and service risks.
Monitoring Engineering & Automation
Develop custom monitoring solutions and Dynatrace Extensions 2.0.
Create and maintain custom SQL, Oracle, and enterprise application monitoring extensions.
Build dashboards, metrics, health checks, and alerting policies.
Automate monitoring deployment and configuration processes.
Support infrastructure-as-code and configuration management initiatives where applicable.
Incident Response & Root Cause Analysis
Participate in major incident response and war room activities.
Perform root cause analysis for application, infrastructure, and monitoring-related incidents.
Provide monitoring expertise during outage investigations.
Develop corrective actions and preventative monitoring improvements.
ServiceNow & Event Management Integration
Support integration between Dynatrace, ServiceNow, and enterprise event management platforms.
Define monitoring requirements needed for automated incident creation.
Assist with event correlation, CI alignment, and ticket automation workflows.
Partner with operations teams to improve incident response processes.
Cloud & Modern Application Monitoring
Support monitoring solutions for AWS, Azure, SaaS, containerized, and microservices-based applications.
Deploy and support ActiveGate infrastructure.
Assist application teams with onboarding modern cloud workloads into enterprise monitoring platforms.
Support SSO, authentication, synthetic transaction monitoring, and end-user experience monitoring.
Customer Engagement & Consulting
Consult with application owners and infrastructure teams regarding monitoring requirements.
Lead onboarding activities for new applications and services.
Conduct monitoring requirement workshops and technical discovery sessions.
Provide guidance and best practices across the enterprise.
Governance, Compliance & Documentation
Maintain monitoring architecture documentation, procedures, and standards.
Support NERC/CIP and other regulatory monitoring requirements.
Participate in change management, CAB reviews, and operational governance processes.
Develop and maintain operational runbooks and support documentation.
Required Qualifications
Education
Bachelor’s Degree in Computer Science, Information Technology, Information Systems, Engineering, or related technical field; or equivalent work experience.
Experience
5+ years supporting enterprise monitoring or observability platforms.
5+ years supporting large-scale enterprise infrastructure environments.
Experience administering Dynatrace or comparable monitoring solutions.
Experience working in regulated utility, critical infrastructure, energy, or large enterprise environments preferred.
Required Technical Skills
Dynatrace Administration and Engineering
Application Performance Monitoring (APM)
Observability Engineering
Synthetic Monitoring
Real User Monitoring (RUM)
Windows Server Administration
Linux Administration
VMware Monitoring
AWS and Azure Monitoring
ActiveGate Deployment and Administration
Monitoring Dashboard Development
Alert Management and Event Correlation
ServiceNow Integration
Root Cause Analysis
Enterprise Incident Management
TCP/IP, DNS, SSL, Load Balancers, and Networking Fundamentals
All Job Ads are subject to GrabJobs’s Terms of Service. We allow users to flag postings that may be in violation of those terms. Job Ads may also be flagged by GrabJobs moderation team. However, no moderation system is perfect, and flagging a posting does not ensure that it will be removed.
Be the first to receive the latest Others Full-Time Jobs in the US.
Setup your job alert:
By activating job alerts, I agree to GrabJobs Terms & Privacy Policy. I can unsubscribe to job alerts anytime.
Skip
GrabJobs is the no1 job portal in the US, connecting you to thousands of jobs fast!
Find the best jobs in the US, apply in 1 click and get a job today!