Infrastructure Engineer Staff - Dynatrace SME
American Electric Power
- Location
- Gahanna, OH
- Work model
- On-Site
- Level
- Staff
- H-1B history
- 16 approvals (FY2023)
- Posted
- Sep 4, 2026
Skills
About this role
Job Posting End Date 09-19-2026 Please note the job posting will close on the day before the posting end date.
Job Summary
Part of a larger team delivering high quality Computer, Network, Storage and End-User Infrastructure technology solutions, on-going support to the business. Independently completes and leads the largest and most complex infrastructure project assignments. Plan, research, evaluate, design, and engineer the enterprise's technology infrastructure. Provide technical support and troubleshooting, cost estimates, justifications, and recommendations. Produces technical documentation, support and configuration. Helps manage, plan, and maintain technical platforms including upgrading systems. Monitor system performance, and install and configure hardware. Responsible for collaborating with other Job Families such as Project Managers, Architects, Solution Engineers, Technicians, Business Analysts to deliver consistent, reliable technology solutions that leverage AEP's technology standards, architectures and best practices.
Job Description
AEP is seeking an experienced Enterprise Observability and Monitoring Engineer to serve as a subject matter expert for enterprise monitoring technologies including Dynatrace and related monitoring platforms. This position is responsible for the design, implementation, administration, enhancement, and operational support of AEP's enterprise monitoring ecosystem. The engineer will partner with application, infrastructure, cloud, cybersecurity, and operations teams to deliver proactive monitoring, performance management, automation, and observability solutions across critical business applications and infrastructure. This role requires strong technical expertise, excellent communication skills, and the ability to translate business requirements into actionable monitoring and observability strategies. What you’ll do: Essential Job Functions & Tasks Enterprise Monitoring & Observability Administer and support Dynatrace enterprise monitoring environments. Design and implement monitoring solutions for business-critical applications and infrastructure. Develop standards for monitoring, alerting, dashboards, and operational observability. Continuously improve enterprise monitoring coverage and accuracy. Tune alerting thresholds and event correlation to reduce false positives while ensuring timely incident detection. Support enterprise observability initiatives across on-premises, cloud, and hybrid platforms. Application Performance Monitoring (APM) Deploy and administer Dynatrace OneAgent technologies. Onboarding large scale applications Monitor application health, service dependencies, user experience, and transaction performance. Support Real User Monitoring (RUM) and Synthetic Monitoring implementations. Analyze application performance issues and identify performance bottlenecks. Provide end-to-end visibility across application ecosystems. Infrastructure Monitoring Monitor Windows, Linux, VMware, OpenShift, and cloud-hosted environments. Support monitoring for databases, middleware, network infrastructure, storage systems, domain services, load balancers, and enterprise applications. Assist infrastructure teams with capacity planning and performance optimization. Proactively identify monitoring gaps and service risks. Monitoring Engineering & Automation Develop custom monitoring solutions and Dynatrace Extensions 2.0. Create and maintain custom SQL, Oracle, and enterprise application monitoring extensions. Build dashboards, metrics, health checks, and alerting policies. Automate monitoring deployment and configuration processes. Support infrastructure-as-code and configuration management initiatives where applicable. Incident Response & Root Cause Analysis Participate in major incident response and war room activities. Perform root cause analysis for application, infrastructure, and monitoring-related incidents. Provide monitoring expertise during outage investigations.