Monitoring Command Center Engineer
Dallas, IA - USA
Job Summary
Position Title: Monitoring / Command Center Engineer
Location: Remote
Duration: Full time with Client - Direct Hire
ROLE OVERVIEW
The Monitoring / Command Center Engineer is responsible for proactive 24x7 monitoring incident detection analysis and operational support across onpremise and Azure IaaS/PaaS environments. The role ensures high availability performance and reliability of critical infrastructure and applications using enterprise monitoring tools strong troubleshooting skills and ITILaligned operational practices.
Job Roles and Responsibilities:
- Monitor network system and application performance across onpremise and Azure environments using tools such as Dynatrace LogicMonitor Databricks and other integrated monitoring platforms.
- Analyze monitoring data by building dashboards reviewing trends querying logs and researching historical data to proactively identify and prevent potential outages.
- Investigate alerts and anomalies perform detailed root cause analysis and determine appropriate corrective or preventive actions.
- Manage incidents and service requests using ServiceNow or similar ITSM tools including assignment tracking escalation and closure within SLA timelines.
- Identify recurring issues and collaborate with infrastructure network application and cloud teams to implement automation selfhealing scripts and monitoring improvements.
- Create maintain and update documentation including Standard Operating Procedures (SOPs) Operations Playbooks and Knowledge Base articles using consistent standards.
- Troubleshoot complex issues involving networking DNS DHCP SMTP web/application servers and Azure infrastructure and PaaS services.
- Develop and enhance PowerShell scripts or equivalent automation to support monitoring remediation and operational efficiency.
- Participate in major incident bridges service interruption analysis and recovery activities ensuring clear communication of technical and business impact.
- Support continuous improvement initiatives aligned to ITIL 4 practices and operational excellence.
Job Roles and Responsibilities:
- 2 years of handson systems or infrastructure engineering experience including Windows Linux and Azure cloud environments.
- Strong analytical and criticalthinking skills with the ability to diagnose and resolve complex technical issues.
- Experience managing incidents in fastpaced highavailability or 24x7 Command Center / NOC environments.
- Proficiency with Dynatrace or comparable enterprise monitoring tools (mandatory).
- Experience working with ServiceNow or similar ITSM platforms.
- PowerShell scripting experience or equivalent automation skills preferred.
- Strong troubleshooting expertise across network system infrastructure and application layers with the ability to understand endtoend impact.
- Excellent communication skills including the ability to translate technical information for nontechnical stakeholders and articulate business impact during incidents.
- Solid understanding of application and infrastructure design principles and structured problemsolving approaches.
- Familiarity with ITIL 4 foundational practices.
Nice to have:
- Experience in BFSI or other regulated environments
- Exposure to automation AIOps predictive monitoring or selfhealing solutions
- Experience supporting global customers and followthesun operational models
Required Skills:
Monitoring / Command Center Engineersystems or infrastructure engineering