Technical Support Advisor, IT Production & Incident Management
Overview
You will own triage and resolution for Sev1 and Sev2 incidents across critical production applications, protecting stability for millions of users. You will coordinate incident bridges, drive root cause analysis, and partner with engineering, infrastructure, and business teams. This role sits in a 24x7 high-availability environment, using Splunk, Dynatrace, and ServiceNow to turn production challenges into stronger systems. You will also identify automation opportunities and improve runbooks, making a direct impact on operational reliability.
What You'll Do7
- 1Lead triage and resolution for Sev1 and Sev2 incidents, restoring service with minimal business impact
- 2Coordinate incident bridges, document key decisions, and update business, infrastructure, database, network, and development partners
- 3Analyze application logs, alerts, batch jobs, and performance trends using SQL, Splunk, and Dynatrace to find root cause and workarounds
- 4Monitor application health and stability, using trend analysis to prevent repeat issues
- 5Partner with engineering teams on code issues, release readiness, and production deployments
- 6Maintain runbooks, standard operating procedures, and knowledge articles for rapid response
- 7Identify automation and monitoring improvements that reduce manual effort and strengthen service performance
Requirements11
- 14+ years in IT production support, application support, or technical operations
- 2Experience supporting business-critical applications in a 24x7 high-availability environment
- 3Hands-on incident management, including outage communications and root cause analysis
- 4Working knowledge of ITIL Incident, Problem, and Change Management practices
- 5Strong troubleshooting with log review and data validation using SQL or similar query methods
- 6Experience with monitoring and ticketing tools like Splunk, Dynatrace, Kibana, ELK, ServiceNow, or Jira
- 7Ability to translate technical issues into business-friendly language and manage multiple priorities
- 8Demonstrated collaboration, ownership, and customer-focused judgment
- 9Preferred: PowerShell, Python, or Shell scripting for diagnostics or automation
- 10Preferred: ITIL Foundation certification or experience as an Incident Manager
- 11Preferred: Exposure to cloud platforms, CI/CD, DevOps, microservices, containers, or message queues
Salary Insight
Salary not disclosed in listing
Similar open positions
Explore active roles that match your skills and interests.

Hexaware Technologies, Inc
VerifiedApplication Support Lead, IT Operations & Incident Management
You will own application support for a global client, ensuring uptime and rapid incident resolution across Windows Server, SQL Server, and ITIL-based processes. You will lead a team of support engineers, coordinate with development and infrastructure teams, and maintain service levels in a 24/7 environment. This role requires deep troubleshooting skills and a track record of driving root cause analysis in high-stakes production systems.
Manifest Solutions
VerifiedSenior Enterprise Monitoring Engineer, Dynatrace
You will own the enterprise monitoring and observability platform using Dynatrace, supporting critical applications across on-prem, cloud, and hybrid environments. You will design monitoring standards, tune alerting, and drive incident response for a large-scale infrastructure. The role sits within the infrastructure engineering team, working with application owners, cloud architects, and operations. You will lead onboarding and automation, reducing false positives and improving visibility. This role stands out due to its impact on regulated utilities and NERC/CIP compliance.

Radiantze
VerifiedIncident Manager, ITIL & Major Incident Response
You will own the end-to-end incident lifecycle for a high-volume financial services environment, restoring IT services within aggressive SLAs. You will coordinate cross-functional teams across ServiceNow and Jira during major incidents, driving root cause analysis to prevent recurrence. Your command of ITIL practices will ensure every incident follows a disciplined path from detection to closure. This role thrives on chaos, turning outages into measurable improvements in system resilience.
Bertrandt
VerifiedApplication Support II
Drive support for enterprise applications at scale. Own resolution of functional configuration and integration issues while partnering with IT Service Desk and stakeholders. Deliver measurable improvements within 90 days.
Lubeco
VerifiedLead Senior Support Engineer, Enterprise IT & Manufacturing
Lead Senior Support Engineer with 7-10 years of IT support experience. Own the escalation path for complex technical issues across manufacturing and office environments. Lead a team of support engineers while handling hands-on troubleshooting. Drive operational stability for Microsoft 365, Azure, and Active Directory. Mentor junior staff and improve support processes. Your decisions directly impact production uptime and team growth.
Lightedge
VerifiedTriage Technician Lightedge Kansas City Missouri
The Triage Technician leads inbound customer requests and system alerts through ticket queues and phone systems. They route escalations to technical resources while reporting to the Triage and Support Supervisor. This role focuses on delivering exceptional customer experience.