Job Description
Key Responsibilities
Incident & Problem Management:
Troubleshoot and resolve complex production issues related to Unix/Linux environments, databases, and scheduled batch jobs. Root Cause Analysis (RCA):
Perform in-depth investigations into recurring failures, implement permanent fixes, and document findings. Batch Monitoring & Scheduling:
Manage, monitor, and automate routine operational workflows using job scheduling tools like Autosys or Control-M. System Monitoring:
Oversee application health and server metrics using enterprise monitoring tools like Splunk, Dynatrace, ITRS, or AppDynamics. Deployment & Release Support:
Assist with code deployments, configuration changes, system patch validation, and rollback procedures during maintenance windows. Stakeholder Communication:
Coordinate with L3 development teams, business users, and infrastructure groups during high-severity (P1/P2) outages.
Troubleshoot and resolve complex production issues related to Unix/Linux environments, databases, and scheduled batch jobs. Root Cause Analysis (RCA):
Perform in-depth investigations into recurring failures, implement permanent fixes, and document findings. Batch Monitoring & Scheduling:
Manage, monitor, and automate routine operational workflows using job scheduling tools like Autosys or Control-M. System Monitoring:
Oversee application health and server metrics using enterprise monitoring tools like Splunk, Dynatrace, ITRS, or AppDynamics. Deployment & Release Support:
Assist with code deployments, configuration changes, system patch validation, and rollback procedures during maintenance windows. Stakeholder Communication:
Coordinate with L3 development teams, business users, and infrastructure groups during high-severity (P1/P2) outages.
About this job listing
This job opportunity is provided through our
external job listing network. MyJobAlerts helps
you discover job opportunities and redirects you
to the original listing to apply.