
Site Reliability Engineer Shift Man...
Full Job Description
About the Role
We are seeking a Site Reliability Engineer (I) for our Global Command Center (GCC). This role ensures availability, reliability, performance, and operational stability of business-critical applications in Chennai, India. You will serve as a key member providing proactive monitoring, incident response, automation, and operational excellence.
Key Responsibilities
- Reliability: Monitor enterprise apps and infrastructure to ensure SLA compliance; identify trends and drive improvements for system resiliency.
- Observability: Configure monitoring platforms (AppDynamics, Dynatrace, Datadog, Splunk, Azure Monitor) and reduce alert fatigue through optimization.
- Automation: Develop scripts using PowerShell, Python, and Bash to automate processes and enable self-healing capabilities.
- Governance: Produce incident reports, maintain documentation, support audits, and track reliability metrics.
Requirements
We require 2+ years of experience in Site Reliability Engineering or DevOps within mission-critical environments. Strong understanding of ITSM (ServiceNow), Incident/Problem Management is essential.
Tech Stack: Linux, Windows Server; Azure, AWS, GCP; AppDynamics, Datadog, Splunk.
Company
NCR Voyix
NCR Voyix Corporation (NYSE: VYX) stands as a global leader in unified commerce, serving over 35 countries worldwide with headquarters in Atlanta, Georgia.The company empowers retailers and restaurant...