Reliability Operations Engineer I
Insight GlobalHyderabad, Telangana
it-jobs
Job Description
Job Description The Reliability Operations Engineer I serves as the first line of operational support, monitoring and responding to alerts across infrastructure, network, and Kubernetes environments. This role follows established runbooks, manages incidents and tickets, and escalates complex issues while ensuring platform stability and uptime. This position supports a 24x7 operations environment. Engineers are typically assigned to a day, mid, or night shift; however, flexibility is required to occasionally provide coverage across other shifts based on business needs. This role requires onsite presence in Hyderabad five days per week. Responsibilities Monitor and triage infrastructure, network, and platform alerts. Execute operational runbooks and standard procedures. Create, update, and manage incident tickets. Escalate issues that require engineering intervention. Support incident response and operational health checks. Collaborate with senior engineers to resolve recurring issues. We are a company committed to creating diverse and inclusive environments where people can bring their full, authentic selves to work every day. We are an equal opportunity/affirmative action employer that believes everyone matters. Qualified candidates will receive consideration for employment regardless of their race, color, ethnicity, religion, sex (including pregnancy), sexual orientation, gender identity and expression, marital status, national origin, ancestry, genetic factors, age, disability, protected veteran status, military or uniformed service member status, or any other status or characteristic protected by applicable laws, regulations, and ordinances. If you need assistance and/or a reasonable accommodation due to a disability during the application or recruiting process, please send a request to HR@insightglobal.com.To learn more about how we collect, keep, and process your private information, please review Insight Global's Workforce Privacy Policy: https://insightglobal.com/workforce-privacy-policy/. Skills and Requirements Required Qualifications 2-4 years of experience in NOC, cloud, infrastructure, or network operations. Knowledge of Kubernetes/OpenShift fundamentals (pods, deployments, manifests) and/or networking fundamentals (IP routing, BGP). Experience with incident management and ticketing systems. Linux troubleshooting fundamentals. Strong communication and problem-solving skills. Ability to support a shift-based operations environment. Preferred Qualifications Exposure to Cisco SDN/SDA. Familiarity with OpenShift operations, GitLab, GitOps, and observability tools.
Get AI-Matched to This Job
Upload your resume and our AI will score how well you match this and thousands of similar roles.