Sr Site Reliability Engineer
Oracle · Bengaluru, Karnataka, India
Free to search · AI fit score against your CV · tailor your résumé in one click
Oracle · Bengaluru, Karnataka, India
Our primary objectives are: Ensure maximum possible service availability and performance Deliver premier customer service Provide comprehensive support to our Engineering and other Operational and technical teams These objectives translate into a broad and dynamic scope of responsibilities for the GNOC team. Engineers will have the capability to centrally manage OCI s networks and implement automated solutions to address common operational challenges efficiently. NETWORK OPERATIONS Fault handling on incident tickets. provide break-fix support and escalation for event remediation, including leading root cause analysis (RCA) efforts Work closely with shift lead to ensure tickets are handled in a timely manner Using existing procedures and tooling, develop and safely complete network changes Mentor junior engineers as needed Participate in operational rotations providing break-fix support Possess analytical skills in resolving network issues with advanced troubleshooting and coordination with onsite support teams and vendors Identifying actionable incidents using a monitoring system, strong analytical problem-solving skills to mitigate network events/incidents, and following up on routine root cause analysis (RCA), coordinating with support teams and vendors Join major events/incidents calls, use technical and analytical skills to resolve network issues that impact Oracle customer/service, coordinate with SMEs, and provide RCA document Fault handling and escalation (identifying and responding to faults on OCI s systems and networks, collaborating closely with 3rd party suppliers, handling escalation through to resolution) Experience with Incident Response plans and strategies in cloud computing environments Have worked with large enterprise network infrastructure and cloud computing environments, supporting 24/7 and willing to work in rotational shifts in a network operations role Provide on-call support services as needed, job duties are varied and complex, needing independent judgment AUTOMATIONS/SCRIPTING The role includes collaborating with networking automation services to integrate support tooling and frequently developing scripts to automate routine tasks Preference for individuals with experience in scripting and network automation - Python, Puppet, SQL, and/or Ansible You will use automation to complete work and develop scripts for routine tasks PROJECT MANAGEMENT Ability to act in a project lead role as needed TECHNICAL QUALIFICATIONS: NETWORKING Experience working in a large ISP or cloud provider environment Exposure to commodity Ethernet hardware (Broadcom/Mellanox) Protocol experience with BGP/OSPF/IS-IS, TCP, IPv4, IPv6, DNS, DHCP, MPLS Experience with networking protocols such as TCP/IP, VPN, DNS, DHCP, and SSL Experience supporting network technologies, especially Juniper, Cisco, Arista, firewalls, and switches Strong analytical skills and ability to collate and interpret data from various sources Ability to diagnose network alerts to assess and prioritize faults and respond or escalate accordingly Cisco and Juniper certifications are desired SOFT SKILLS: Highly motivated and self-starter Bachelor s degree is preferred with at least 3-5 years of network-related experience Strong oral and written communication skills Excellent time management and organization skills Comfortable and able to deal with a wide range of issues in a fast-paced environment Excellent organizational, verbal, and written communication skills Minimum Job Qualifications Education and/or Experience: 8 years of experience in software engineering, infrastructure management, or related field OR Bachelors Degree in Computer Science, Engineering, or related field AND 4 years of experience in software engineering, infrastructure management, or related field OR Masters Degree in Computer Science, Engineering, or related field AND 2 year of experience in software engineering, infrastructure management, or related field. OR Doctorate in Computer Science, Engineering, or related field Job Skills: Same skills as prior level plus; Operating Systems Demonstrated ability in or knowledge of operating systems, including installing, upgrading, and troubleshooting various operating environments. Automation Experience: 3 years of experience in automation. Programming Experience: 3 years of experience in programming and/or scripting. Preferred Job Qualifications Education and/or Experience: 9 years of experience in software engineering, infrastructure management, or related field OR Bachelors Degree in Computer Science, Engineering, or related field AND 5 years of experience in software engineering, infrastructure management, or related field OR Masters Degree in Computer Science, Engineering, or related field AND 3 years of experience in software engineering, infrastructure management, or related field OR Doctorate in Computer Science, Engineering, or related field AND 1 year of experience in software engineering, infrastructure management, or related field. Automation Experience: 5 years of experience in automation. Programming Experience: 5 years of experience in programming and/or scripting. Key Responsibilities Capacity Ingestion and Management: -Takes proactive steps to design and architect infrastructure and/or service according to terms for reliability and functionality. -Forecasts demands for infrastructure and responds to capacity needs, ensuring systems have sufficient resources to handle current and future workloads. -Collaborates with the software development team to develop infrastructures and features that are reliable and scalable according to deployment requirements. -Independently ide